Visible evidence
See which project or session signal supports each finding.
Better Harness · Open-source insights for the Agent Work Loop
Better Harness turns project and session evidence into loop-level insights, prioritized improvements, and verifiable next steps—inside the coding agent you already use.
Evidence, impact, bounded repair, and acceptance checks in one reviewable report.Ten host adapters are supported. Six have verified setup paths; Pi, Kimi Code, WorkBuddy, and Grok link to their current support boundaries.
Add the repository marketplace, then install the plugin.
View setupChoose Desktop or CLI for the correct entrypoint.
View setupBuilt into Qoder Desktop; Qoder CLI can reuse it or install separately.
View setupLoad the source-local plugin with --plugin-dir.
View setupInstall as a Qwen Code extension.
View setupAdd the marketplace and install the plugin.
View setupInstall and evidence adapters are available; a full interactive report smoke remains pending.
View support detailsPlugin, evidence, and report adapters are available; a full interactive report smoke remains pending.
View support detailsEvidence and report adapters are available; installation stays on WorkBuddy-owned paths.
View support detailsEvidence and report adapters are available; install by symlinking the skill into ~/.grok/skills.
View support detailsHarness Inspector traces product intent through agent activity, sessions, files, and commits in one read-only workspace, keeping evidence strength and limitations visible.
The interactive sample uses fictional English data. It does not read your workspace, Git history, or coding-agent sessions.
Better Harness keeps unsupported claims out of the score and turns observed workflow gaps into findings a team can inspect, discuss, and verify.
See which project or session signal supports each finding.
Start with the workflow gap that matters most.
Keep the proposed change scoped to the observed problem.
Know what evidence would make the improvement reviewable.
Explore the self-contained English sample report

This static final frame summarizes historical Harness reports. It shows recorded trends, not causal proof of improvement.
Better Harness combines feedforward guides (AGENTS.md, specs, Skills, acceptance criteria) with feedback sensors (linters, tests, Hooks, evaluation agents), and evaluates five parts of delivery—the Agent Work Loop:
Does the agent know the goal and what “done” means?
Is the work on supported, repeatable paths?
Is there evidence the change actually works?
Does AI speed bypass quality checks or acceptance?
Does the next task benefit from this one?
Ten capability-level host adapters feed the same evidence pipeline. Six have verified Quickstart paths; Pi, Kimi Code, WorkBuddy, and Grok keep their current adapter-support boundaries explicit.