← Files QAMapARCHIVED FILE
docs/eval.md
2.75 KB · Oct 2, 2026 · 00:32 UTC
# QAMap Eval `qamap eval` scores whether a branch is ready for human review in an AI-assisted workflow. It does not try to prove that the code is correct. Instead, it checks whether reviewers have enough evidence to trust, challenge, and verify the change without absorbing unnecessary cognitive load. For a complete PR-facing report that also includes QAMap review findings and suggested domain tests, use `qamap verify`. ## Usage ```sh qamap verify . --base origin/main --head HEAD --pr-body-file pr-body.md qamap eval . --base origin/main --head HEAD --format markdown qamap eval . --base origin/main --head HEAD --pr-body-file pr-body.md qamap eval services/listing --workspace-root . --base origin/main --head HEAD --include-working-tree ``` The GitHub Action can append the same report to the PR comment. On pull request events it reads the PR body from `GITHUB_EVENT_PATH`. ## Gates Each gate is scored from `0` to `2`. | Gate | What it checks | | --- | --- | | Validation commands | The project exposes a real test command and supporting commands such as typecheck, lint, build, e2e, tox, Ruff, mypy, `go vet`, clippy, or Maven verify. | | Changed test coverage | Source changes are paired with changed test files, or at least a runnable test command exists. | | Intent capture | The PR body, decision docs, or PR template capture the problem, rationale, context, alternatives, or tradeoffs. | | Risk explanation | Risky surfaces such as config, workflows, API contracts, auth, billing, migrations, and env files include risk or rollback context. | | Domain test plan | Changed files can be mapped to focused domain verification scenarios. | | Review size | The branch is small enough to review without unnecessary verification tax. | `validationCommands` in `qamap.config.json` can supply commands for custom stacks or monorepos when standard project files are not enough. ## Ratings | Rating | Meaning | | --- | --- | | `strong` | The branch has clear verification evidence and should be easy to review. | | `ready` | The branch is probably reviewable, with some follow-up still useful. | | `needs-work` | Reviewers should ask for more tests, context, risk notes, or a smaller diff. | | `high-risk` | The branch is expensive or risky to verify; add evidence before relying on it. | ## Why This Exists AI-assisted code often looks plausible before it is truly understood. That creates verification tax, cognitive debt, and intent debt: - verification tax: humans must spend extra time proving the generated result is safe - cognitive debt: reviewers receive more code than they can meaningfully understand - intent debt: the reason for a change disappears into a prompt or chat session `qamap eval` turns those risks into a small, static, explainable checklist that can run locally or in CI.
SHA-256: f2df5921052ca59c2d34916b7990729d8f39779210dc1fee59b0a94029a15fa8