BettingProduct.ai

AI Agent

Agent Review

Catches what's wrong with a change before a human has to, without slowing delivery down to a crawl.

What is it?

A second agent that audits every implementation branch against the original task's acceptance criteria and the codebase's own rules, cites the exact file and line for anything it flags, and is explicitly barred from speculating about runtime behaviour it can't see in the diff.

What does it do?

  • -A visibility boundary that's actually enforced: it judges the diff, the pre-validation output, and rule compliance, and nothing else, no guessing about what might break in production.
  • -A held-out test suite the implementation never saw runs automatically before the model even looks at the diff, and it can tell a pre-existing failure from one this change actually caused.
  • -Every FAIL has to cite file, line, and rule, or it gets downgraded, no vague rejections.
  • -A clean pass opens the pull request itself: review isn't a gate someone has to remember to run, it's the thing that makes the PR appear.

Why does it matter?

An agent reviewing its own work is not real review. Verification has to be a separate pass, with a narrower mandate than the one that wrote the code, and it has to be able to say no.

Who is it for?

Engineering Leaders

How it works

Claude Sonnet, given the PR diff, the changed file contents, the original task, and the PR description, run with external tool access disabled for determinism. Outputs a structured PASS, HITL, or FAIL verdict as a comment; PASS triggers PR creation directly, anything else triggers the Resumption agent.

How can I get it?