Adversarial Review tests structured disagreement for agentic code review
A new arXiv paper proposes a three-agent code-review protocol in which a reviewer evaluates an agent’s code and a critic audits that review before edits are made. The authors report improved benchmark results over tested baselines, while also identifying false consensus as a failure mode.