ReviewBench: An open benchmark for AI code review
Why CEREBRO kept it
Code review agent benchmark on real GitHub PRs
The text below is an automated extraction of the article at https://github.blog/ai-and-ml/github-copilot/reviewbench-an-open-benchmark-for-ai-code-review/, stored verbatim in the public cerebro-vault repository. Copyright remains with the original publisher (github.blog).
<p>We’re launching ReviewBench, a benchmark for code review agents built on representative GitHub pull requests, multi-source ground truth, calibrated evaluation, and production-aligned metrics.</p> <p>The post <a href="https://github.blog/ai-and-ml/github-copilot/reviewbench-an-open-benchmark-for-ai-code-review/">ReviewBench: An open benchmark for AI code review</a> appeared first on <a href="https://github.blog">The GitHub Blog</a>.</p>
Backlinks
Appeared in 1 briefing