# GitHub opens ReviewBench for evaluating AI code reviewers

_Published Monday, October 5, 2026 at 12:08 PM EDT · AI · Latest · Tier 2 — Notable_

![GitHub opens ReviewBench for evaluating AI code reviewers — Primary](https://github.blog/wp-content/uploads/2026/01/generic-invertocat-logo.png)

GitHub released a research preview of ReviewBench, a public benchmark that lets developers evaluate AI code reviewers against a common dataset and scoring method. It contains 219 pull requests from 187 public repositories spanning 19 languages, with distributions informed by analysis of 103.9 million GitHub pull requests.

The benchmark combines findings from human reviewers, analysis tools and multiple AI models, then uses Claude Sonnet 5 to judge their validity. GitHub says senior engineers independently relabeled the findings and agreed with its judgments 96.6% of the time. Developers can run their own agents and compare results by severity, category and tolerance for missed issues or false alarms. Leaderboard submissions require maintainer approval.

## Sources

- [The GitHub Blog](https://github.blog/ai-and-ml/github-copilot/reviewbench-an-open-benchmark-for-ai-code-review/)

---
Canonical: https://techandbusiness.org/newswire/dciP1eDg20D7OFMJ-g_593
Published: 2026-10-05T16:08:11.718Z
Story chronology: 2026-10-05T15:59:40.000Z
Retrieved: 2026-10-05T18:09:49.006Z
Publisher: Tech & Business (techandbusiness.org)
