EasyHubExplore
Explore
EN

Site appearance

Your color. Your style.

Accent colorRose
Visual styleSame content, fresh look

Soft gradients, dimensional icons

Applied: Rose · Studio. Saved in this browser.

Development·EASYHUB JOURNAL

GitHub launches ReviewBench, an open AI code-review agent benchmark built from 219 real pull requests across 19 languages

GitHubSource published
Editorial illustration: GitHub launches ReviewBench, an open AI code-review agent benchmark built from 219 real pull requests across 19 languages
EasyHub editorial illustration · Not a source photograph

What changed

GitHub launched ReviewBench on October 5 as an offline benchmark for AI code-review agents. GitHub analyzed distributions across 103.9 million pull requests, then selected 219 pull requests from 187 public open-source repositories spanning 19 languages. Its golden set combines human review, follow-up fixes, static analysis and frontier-model findings under a common rubric, with independent senior-engineer relabeling reporting 96.6% agreement. GitHub is publishing the dataset, methodology, judge prompt/configuration and a self-serve runner for third-party agents and leaderboard submissions.

Read original source
Original title
ReviewBench: An open benchmark for AI code review
Source
GitHub · github.blog
Topic
Development
Source month
2026-10

This is a concise EasyHub summary of the linked source, not the full report or original reporting. Availability and preview conditions are described in the summary and original.

Summary page published · Editorial information

Back to the news timeline