GitHub has introduced ReviewBench, an open benchmark for evaluating AI code-review agents against pull-request distributions modelled on more than 100 million G
GitHub has introduced ReviewBench, an open benchmark for evaluating AI code-review agents against pull-request distributions modelled on more than 100 million GitHub pull requests. Why it matters The benchmark uses a multi-source golden set
Open discussion →