kapynResearch

ReviewBench: An open benchmark for AI code review

ReviewBench is a new open benchmark for evaluating AI code review agents. Built on real GitHub pull requests, it provides calibrated evaluation and production-aligned metrics to measure agent performance. Developers can use it to accurately test code generation and review capabilities against standardized ground truth.

GitHub Blog·Oct 5, 2026

Opening Kapyn…