From Static to Dynamic: Benchmarking Real-World Code Review with MCR-Bench
In real-world software development, code review typically involves iterative interactions between developers and reviewers to improve software quality, making the process costly and time-consuming. Although recent work explores large language models (LLMs) for automated code review, most approaches oversimplify code review into a single-round, static decision task, which fails to capture the multi-round interactive nature and the complex problem-solving processes inherent in realistic review scenarios. To bridge this gap, we introduce MCR-Bench, the first defect state-aware benchmark designed
Lineage graph
Paper → model → repo connections mined from source citations (Tier-1 exact match).
Why these links exist
Every edge carries a method, confidence, and the source snippet that justified it — so bad links are debuggable.
- FuzzyOverlapping authors or contributors · 62%modular/modular →
“Shared author/contributor keys: liu”
- FuzzyOverlapping authors or contributors · 62%bytedance/deer-flow →
“Shared author/contributor keys: wang”
- FuzzyOverlapping authors or contributors · 62%ray-project/ray →
“Shared author/contributor keys: wang”
- FuzzySimilar title/name (fuzzy) · 59%tirth8205/code-review-graph →
“Fuzzy title match (0.73): “From Static to Dynamic: Benchmarking Real-World Code Review ” ≈ “tirth8205/code-review-graph””
- LinkedLinked via arxiv author · 85%Dewu Zheng →
“From Static to Dynamic: Benchmarking Real-World Code Review with MCR-Bench”
- LinkedLinked via arxiv author · 85%Yanlin Wang →
“From Static to Dynamic: Benchmarking Real-World Code Review with MCR-Bench”
- LinkedLinked via arxiv author · 85%Xiwen Wang →
“From Static to Dynamic: Benchmarking Real-World Code Review with MCR-Bench”
- LinkedLinked via arxiv author · 85%Kefeng Duan →
“From Static to Dynamic: Benchmarking Real-World Code Review with MCR-Bench”
