Preprint
Aug 2026
Evaluating Agentic Code Repair Capabilities in Distributed Systems
DDBench is introduced, a code-repair benchmark of 60 historical bugs mined from 13 open-source distributed systems, partitioned into three difficulty tiers, isolating the effect of debugging context from model capability.
Yibo Yan, Huijuan Wang, Junzhou He et al.
· 0 citations