Review
Jul 2026
ExplainBench: Evaluating Code Explanations from Agents
This work proposes ExplainBench, a benchmark to automatically evaluate explanations from coding agents, based on the intuition that informative explanations should enable an LLM to correctly answer questions, allowing quantitative comparison of explanation quality between agents.
Zhiyuan Pan, Sungmin Kang, Imam Nur Bani Yusuf et al.
· 0 citations