Coding agents are increasingly deployed for iterative development on real repositories, yet existing evaluation barely answers a basic question: \emph{do coding agents reuse existing code or reinvent the wheel?} The question matters: every duplicated implementation is a fix applied twice and agents produce code far fas...
Dong-Sheng Ma, Si-Zhe Wang, Xin-Yi Huang et al.· 0 citations
SA is proposed, a Stable Advantage Fusion framework that avoids entropy collapse and consistently outperforms fixed-coefficient GRPO+OPD fusion, improving the aggregate score by 0.70% across all six model-domain settings while achieving more stable training.
Yifan Ding, Xin Wei, Yoshua Y. Li et al.· arXiv.org· 3 citations
DagEvo is introduced, whose diagnostician extracts recurring error causes from this history and stores them in a hierarchical error-cause memory, whose diagnostician extracts recurring error causes from this history and stores them in a hierarchical error-cause memory.
Xin Wei, Yinze Ding, Yoshua Y. Li et al.· 0 citations
Decompose--Enhance--Correct (Decompose--Enhance--Correct), a visual-consistency-guided agentic framework that improves frozen table parsers without retraining is proposed, which derives a 1,977-table Consensus-Hard Set from 4,556 candidates through offline metrics and cross-model consensus.
Jutao Xiao, Yuan Qu, Dongsheng Ma et al.· 0 citations
We use cookies to run the site and, with your consent, for analytics and to show ads.
See our Cookie Policy.