Skill retrieval has recently emerged as a promising paradigm for identifying the desirable execution guidelines from the skill gallery, thus equipping large language model (LLM) agents with the procedural knowledge to accomplish the specified task. To this end, most existing methods customize the retrieval model or rec...
Shuo Liu, Yu-Tong Yang, Haonan Xiao et al.· 0 citations
This work proposes a novel difficulty-aware task formulation pipeline with a dual-track evaluation framework, facilitating comprehensive evaluation of proactive bug-fixing capability, and proposes a novel difficulty-aware task formulation pipeline with a dual-track evaluation framework.
Hao-Bin Li, Ping Deng, Weizhong Qian et al.· 0 citations
RA-CAD (ReAct Agent for CAD), a state-aware agent that interacts with the CAD environment through a Generate--Execute--Critique--Rewrite loop achieves state-of-the-art execution validity and geometric quality compared with existing methods and strong proprietary language models, demonstrating the effectiveness of the p...
Shu-Hao Yan, Changhao He, Peng Hu et al.· 0 citations
We use cookies to run the site and, with your consent, for analytics and to show ads.
See our Cookie Policy.