General-purpose agents increasingly write code, use tools, and complete complex digital tasks, raising the question of how far these capabilities carry into the physical world. To investigate this, we introduce RobotWorld, a challenging simulation testbed for robot use: turning instructions and observations into physic...
Zhiqin Yang, Chen-Xin Li, Xiao-Meng Hu et al.· 0 citations
A taxonomy of reasoning enhancement techniques is proposed, categorized into training-time strategies (e.g., supervised fine-tuning, reinforcement learning) and test-time mechanisms (e.g., prompt engineering, multi-agent systems), and outlining future directions toward building efficient, robust, and sociotechnically r...
Zi-Zhan Ma, Wen-Xuan Wang, Meidan Ding et al.· 16 citations
We use cookies to run the site and, with your consent, for analytics and to show ads.
See our Cookie Policy.