Active perception and manipulation are crucial for robots to interact with complex scenes. Existing benchmarks struggle to evaluate how robots effectively acquire and maintain information in memory in an active manner. To this end, we introduce ActiveArena-Sim, an active-perception simulator with controllable viewpoint...
Yi-Bo Li, En-Shen Zhou, Rui Chen et al.· 1 citation
This work proposes RealPref, a benchmark for evaluating natural preference-following in personalized user-LLM interactions, and finds that LLM performance drops significantly as context length grows and preference expression becomes more implicit, and that generalizing user preference understanding to unseen scenarios...
Qian-Yu Guo, Yi-Bo Li, Yue Liu et al.· 7 citations· ⚡1
This paper proposes PhysAgent, a reflective agentic framework that closes the loop among physical program generation, physics simulation, stage-specific verification, and targeted program repair, and design a set of physics-control APIs to support more stable and complex motion behaviors.
Qirui Li, Jinkun Hao, Yibo Li et al.· arXiv.org· 1 citation
We use cookies to run the site and, with your consent, for analytics and to show ads.
See our Cookie Policy.