General-purpose robot agents must learn from experience, transfer to new tasks, and act efficiently. Code as Policies (CaP) methods generate and repair programs at runtime, incurring latency and entangling reusable mechanisms with task-specific decisions. We introduce RACaP, an agentic framework that moves coding to ev...
Ze-Xi Li, Ye-Hang Zhang, Hao-Jian Huang et al.· 0 citations
General-purpose robots must infer what a new task requires and translate that understanding into appropriate physical action. In-context learning (ICL) for robots supports this process by using demonstrations and interaction to direct existing competence with neural parameters held fixed during deployment. We organize...
Hao-Jian Huang, Ze-Xi Li, Ju-Hao Guo et al.· 0 citations
Robo-Harness K1 is introduced, a robot-use agent (RUA) framework that exposes perception as tools that makes 3D geometry accessible without changing the VLM architecture or training a depth encoder, and suggests that perception-augmented RUAs offer a promising route to sample-efficient, generalizable robotic policies t...
Ze-Xi Li, Ye-Hang Zhang, Wen-Qian Li et al.· 0 citations
World Action Agent is presented, a multi-agent harness through which VLMs pilot robots with basic tools, making every decision within a visual action workspace, outperforming end-to-end VLAs, code-as-policy agents, and a visual-harness baseline with the same backbone.
Ye-Hang Zhang, Hao-Jian Huang, Yi-Fan Chang et al.· 0 citations
AdaFusion is presented, a lightweight adaptive fusion framework that integrates complementary signals from multiple frozen PFMs through low-dimensional feature compression and a sample-conditioned gating module that reweights model-wise (and optionally channel-wise) contributions.
Yu Xiao, Yang Hu, Bin Li et al.· 0 citations
We use cookies to run the site and, with your consent, for analytics and to show ads.
See our Cookie Policy.