Vision-Language-Action (VLA) models have demonstrated strong performance across diverse robotic manipulation tasks, yet their predominantly vision-centric perception and position-controlled execution remain insufficient for contact-rich manipulation. Visual observations alone often provide limited evidence of contact o...
Bo-Han Gan, Xuan Wen, Yong-Shen Zhao et al.· 0 citations
A symmetric Dual-Arm Expert (DAE) architecture built upon a shared Vision-Language Model (VLM) backbone with decoupled, arm-specific expert towers is proposed, providing preliminary evidence of emergent skill generalization from single- to dual-arm tasks (as well as the reverse), together with cross-arm motion-domain s...
Yong-Shen Zhao, Han Gao, Bao-Ping Cheng et al.· 0 citations
Vision-Language-Action (VLA) models have recently emerged as a promising paradigm for embodied intelligence by unifying perception, language grounding, and action generation in an end-to-end framework. Despite their strong performance, their deployment in physical control loops introduces new security risks: small pert...
Kaizheng Liu, Yuntao Hu, Jiaxing Li et al.· International Conference on...· 0 citations
We use cookies to run the site and, with your consent, for analytics and to show ads.
See our Cookie Policy.