Skip to content

Author

Yi-Fan Xie

We have 5 of 11 papers

We haven’t gathered this author’s papers yet. Follow them and we’ll fetch their work.

Not the right person? Other researchers publish under this name.

Preprint Sep 2026

MM-ABC: Towards Generalist Mobile Manipulation via Seeing, Coordinating and Imagining

Mobile manipulation extends robot interaction beyond a fixed kinematic workspace by making the reachable region itself controllable. This flexibility introduces two central challenges: spatially grounded perception under continuous ego-motion and coordinated control of heterogeneous arm and base actions. Existing appro...

Qi-Wei Liang, Guang-Yu Chen, Shao-Long Zhu et al. · 0 citations
Preprint Sep 2026

MoPA: Coordinated Mobile Manipulation via Subsystem-Specific Perception Alignment

MoPA is presented, a framework that aligns perceptual conditioning with mobility and manipulation while preserving coordination at the action level, and achieves state-of-the-art performance across all three task suites.

Guang-Yu Chen, Qi-Wei Liang, Shao-Long Zhu et al. · 2 citations
Preprint Sep 2026

GeoLAM: Learning Geometry-Grounded Latent Actions from Unlabeled Human Videos

GeoLAM, a framework for learning geometry-grounded latent actions from action-free human videos, combines future-frame reconstruction through a frozen geometric feature hierarchy with motion supervision from a training-only 4D geometry teacher and requires neither the geometry teacher nor future-video generation.

Yi-Fan Xie, He-Kun Tian, Jin-Kun Liu et al. · 0 citations
Jul 2026

CheckVLA: Execution-Time Verification with Action-Conditioned World Model for Long-Horizon Mobile Manipulation

Vision-language-action (VLA) policies commonly execute long-horizon mobile manipulation through open-loop action chunks, issuing multiple actions without receiving new high-level visual input. A committed chunk therefore implies how observations should evolve, but accidental deviations can violate this expectation whil...

Yu-Shan Liu, Pei-Bo Sun, Xin-Tao Chao et al. · 2 citations
Jul 2026

Xiaomi-Robotics-1: Scaling Vision-Language-Action Models with over 100K Hours of Real-World Trajectories

Xiao-Robotics-1 serves as a strong robot foundation policy that can be efficiently fine-tuned on complex, dexterous tasks with high data efficiency and across multiple simulation benchmarks, Xiaomi-Robotics-1 outperforms state-of-the-art methods.

Xiaomin Guo, Piao-Piao Jin, Jason Li et al. · 16 citations · ⚡2

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.