Skip to content

Author

Wenbo Ding

5 papers indexed here

We haven’t gathered this author’s papers yet. Follow them and we’ll fetch their work.

Not the right person? Other researchers publish under this name.

Preprint Sep 2026

MM-ABC: Towards Generalist Mobile Manipulation via Seeing, Coordinating and Imagining

Mobile manipulation extends robot interaction beyond a fixed kinematic workspace by making the reachable region itself controllable. This flexibility introduces two central challenges: spatially grounded perception under continuous ego-motion and coordinated control of heterogeneous arm and base actions. Existing appro...

Qi-Wei Liang, Guang-Yu Chen, Shao-Long Zhu et al. · 0 citations
Preprint Sep 2026

MoPA: Coordinated Mobile Manipulation via Subsystem-Specific Perception Alignment

MoPA is presented, a framework that aligns perceptual conditioning with mobility and manipulation while preserving coordination at the action level, and achieves state-of-the-art performance across all three task suites.

Guang-Yu Chen, Qi-Wei Liang, Shao-Long Zhu et al. · 2 citations
Preprint Sep 2026

GeoLAM: Learning Geometry-Grounded Latent Actions from Unlabeled Human Videos

GeoLAM, a framework for learning geometry-grounded latent actions from action-free human videos, combines future-frame reconstruction through a frozen geometric feature hierarchy with motion supervision from a training-only 4D geometry teacher and requires neither the geometry teacher nor future-video generation.

Yi-Fan Xie, He-Kun Tian, Jin-Kun Liu et al. · 0 citations
Preprint Sep 2026

GLAM: Training a latent world model over global spatiotemporal memory for active exploration and navigation

Active exploration and semantic navigation require an embodied agent to build memory from partial observations, predict how the evolution of observed spatial memory may support future motion, and convert that prediction into actionable plans. We present GLAM, a goal-conditioned latent world model trained over global sp...

I-Tak Ieong, Rui-Zhi Feng, Zhao-Yang Lu et al. · 0 citations
Jul 2026

CheckVLA: Execution-Time Verification with Action-Conditioned World Model for Long-Horizon Mobile Manipulation

Vision-language-action (VLA) policies commonly execute long-horizon mobile manipulation through open-loop action chunks, issuing multiple actions without receiving new high-level visual input. A committed chunk therefore implies how observations should evolve, but accidental deviations can violate this expectation whil...

Yu-Shan Liu, Pei-Bo Sun, Xin-Tao Chao et al. · 2 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.