Skip to content

Author

Yi-Ming Li

5 papers indexed here

We haven’t gathered this author’s papers yet. Follow them and we’ll fetch their work.

Not the right person? Other researchers publish under this name.

Preprint Sep 2026

DAVIS: A Depth-Only End-to-End Active-Vision Framework for Humanoid Soccer Skills

Humanoid soccer contact skills require more than producing high-impact foot-ball contacts: the robot must close the loop over perception, approach, alignment, impact, and recovery while its own motion induces substantial viewpoint changes, frequent loss of the ball from view, and uncertain contact outcomes. In this wor...

Jia-Kang Jin, Yi-Xiao Huo, Peng-Yuan Wang et al. · 0 citations
Preprint Sep 2026

SkillX: Unified Multi-Skill Policy Learning for Humanoid Soccer

SkillX is presented, a unified reinforcement learning framework that learns and composes multiple atomic soccer skills through a single command-conditioned policy, enabling the robot to execute atomic skills and transition among them such as dribbling, trapping, and shooting.

Zhang-Chen Ye, En-Xuan Ruan, Yi-Fei Bao et al. · 2 citations
#machine learning Preprint Sep 2026

Acting in Meters: Learning Metric Interactions for Precise Robotic Manipulation

Vision-Language-Action models and World-Action Models have advanced language-conditioned robotic manipulation, yet often leave metric relations among actions, objects, and scene geometry implicit. Human manipulation combines semantic understanding of task-relevant objects with spatial feedback that guides hand motion r...

Li-Jie Wang, Zheng Lu, Yi-Ming Wang et al. · 0 citations
#machine learning Preprint Sep 2026

Miles v0.1: Production-Level Post-Training

We present Miles v0.1, a full-stack, production-ready system for frontier post-training. Building upon the clean design of slime, Miles designs each stage of the reinforcement-learning (RL) training loop around a single principle: components should be verified, clean, and customizable. With accuracy, efficiency, reliab...

R. Chen, Ma Cheng, Shi Dong et al. · 0 citations
#machine learning Preprint Sep 2026

TrojanWorld: Backdooring World-Model Agents via Imagination Steering

To achieve effective, stealthy, and persistent control, TrojanWorld combines Decision-Reflective Induction to steer trigger-conditioned imagination toward attacker-specified actions using decision feedback, Clean Behavior Anchoring to preserve trigger-free predictive and behavioral fidelity, and Causal Propagation to s...

Wen-Kai Huang, Si-Yuan Liang, Gaolei Li et al. · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.