Skip to content

Author

Xian-Da Guo

We have 6 of 26 papers

We haven’t gathered this author’s papers yet. Follow them and we’ll fetch their work.

Not the right person? Other researchers publish under this name.

Preprint Aug 2026

DriveCache: Action-Aware Caching for Driving World Model Inference

DriveCache is proposed, a training-free, action-aware controller that uses planned motion to allocate reuse across scenes and dynamic programming to place it across denoising steps under a calibrated response budget, which improves the overall fidelity-efficiency trade-off over evaluated cache methods.

Jianchun Yang, Jian Liang, Xian-Da Guo et al. · 0 citations
Jul 2026

EA-Nav: Learning Safe Visual Navigation Policies with Embodiment Awareness

Cross-embodiment navigation is a key challenge in embodied intelligence. Due to differences in embodiment, the same visual observation may imply different actions for different agents, making prediction ambiguous when relying solely on vision. Existing studies mainly rely on reinforcement learning, which requires large...

Jialu Zhang, Yong Du, Xianda Guo et al. · 0 citations
#natural language process... Open access Sep 2026

On-Policy Distillation for Vision-Language Model Adaptation, an Effective Paradigm on Low-Quality Multimodal Data

OnPoKD is the first framework that applies on-policy distillation to vision-language model adaptation by learning target construction as a policy decision, and is the first framework that applies on-policy distillation to vision-language model adaptation by learning target construction as a policy decision.

Hong-Yuan Zhang, Xian-Da Guo, Yan-Lun Peng et al. · 0 citations
Jul 2026

SkillNav: Score-Level Skill Intervention for Zero-Shot Object Goal Navigation

Vision-Language Model (VLM) agents have advanced zero-shot object-goal navigation, yet single-frame reasoning leaves them without the cross-step behavioral awareness an embodied navigator requires, producing recurring failures such as dead-end stalls, in-room loops, and circuitous approaches to detected targets. Prompt...

Rui Sang, Yiqun Duan, Pinhan Fu et al. · 0 citations
Preprint Aug 2026

DreamWAM: Beyond RGB Future Prediction for World Action Models

DreamWAM is introduced, which reformulates future prediction as structured world modeling beyond RGB, representing future states through complementary views of appearance, motion, geometry, and semantics, showing that robust world-action learning depends not only on predicting the future, but on representing it in a fo...

Shanglin Yuan, Weiheng Zhao, Xin Shi et al. · 4 citations
Jul 2026

TrustVLA: Mechanism-Guided Inference-Time Defense Against Vision-Language-Action Backdoors

TrustVLA is introduced, a mechanism-guided inference-time defense that adapts the Dirichlet evidence framework from trusted classification to monitor per-token, per-layer epistemic uncertainty in VLA policies, providing a retraining-free, mechanism-guided defense for visual-triggered VLA backdoors.

Pin-Han Fu, Xian-Da Guo, Xue-Tao Li et al. · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.