Skip to content

Author

Yutong Xie

2 papers indexed here

We haven’t gathered this author’s papers yet. Follow them and we’ll fetch their work.

Not the right person? Other researchers publish under this name.

Preprint Aug 2026

Counterfactual Anatomy-guided Spatial-Temporal Decoding for Annotation-Free Hallucination Mitigation in Medical VLMs

Experiments on the SLAKE and MIMIC-CXR datasets demonstrate that CAST consistently outperforms strong baselines and surpasses decoding strategies reliant on ground truth, and indicates that compact, automatically selected regions provide highly effective contrastive guidance without expert annotations.

Yifan Lu, A. Dukre, Abhijit Das et al. · 0 citations
#small language model Preprint Aug 2026

Grounding Isn't Knowing: Do VLMs Need Object Localization for Spatial Reasoning?

This work investigates two representative model families, LLaVA-1.5 and Qwen2.5, and provides a token-, layer-, and head-level account of how VLMs transform object grounding into spatial relations, showing that knowing where objects are is not equivalent to knowing how they relate.

Xiwei Liu, Yulong Li, Xinlin Zhuang et al. · 0 citations