Vision-language-action (VLA) models often treat main-view and wrist-view observations as parallel visual inputs, overlooking their distinct roles in robot manipulation. Fine-grained manipulation, however, benefits from anticipating how wrist-local interactions may evolve under the global task context. To address this l...
Yuhao Pan, Haosong Peng, Zhengsheng Zhang et al.· 0 citations
Accurate outdoor localization in Non-Line-of-Sight (NLoS) environments remains a critical challenge for wireless communication and sensing systems. Existing methods, including positioning based on the Global Navigation Satellite System (GNSS) and triple Base Stations (BSs) techniques, cannot provide reliable performanc...
Jia-Jie Xu, Yi-Fan Guo, Xiu-Cheng Wang et al.· 2026 IEEE/CIC International...· 0 citations
BeamRMX is proposed, which is the first dedicated framework to treat the spatial radiation pattern as the primary BeamRM query and learn how scene geometry transforms it into the received power field, and shows consistent gains over deterministic and diffusion baselines.
Yue Zhang, Xiucheng Wang, Wenshuo Chen et al.· 0 citations
RadioDiff-v2, a dual-branch one-dimensional diffusion transformer trained with flow matching, is proposed, a dual-branch one-dimensional diffusion transformer trained with flow matching that leads every baseline on every metric.
Xiucheng Wang, Jun Huang, Nan Cheng· arXiv.org· 0 citations
RadioVIL is proposed, an efficient two-stage framework that reformulates joint radio map inpainting and zero-shot vehicle localization as a prior-guided physical inverse problem and unlocks accurate zero-shot vehicle localization directly from sparse radio maps, paving a robust way for ISAC at the 6G edge.
Ruixin Zhao, Xiucheng Wang, Qiming Zhang et al.· 0 citations
We use cookies to run the site and, with your consent, for analytics and to show ads.
See our Cookie Policy.