Multi-modal object Re-Identification (ReID) benefits from complementary information across heterogeneous imaging modalities. To further enrich semantic representation, text descriptions have recently been incorporated as an additional modality. However, recent vision-language approaches often treat text descriptions as...
Wei-Xiang Zhou, Yu-Hao Wang, Xing-Guo Xu et al.· 0 citations
Multi-modal object Re-Identification (ReID) aims to retrieve specific objects by integrating complementary information from multiple modalities. However, existing multi-modal ReID methods do not effectively address background interference suppression or achieve tri-modal alignment, instead focusing on pairwise feature...
Weixiang Zhou, Jiabei Zuo, Yuhao Wang et al.· IEEE Transactions on Image P...· 0 citations
A robust multi-modal ReID framework with dual semantic guidance and global-local mutual modulation, which mainly consists of three key components, namely the Text-Semantic Injector (TSI), the Masked Global-Local Modulator (MGLM), and the Hierarchical MoE Fusion (HMF).
Wei-Xiang Zhou, Xing-Guo Xu, Yu-Hao Wang et al.· IEEE transactions on circuit...· 0 citations
We use cookies to run the site and, with your consent, for analytics and to show ads.
See our Cookie Policy.