Multi-modal object Re-Identification (ReID) benefits from complementary information across heterogeneous imaging modalities. To further enrich semantic representation, text descriptions have recently been incorporated as an additional modality. However, recent vision-language approaches often treat text descriptions as...
Wei-Xiang Zhou, Yu-Hao Wang, Xing-Guo Xu et al.· 0 citations
Low-light image super-resolution aims to recover normal-light high-resolution images from dark low-resolution observations captured by image sensors, in which illumination attenuation, sensor noise, blur, and low resolution are entangled, making it more challenging than conventional super-resolution. Diffusion-based me...
Zi-Yu Yue, Jun-Ran Zhang, Zhi-Xun Su· Italian National Conference...· 0 citations
A robust multi-modal ReID framework with dual semantic guidance and global-local mutual modulation, which mainly consists of three key components, namely the Text-Semantic Injector (TSI), the Masked Global-Local Modulator (MGLM), and the Hierarchical MoE Fusion (HMF).
Wei-Xiang Zhou, Xing-Guo Xu, Yu-Hao Wang et al.· IEEE transactions on circuit...· 0 citations
We use cookies to run the site and, with your consent, for analytics and to show ads.
See our Cookie Policy.