A radiation, rotation, and scale invariant (RRSI) feature descriptor that enables feature encoding, interaction, and fusion across intra-modal, dual-head sampled, and inter-modal regions, and introduces a bidirectional cross-modal generative reconstruction constraint during training.
Abstract
Multimodal image matching is a fundamental task for multi-source information fusion. However, geometric distortions and nonlinear radiometric differences (NRD) severely limit performance, especially under radiometric, rotation, and scale variations. To address this issue, we propose a radiation, rotation, and scale invariant (RRSI) feature descriptor. First, a dual-head regional sampling (DHRS) module simultaneously performs Cartesian and Log-Polar sampling on keypoint neighborhoods, retaining spatial structural properties while enhancing robustness to rotation and scale variations. We then jointly encode geometric and radiometric relations between multimodal images in a unified deep feature space, enabling feature encoding, interaction, and fusion across intra-modal, dual-head sampled, and inter-modal regions. Furthermore, we introduce a bidirectional cross-modal generative reconstruction constraint during training. By decoding implicit features into structural patches of the counterpart modality, this mechanism anchors modality-invariant geometric topologies without additional inference overhead. Experiments on optical-infrared and optical-SAR datasets demonstrate highly competitive matching performance and strong robustness to rotation and scale variations. RRSI supports the full rotation range from 0 to 360 degrees and scale factors up to four. Its generalization ability is further validated on multimodal images from computer vision, remote sensing, and medical imaging. The implementation will be made publicly available at https://github.com/yeyuanxin110/RRSI .
This paper proposes a robust feature-based matching framework that reduces reliance on intensity information while enhancing structural representation and demonstrates that MIHOG can provide dense and reliable correspondences under complex cross-modal radiometric and geometric variations.
To address the difficulty of accurately aligning target regions in visible and infrared images caused by differences in imaging mechanisms and inconsistent radiometric discrepancies, a multi-stage registration method based on a deep geometric similarity evaluation network is proposed. The method constructs a structure-...
Infrared and visible image fusion is pivotal for robust visual perception across all weather conditions and scenes. Although deep learning-based methods have made notable progress, most either assume pre-aligned inputs or rely on implicit feature-space alignment, which fails to fundamentally address the amplification o...
Jin-Yuan Liu, Zengxi Zhang, Jiahao Zhang et al.· IEEE Transactions on Pattern...· 1 citation
Most advances in keypoint descriptions address monomodal settings, where image variations arise from viewpoint, illumination, or contrast changes. Multimodal scenarios involve images produced by fundamentally different sensing processes, such as multispectral imaging, RGB-depth, satellite imagery, or medical imaging, c...
Paul-Werner Schneider, Nazim Haouchine· 0 citations
Infrared-visible image fusion facilitates robust multimodal perception by integrating complementary textural nuances from visible sensors with thermal signatures from infrared systems. Due to the task's inherently ill-posed nature, existing methods heavily rely on structural priors but typically enforce rotation equiva...
Jia-Bao Wang, Wen-Jian Liu, Yao-Ming Cai et al.· 0 citations
Visible-infrared object detection relies on complementary RGB and thermal cues, but its performance is often degraded by cross-modal spatial misalignment. Most existing methods rely on implicit feature adaptation to handle weakly misaligned scenarios, while large-offset geometric discrepancies remain insufficiently add...
Ming Qi, Yuyang Wang, Ming-Jing Zhao et al.· 0 citations
We use cookies to run the site and, with your consent, for analytics and to show ads.
See our Cookie Policy.