Aug 2026
Multimodal sentiment analysis with multi-level representation learning and global tri-modality unified fusion.
A Global Tri-Modality Transformer (GTMT) that first performs parallel fusion of the three modalities and then conducts deep integration guided by the textual modality to achieve the cross-modal semantic alignment and correlation, significantly improving the global unified fusion effectiveness of the tri-modality information.
Zelong Li, Shuhua Lu, Fang Cui et al.
· Neural Networks · 0 citations