Skip to content

Mamba-CrossMod: a multimodal affective analysis framework based on selective state space model

Sep 2026 · Pattern Analysis and Applications · Vol 29 · 0 citations · 38 references

TL;DR

The proposed Mamba-CrossMod is a novel multimodal feature fusion framework that introduces Mamba-ATT, an enhanced attention mechanism based on a selective state-space model for capturing long-range dependencies with theoretically linear complexity.

View source

Similar papers

#artificial intelligence Preprint Sep 2026

Cross-Modal Emotion Understanding: A Transformer-GAT Approach for Dialogue Emotion Recognition

Transformer-GAT is proposed, a hybrid framework that combines Transformer and the Graph Attention Network to enable cross-modal emotion understanding and effectively integrates multimodal features, balances global and local contexts, and provides deeper emotional insights, offering new directions for multimodal emotion...

Jia-Qi Qiao, Yifan Lyu, Xiu-Juan Xu · 0 citations
Open access Aug 2026

Learning Adaptive Cross-Modal Interactions for Multimodal Sentiment Analysis

A framework for learning adaptive cross-modal interactions for multimodal sentiment analysis that consistently outperforms previous methods and enhances multimodal representation capability for sentiment classification is proposed.

Chuhan Cheng, Hangcheng Wu, Jun-Qiao Wang et al. · 0 citations
Sep 2026

CLMER: A Framework for Contrastive Learning-Based Multimodal Emotion Recognition.

Experimental evaluations on two public datasets DEAP, AMIGOS and a private dataset MAN-II demonstrate that CLMER significantly outperforms unimodal and traditional fusion approaches, achieving state-of-the-art performance in emotion classification tasks.

Shuang Niu, Jian He, Yu Liang et al. · 0 citations
Sep 2026

A hierarchical multistage semantic fusion framework for multimodal sentiment analysis

Multimodal sentiment analysis aims to improve cross-modal fusion to understand sentiment better. Most existing methods rely on single-stage fusion or local cross-modal interactions, making it difficult to fully capture relationships among modalities, thereby limiting their sentiment representation capabilities and...

Guang-Yu Mu, Jia-Xiu Dai, Yuan-Yuan Yue et al. · 0 citations
Conference Aug 2026

Multimodal Emotion Recognition with Emotion-Specific Cross-Modal Attention Blocks

Multimodal emotion recognition is increasingly important for healthcare, education, and human-computer interaction. However, many existing systems learn a single shared representation for all emotions, which can blur subtle class-specific cues. This paper proposes an emotion-specific multimodal architecture that combin...

Gnanaseelan Dharshika, A. Ramanan · 0 citations
Review Open access Aug 2026

A Systematic Review of Emotion Recognition: From Unimodal Signals to Multimodal Integration

These findings reveal that multimodal systems, which fuse visual, acoustic, linguistic, linguistic, and physiological signals, consistently outperform unimodal counterparts, achieving accuracy levels above 85% on benchmark datasets.

B. Bashir, Zayyanu Yunusa · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.