Skip to content

Attention-based Bi-LSTM model with hybrid wavelet–MFCC features for speech emotion recognition

Aug 2026 · International Journal of Speech Technology · Vol 29 · 0 citations · 31 references

TL;DR

The obtained experimental results prove the superiority of the proposed hybrid representation over the single Wavelet and MFCC features, achieving the overall recognition accuracy of 99% and average accuracy of 94%.

View source

Similar papers

Conference Aug 2026

Speech Emotion Recognition Using Transformer-Based Architectures with Self-Attention Mechanisms

Speech Emotion Recognition (SER) has become a key aspect in human-computer interaction, and affective computing, yet, current methods are faced with the challenge of modeling long-range context and fine-grain emotional expressions in speech signals. This paper has countered these shortcomings, giving a Transformer-base...

Dalphin Mary F, Binu Siva Singh S. K · 0 citations
Open access Aug 2026

Intelligent Audio-based Emotion Recognition in Speech by Deep Learning and Feature Engineering Techniques

This work introduces ExpressNet, an optimum Multi-Layer Perceptron (MLP)-based SER model aimed to solve issues by leveraging a wide range of prosodic and spectral qualities incorporating Mel-Frequency Cepstral Coefficients (MFCCs), spectral contrast, and pitch variations.

Ramakrishna Gandi, A. Geetha, B. R. Reddy · 0 citations
Conference Open access Sep 2026

A Multimodal Speech Emotion Recognition Framework for Malayalam through Audio-Text Fusion

Speech emotion recognition (SER) is used in many domains, such as translation, intelligent assistants, healthcare monitoring, large language models, and human-computer interaction. Emotion recognition in Malayalam, however, remains challenging because of the language's rich morphological structure. This work introduces...

Athira Raj, Christy James Jose, K. Biju · 0 citations
Open access 2019

Speech Emotion Recognition using Convolutional Neural Networks and Recurrent Neural Networks with Attention Model

Speech emotion recognition is an upcoming subfield of automatic speech recognition that shares multiple similarities with mood recognition in music signals. Audio signals containing human speech are used as input to classification algorithms trained to recognize emotions in the form of audio features. This thesis outline...

G. Tomas, S. Weinzierl, Athanasios Lykartsis · 2 citations
Sep 2026

Speech Emotion Recognition Using Hybrid VMD and EWT Based Cepstral Feature Extraction.

OBJECTIVE Speech Emotion Recognition (SER) has gained significant research attention over the past three decades owing to its diverse real-world applications, including human-computer interaction, healthcare, call centers, automotive systems, education, and security. The primary goal of SER is to accurately identify hu...

S. Mishra, Pankaj Warule, S. S. Nayak et al. · 0 citations
Open access Aug 2026

Domain Adaptation Techniques for Cross-Subject EEG-Based Emotion Recognition

Electroencephalogram (EEG) signals used for emotion classification have gained a lot of research interest. However, improving the efficacy of emotion recognition across individuals is difficult. Due to the weak generalizability of characteristics across individuals, it has always been challenging to identify cross-subj...

Shreyashi Dhar, Dharmpal Singh · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.