A Robust Framework for Speech Emotion Recognition Using Attention Based Convolutional Peephole LSTM

Ramya Paramasivam; K. Lavanya; Parameshachari Bidare Divakarachari; David Camacho

Author	Ramya Paramasivam K. Lavanya Parameshachari Bidare Divakarachari David Camacho
Keywords	Attention Mechanisms Convolutional Peephole Long Short-Term Memory Feature Selection Improved Jellyfish Optimization Algorithm Speech Emotion Recognition
Abstract	Speech Emotion Recognition (SER) plays an important role in emotional computing which is widely utilized in various applications related to medical, entertainment and so on. The emotional understanding improvises the user machine interaction with a better responsive nature. The issues faced during SER are existence of relevant features and increased complexity while analyzing of huge datasets. Therefore, this research introduces a wellorganized framework by introducing Improved Jellyfish Optimization Algorithm (IJOA) for feature selection, and classification is performed using Convolutional Peephole Long Short-Term Memory (CP-LSTM) with attention mechanism. The raw data acquisition takes place using five datasets namely, EMO-DB, IEMOCAP, RAVDESS, Surrey Audio-Visual Expressed Emotion (SAVEE) and Crowd-sourced Emotional Multimodal Actors Dataset (CREMA-D). The undesired partitions are removed from the audio signal during pre-processing and fed into phase of feature extraction using IJOA. Finally, CP LSTM with attention mechanisms is used for emotion classification. As the final stage, classification takes place using CP-LSTM with attention mechanisms. Experimental outcome clearly shows that the proposed CP-LSTM with attention mechanism is more efficient than existing DNN-DHO, DH-AS, D-CNN, CEOAS methods in terms of accuracy. The classification accuracy of the proposed CP-LSTM with attention mechanism for EMO-DB, IEMOCAP, RAVDESS and SAVEE datasets are 99.59%, 99.88%, 99.54% and 98.89%, which is comparably higher than other existing techniques.
Year of Publication	2025
Journal	International Journal of Interactive Multimedia and Artificial Intelligence
Volume	9
Start Page	45
Issue	Regular Issue
Number	4
Number of Pages	45-58
Date Published	09/2025
ISSN Number	1989-1660
URL	https://www.ijimai.org/journal/bibcite/reference/3532
DOI	10.9781/ijimai.2025.02.002
	DOI Google Scholar BibTeX EndNote X3 XML EndNote 7 XML Endnote tagged Marc RIS
Attachment	ijimai9_4_4.pdf685.85 KB
Acknowledgment	This work has been supported by the project PCI2022-134990-2 (MARTINI) of the CHISTERA IV Cofund 2021 program; by MCIN/AEI/10.13039/501100011033/ and European Union NextGenerationEU/PRTR for XAI-Disinfodemics (PLEC 2021-007681) grant, by European Comission under IBERIFIER Plus - Iberian Digital Media Observatory (DIGITAL-2023-DEPLOY-04-EDMO-HUBS 101158511); and by EMIF managed by the Calouste Gulbenkian Foundation, in the project MuseAI.