返回
Optimal feature selection based speech emotion recognition using two-stream deep convolutional neural network
DOI:10.1002/int.22505.png)
摘要
En 中文
Speech signal processing is an active area of research, the most dominant source of exchanging information among human beings, and the best way for human-computer interaction (HCI). Human behavior assessments and emotion recognition from a speech signal, such as speech emotion recognition (SER) is an emerging HCI area of exploration with various real time claims. The performance of an efficient SER system depends on feature learning, which include salient and discriminative information such as high-level deep features. In this paper, we proposed a two-stream deep convolutional neural network with an iterative neighborhood component analysis (INCA) to learn mutually spatial-spectral features and select the most discriminative optimal features for the final prediction. Our model is composed of two channels, and each channel is associated with the convolutional neural network structure to extract cues from the oral signals. The first channel extracts feature from the spectral domain, and the second channel extracts features from the spatial domain, which are then fused and fed to the INCA to remove the severance and select the optimal features for the final model training. The joint refine features are passed from the fully connected network with a softmax classifier to yield the predictions of the different emotions. We trained our proposed system using three benchmarks, which included the EMO-DB, SAVEE, and RAVDESS emotional speech corpora, and we tested the prediction performance to secure 95%, 82%, and 85% recognition rates. The performance of the system shows the effectiveness and significance of the proposed system.
Keyword:
affective computing
deep convolutional neural network
feature selection
multifeature learning
speech emotion recognition
AI总结
对已上传原文的论文进行重点信息的提取,主要内容包括:简要概述、研究摘要、背景介绍、关键亮点、图文解析、展望与总结。
期刊
IF:
3.7
论文数:
3.1K
被引数:
8.1K
机构
引用论文
Att-Net: Enhanced emotion recognition system using lightweight self-attention moduleAtt-Net: 使用轻量级自我注意模块的增强型情感识别系统
Emotion Recognition from Chinese Speech for Smart Affective Services Using a Combination of SVM and DBN使用SVM和DBN的组合从中文语音中识别情感以实现智能情感服务
SENSORS
IF3.5
Clustering-Based Speech Emotion Recognition by Incorporating Learned Features and Deep BiLSTM
IEEE ACCESS
IF3.6

