arrow
返回

Optimal feature selection based speech emotion recognition using two-stream deep convolutional neural network

delete2021-05-26
delete70
delete
OA
AI
S
Soonil Kwon *
DOI:10.1002/int.22505delete
delete原文链接
delete分享
delete收藏
查看原文
摘要

摘要

En 中文
Speech signal processing is an active area of research, the most dominant source of exchanging information among human beings, and the best way for human-computer interaction (HCI). Human behavior assessments and emotion recognition from a speech signal, such as speech emotion recognition (SER) is an emerging HCI area of exploration with various real time claims. The performance of an efficient SER system depends on feature learning, which include salient and discriminative information such as high-level deep features. In this paper, we proposed a two-stream deep convolutional neural network with an iterative neighborhood component analysis (INCA) to learn mutually spatial-spectral features and select the most discriminative optimal features for the final prediction. Our model is composed of two channels, and each channel is associated with the convolutional neural network structure to extract cues from the oral signals. The first channel extracts feature from the spectral domain, and the second channel extracts features from the spatial domain, which are then fused and fed to the INCA to remove the severance and select the optimal features for the final model training. The joint refine features are passed from the fully connected network with a softmax classifier to yield the predictions of the different emotions. We trained our proposed system using three benchmarks, which included the EMO-DB, SAVEE, and RAVDESS emotional speech corpora, and we tested the prediction performance to secure 95%, 82%, and 85% recognition rates. The performance of the system shows the effectiveness and significance of the proposed system.
Keyword:
affective computing
deep convolutional neural network
feature selection
multifeature learning
speech emotion recognition
AI总结

AI总结

对已上传原文的论文进行重点信息的提取,主要内容包括:简要概述、研究摘要、背景介绍、关键亮点、图文解析、展望与总结。

期刊

International Journal of Intelligent Systems 封面图
International Journal of Intelligent Systems
IF:
3.7
论文数:
3.1K
被引数:
8.1K

机构

S
Sejong University
学者数:
8.3K
论文数: 1.1W
被引数: 1.5W
引用论文

引用论文

err分享
err收藏
A Review of Emotion Recognition Using Physiological Signals基于生理信号的情绪识别研究综述
errSENSORS
IF3.5
err2018-06-28
err518
errOAAI
errShu, Lin; Xie, Jinyan; Yang, Mingyue; Li, Ziyi; Li, Zhenqi; Liao, Dan; Xu, Xiangmin; Yang, Xinyi
err分享
err收藏
Speech emotion recognition using amplitude modulation parameters and a combined feature selection procedure
err2014-06-01
err68
PREAI
errMencattini, Arianna; Martinelli, Eugenio; Costantini, Giovanni; Todisco, Massimiliano; Basile, Barbara; Bozzali, Marco; Di Natale, Corrado
err分享
err收藏
err分享
err收藏
err分享
err收藏
err分享
err收藏
Speech Emotion Recognition From 3D Log-Mel Spectrograms With Deep Learning Network
err2019-01-01
err203
errOAAI
errMeng, Hao; Yan, Tianhao; Yuan, Fei; Wei, Hongwei
err分享
err收藏
学者 查看更多内容