arrow
返回

Human-Robot Collaboration Using Sequential-Recurrent-Convolution-Network-Based Dynamic Face Emotion and Wireless Speech Command Recognitions

delete2023-01-01
delete5
delete
OA
AI
C
Chih‐Lyang Hwang *
Y
Yuchen Deng
S
Shih‐En Pu
DOI:10.1109/ACCESS.2022.3228825delete
delete原文链接
delete原文求助
delete分享
delete收藏
摘要

摘要

En 中文
The proposed sequential recurrent convolution network (SRCN) includes two parts: one convolution neural network (CNN) and a sequence of long short-term memory (LSTM) models. The CNN is to achieve the feature vector of face emotion or speech command. Then, a sequence of LSTM models with the shared weight reflects a sequence of inputs provided by a (pre-trained) CNN with a sequence of input sub-images or spectrograms corresponding to face emotion and speech command, respectively. Simply put, one SRCN for dynamic face emotion recognition (SRCN-DFER) and another SRCN for wireless speech command recognition (SRCN-WSCR) are developed. The proposed approach not only effectively tackles the recognitions of dynamic mapping of face emotion and speech command with average generalized recognition rate of 98% and 96.7% but also prevents the overfitting problem in a noisy environment. The comparisons among mono and stereo visions, Deep CNN, and ResNet50 confirm the superiority of the proposed SRCN-DFER. The comparisons among SRCN-WSCR with noise-free data, SRCN-WSCR with noisy data, and multiclass support vector machine validate its robustness. Finally, the human-robot collaboration (HRC) using our developed omnidirectional service robot, including human and face detections, trajectory tracking by the previously designed adaptive stratified finite-time saturated control, face emotion and speech command recognitions, and music play, validates the effectiveness, feasibility, and robustness of the proposed method.
Keyword:
Face recognition
Speech recognition
Emotion recognition
Service robots
Human-robot interaction
Image recognition
Collaboration
Human--robot collaboration
CNN
LSTM
human and face detection
dynamic face emotion recognition
wireless speech command recognition
omnidirectional service robot
visual searching and tracking
adaptive stratified finite-time saturated control

期刊

IEEE Access 封面图
IEEE Access
IF:
3.6
论文数:
9.8W
被引数:
29.4W

机构

N
national taiwan university of science & technology
学者数:
8.8K
论文数: 8.7K
被引数: 9
引用论文

引用论文

diffGrad: An Optimization Method for Convolutional Neural NetworksdiffGrad: 一种卷积神经网络的优化方法
err2020-11-01
err149
errOAAI
errDubey, Shiv Ram; Chakraborty, Soumendu; Roy, Swalpa Kumar; Mukherjee, Snehasis; Singh, Satish Kumar; Chaudhuri, Bidyut Baran
err分享
err收藏
Three-Layer Weighted Fuzzy Support Vector Regression for Emotional Intention Understanding in Human Robot Interaction
err2018-10-01
err60
PREAI
errChen, Luefeng; Zhou, Mengtian; Wu, Min; She, Jinhua; Liu, Zhentao; Dong, Fangyan; Hirota, Kaoru
err分享
err收藏
Raman Scattering in Three-Dimensional Photonic Crystals
err2005-05-01
err0
PREAI
errV. S. Gorelik; L. I. Zlobina; P. P. Sverbil’; A. B. Fadyushin; A. V. Chervyakov
err分享
err收藏
学者 查看更多内容