arrow
Return

Human-Robot Collaboration Using Sequential-Recurrent-Convolution-Network-Based Dynamic Face Emotion and Wireless Speech Command Recognitions

delete2023-01-01
delete5
delete
OA
AI
C
Chih‐Lyang Hwang *
Y
Yuchen Deng
S
Shih‐En Pu
DOI:10.1109/ACCESS.2022.3228825delete
deleteOriginal
deleteOriginal request for help
deleteShare
deleteSave
Abstract

Abstract

En 中文
The proposed sequential recurrent convolution network (SRCN) includes two parts: one convolution neural network (CNN) and a sequence of long short-term memory (LSTM) models. The CNN is to achieve the feature vector of face emotion or speech command. Then, a sequence of LSTM models with the shared weight reflects a sequence of inputs provided by a (pre-trained) CNN with a sequence of input sub-images or spectrograms corresponding to face emotion and speech command, respectively. Simply put, one SRCN for dynamic face emotion recognition (SRCN-DFER) and another SRCN for wireless speech command recognition (SRCN-WSCR) are developed. The proposed approach not only effectively tackles the recognitions of dynamic mapping of face emotion and speech command with average generalized recognition rate of 98% and 96.7% but also prevents the overfitting problem in a noisy environment. The comparisons among mono and stereo visions, Deep CNN, and ResNet50 confirm the superiority of the proposed SRCN-DFER. The comparisons among SRCN-WSCR with noise-free data, SRCN-WSCR with noisy data, and multiclass support vector machine validate its robustness. Finally, the human-robot collaboration (HRC) using our developed omnidirectional service robot, including human and face detections, trajectory tracking by the previously designed adaptive stratified finite-time saturated control, face emotion and speech command recognitions, and music play, validates the effectiveness, feasibility, and robustness of the proposed method.
Keywords:
Face recognition
Speech recognition
Emotion recognition
Service robots
Human-robot interaction
Image recognition
Collaboration
Human--robot collaboration
CNN
LSTM
human and face detection
dynamic face emotion recognition
wireless speech command recognition
omnidirectional service robot
visual searching and tracking
adaptive stratified finite-time saturated control

Journal

IEEE Access cover
IEEE Access
IF:
3.6
Papers:
9.8W
Citations:
29.4W

Organization

N
national taiwan university of science & technology
Scholars:
8.8K
Papers: 8.7K
Citations: 9
Cited Papers

Cited Papers

Humanoid Robot's Visual Imitation of 3-D Motion of a Human Subject Using Neural-Network-Based Inverse Kinematics
err2016-06-01
err14
PREAI
errHwang, Chih-Lyang; Chen, Bo-Lin; Syu, Huei-Ting; Wang, Chao-Kuei; Karkoub, Mansour
errShare
errSave
diffGrad: An Optimization Method for Convolutional Neural Networks
err2020-11-01
err149
errOAAI
errDubey, Shiv Ram; Chakraborty, Soumendu; Roy, Swalpa Kumar; Mukherjee, Snehasis; Singh, Satish Kumar; Chaudhuri, Bidyut Baran
errShare
errSave
Three-Layer Weighted Fuzzy Support Vector Regression for Emotional Intention Understanding in Human Robot Interaction
err2018-10-01
err60
PREAI
errChen, Luefeng; Zhou, Mengtian; Wu, Min; She, Jinhua; Liu, Zhentao; Dong, Fangyan; Hirota, Kaoru
errShare
errSave
errShare
errSave
Raman Scattering in Three-Dimensional Photonic Crystals
err2005-05-01
err0
PREAI
errV. S. Gorelik; L. I. Zlobina; P. P. Sverbil’; A. B. Fadyushin; A. V. Chervyakov
errShare
errSave
researcher View more