返回
Enhanced spatial-temporal learning network for dynamic facial expression recognition
DOI:10.1016/j.bspc.2023.105316.png)
摘要
En 中文
The recognition of dynamic facial expressions has received increasing attention since they can better reflect the real expression process of emotion than a static image. However, due to various factors such as subtle variation differences, pose, occlusion, and illumination, it has been a challenging vision task to obtain discriminative expression features in dynamic facial expression recognition. Traditional CNN-based deep learning networks lack global and temporal contextual expression understanding, which tends to affect the final recognition of dynamic expressions. Therefore, we propose an enhanced spatial-temporal learning network (ESTLNet) for more robust dynamic facial expression recognition, which consists of a spatial fusion learning module (SFLM) and a temporal transformer enhancement module (TTEM). First, the SFLM obtains a more expressive spatial feature representation through a dual-channel feature fusion learning module. Then, the TTEM extracts more valid temporal contextual expression features based on the above spatial features through an encoder constructed by cascading a self-attention learning network and an effective gated feed-forward network. Finally, the co-enhanced spatial-temporal model approach is assessed on the four broadly used dynamic expression datasets (DFEW, AFEW, CK+, and Oulu-CASIA). Extensive experimental outcomes demonstrate that our approach surpasses several existing state-of-the-art methods, leading to notable enhancements in performance.
Keyword:
Dynamic facial expression recognition
Enhanced spatial -temporal learning
Feature enhancement
Transformer
期刊
IF:
4.9
论文数:
1.0W
被引数:
2.4W
机构
引用论文
Deeper cascaded peak-piloted network for weak expression recognition用于弱表情识别的更深级联峰值引导网络
VISUAL COMPUTER
IF2.9
Caw’s Walking State Recognition Based on Accelerometers and Gyroscopes Installed on Ear-Tags and Collar-Tags基于安装在耳标和项圈上的加速度计和陀螺仪的Caw步行状态识别
Collaborative discriminative multi-metric learning for facial expression recognition in video基于协同判别多度量学习的视频人脸表情识别
PATTERN RECOGNITION
IF7.6

