arrow
返回

Deep operational audio-visual emotion recognition

delete2024-07-01
delete0
PRE
AI
K
Kaan Aktürk
A
Ali Seydi Keçeli *
DOI:10.1016/j.neucom.2024.127713delete
delete原文链接
delete原文求助
delete分享
delete收藏
摘要

摘要

En 中文
Emotions play a large role in interpersonal communication, marketing, healthcare and the service industry. For this reason, much research has been carried out on emotion classification until today. Audio-visual emotion recognition is a field within artificial intelligence and machine learning that focuses on recognizing and understanding human emotions from both visual and audio cues. It combines computer vision and audio processing techniques to analyze and interpret emotional states expressed by individuals. This paper presents a deep learning model developed over an operational neural network using multiple inputs and aimed at audio-visual emotion recognition. The proposed network utilizes both visual and audio information in an end to end approach. The primary objective of this work is to demonstrate that multi-input models can produce more efficient outcomes compared to single-input models in emotion classification. Another objective is to demonstrate the superior performance of weight calculation methods employed in operational neural networks compared to the conventional weight calculation methods used in convolutional neural networks. Therefore, we want to demonstrate that substituting convolutional neural network approaches with operational neural network methods can yield superior outcomes in emotion categorization models. In the proposed architecture, regular convolutional layers are replaced with operational layers. The experimental results demonstrate that the operational convolutional architecture performs better compared to the classical convolutional neural network architecture.
Keyword:
Audio-visual emotion classification
Operational neural network
Visual geometry group
Multi-input classification

期刊

Neurocomputing 封面图
Neurocomputing
IF:
6.5
论文数:
2.5W
被引数:
6.5W

机构

H
Hacettepe University
学者数:
1.2W
论文数: 1.0W
被引数: 11
引用论文

引用论文

Improving the Performance of VGG Through Different Granularity Feature Combinations
err2021-01-01
err11
errOAAI
errZhou, Yuepeng; Chang, Huiyou; Lu, Yonghe; Lu, Xili; Zhou, Ruqi
err分享
err收藏
Speech Emotion Recognition Using Convolution Neural Networks and Multi-Head Convolutional Transformer基于卷积神经网络和多头卷积变换器的语音情感识别
errSENSORS
IF3.5
err2023-07-07
err15
errOAAI
errUllah, Rizwan; Asif, Muhammad; Shah, Wahab Ali; Anjam, Fakhar; Ullah, Ibrar; Khurshaid, Tahir; Wuttisittikulkij, Lunchakorn; Shah, Shashi; Ali, Syed Mansoor; Alibakhshikenari, Mohammad
err分享
err收藏
Molecular Upconversion Nanoparticles for Live-Cell Imaging活细胞成像的分子上转换纳米颗粒
err2025-02-12
err0
PREAI
errHaye, L; Pini, F; Soro, LK; Knighton, RC; Fayad, N; Benard, M; Gagliazzo, F; Light, ME; Natile, MM; Charbonnière, LJ; Hildebrandt, N; Reisch, A
err分享
err收藏
Multimodal Emotion Recognition With Transformer-Based Self Supervised Feature Fusion基于Transformer自监督特征融合的多模态情感识别
err2020-01-01
err89
errOAAI
errSiriwardhana, Shamane; Kaluarachchi, Tharindu; Billinghurst, Mark; Nanayakkara, Suranga
err分享
err收藏
学者 查看更多内容