返回
Audio based depression detection using Convolutional Autoencoder
DOI:10.1016/j.eswa.2021.116076.png)
摘要
En 中文
Depression is a serious and common psychological disorder that requires early diagnosis and treatment. In severe episodes the condition may result in suicidal thoughts. Recently, the need for building an effective audio-based Automatic Depression Detection (ADD) system has sparked the interest of the research community. To date, most of the reported approaches to recognize depression rely on hand-crafted feature extraction for audio data representation. They combine wide variety of audio-related features to improve the classification performance. However, combining many hand-crafted features including relevant and less-relevant can enlarge the feature space which can lead to high-dimensionality issues as not all the features would carry significant information regarding depression. Having high number of features can make the pattern recognition more difficult and increase the risk of overfitting. To overcome these limitations, an audio-based framework of depression detection which includes an adaptation of a deep learning (DL) technique is proposed to automatically extract the highly relevant and compact feature set. This proposed framework uses an end-to-end Convolutional Neural Network based Autoencoder (CNN AE) technique to learn the highly relevant and discriminative features from raw sequential audio data, and hence to detect depressed people more accurately. In addition, to address the sample imbalance problem we use a cluster-based sampling technique which highly reduces the risk of bias towards the major class (non-depressed). To evaluate the performance and effectiveness of the proposed pipeline, we perform the experiments on Distress Analysis Interview Corpus-Wizard of Oz (DAIC-WOZ) dataset and compare them with the hand-crafted feature extraction methods and other outstanding studies in this domain. The results show that proposed method outperforms other well-known audio-based ADD models with at least 7% improvement in F-measure for classifying depression.
Keyword:
Audio depression detection
Semi-supervised learning
Convolutional Autoencoder
Early depression detection
期刊
IF:
7.5
论文数:
2.9W
被引数:
10.2W
机构
引用论文
Market Impacts of a Transmission Investment: Evidence from the ERCOT Competitive Renewable Energy Zones Project
Energies
IF0
Projections of global mortality and burden of disease from 2002 to 20302030年全球死亡率和疾病负担2002年预测
PLOS MEDICINE
IF9.9
EAMA: Empirically adjusted meta-analysis for large-scale simultaneous hypothesis testing in genomic experiments
PLOS ONE
IF0
Automatic Emotion Recognition Using Temporal Multimodal Deep Learning基于时间多模态深度学习的自动情感识别
IEEE ACCESS
IF3.6
Long Short Term Memory Hyperparameter Optimization for a Neural Network Based Emotion Recognition Framework
IEEE ACCESS
IF3.6
Multimodal Depression Detection: Fusion of Electroencephalography and Paralinguistic Behaviors Using a Novel Strategy for Classifier Ensemble多模态抑郁症检测: 使用新的分类器集成策略融合脑电图和副语言行为

