arrow
返回

Machine learning-based infant crying interpretation

delete2024-02-08
delete1
delete
OA
AI
M
Mohammed Hammoud
M
Melaku N. Getahun
A
Anna Baldycheva *
A
Andrey Somov
DOI:10.3389/frai.2024.1337356delete
delete原文链接
delete分享
delete收藏
查看原文
摘要

摘要

En 中文
Crying is an inevitable character trait that occurs throughout the growth of infants, under conditions where the caregiver may have difficulty interpreting the underlying cause of the cry. Crying can be treated as an audio signal that carries a message about the infant's state, such as discomfort, hunger, and sickness. The primary infant caregiver requires traditional ways of understanding these feelings. Failing to understand them correctly can cause severe problems. Several methods attempt to solve this problem; however, proper audio feature representation and classifiers are necessary for better results. This study uses time-, frequency-, and time-frequency-domain feature representations to gain in-depth information from the data. The time-domain features include zero-crossing rate (ZCR) and root mean square (RMS), the frequency-domain feature includes the Mel-spectrogram, and the time-frequency-domain feature includes Mel-frequency cepstral coefficients (MFCCs). Moreover, time-series imaging algorithms are applied to transform 20 MFCC features into images using different algorithms: Gramian angular difference fields, Gramian angular summation fields, Markov transition fields, recurrence plots, and RGB GAF. Then, these features are provided to different machine learning classifiers, such as decision tree, random forest, K nearest neighbors, and bagging. The use of MFCCs, ZCR, and RMS as features achieved high performance, outperforming state of the art (SOTA). Optimal parameters are found via the grid search method using 10-fold cross-validation. Our MFCC-based random forest (RF) classifier approach achieved an accuracy of 96.39%, outperforming SOTA, the scalogram-based shuffleNet classifier, which had an accuracy of 95.17%.
Keyword:
time-series classification
Mel-frequency cepstral coefficient
spectrogram
machine learning
audio processing
time-series imagining
AI总结

AI总结

对已上传原文的论文进行重点信息的提取,主要内容包括:简要概述、研究摘要、背景介绍、关键亮点、图文解析、展望与总结。

期刊

F
Frontiers in Artificial Intelligence
IF:
4.7
论文数:
2.4K
被引数:
4.4K

机构

S
skolkovo institute of science & technology
学者数:
3.3K
论文数: 2.3K
被引数: 1
U
University of Exeter
学者数:
2.0W
论文数: 2.1W
被引数: 3.6W
引用论文

引用论文

Infant Cry Language Analysis and Recognition: An Experimental Approach
err2019-05-01
err29
PREAI
errLiu, Lichuan; Li, Wei; Wu, Xianwen; Zhou, Benjamin X.
err分享
err收藏
How can cry acoustics associate newborns' distress levels with neurophysiological and behavioral signals?
err2023-09-20
err4
errOAAI
errLaguna, Ana; Pusil, Sandra; Acero-Pousa, Irene; Zegarra-Valdivia, Jonathan Adrian; Paltrinieri, Anna Lucia; Bazan, Angel; Piras, Paolo; Perera, Claudia Palomares i; Garcia-Algar, Oscar; Orlandi, Silvia
err分享
err收藏
Excitation of Nd3+ and Tm3+ by the energy transfer from Si nanocrystals
err2002-03-01
err0
PREAI
errKei Watanabe; Hiroyuki Tamaoka; Minoru Fujii; Kazuyuki Moriwaki; Shinji Hayashi
err分享
err收藏
Combining PHM information and system architecture to support aircraft maintenance planning结合PHM信息与系统架构以支持飞机维护计划
err2013-04-01
err0
PREAI
errFelipe Augusto Sviaghin Ferri; Leonardo Ramos Rodrigues; Joao Paulo Pordeus Gomes; Ivo Paixao de Medeiros; Roberto Kawakami Harrop Galvao; Cairo Lucio Nascimento
err分享
err收藏
学者 查看更多内容