arrow
返回

Enhancing human computer interaction with coot optimization and deep learning for multi language identification

delete2024-10-03
delete3
delete
OA
AI
E
Elvir Akhmetshin
G
Galina Vladimirovna Meshkova
M
Maria V. Mikhailova
R
Rustem Shichiyakh
J
Joshi, Gyanendra Prasad *
W
Woong Cho *
DOI:10.1038/s41598-024-74327-2delete
delete原文链接
delete分享
delete收藏
查看原文
摘要

摘要

En 中文
Human-Computer Interaction (HCI) is a multidisciplinary field focused on designing and utilizing computer technology, underlining the interaction interface between computers and humans. HCI aims to generate systems that allow consumers to relate to computers effectively, efficiently, and pleasantly. Multiple Spoken Language Identification (SLI) for HCI (MSLI for HCI) denotes the ability of a computer system to recognize and distinguish various spoken languages to enable more complete and handy interactions among consumers and technology. SLI utilizing deep learning (DL) involves using artificial neural networks (ANNs), a subset of DL models, to automatically detect and recognize the language spoken in an audio signal. DL techniques, particularly neural networks (NNs), have succeeded in various pattern detection tasks, including speech and language processing. This paper develops a novel Coot Optimizer Algorithm with a DL-Driven Multiple SLI and Detection (COADL-MSLID) technique for HCI applications. The COADL-MSLID approach aims to detect multiple spoken languages from the input audio regardless of gender, speaking style, and age. In the COADL-MSLID technique, the audio files are transformed into spectrogram images as a primary step. Besides, the COADL-MSLID technique employs the SqueezeNet model to produce feature vectors, and the COA is applied to the hyperparameter range of the SqueezeNet method. The COADL-MSLID technique exploits the SLID process's convolutional autoencoder (CAE) model. To underline the importance of the COADL-MSLID technique, a series of experiments were conducted on the benchmark dataset. The experimentation validation of the COADL-MSLID technique exhibits a greater accuracy result of 98.33% over other techniques.
Keyword:
Spoken language identification
Human-computer interface
Coot optimization algorithm
SqueezeNet
Deep learning
AI总结

AI总结

对已上传原文的论文进行重点信息的提取,主要内容包括:简要概述、研究摘要、背景介绍、关键亮点、图文解析、展望与总结。

期刊

Scientific Reports 封面图
Scientific Reports
IF:
3.9
论文数:
28.0W
被引数:
83.5W

机构

S
sechenov first moscow state medical university
学者数:
6.0K
论文数: 3.5K
被引数: 15
K
Kangwon National University
学者数:
1.0W
论文数: 9.4K
被引数: 13
K
Kazan Federal University
学者数:
4.4K
论文数: 2.6K
被引数: 4.0K
K
Kuban State Agrarian University
学者数:
49
论文数: 52
被引数: 29
B
Bauman Moscow State Technical University
学者数:
1.5K
论文数: 761
被引数: 537
学者 查看更多机构
引用论文

引用论文

PI3‐kinase pathway biomarkers in oral cancer and tumor immune cellsPI3‐kinase通路生物标志物在口腔癌及肿瘤免疫细胞中的作用
err2018-12-16
err0
errOAAI
errMohammad Y. Ibrahim; Maria I. Nunez; Nusrat Harun; J. Jack Lee; Adel K. El‐Naggar; Renata Ferrarotto; Ignacio Wistuba; Jeffrey Myers; Bonnie S. Glisson; William N. William
err分享
err收藏
Decoding lip language using triboelectric sensors with deep learning
err2022-03-17
err136
errOAAI
errLu, Yijia; Tian, Han; Cheng, Jia; Zhu, Fei; Liu, Bin; Wei, Shanshan; Ji, Linhong; Wang, Zhong Lin
err分享
err收藏
err分享
err收藏
The unexpected effect of cyclosporin A on CD56+CD16− and CD56+CD16+ natural killer cell subpopulations
err2007-09-01
err0
errOAAI
errHongbo Wang; Bartosz Grzywacz; David Sukovich; Valarie McCullar; Qing Cao; Alisa B. Lee; Bruce R. Blazar; David N. Cornfield; Jeffrey S. Miller; Michael R. Verneris
err分享
err收藏
Toward a Vision-Based Intelligent System: A Stacked Encoded Deep Learning Framework for Sign Language Recognition迈向基于视觉的智能系统: 用于手语识别的堆栈编码深度学习框架
errSENSORS
IF3.5
err2023-11-09
err8
errOAAI
errIslam, Muhammad; Aloraini, Mohammed; Aladhadh, Suliman; Habib, Shabana; Khan, Asma; Alabdulatif, Abduatif; Alanazi, Turki M.
err分享
err收藏
学者 查看更多内容