arrow
Return

Deep learning in auditory attention decoding: A systematic review

delete2026-03-12
delete0
delete
OA
AI
J
Jiashu Yang *
M
Mengjie Huang *
杨瑞 cover
杨瑞 (Rui Yang) *
DOI:10.1080/21642583.2026.2640250delete
deleteOriginal
deleteOriginal request for help
deleteShare
deleteSave
Abstract

Abstract

En 中文
Auditory attention decoding (AAD) utilizes electroencephalography (EEG) to detect an individual's attentional focus in noisy, multi-speaker environments, serving as a cornerstone for next-generation neuro-steered hearing aids. Although deep learning has recently surpassed traditional linear models in decoding accuracy, a holistic overview connecting preprocessing pipelines with advanced model architectures is lacking. This systematic review addresses this deficiency by analyzing 82 studies selected via PRISMA guidelines. We comprehensively dissect the AAD framework, from data processing to the evolution of architectures like convolutional neural networks (CNNs), graph neural networks (GNNs), and Transformers. Beyond reviewing trends, we synthesize practical design guidance, recommending the alignment of preprocessing complexity with model capacity and the use of hybrid architectures to capture spatiotemporal dynamics. Crucially, this study highlights persistent challenges impeding real-world transferability, including the reliance on high-density EEG montages incompatible with wearables, the computational latency on edge devices, and the lack of realistic acoustic scenarios. By identifying these bottlenecks and synthesizing effective design choices, this review offers a structured reference for developing robust, low-latency, and subject-independent auditory attention detection systems suitable for practical brain-computer interface applications.
Keywords:
Auditory attention decoding
deep learning
electroencephalography

Journal

S
Systems Science and Control Engineering
IF:
0
Papers:
1
Citations:
0

Organization

X
Xi'an Jiaotong-Liverpool University
Scholars:
443
Papers: 257
Citations: 5.4K