1
Return

TPA-ConvNeXt: Trigonometric Phase Attention for Robust Retinal Disease Classification Across Fundus and OCT

delete2026-07-08
delete0
delete
OA
AI
N
Nebras Sobahi
O
Orhan Atila *
S
Salih Taha Alperen Özçelik
A
Abdulkadir Şengür
DOI:10.3390/bioengineering13070781delete
deleteOriginal
deleteShare
deleteSave
View PDF
Abstract

Abstract

En 中文
This study proposes TPA-ConvNeXt, a ConvNeXt-Small-based deep learning architecture for retinal image classification using a phase-based trigonometric attention mechanism. The proposed Trigonometric Phase Attention (TPA) block recalibrates feature maps through learnable phase and amplitude modulation derived from channel-wise and spatial context. In addition, a stable fine-tuning strategy combining layer-wise learning rate decay (LLRD), linear warmup, and cosine annealing is employed to adapt pretrained backbones to medical image data. The method was evaluated on both fundus and optical coherence tomography (OCT) datasets. On the HYAMD fundus dataset, 5-fold cross-validation yielded a mean Macro F1 of 0.8940 ± 0.0403 and an ROC-AUC of 0.9449. External fundus validation further demonstrated cross-dataset robustness, achieving 0.9312 Macro F1 and 0.9803 ROC-AUC when trained on HYAMD and tested on AMDNet23. In OCT experiments, the model achieved Macro F1 scores of 0.9790 on MAK1_OCT, 0.9387 on OCTDL, and 0.9989 on the cleaned OCT2017 benchmark test set after MD5 duplicate removal. Ablation results showed that warmup contributed most strongly to stable optimization, while the proposed TPA block provided a smaller but consistent performance gain. Additional statistical analysis across the five HYAMD folds showed that TPA-ConvNeXt achieved comparable performance to the ConvNeXt-Small baseline, with a small numerical accuracy difference that did not reach statistical significance. Grad-CAM visualizations indicated that the model focused on clinically relevant retinal regions, and complexity analysis showed that the full model required 51.82 M parameters and 17.41 GFLOPs, adding only 2.36 M parameters and 0.02 GFLOPs over the ConvNeXt-Small baseline. These findings suggest that TPA-ConvNeXt provides a robust and generalizable framework for retinal image classification across both fundus and OCT modalities.
Keywords:
Trigonometric Phase Attention
ConvNeXt
layer-wise learning rate decay
fundus images
OCT images
HYAMD
OCT2017
AI Summary

AI Summary

Key information extracted from the uploaded paper, including a brief overview, abstract, background, key highlights, visual analysis, and future outlook.

Journal

B
Bioengineering
IF:
3.7
Papers:
5.9K
Citations:
1.3W

Organization

F
Firat University
Scholars:
3.9K
Papers: 3.8K
Citations: 43
K
King Abdulaziz University
Scholars:
1.9W
Papers: 1.9W
Citations: 3.3W
Bingöl University cover
Bingöl University
Scholars:
116
Papers: 84
Citations: 940
Cited Papers

Cited Papers

Citing Papers

Citing Papers