arrow
Return

Split-net: Dual transformer encoder with splitting scene text image for script identification

delete2025-06-04
delete0
PRE
AI
A
Ayush Roy
P
Palaiahnakote Shivakumara
U
Umapada Pal
刘程琳 (Cheng‐Lin Liu)
DOI:10.1016/j.patrec.2025.05.026delete
deleteOriginal
deleteOriginal request for help
deleteShare
deleteSave
Abstract

Abstract

En 中文
• Proposing a dual encoder for generating local features. • A novel Attention Module to integrate the transformer blocks of M-ViT encoders. • Gradient-based feature fusion to integrate the extracted features. • Experimental results show that the proposed method is better than the state-of-the-art.

Journal

Pattern Recognition Letters cover
Pattern Recognition Letters
IF:
3.3
Papers:
7.8K
Citations:
1.6W

Organization

No organization information available