Return
Split-net: Dual transformer encoder with splitting scene text image for script identification
DOI:10.1016/j.patrec.2025.05.026.png)
Abstract
En 中文
• Proposing a dual encoder for generating local features. • A novel Attention Module to integrate the transformer blocks of M-ViT encoders. • Gradient-based feature fusion to integrate the extracted features. • Experimental results show that the proposed method is better than the state-of-the-art.
Journal
IF:
3.3
Papers:
7.8K
Citations:
1.6W
Organization
No organization information available

