arrow
Return

Deep learning based text detection using resnet for feature extraction

delete2023-05-03
delete2
PRE
AI
L
Li-Kun Huang
H
Hsiao‐Ting Tseng
C
Chen-Chiung Hsieh *
C
Chih-Sin Yang
DOI:10.1007/s11042-023-15449-zdelete
deleteOriginal
deleteOriginal request for help
deleteShare
deleteSave
Abstract

Abstract

En 中文
Popular deep learning models for text segmentation include CTPN, EAST, and PixelLink. However, they are not very well capable of dealing with the images containing densely distributed characters, and those characters may be connected. For these problems, the ResNet with excellent sensitivity for feature extraction is used to replace those embedded convolution neural networks in the main structures of CTPN and EAST. The experimental results showed that a better feature extraction network could significantly improve the precision of text localization. Noteworthy, the results indicate that the accuracy of modified EAST with ResNet101 would be the highest with a deeper depth and larger width of ResNet. The accuracy of text segmentation on ICDAR 2015 is 83.4% which is 7% higher than the original PVANET-EAST. The text detection accuracy is 83.9% on the untrained scanned document. Also, it achieved an accuracy of 86.3% when applied to self-collected Chinese calligraphy. Those results demonstrated that text detection using ResNet is a better improvement for OCR applications.
Keywords:
Text segmentation
Deep learning models
Convolutional neural network
Feature extraction
Optical character recognition

Journal

Multimedia Tools and Applications cover
Multimedia Tools and Applications
IF:
3
Papers:
1.9W
Citations:
3.2W

Organization

N
National Tsing Hua University
Scholars:
1.6W
Papers: 1.4W
Citations: 1.7W
N
National Central University
Scholars:
1.0W
Papers: 8.5K
Citations: 6.4K