arrow
返回

TextMountain: Accurate scene text detection via instance segmentation

delete2021-02-01
delete76
delete
OA
AI
Y
Yixing Zhu
杜
杜俊 (Jun Du) *
DOI:10.1016/j.patcog.2020.107336delete
delete原文链接
delete分享
delete收藏
查看原文
摘要

摘要

En 中文
In this paper, we propose a novel scene text detection method named TextMountain. The key idea of TextMountain is making full use of border-center information. Different from previous works that treat center-border as a binary classification problem, we predict text center-border probability (TCBP) and text center-direction (TCD). The TCBP is just like a mountain whose top is text center and foot is text border. The mountaintop can separate text instances which cannot be easily achieved using semantic segmentation map and its rising direction can plan a road to top for each pixel on mountain foot at the group stage. The TCD helps TCBP learning better. Our label rules will not lead to the ambiguous problem with the transformation of angle, so the proposed method is robust to multi-oriented text and can also handle well curved text. In inference stage, each pixel at the mountain foot needs to search the path to the mountaintop and this process can be efficiently completed in parallel, yielding the efficiency of our method compared with others. The experiments on MLT, ICDAR2015, RCTW-17 and SCUT-CTW150 0 datasets demonstrate that the proposed method achieves better or comparable performance in terms of both accuracy and efficiency. It is worth mentioning our method achieves an F-measure of 76.85% on MLT which outperforms the previous methods by a large margin. Code will be made available. (c) 2020 Elsevier Ltd. All rights reserved.
Keyword:
Scene text detection
Curved text
Multi-oriented text
CNN
Deep learning
AI总结

AI总结

对已上传原文的论文进行重点信息的提取,主要内容包括:简要概述、研究摘要、背景介绍、关键亮点、图文解析、展望与总结。

期刊

Pattern Recognition 封面图
Pattern Recognition
IF:
7.6
论文数:
1.3W
被引数:
4.5W

机构

C
chinese academy of sciences
学者数:
56.7W
论文数: 45.0W
被引数: 704
引用论文

引用论文

err分享
err收藏
Semantic Understanding of Scenes Through the ADE20K Dataset通过ADE20K数据集对场景进行语义理解
err2018-12-07
err934
PREAI
errZhou, Bolei; Zhao, Hang; Puig, Xavier; Xiao, Tete; Fidler, Sanja; Barriuso, Adela; Torralba, Antonio
err分享
err收藏
SegLink plus plus : Detecting Dense and Arbitrary-shaped Scene Text by Instance-aware Component Grouping
err2019-12-01
err115
PREAI
errTang, Jun; Yang, Zhibo; Wang, Yongpan; Zheng, Qi; Xu, Yongchao; Bai, Xiang
err分享
err收藏
Text/non-text image classification in the wild with convolutional neural networks
err2017-06-01
err76
PREAI
errBai, Xiang; Shi, Baoguang; Zhang, Chengquan; Cai, Xuan; Qi, Li
err分享
err收藏
A blind deconvolution model for scene text detection and recognition in video
err2016-06-01
err37
PREAI
errKhare, Vijeta; Shivakumara, Palaiahnakote; Raveendran, Paramesran; Blumenstein, Michael
err分享
err收藏
Curved scene text detection via transverse and longitudinal sequence connection
err2019-06-01
err187
errOAAI
errLiu, Yuliang; Jin, Lianwen; Zhang, Shuaitao; Luo, Canjie; Zhang, Sheng
err分享
err收藏
err分享
err收藏
The Pascal Visual Object Classes (VOC) ChallengePascal视觉对象课程 (VOC) 挑战
err2009-09-09
err9.0K
PREAI
errEveringham, Mark; Van Gool, Luc; Williams, Christopher K. I.; Winn, John; Zisserman, Andrew
err分享
err收藏
学者 查看更多内容