arrow
返回

OCR post-correction for detecting adversarial text images

delete2022-05-01
delete4
delete
OA
AI
N
Niddal Imam *
V
Vassilios G. Vassilakis
DOI:10.1016/j.jisa.2022.103170delete
delete原文链接
delete分享
delete收藏
查看原文
摘要

摘要

En 中文
The amount of images with embedded text shared on Online Social Networks (OSNs), such as Twitter orFacebook has been growing in recent years. It is becoming important to analyse the images uploaded into theseplatforms, as adversaries may spread images with toxic content or misinformation (i.e. spam). Optical characterrecognition (OCR) systems have been used to detect images with malicious content, where the embedded textgets extracted and classified using machine learning algorithms. However, most existing OCR-based systemsare adversary-agnostic models, in which the extracted text from an image is not checked by humans before theclassification. Consequently, these fully automated models become vulnerable to minor modifications of images'pixels or textual content (e.g.,character-levelperturbations), which do not affect human understanding, but couldcause the OCR systems to misrecognise the embedded text. In this paper, we propose an OCR post-correctionalgorithm to improve the robustness of OCR-based systems against images with perturbed embedded texts.Experimental results showed that our proposed algorithm improves the robustness of three state-of-the-art OCRmodels with at least 10% against adversarial text images, and it outperforms five spellcheckers in correctingadversarial text. Also, we evaluated the perceptibility of our adversarial images, and this study showed that91% of the participants were able to correctly recognise the adversarial text images. Additionally, we developedan adversary-aware OCR-based system for detecting adversarial text images using the proposed algorithm, andour evaluation results showed considerable improvement in the performance of an OCR-based system.
Keyword:
Deep learning
Spam image
OCR
Text recognition
Text classification
Adversarial text attack
AI总结

AI总结

对已上传原文的论文进行重点信息的提取,主要内容包括:简要概述、研究摘要、背景介绍、关键亮点、图文解析、展望与总结。

期刊

Journal of Information Security and Applications 封面图
Journal of Information Security and Applications
IF:
3.7
论文数:
2.0K
被引数:
4.9K

机构

U
university of york - uk
学者数:
1.5W
论文数: 1.5W
被引数: 15
引用论文

引用论文

“It's almost like they're trying to hide it”: How User-Provided Image Descriptions Have Failed to Make Twitter Accessible
err2019-05-13
err0
errOAAI
errCole Gleason; Patrick Carrington; Cameron Cassidy; Meredith Ringel Morris; Kris M. Kitani; Jeffrey P. Bigham
err分享
err收藏
err分享
err收藏
err分享
err收藏
SSVEP signatures of binocular rivalry during simultaneous EEG and fMRI
err2015-03-01
err0
errOAAI
errKeith W. Jamison; Abhrajeet V. Roy; Sheng He; Stephen A. Engel; Bin He
err分享
err收藏
err分享
err收藏
学者 查看更多内容