arrow
返回

Automatic Data Augmentation from Massive Web Images for Deep Visual Recognition

delete2018-07-24
delete2
PRE
AI
Y
Yalong Bai *
K
Kuiyuan Yang
T
Tao Mei
W
Wei‐Ying Ma
T
Tiejun Zhao
DOI:10.1145/3204941delete
delete原文链接
delete原文求助
delete分享
delete收藏
摘要

摘要

En 中文
Large-scale image datasets and deep convolutional neural networks (DCNNs) are the two primary driving forces for the rapid progress in generic object recognition tasks in recent years. While lots of network architectures have been continuously designed to pursue lower error rates, few efforts are devoted to enlarging existing datasets due to high labeling costs and unfair comparison issues. In this article, we aim to achieve lower error rates by augmenting existing datasets in an automatic manner. Our method leverages both the web and DCNN, where the web provides massive images with rich contextual information, and DCNN replaces humans to automatically label images under the guidance of web contextual information. Experiments show that our method can automatically scale up existing datasets significantly from billions of web pages with high accuracy. The performance on object recognition tasks and transfer learning tasks have been significantly improved by using the automatically augmented datasets, which demonstrates that more supervisory information has been automatically gathered from the web. Both the dataset and models trained on the dataset have been made publicly available.
Keyword:
Dataset construction
deep convolutional neural network
dataset augmentation
AI总结

AI总结

对已上传原文的论文进行重点信息的提取,主要内容包括:简要概述、研究摘要、背景介绍、关键亮点、图文解析、展望与总结。

期刊

ACM Transactions on Multimedia Computing Communications and Applications 封面图
ACM Transactions on Multimedia Computing Communications and Applications
IF:
6
论文数:
2.0K
被引数:
5.4K

机构

H
harbin institute of technology
学者数:
8.0W
论文数: 6.6W
被引数: 66