arrow
返回

Empowering Visual Categorization With the GPU

delete2011-02-01
delete58
PRE
AI
K
Koen E. A. van de Sande *
T
Theo Gevers
C
Cees G. M. Snoek
DOI:10.1109/TMM.2010.2091400delete
delete原文链接
delete原文求助
delete分享
delete收藏
摘要

摘要

En 中文
Visual categorization is important to manage large collections of digital images and video, where textual metadata is often incomplete or simply unavailable. The bag-of-words model has become the most powerful method for visual categorization of images and video. Despite its high accuracy, a severe drawback of this model is its high computational cost. As the trend to increase computational power in newer CPU and GPU architectures is to increase their level of parallelism, exploiting this parallelism becomes an important direction to handle the computational cost of the bag-of-words approach. When optimizing a system based on the bag-of-words approach, the goal is to minimize the time it takes to process batches of images. In this paper, we analyze the bag-of-words model for visual categorization in terms of computational cost and identify two major bottlenecks: the quantization step and the classification step. We address these two bottlenecks by proposing two efficient algorithms for quantization and classification by exploiting the GPU hardware and the CUDA parallel programming model. The algorithms are designed to 1) keep categorization accuracy intact, 2) decompose the problem, and 3) give the same numerical results. In the experiments on large scale datasets, it is shown that, by using a parallel implementation on the Geforce GTX260GPU, classifying unseen images is 4.8 times faster than a quad-core CPU version on the Core i7 920, while giving the exact same numerical results. In addition, we show how the algorithms can be generalized to other applications, such as text retrieval and video retrieval. Moreover, when the obtained speedup is used to process extra video frames in a video retrieval benchmark, the accuracy of visual categorization is improved by 29%.
Keyword:
Bag-of-words
computational efficiency
General-Purpose computation on Graphics Processing Units (GPGPU)
image classification
image/video retrieval
multi-core processing
parallel processing
support vector machines
AI总结

AI总结

对已上传原文的论文进行重点信息的提取,主要内容包括:简要概述、研究摘要、背景介绍、关键亮点、图文解析、展望与总结。

期刊

IEEE Transactions on Multimedia 封面图
IEEE Transactions on Multimedia
IF:
9.7
论文数:
4.5K
被引数:
2.4W

机构

U
university of amsterdam
学者数:
6.0W
论文数: 5.1W
被引数: 94
引用论文

引用论文

A novel endoesophageal magnetic device to prevent gastroesophageal reflux
err2008-12-31
err0
PREAI
errMauro Bortolotti; Annamaria Grandis; Giosuè Mazzero
err分享
err收藏
Optimal paths for a bimolecular, light-driven engine
err2008-05-16
err0
PREAI
errS. J. Watowich; K. H. Hoffmann; R. S. Berry
err分享
err收藏
IMPLICATIONS OF GLOBAL PRICING POLICIES ON ACCESS TO INNOVATIVE DRUGS: THE CASE OF TRASTUZUMAB IN SEVEN LATIN AMERICAN COUNTRIES
err2015-05-20
err0
PREAI
errAndres Pichon-Riviere; Osvaldo Ulises Garay; Federico Augustovski; Carlos Vallejos; Leandro Huayanay; Maria del Pilar Navia Bueno; Alarico Rodriguez; Carlos José Coelho de Andrade; Jefferson Antonio Buendía; Michael Drummond
err分享
err收藏
Anatomy and Histology of the Lacrimal Fluid Drainage System
err2000-01-01
err0
errOAAI
errRieko KOMINAMI; Satoru YASUTAKA; Yutaka TANIGUCHI; Harumichi SHINOHARA
err分享
err收藏
Etiology of Epiphora
err2021-10-05
err0
errOAAI
errJeong Min Lee; Ji Sun Baek
err分享
err收藏
err分享
err收藏
学者 查看更多内容