arrow
返回

High dimensional nearest neighbor searching

delete2006-09-01
delete22
PRE
AI
H
Hakan Ferhatosmanoğlu *
E
Ertem Tuncel
D
Divyakant Agrawal
A
Amr El Abbadi
DOI:10.1016/j.is.2005.01.001delete
delete原文链接
delete原文求助
delete分享
delete收藏
摘要

摘要

En 中文
As databases increasingly integrate different types of information such as time-series, multimedia and scientific data, it becomes necessary to support efficient retrieval of multi-dimensional data. Both the dimensionality and the amount of data that needs to be processed are increasing rapidly. As a result of the scale and high dimensional nature, the traditional techniques have proven inadequate. In this paper, we propose search techniques that are effective especially for large high dimensional data sets. We first propose VA(+)-file technique which is based on scalar quantization of the data. VA(+)-file is especially useful for searching exact nearest neighbors (NN) in non-uniform high dimensional data sets. We then discuss how to improve the search and make it progressive by allowing some approximations in the query result. We develop a general framework for approximate NN queries, discuss various approaches for progressive processing of similarity queries, and develop a metric for evaluation of such techniques. Finally, a new technique based on clustering is proposed, which merges the benefits of various approaches for progressive similarity searching. Extensive experimental evaluation is performed on several real-life data sets. The evaluation establishes the superiority of the proposed techniques over the existing techniques for high dimensional similarity searching. The techniques proposed in this paper are effective for real-life data sets, which are typically non-uniform, and they are scalable with respect to both dimensionality and size of the data set. (C) 2005 Elsevier B.V. All rights reserved.
Keyword:
high dimensional data
nearest neighbor queries
indexing
similarity search
approximate and progressive search
nonuniform data
scalability
performance
AI总结

AI总结

对已上传原文的论文进行重点信息的提取,主要内容包括:简要概述、研究摘要、背景介绍、关键亮点、图文解析、展望与总结。

期刊

Enterprise Information Systems 封面图
Enterprise Information Systems
IF:
3.9
论文数:
2.8K
被引数:
1.8K

机构

暂无机构信息
引用论文

引用论文

Multidimensional access methods
err1998-06-01
err895
errOAAI
errGaede, V; Gunther, O
err分享
err收藏
Towards energy-autonomous wake-up receiver using Visible Light Communication
err2016-01-01
err0
errOAAI
errJoyce Sariol Ramos; Ilker Demirkol; Josep Paradells; Daniel Vossing; Karim M. Gad; Martin Kasemann
err分享
err收藏
The PHEV Charging Infrastructure Planning (PCIP) Problem
err2010-06-27
err0
PREAI
errYogesh Dashora; John W Barnes; Rekha S Pillai; Todd E Combs; Michael Hilliard; Madhu S Chinthavali
err分享
err收藏
Quantitative analysis of seismogenic shear-induced turbulence in lake sediments
err2010-04-01
err0
PREAI
errNadav Wetzler; Shmuel Marco; Eyal Heifetz
err分享
err收藏
Gd(III)–Gd(III) Relaxation-Induced Dipolar Modulation Enhancement for In-Cell Electron Paramagnetic Resonance Distance Determination
err2019-03-13
err0
errOAAI
errMykhailo Azarkh; Anna Bieber; Mian Qi; Jörg W. A. Fischer; Maxim Yulikov; Adelheid Godt; Malte Drescher
err分享
err收藏
学者 查看更多内容