arrow
返回

Efficient Query Processing for Scalable Web Search

delete2018-01-01
delete40
delete
OA
AI
N
Nicola Tonellotto *
C
Craig Macdonald
I
Iadh Ounis
DOI:10.1561/1500000057delete
delete原文链接
delete分享
delete收藏
查看原文
摘要

摘要

En 中文
Search engines are exceptionally important tools for accessing information in today's world. In satisfying the information needs of millions of users, the effectiveness (the quality of the search results) and the efficiency (the speed at which the results are returned to the users) of a search engine are two goals that form a natural trade-off, as techniques that improve the effectiveness of the search engine can also make it less efficient. Meanwhile, search engines continue to rapidly evolve, with larger indexes, more complex retrieval strategies and growing query volumes. Hence, there is a need for the development of efficient query processing infrastructures that make appropriate sacrifices in effectiveness in order to make gains in efficiency. This survey comprehensively reviews the foundations of search engines, from index layouts to basic term-at-a-time (TAAT) and document-at-a-time (DAAT) query processing strategies, while also providing the latest trends in the literature in efficient query processing, including the coherent and systematic reviews of techniques such as dynamic pruning and impact-sorted posting lists as well as their variants and optimisations. Our explanations of query processing strategies, for instance the WAND and BMW dynamic pruning algorithms, are presented with illustrative figures showing how the processing state changes as the algorithms progress. Moreover, acknowledging the recent trends in applying a cascading infrastructure within search systems, this survey describes techniques for efficiently integrating effective learned models, such as those obtained from learning-to-rank techniques. The survey also covers the selective application of query processing techniques, often achieved by predicting the response times of the search engine (known as query efficiency prediction), and making per-query tradeoffs between efficiency and effectiveness to ensure that the required retrieval speed targets can be met. Finally, the survey concludes with a summary of open directions in efficient search infrastructures, namely the use of signatures, real-time, energy-efficient and modern hardware and software architectures.
Keyword:
INVERTED FILES
DOCUMENT-RETRIEVAL
TEXT
COMPRESSION
PERFORMANCE
RANKING
MODELS
IMPLEMENTATIONS
STRATEGIES
RELEVANCE
AI总结

AI总结

对已上传原文的论文进行重点信息的提取,主要内容包括:简要概述、研究摘要、背景介绍、关键亮点、图文解析、展望与总结。

期刊

F
Foundations and Trends in Information Retrieval
IF:
12.9
论文数:
50
被引数:
824

机构

U
university of glasgow
学者数:
3.5W
论文数: 3.1W
被引数: 37
C
consiglio nazionale delle ricerche (cnr)
学者数:
6.2W
论文数: 5.7W
被引数: 48
引用论文

引用论文

暂无论文信息