arrow
返回

Automatic Web-based relational data imputation

delete2018-02-07
delete0
PRE
AI
H
Hailong Liu *
Z
Zhanhuai Li
C
Chen, Qun
C
Chen Zhao-qiang
DOI:10.1007/s11704-016-6319-3delete
delete原文链接
delete原文求助
delete分享
delete收藏
摘要

摘要

En 中文
Data incompleteness is one of the most important data quality problems in enterprise information systems. Most existing data imputing techniques just deduce approximate values for the incomplete attributes by means of some specific data quality rules or some mathematical methods. Unfortunately, approximation may be far away from the truth. Furthermore, when observed data is inadequate, they will not work well. The World Wide Web (WWW) has become the most important and the most widely used information source. Several current works have proven that using Web data can augment the quality of databases. In this paper, we propose a Web-based relational data imputing framework, which tries to automatically retrieve real values from the WWW for the incomplete attributes. In the paper, we try to take full advantage of relations among different kinds of objects based on the idea that the same kind of things must have the same kind of relations with their relatives in a specific world. Our proposed techniques consist of two automatic query formulation algorithms and one graph-based candidates extraction model. Several evaluations are proposed on two high-quality real datasets and one poor-quality real dataset to prove the effectiveness of our approaches.
Keyword:
data incompleteness
imputation
World Wide Web
query formulation
candidate selection
semantic relation
AI总结

AI总结

对已上传原文的论文进行重点信息的提取,主要内容包括:简要概述、研究摘要、背景介绍、关键亮点、图文解析、展望与总结。

期刊

Frontiers of Computer Science 封面图
Frontiers of Computer Science
IF:
4.6
论文数:
1.6K
被引数:
2.8K

机构

N
Northwestern Polytechnical University
学者数:
4.6W
论文数: 3.7W
被引数: 5.3W