arrow
Return

Proxy-Enriched Imputation on Contextually Incomplete Web Tables

delete2026-01-01
delete0
PRE
AI
D
Denis Nagel *
J
Jonas MeiSSner
N
Niklas Kiehne
W
Wolf‐Tilo Balke
DOI:10.1007/978-3-032-09527-5_12delete
deleteOriginal
deleteOriginal request for help
deleteShare
deleteSave
Abstract

Abstract

En 中文
Structured data in the form of Web tables and open data repositories lays the foundation for the Semantic Web.However, quality concerns are often raised, mostly with respect to accuracy and completeness. While missing value imputation is the go-to solution to fill in blanks, (1) it can only approximate known unknowns and (2) struggles if values are missing not at random. As a remedy, this paper combines semantically rich narratives with proven-quality knowledge graphs to dynamically assess the completeness of individual data sets while accounting for the user's intent, too. Having determined which values are actually missing, we leverage state-of-the-art NLP techniques to identify functionally dependent attributes as proxies for later value imputation. Being functionally dependent (at least to some degree), these attributes provide the necessary context in the sense of relatedness allowing for more sophisticated imputation techniques. As a proof of concept we demonstrate our approach's benefits in a real world setting using real-life narratives on the open data repository of the World Health Organization.
Keywords:
Narrative Intelligence
Open Data
Web Tables
Imputation

Journal

S
SEMANTIC WEB-ISWC 2025, PT I
IF:
0
Papers:
30
Citations:
0

Organization

B
Braunschweig University of Technology
Scholars:
7.7K
Papers: 6.6K
Citations: 19