vix.ing · top · new · best · stats · spec

InfoGather+

2013/06/22 by Meihui Zhang, Kaushik Chakrabarti · 2 citations
Decision Sciences · Computer Science · #Data Quality and Management #Web Data Mining and Analysis #Data Management and Algorithms

paper · doi:10.1145/2463676.2465276

openalex publication_date 2013/06/22 · openalex created_date 2016/06/24 · openalex updated_date 2026/07/29

Abstract

Users often need to gather information about "entities" of interest. Recent efforts try to automate this task by leveraging the vast corpus of HTML tables; this is referred to as "entity augmentation". The accuracy of entity augmentation critically depends on semantic relationships between web tables as well as semantic labels of those tables. Current techniques work well for string-valued and static attributes but perform poorly for numeric and time-varying attributes.

Cited by

Related