Puttingweb tables into context

Publikation: Beitrag in Buch/Konferenzbericht/Sammelband/GutachtenBeitrag in KonferenzbandBeigetragenBegutachtung

Beitragende

  • Katrin Braunschweig - , Technische Universität Dresden (Autor:in)
  • Maik Thiele - , Technische Universität Dresden (Autor:in)
  • Elvis Koci - , Technische Universität Dresden (Autor:in)
  • Wolfgang Lehner - , Professur für Datenbanken (Autor:in)

Abstract

Web tables are a valuable source of information used in many application areas. However, to exploit Web tables it is necessary to understand their content and intention which is impeded by their ambiguous semantics and inconsistencies. Therefore, additional context information, e.g.Text in which the tables are embedded, is needed to support the table understanding process. In this paper, we propose a novel contextualization approach that 1) splits the table context in topically coherent paragraphs, 2) provides a similarity measure that is able to match each paragraph to the table in question and 3) ranks these paragraphs according to their relevance. Each step is accompanied by an experimental evaluation on real-world data showing that our approach is feasible and effectively identifies the most relevant context for a given Web table.

Details

OriginalspracheEnglisch
TitelKDIR 2016 - 8th International Conference on Knowledge Discovery and Information Retrieval
Redakteure/-innenAna Fred, Jan Dietz, David Aveiro, Kecheng Liu, Jorge Bernardino, Joaquim Filipe, Joaquim Filipe
Herausgeber (Verlag)SCITEPRESS - Science and Technology Publications
Seiten158-165
Seitenumfang8
ISBN (elektronisch)9789897582035
PublikationsstatusVeröffentlicht - 2016
Peer-Review-StatusJa

Konferenz

Titel8th International Joint Conference on Knowledge Discovery, Knowledge Engineering and Knowledge Management, IC3K 2016
Dauer9 - 11 November 2016
StadtPorto
LandPortugal

Externe IDs

ORCID /0000-0001-8107-2775/work/142253534

Schlagworte

ASJC Scopus Sachgebiete

Schlagwörter

  • Information Extraction, Similarity Measures, Text Tiling, Web Tables