Wiki-TabNER:Advancing Table Interpretation Through Named Entity Recognition (2403.04577v1)

Published 7 Mar 2024 in cs.AI and cs.CL

Abstract: Web tables contain a large amount of valuable knowledge and have inspired tabular LLMs aimed at tackling table interpretation (TI) tasks. In this paper, we analyse a widely used benchmark dataset for evaluation of TI tasks, particularly focusing on the entity linking task. Our analysis reveals that this dataset is overly simplified, potentially reducing its effectiveness for thorough evaluation and failing to accurately represent tables as they appear in the real-world. To overcome this drawback, we construct and annotate a new more challenging dataset. In addition to introducing the new dataset, we also introduce a novel problem aimed at addressing the entity linking task: named entity recognition within cells. Finally, we propose a prompting framework for evaluating the newly developed LLMs on this novel TI task. We conduct experiments on prompting LLMs under various settings, where we use both random and similarity-based selection to choose the examples presented to the models. Our ablation study helps us gain insights into the impact of the few-shot examples. Additionally, we perform qualitative analysis to gain insights into the challenges encountered by the models and to understand the limitations of the proposed dataset.

PDF HTML Abstract

Summarize PDF Markdown Bookmark Chat (Pro)

References (42)

Authors (5)

Aneta Koleva (6 papers)
Martin Ringsquandl (14 papers)
Ahmed Hatem (3 papers)
Thomas Runkler (34 papers)
Volker Tresp (158 papers)

Citations (1)

View on Semantic Scholar

Wiki-TabNER:Advancing Table Interpretation Through Named Entity Recognition (2403.04577v1)

Related Papers