Conference paper Open Access

Mining and Leveraging Background Knowledge for Improving Named Entity Linking

Weichselbraun, Albert; Kuntschik, Philipp; Braşoveanu, Adrian M. P.

Knowledge-rich Information Extraction (IE) methods aspire towards combining classical IE with background knowledge obtained from third-party resources. Linked Open Data repositories that encode billions of machine readable facts from sources such as Wikipedia play a pivotal role in this development. The recent growth of Linked Data adoption for Information Extraction tasks has shed light on many data quality issues in these data sources that seriously challenge their usefulness such as completeness, timeliness and semantic correctness. Information Extraction methods are, therefore, faced with problems such as name variance and type confusability. If multiple linked data sources are used in parallel, additional concerns regarding link stability and entity mappings emerge. This paper develops methods for integrating Linked Data into Named Entity Linking methods and addresses challenges in regard to mining knowledge from Linked Data, mitigating data quality issues, and adapting algorithms to leverage this knowledge. Finally, we apply these methods to Recognyze, a graph-based Named Entity Linking (NEL) system, and provide a comprehensive evaluation which compares its performance to other well-known NEL systems, demonstrating the impact of the suggested methods on its own entity linking performance.

Files (742.3 kB)
Name Size
WIMS2018_NEL_Weichselbraun.pdf
md5:2aad98af4b39d2c91d1fa955eae16a42
742.3 kB Download
18
25
views
downloads
Views 18
Downloads 25
Data volume 18.6 MB
Unique views 16
Unique downloads 22

Share

Cite as