Skip to main content

Showing 1–1 of 1 results for author: Jóhannesson, B G

Searching in archive cs. Search in all archives.
.
  1. arXiv:2206.05014  [pdf, other

    cs.CL

    Building an Icelandic Entity Linking Corpus

    Authors: Steinunn Rut Friðriksdóttir, Valdimar Ágúst Eggertsson, Benedikt Geir Jóhannesson, Hjalti Daníelsson, Hrafn Loftsson, Hafsteinn Einarsson

    Abstract: In this paper, we present the first Entity Linking corpus for Icelandic. We describe our approach of using a multilingual entity linking model (mGENRE) in combination with Wikipedia API Search (WAPIS) to label our data and compare it to an approach using WAPIS only. We find that our combined method reaches 53.9% coverage on our corpus, compared to 30.9% using only WAPIS. We analyze our results and… ▽ More

    Submitted 10 June, 2022; originally announced June 2022.

    Comments: 9 pages, 5 figures, submitted to Dataset Creation for Lower-Resourced Languages, an LREC 2022 Workshop, 9am-1pm June 24th, 2022