<!DOCTYPE article PUBLIC "-//NLM//DTD JATS (Z39.96) Journal Archiving and Interchange DTD v1.0 20120330//EN" "JATS-archivearticle1.dtd">
<article xmlns:xlink="http://www.w3.org/1999/xlink">
  <front>
    <journal-meta />
    <article-meta>
      <title-group>
        <article-title>Machine Translation vs. Multilingual Approaches for Entity Linking</article-title>
      </title-group>
      <contrib-group>
        <contrib contrib-type="author">
          <string-name>Henry Rosales-Mendez</string-name>
          <xref ref-type="aff" rid="aff0">0</xref>
        </contrib>
        <contrib contrib-type="author">
          <string-name>Aidan Hogan</string-name>
          <xref ref-type="aff" rid="aff0">0</xref>
        </contrib>
        <contrib contrib-type="author">
          <string-name>Barbara Poblete</string-name>
          <xref ref-type="aff" rid="aff0">0</xref>
        </contrib>
        <aff id="aff0">
          <label>0</label>
          <institution>IMFD Chile &amp; Department of Computer Science, University of Chile</institution>
        </aff>
      </contrib-group>
      <abstract>
        <p>Entity Linking (EL) associates the entities mentioned in a given input text with their corresponding knowledge-base (KB) entries. A recent EL trend is towards multilingual approaches. However, one may ask: are multilingual EL approaches necessary with recent advancements in machine translation? Could we not simply focus on supporting one language in the EL system and translate the input text to that language? We present experiments along these lines comparing multilingual EL systems with their results over machine translated text.</p>
      </abstract>
    </article-meta>
  </front>
  <body>
    <sec id="sec-1">
      <title>Introduction</title>
      <p>Entity Linking (EL) associates the entities mentioned in a given input text with
their corresponding knowledge-base (KB) identi ers; e.g., taking Wikidata as a
target KB, for the input text \Michael Jackson was born in Gary, Indiana ",
we can link Michael Jackson with the Wikidata identi er wd:Q2831. However,
multiple KB entities may have the same label; e.g., wd:Q167877, wd:Q6831554,
and wd:Q3856193 are all identi ers for people called Michael Jackson in
Wikidata. On the other hand, the same entity can be mentioned multiple ways, e.g.,
\Michael J. Jackson", \Jackson", \King of Pop", etc., can refer to wd:Q2831.</p>
      <p>
        Another practical challenge is being able to cope with input texts from
various languages. While many EL approaches have been proposed down through
the years, only recently have multilingual EL approaches { con gurable for
various input languages { become more popular (e.g., [
        <xref ref-type="bibr" rid="ref1 ref2 ref4 ref7">2,1,4,7</xref>
        ]). Despite this trend,
there are few studies evaluating multilingual EL. Hence, in our paper accepted
for the Resource Track at ISWC [
        <xref ref-type="bibr" rid="ref5">5</xref>
        ], we propose a multilingual EL benchmark
and use it to perform experiments in order to study the behaviour of
state-of-theart multilingual approaches for ve languages: English, French, German, Italian,
and Spanish. We call our dataset VoxEL; a particular design goal of the dataset
is to have (insofar as possible) the same text in di erent languages, and in
particular, the same annotations per sentence across languages. Thus performance
across languages { not just systems { can be compared directly. We also
compared the results of multilingual EL systems with what would be possible using
a state-of-the-art machine translation approach (Google Translate) to translate
the text to English (the primary language supported by most tools). We refer
the reader to our Resource Track paper [
        <xref ref-type="bibr" rid="ref5">5</xref>
        ] for more details.
      </p>
      <p>In this poster, we present some additional results omitted from the full paper
for reasons of space. More speci cally, the poster focuses on the question of
how an a priori machine translation process compares with multilingual EL
approaches, contributing novel results using VoxEL to evaluate EL performance
using machine translation of the input to languages other than English. More
generally, in the poster session, we would like to discuss with interested attendees
the interplay between multilingual EL and machine translation.
2</p>
    </sec>
    <sec id="sec-2">
      <title>Evaluating Multilingual Entity Linking Approaches</title>
      <p>
        A multilingual EL system is characterised by being con gurable for multiple
input languages. In this work, we evaluate four multilingual EL systems with public
APIs, namely Babelfy [
        <xref ref-type="bibr" rid="ref4">4</xref>
        ], DBpedia Spotlight [
        <xref ref-type="bibr" rid="ref1">1</xref>
        ], FREME [
        <xref ref-type="bibr" rid="ref7">7</xref>
        ] and TagME [
        <xref ref-type="bibr" rid="ref2">2</xref>
        ]. For
reasons of space, we refer to our previous work [
        <xref ref-type="bibr" rid="ref5 ref6">6,5</xref>
        ] for further details on these
systems and other multilingual EL systems proposed in the literature.
      </p>
      <p>
        Evaluating multilingual EL systems requires benchmark datasets with texts
in various languages. To further compare the quality of EL results across
languages { not just systems { we need (insofar as possible) the same text and
annotations in the di erent languages. Only a few such datasets have been proposed:
TAC KBP1, SemEval2, and MEANTIME [
        <xref ref-type="bibr" rid="ref3">3</xref>
        ]. However, MEANTIME [
        <xref ref-type="bibr" rid="ref3">3</xref>
        ] is the
only publicly available dataset; SemEval is published by a third-party whereas
the TAC KBP dataset we could not acquire. Furthermore, we found that these
datasets exhibit di erences in their annotations for di erent languages. For a
more detailed explanation of multilingual benchmark datasets see [
        <xref ref-type="bibr" rid="ref5">5</xref>
        ] and for
results comparing various EL systems over the SemEval dataset, see [
        <xref ref-type="bibr" rid="ref6">6</xref>
        ].
      </p>
      <p>
        To support multilingual EL evaluation, in [
        <xref ref-type="bibr" rid="ref5">5</xref>
        ] we proposed VoxEL: a curated
text extracted from the multilingual VoxEurop news site3 and manually
annotated for EL benchmarking. This dataset contains 15 documents for each of the
ve supported languages: Germany, English, Spanish, French and Italian. To
support comparison across languages, VoxEL was edited to ensure the same
annotations per sentence across languages, normalising variances across languages.
Given a lack of consensus on the de nition of \entity", VoxEL features two
annotated version of the documents for each language: one strict that includes entities
referring to people, places and organisations, and one relaxed that includes links
to all unambiguous pages of Wikipedia. Per language, VoxEL contains 204 and
674 annotations in the strict and relaxed version respectively.
3
      </p>
    </sec>
    <sec id="sec-3">
      <title>Experiments</title>
      <p>We conduct experiments using VoxEL to compare the behaviour of the four
aforementioned multilingual EL systems for the ve di erent languages o ered
by the dataset: German (DE), English (EN), Spanish (ES), French (FR) and
Italian (IT). All systems were con gured with their default parameters, except
1 https://tac.nist.gov/2017/KBP/; June 1st, 2018.
2 http://alt.qcri.org/semeval2018/; June 1st, 2018.
3 http://www.voxeurop.eu; June 1st, 2018.
Babelfy, which allows to select a more strict or more relaxed notion of entity; we
study the performance of both, denoted henceforth as BabelfyS and BabelfyR
respectively. Aside from testing EL over the text in its native language, we also
include results for EL applying machine translation { namely Google Translate4
{ from each language of VoxEL to the other four languages; the purpose of this
approach is to simulate an EL approach supporting one language and see if EL
performs competitively when input text is translated from other languages.</p>
      <p>The results are given in Table 1, where we present the F1-measure for various
con gurations. On the left we present the system and language con gured. At the
top of the table we present the Relaxed and Strict versions of the dataset, where
for each version, we present the language of the input text, which is machine
translated to the con gured language; for example, row ! ES, column DE !,
gives the result for a German input text translated to Spanish (DE ! ES ) and
processed by the given EL systems con gured for Spanish. Where input and
translated languages coincide, we use the input text directly (such results are
indicated with boxes). The best result per column for each dataset version and
system is presented in bold. TagME supports English and German only.
4 https://translate.google.com; June 1st, 2018.</p>
    </sec>
    <sec id="sec-4">
      <title>Discussion</title>
      <p>In Table 1, we see that DBpedia Spotlight, FREME and TagMe often perform
markedly better when the input text is either in English, or translated to
English; the one exception to this trend is that FREME performs slightly better
over the untranslated Italian text in the Strict version of the dataset than over
the translated English text. On the other hand, Babelfy generally performs best
for (translated) Spanish texts in the Relaxed version, and (translated) Italian
texts in the Strict version, though performance across languages is more
balanced in general than for the former systems. These results suggest that prior
machine translation makes little di erence in the case of Babelfy, but markedly
improves the performance of other systems when dealing with non-English texts;
the reasons for this may include the quality of language-speci c components, the
richness of KB information available for a particular language, etc.</p>
      <p>It is important to highlight in such cases that the output of the EL process
after translation is still in the translated language; e.g., if we process text in
French by translating it to English and performing EL con gured for English,
we may get better results, but the output text is in English, not French. But
we put forward that given (1) a high(er) quality annotation in the translated
English text, (2) a sentence-to-sentence correspondence between the French and
translated English text, and (3) cross-language links provided by KBs; it would
not be di cult to \transfer" the annotations back to the original French text.</p>
      <p>In any case, these results raise the question of what role machine translation
should play for EL, and indeed, in what circumstances it makes sense to develop
multilingual EL systems, and in what circumstances it makes sense to develop
monolingual EL systems with a priori translation.</p>
      <p>Acknowledgements Henry Rosales-Mendez was supported by
CONICYT-PCHA/Doctorado Nacional/2016-21160017. The work was also supported by the Millennium
Institute for Foundational Research on Data (IMFD) and by Fondecyt Grant No. 1181896.</p>
    </sec>
  </body>
  <back>
    <ref-list>
      <ref id="ref1">
        <mixed-citation>
          1.
          <string-name>
            <surname>Daiber</surname>
            ,
            <given-names>J.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Jakob</surname>
            ,
            <given-names>M.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Hokamp</surname>
            ,
            <given-names>C.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Mendes</surname>
            ,
            <given-names>P.N.</given-names>
          </string-name>
          :
          <article-title>Improving e ciency and accuracy in multilingual entity extraction</article-title>
          .
          <source>In: I-SEMANTICS, ACM</source>
          (
          <year>2013</year>
          )
          <volume>121</volume>
          {
          <fpage>124</fpage>
        </mixed-citation>
      </ref>
      <ref id="ref2">
        <mixed-citation>
          2.
          <string-name>
            <surname>Ferragina</surname>
            ,
            <given-names>P.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Scaiella</surname>
            ,
            <given-names>U.</given-names>
          </string-name>
          :
          <article-title>Tagme: on-the- y annotation of short text fragments (by Wikipedia entities)</article-title>
          .
          <source>In: CIKM</source>
          ,
          <string-name>
            <surname>ACM</surname>
          </string-name>
          (
          <year>2010</year>
          )
          <volume>1625</volume>
          {
          <fpage>1628</fpage>
        </mixed-citation>
      </ref>
      <ref id="ref3">
        <mixed-citation>
          3.
          <string-name>
            <surname>Minard</surname>
            ,
            <given-names>A. L.</given-names>
          </string-name>
          , et al.
          <article-title>MEANTIME, the NewsReader multilingual event and time corpus</article-title>
          .
          <source>LREC-ELRA</source>
          (
          <year>2016</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref4">
        <mixed-citation>
          4.
          <string-name>
            <surname>Moro</surname>
            ,
            <given-names>A.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Raganato</surname>
            ,
            <given-names>A.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Navigli</surname>
          </string-name>
          , R.:
          <article-title>Entity linking meets word sense disambiguation: a uni ed approach</article-title>
          .
          <source>Trans. of the ACL 2</source>
          (
          <year>2014</year>
          )
          <volume>231</volume>
          {
          <fpage>244</fpage>
        </mixed-citation>
      </ref>
      <ref id="ref5">
        <mixed-citation>
          5.
          <string-name>
            <surname>Rosales-Mendez</surname>
            ,
            <given-names>H.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Hogan</surname>
            <given-names>A.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Poblete</surname>
            <given-names>B.</given-names>
          </string-name>
          <article-title>VoxEL: A Benchmark Dataset for Multilingual Entity Linking</article-title>
          . In ISWC (
          <year>2018</year>
          )
          <article-title>(to appear)</article-title>
        </mixed-citation>
      </ref>
      <ref id="ref6">
        <mixed-citation>
          6.
          <string-name>
            <surname>Rosales-Mendez</surname>
            ,
            <given-names>H.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Poblete</surname>
            ,
            <given-names>B.</given-names>
          </string-name>
          , and
          <string-name>
            <surname>Hogan</surname>
            ,
            <given-names>A.</given-names>
          </string-name>
          <string-name>
            <surname>Multilingual Entity</surname>
          </string-name>
          <article-title>Linking: Comparing English and Spanish</article-title>
          . In LD4IE@ISWC (
          <year>2017</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref7">
        <mixed-citation>
          7.
          <string-name>
            <surname>Sasaki</surname>
            ,
            <given-names>F.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Dojchinovski</surname>
            ,
            <given-names>M.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Nehring</surname>
            ,
            <given-names>J.</given-names>
          </string-name>
          <string-name>
            <surname>Chainable</surname>
          </string-name>
          and
          <string-name>
            <surname>Extendable Knowledge Integration Web Services. In</surname>
            <given-names>ISWC</given-names>
          </string-name>
          , (
          <year>2016</year>
          )
          <volume>89</volume>
          {
          <fpage>101</fpage>
        </mixed-citation>
      </ref>
    </ref-list>
  </back>
</article>