<!DOCTYPE article PUBLIC "-//NLM//DTD JATS (Z39.96) Journal Archiving and Interchange DTD v1.0 20120330//EN" "JATS-archivearticle1.dtd">
<article xmlns:xlink="http://www.w3.org/1999/xlink">
  <front>
    <journal-meta />
    <article-meta>
      <title-group>
        <article-title>NLP for Interlinking Multilingual LOD</article-title>
      </title-group>
      <contrib-group>
        <contrib contrib-type="author">
          <string-name>INRIA</string-name>
        </contrib>
        <contrib contrib-type="author">
          <string-name>Grenoble</string-name>
        </contrib>
      </contrib-group>
      <abstract>
        <p>Nowadays, there are many natural languages on the Web, and we can expect that they will stay there even with the development of the Semantic Web. Though the RDF model enables structuring information in a uni ed way, the resources can be described using di erent natural languages. To nd information about the same resource across di erent languages, we need to link identical resources together. In this paper we present an instance-based approach for resource interlinking. We also show how a problem of graph matching can be converted into a document matching for discovering cross-lingual mappings across RDF data sets.</p>
      </abstract>
      <kwd-group>
        <kwd>Multilingual Mappings</kwd>
        <kwd>Cross-Lingual Link Discovery</kwd>
        <kwd>CrossLingual RDF Data Set Linkage</kwd>
      </kwd-group>
    </article-meta>
  </front>
  <body>
    <sec id="sec-1">
      <title>-</title>
      <p>Due to the Resource Description Framework (RDF), the information on the Web
can be turned from the unstructured mass into the structured data represented
in the form of triples. The Linked Open Data (LOD) cloud containing billions
of triples is constantly growing. Since data sets are created independently, there
can be several Uniform Resource Identi ers (URIs) denoting the same entity
across di erent RDF data sets. As a result, one needs to address the problem
of entity resolution: identify and interlink the same entity across multiple data
sources.</p>
      <p>The RDF syntax is relatively simple and unambiguous: RDF = graph +
identi ers (labels). This is what the identi cation of resources can be based on.
However, this problem can become particularly di cult when there are
multilingual elements in a graph as a simple string matching technique is doomed
to fail. Hence, speci c Natural Language Processing (NLP) techniques must be
considered.</p>
      <p>Our research problem is to nd out methods for linking the same resource
located in several RDF data sets and described in various natural languages and
study the impact of available NLP techniques on the interlinking procedure.</p>
    </sec>
    <sec id="sec-2">
      <title>Relevancy</title>
      <p>Internet is a multilingual system, and we believe that it will continue to
accommodate a diversity of natural languages despite the development of the Semantic
Web. Even though there are many resources in English, some other languages
occupy a decent portion of the Web space as well (see Internet world users by
language statistics 1). And we expect the necessity to tackle the
multilinguality problem to persist. There are many resources that could be interlinked. At
present, the number of languages 2 of RDF data sets amounts to 503.</p>
      <p>The importance of cross-lingual mappings has been discussed in several works
[1{3].</p>
      <p>Recently a Best Practices for Multilingual Linked Open Data Community
Group 3 has been created to elaborate a large spectrum of practices with regard
to multilingual LOD.</p>
      <p>
        Availability of the cross-lingual links is imperative for several neighboring
research areas. For example, to overcome the problem of ontology
heterogeneity, some research has been done on monolingual ontology integration based on
instances interlinked by owl:sameAs [
        <xref ref-type="bibr" rid="ref4">4</xref>
        ]. If owl:sameAs links could be provided
between instances expressed in di erent languages, other experiments on
integrating underlying ontologies could be conducted.
      </p>
      <p>
        The owl:sameAs links between instances can be also valuable in other
applications such as Question Answering over multilingual structured knowledge-base
[
        <xref ref-type="bibr" rid="ref5">5</xref>
        ] since a system can take advantage of the information presented in a language
di erent from a language that is being queried.
      </p>
      <p>Thus, the growing number of data sources in RDF format with
multilingual labels and the importance of cross-lingual links for other Semantic Web
applications motivate our interest in cross-lingual link discovery.
3</p>
    </sec>
    <sec id="sec-3">
      <title>Related Work</title>
      <p>
        The problem of searching for the same entity across multiple sources has been
investigated in several research elds. In database community, it is known as
instance identi cation, record linkage or record matching problem. In [
        <xref ref-type="bibr" rid="ref6">6</xref>
        ], the
authors use the term "duplicate record detection" and provide a thorough survey
on the matching techniques. Though the work done in record linkage is similar
to our research, it does not contain cross-lingual aspect and RDF semantics.
      </p>
      <p>In the eld of Information Retrieval (IR), within the framework of the
CrossLanguage Evaluation Forum (CLEF)4, the Web People Search Evaluation
Campaigns (2007-2010)5 focused on the Web People Search and person name
ambiguity on Web pages and aimed at building a system which could estimate the</p>
      <sec id="sec-3-1">
        <title>1 http://www.internetworldstats.com/stats7.htm</title>
        <p>2 http://stats.lod2.eu/languages
3 http://www.w3.org/community/bpmlod/
4 http://www.clef-initiative.eu/
5 http://nlp.uned.es/weps/weps-3
number of referents and cluster Web pages that refer to the same individual into
one group. The research was performed on monolingual data.</p>
        <p>
          Cross-lingual entity linking has been addressed in Knowledge Base
Population track (KBP2011)[
          <xref ref-type="bibr" rid="ref7">7</xref>
          ] within the Text Analysis conference. The task is to link
entity mentions in a text to a knowledge base (Wikipedia). If entity mentions
are not in KB, they should be clustered into a separate group. Experiments were
done both on monolingual (English) and cross-lingual data (Chinese to English).
Authors in [
          <xref ref-type="bibr" rid="ref8">8</xref>
          ] used both language-independent and translation-based methods.
        </p>
        <p>In contrast to the research outlined above, we aim at providing insights into
the problem of cross-lingual interlinking from the point where data are already
in RDF format, and we can vary di erent parameters in order to determine their
impact on the interlinking operation.</p>
        <p>
          In the Semantic Web, interlinking resources that represent the same
realworld object and that are scattered across multiple Linked Data sets is a widely
researched topic. Within the Data Interlinking track (IM@OAEI 2011), several
interlinking systems have been proposed [9{13]. All of the systems were
evaluated on monolingual data sets. Recent developments have been made also in
multilingual ontology matching [
          <xref ref-type="bibr" rid="ref14 ref15">14, 15</xref>
          ].
        </p>
        <p>To the best of our knowledge, there is no interlinking system speci cally
designed to link RDF data sets with multilingual labels.
4</p>
      </sec>
    </sec>
    <sec id="sec-4">
      <title>Research Questions</title>
      <p>
        The goal of our work is to provide methods to link interrelated resources across
multilingual RDF data sets. For now, we restrict ourselves to owl:sameAs link
[
        <xref ref-type="bibr" rid="ref16">16</xref>
        ] as it is a classical type of link that is usually established, and it is also
important for tracking information about the same resource across di erent data
sources. Given two RDF data sets with URIs and literals in di erent natural
languages, the output will be a set of triples of type URI1 owl:sameAs URI2.
      </p>
      <p>Our general research question is: To what extent is it possible to interlink
data sets in di erent languages? To answer this question, within the framework
that we describe in the Proposed Approach section, we need to explore which
parameters in uence this task. More speci cally:
1. How to represent entities from RDF graphs?
{ What is the optimal distance for collecting language elements in
traversal?
{ Is it necessary to preserve the structure of the graph in a virtual
document by weighting the path length?
2. How to make entities described in di erent natural languages comparable?
{ What are the most appropriate Machine Translation techniques
(rulebased, statistical, hybrid)?
{ What is the impact of translating one language into another or pivot
language?
{ How does the output of similarity measures vary according to the
context?
All these parameters will be studied with respect to speci c contexts (language
pairs, data set types, amount of textual data available). We also plan to
experiment with graph matching techniques to see the di erence with a
translationbased approach. Apart from Machine Translation, we will explore techniques
used for word alignment, thesaurus-based word sense disambiguation,
multilingual document ranking, and mapping to multilingual lexicons.
5</p>
    </sec>
    <sec id="sec-5">
      <title>Hypotheses</title>
      <p>We introduce several hypotheses that we would like to test in our research.
1. If two URIs denote the same real-world object, the descriptions of the
properties of this object should overlap with each other.
2. If descriptions are in di erent natural languages, then NLP techniques could
help to decrease uncertainty across a set of resources.
3. If the descriptions of an entity overlap signi cantly, the similarity between
them will be higher than between other entities.
4. If the degree of similarity depends on the available language context for each
entity, then the more language data there are, the better will be the matching
results.
5. If language data can be taken from two sources in RDF graph: property
names and literals; then literals are more important since they are more
informative.
6</p>
    </sec>
    <sec id="sec-6">
      <title>Proposed Approach</title>
      <p>Due to the presence of natural language terms in RDF graphs, we adopt a
language-oriented approach.</p>
      <p>The proposed approach includes several steps (see Figure 1).
1. Given two data sets with a resource representation in di erent natural
language, extract language data for each URI. Thus, we create a "virtual"
document for each URI.
2. Compare virtual documents in pairs from both sets.
3. Find the maximum similarity between two representations of the resource.
4. Establish an owl:sameAs link between the two most similar representations.</p>
      <p>One should mind the following aspects of this approach:
DOCUMENT(ru)
2
1
DOCUMENT(en)</p>
      <p>SIMILARITY
3</p>
      <p>DOCUMENT(en)
translation</p>
      <p>DOCUMENT(zh)</p>
      <p>
        { The idea of creating a "virtual" document has been employed in ontology
matching [
        <xref ref-type="bibr" rid="ref17">17</xref>
        ]. The intuition of converting a graph into a document
representation is that even though the taxonomy (structure) of graphs can be similar,
the possibility to distinguish between two di erent things and identify the
identical ones relies on their comparison. Thus, it is important to take into
account lexical elements in a graph.
{ Once we have documents representing resources, we need to decide how to
dene similarity between these resources. Similarity between documents can be
taken for similarity between resources. Since we have documents in di erent
languages, we can experiment with di erent types of Machine Translation
(statistics-based, rule-based, hybrid). To estimate which strategy yields a
better result, we will run our system by changing the translation component
iteratively. Signi cant di erence in results may signal which translation type
is more bene cial. To enhance scalability, it would be interesting to
translate the whole source corpus once and not to translate each label again and
again. This would also allow for more contextual translation. The choice of
translation techniques can also depend on the language combinations, for
example, for rare languages, for which there does not exist enough parallel
corpora, dictionary-based approaches might help.
{ At the resource comparison step, it is important to reduce a number of
possible comparisons for the sake of time-e ciency. For example, only
comparisons between certain entity types are allowed. In case of using Supervised
Machine Learning, the problem of training data is the most prominent one
since there has been no o cial benchmark. And creating a generic training
set for a heterogeneous amount of Linked Data seems very unrealistic. Then,
instead of training, it would be interesting to test clustering algorithms and
nd appropriate parameters for identity resolution.
{ There are many techniques to compute similarity. A broad overview of them
is given in [
        <xref ref-type="bibr" rid="ref18">18</xref>
        ]. We will use a vector space model [
        <xref ref-type="bibr" rid="ref19">19</xref>
        ] to represent terms in a
"virtual" document as vectors of features. The choice of particular similarity
measures is yet to investigate. When terms are in di erent languages,
document similarity fails. Some similarity measures perform better on long texts.
After transformation of "virtual" document into vectors, similarity metrics
(e.g. Cosine, Euclidean) can be computed.
{ A virtual document per URI shall contain language data in proximity to a
given URI. The hypothesis is that the more textual data we have to
characterize a resource, the easier it would be to identify the identical ones.
      </p>
      <p>There are some complications as to textual data. Two scenarios are possible:
{ URI can be looked up and the textual data extracted (as in case of DBpedia)
{ No extra textual data are available per URI except the data in a graph itself.</p>
      <p>To overcome this lack of context for a particular resource, we propose to
browse a graph up to n+1 hops from the URI under investigation and collect
data along the way. The data carriers are property names and literals. Thus, a
virtual document for a particular URI will be the accumulation of data gathered
during graph traversal.</p>
      <p>This way of collecting a "pro le" for a resource entails a question: does the
di erence of two graph structures a ect the results of interlinking? On the one
hand, taking into account the success of statistical machine translation based
on statistical modelling and probabilities, the order of words is not always that
important. It would be interesting to see whether it holds for RDF interlinking as
well. On the other hand, we can try to preserve the order of collected properties
and literals in a virtual document by putting weight for each language element.
The further it is from the URI at question, the lower the weight. Term weight
can be assigned by computing termf requency in a document or distribution of
terms across a collection of documents known as inverse documentf requency
(IDF). Terms that appear in few documents can be discriminative with regard
to the rest of the documents. Combination of both TF x IDF is widely used in
vector space models.</p>
      <p>Once virtual documents are collected from both graphs, the documents will
be compared and results evaluated.
7</p>
    </sec>
    <sec id="sec-7">
      <title>Re ections</title>
      <p>We believe that we can succeed in nding the solution for our research topic
because we plan to put our research on a solid foundation and combine
different methods to achieve the task. In traditional Web, there has been much
work done on multilingual NLP, i.e. language identi cation, machine
translation, cross-language information retrieval. We are going to conduct series of
experiments and see what works and how we can improve what does not work.
This would allow us to preserve only the best practices and nally crystallize
a solution to the problem. The author of this research proposal is also guided
by the specialists in the domain that will contribute to the right choice of the
research direction.
8</p>
    </sec>
    <sec id="sec-8">
      <title>Evaluation</title>
      <p>
        Evaluation means comparing the retrieved links against some reference. Standard
measures usually serve for evaluation of an interlinking system (Precision, Recall,
F-measure). The biggest challenge for evaluating our system is the absence of
standard benchmark tests. As described in [
        <xref ref-type="bibr" rid="ref20">20</xref>
        ], there are several ways to go
about this challenge. One of them would be to rely on the existing links between
resources in DBpedia. This could be considered as a good alternative if not yet
another hurdle: the existing interlanguage links can be inaccurate [
        <xref ref-type="bibr" rid="ref21">21</xref>
        ]. So, in
our research we plan to experiment with di erent evaluation settings: we may
experiment only with bi-directional links and/or study transitivity in order to
ensure the correctness of test cases. The English, French, Russian versions of
DBpedia 6 and Baidu Baike in Chinese [
        <xref ref-type="bibr" rid="ref22">22</xref>
        ] will be used for our experiments.
We will also try to identify types of entities to focus on.
      </p>
      <sec id="sec-8-1">
        <title>6 http://dbpedia.org/About</title>
      </sec>
    </sec>
  </body>
  <back>
    <ref-list>
      <ref id="ref1">
        <mixed-citation>
          1.
          <string-name>
            <surname>Gracia</surname>
            ,
            <given-names>J.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Montiel-Ponsoda</surname>
            ,
            <given-names>E.</given-names>
          </string-name>
          ,
          <article-title>Go^mez-</article-title>
          <string-name>
            <surname>Perez</surname>
            ,
            <given-names>A.</given-names>
          </string-name>
          :
          <article-title>Cross-lingual Linking on the Multilingual Web of Data</article-title>
          .
          <source>In: Proc. of the 3rd Workshop on the Multilingual Semantic Web (MSW 2012) at ISWC</source>
          <year>2012</year>
          , Boston (USA),
          <source>CEUR-WS ISSN 1613- 0073</source>
          , vol.
          <volume>936</volume>
          (
          <year>2012</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref2">
        <mixed-citation>
          2.
          <string-name>
            <surname>Gracia</surname>
            ,
            <given-names>J.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Montiel-Ponsoda</surname>
            ,
            <given-names>E.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Cimiano</surname>
            ,
            <given-names>P.</given-names>
          </string-name>
          ,
          <article-title>Go^mez-</article-title>
          <string-name>
            <surname>Perez</surname>
            ,
            <given-names>A.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Buitelaar</surname>
            ,
            <given-names>P.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>McCrae</surname>
            ,
            <given-names>J.</given-names>
          </string-name>
          :
          <article-title>Challenges for the multilingual Web of Data</article-title>
          .
          <source>Journal of Web Semantics</source>
          ,
          <volume>11</volume>
          , 63{
          <fpage>71</fpage>
          (
          <year>2012</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref3">
        <mixed-citation>
          3.
          <string-name>
            <surname>Buitelaar</surname>
            ,
            <given-names>P.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Choi</surname>
          </string-name>
          , K.-S.,
          <string-name>
            <surname>Cimiano</surname>
            ,
            <given-names>P.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Hovy</surname>
            ,
            <given-names>H. E.</given-names>
          </string-name>
          :
          <article-title>The Multilingual Semantic Web (Dagstuhl Seminar 12362)</article-title>
          .
          <source>Dagstuhl Reports</source>
          <volume>2</volume>
          (
          <issue>9</issue>
          ), pp.
          <volume>15</volume>
          {
          <issue>94</issue>
          (
          <year>2012</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref4">
        <mixed-citation>
          4.
          <string-name>
            <surname>Zhao</surname>
            ,
            <given-names>L.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Ichise</surname>
          </string-name>
          , R.:
          <article-title>Instance-Based Ontological Knowledge Acquisition</article-title>
          .
          <source>In: Proc.10th International Conference, ESWC 2013</source>
          , Vol.
          <volume>7882</volume>
          , pp.
          <volume>155</volume>
          {
          <fpage>169</fpage>
          . LNCS, Springer Berlin Heidelberg (
          <year>2013</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref5">
        <mixed-citation>
          5.
          <string-name>
            <surname>Cabrio</surname>
            ,
            <given-names>E.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Cojan</surname>
            ,
            <given-names>J.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Gandon</surname>
            ,
            <given-names>F.</given-names>
          </string-name>
          , and
          <string-name>
            <surname>Hallili</surname>
            ,
            <given-names>A.</given-names>
          </string-name>
          :
          <article-title>Querying multilingual DBpedia with QAKiS</article-title>
          .
          <source>In: Proc. 10th International Conference</source>
          ,
          <string-name>
            <surname>ESWC</surname>
          </string-name>
          <year>2013</year>
          .
          <article-title>Demo paper</article-title>
          . Montpellier, France (
          <year>2013</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref6">
        <mixed-citation>
          6.
          <string-name>
            <surname>Elmagarmid</surname>
            ,
            <given-names>A.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Ipeirotis</surname>
            ,
            <given-names>P.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Verykios</surname>
            ,
            <given-names>V.</given-names>
          </string-name>
          :
          <article-title>Duplicate record detection: A survey</article-title>
          .
          <source>IEEE Transactions on Knowledge and Data Engineering</source>
          .
          <volume>19</volume>
          (
          <issue>1</issue>
          ),
          <volume>1</volume>
          {
          <fpage>16</fpage>
          (
          <year>2007</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref7">
        <mixed-citation>
          7.
          <string-name>
            <surname>Ji</surname>
            ,
            <given-names>H.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Grishman</surname>
            ,
            <given-names>R.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Dang</surname>
            ,
            <given-names>H.T.</given-names>
          </string-name>
          :
          <article-title>An Overview of the TAC2011 Knowledge BAse Population Track</article-title>
          .
          <source>In: Proc. Text Analytics Conference (TAC2011)</source>
          (
          <year>2011</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref8">
        <mixed-citation>
          8.
          <string-name>
            <surname>Monahan</surname>
            ,
            <given-names>S.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Lehmann</surname>
            ,
            <given-names>J.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Nyberg</surname>
            ,
            <given-names>T.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Plymale</surname>
            ,
            <given-names>J.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Jung</surname>
            ,
            <given-names>A.</given-names>
          </string-name>
          :
          <article-title>Cross-Lingual CrossDocument Coreference with Entity Linking</article-title>
          .
          <source>In: Proc. TAC2011</source>
          (
          <year>2011</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref9">
        <mixed-citation>
          9.
          <string-name>
            <given-names>Ngonga</given-names>
            <surname>Ngomo</surname>
          </string-name>
          ,
          <string-name>
            <given-names>A.-C.</given-names>
            ,
            <surname>Auer</surname>
          </string-name>
          ,
          <string-name>
            <surname>S.:</surname>
          </string-name>
          <article-title>LIMES - A Time-E cient Approach for LargeScale Link Discovery on the Web of Data</article-title>
          . IJCAI, pp.
          <volume>2312</volume>
          {
          <issue>2317</issue>
          (
          <year>2011</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref10">
        <mixed-citation>
          10.
          <string-name>
            <surname>Nguyen</surname>
            ,
            <given-names>K.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Ichise</surname>
            ,
            <given-names>R.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Le</surname>
            ,
            <given-names>B.</given-names>
          </string-name>
          :
          <article-title>SLINT: A Schema-Independent Linked Data Interlinking System</article-title>
          .
          <source>In: Proc. of the 7th International Workshop on Ontology Matching</source>
          , pp.
          <volume>1</volume>
          {
          <issue>12</issue>
          (
          <year>2012</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref11">
        <mixed-citation>
          11.
          <string-name>
            <surname>Volz</surname>
            ,
            <given-names>J.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Bizez</surname>
            ,
            <given-names>C.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Gaedke</surname>
            ,
            <given-names>M.</given-names>
          </string-name>
          , and
          <string-name>
            <surname>Kobilarov</surname>
          </string-name>
          , G.:
          <article-title>Discovering and maintaining links on the web of data</article-title>
          .
          <source>In: Proc. of ISWC' 09</source>
          , Springer-Verlag Berlin, Heidelberg, pp.
          <volume>650</volume>
          {
          <issue>665</issue>
          ,
          <year>2009</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref12">
        <mixed-citation>
          12.
          <string-name>
            <surname>Araujo</surname>
            ,
            <given-names>S.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Hidders</surname>
            ,
            <given-names>J.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Schwabe</surname>
            ,
            <given-names>D.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Arjen</surname>
          </string-name>
          , P. de Vries:
          <article-title>SERIMI - Resource Description Similarity, RDF Instance Matching and Interlinking</article-title>
          . CoRR, abs/1107.1104 (
          <year>2011</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref13">
        <mixed-citation>
          13.
          <string-name>
            <surname>Niu</surname>
            ,
            <given-names>X.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Rong</surname>
            ,
            <given-names>S.</given-names>
          </string-name>
          , Zhang,
          <string-name>
            <given-names>Y.</given-names>
            , and
            <surname>Wang</surname>
          </string-name>
          , H.:
          <article-title>Zhishi.links results for OAEI 2011</article-title>
          .
          <source>In: Proc. of ISWC' 11 6th Workshop on Ontology Matching</source>
          , pp.
          <volume>220</volume>
          {
          <issue>227</issue>
          (
          <year>2011</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref14">
        <mixed-citation>
          14.
          <string-name>
            <surname>Meilicke</surname>
            ,
            <given-names>C.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Trojahn</surname>
            ,
            <given-names>C.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Svab-Zamazal</surname>
            ,
            <given-names>O.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Ritze</surname>
            ,
            <given-names>D.</given-names>
          </string-name>
          :
          <article-title>Multilingual Ontology Matching Evaluation - a First Report on Using MultiFarm</article-title>
          .
          <source>In: Proc. of the 2d International Workshop on Evaluation of Semantic Technologies</source>
          , pp.
          <volume>1</volume>
          {
          <issue>12</issue>
          ,
          <string-name>
            <surname>Heraklion</surname>
          </string-name>
          , Greece (
          <year>2012</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref15">
        <mixed-citation>
          15.
          <string-name>
            <surname>Meilicke</surname>
            ,
            <given-names>C.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Castro</surname>
            ,
            <given-names>R.G.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Freitas</surname>
          </string-name>
          , F.,
          <string-name>
            <surname>van Hage</surname>
            ,
            <given-names>W.R.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Montiel-Ponsoda</surname>
          </string-name>
          , E.,
          <string-name>
            <surname>de</surname>
            <given-names>Azevedo</given-names>
          </string-name>
          ,
          <string-name>
            <given-names>R.R.</given-names>
            ,
            <surname>Stuckenschmidt</surname>
          </string-name>
          ,
          <string-name>
            <given-names>H.</given-names>
            ,
            <surname>Svab-Zamazal</surname>
          </string-name>
          <string-name>
            <given-names>O.</given-names>
            ,
            <surname>Svatek</surname>
          </string-name>
          ,
          <string-name>
            <given-names>V.</given-names>
            ,
            <surname>Tamilin</surname>
          </string-name>
          ,
          <string-name>
            <given-names>A.</given-names>
            ,
            <surname>Trojahn</surname>
          </string-name>
          ,
          <string-name>
            <given-names>C.</given-names>
            ,
            <surname>Wang</surname>
          </string-name>
          ,
          <string-name>
            <surname>S.:</surname>
          </string-name>
          <article-title>MultiFarm: A Benchmark for Multilingual Ontology Matching</article-title>
          .
          <source>Journal of Web Semantics</source>
          .
          <volume>15</volume>
          ,
          <issue>62</issue>
          {
          <fpage>68</fpage>
          (
          <year>2012</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref16">
        <mixed-citation>
          16.
          <string-name>
            <surname>Halpin</surname>
            ,
            <given-names>H.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Hayes</surname>
            ,
            <given-names>J. P.</given-names>
          </string-name>
          :
          <article-title>When owl: sameAs isn't the Same: An Analysis of Identity Links on the Semantic Web</article-title>
          .
          <source>In: Proc. of the Linked Data on the Web Workshop (LDOW2010)</source>
          , Raleigh, North Carolina, USA, April
          <volume>27</volume>
          ,
          <year>2010</year>
          , CEUR Workshop Proceedings, ISSN
          <volume>1613</volume>
          -0073, online http://ceur-ws.
          <source>org/Vol628/ldow2010 paper09.pdf</source>
        </mixed-citation>
      </ref>
      <ref id="ref17">
        <mixed-citation>
          17.
          <string-name>
            <surname>Qu</surname>
            ,
            <given-names>Y.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Hu</surname>
          </string-name>
          , W., Cheng, G.:
          <article-title>Constructing virtual documents for ontology matching</article-title>
          .
          <source>In: Proc. of the 15th International Conference of World Wide Web</source>
          , pp.
          <volume>23</volume>
          {
          <issue>31</issue>
          (
          <year>2006</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref18">
        <mixed-citation>
          18.
          <string-name>
            <surname>Euzenat</surname>
            ,
            <given-names>J.</given-names>
          </string-name>
          and
          <string-name>
            <surname>Shvaiko</surname>
            ,
            <given-names>P.</given-names>
          </string-name>
          : Ontology Matching. Springer-Verlag, Heidelberg (
          <year>2007</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref19">
        <mixed-citation>
          19.
          <string-name>
            <surname>Salton</surname>
            ,
            <given-names>G.</given-names>
          </string-name>
          :
          <article-title>The SMART Retrieval System: Experiments in Automatic Document Processing</article-title>
          . Prentice Hall, Englewood Cli s,
          <source>NJ</source>
          (
          <year>1971</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref20">
        <mixed-citation>
          20.
          <string-name>
            <surname>Lesnikova</surname>
            ,
            <given-names>T.</given-names>
          </string-name>
          :
          <article-title>Interlinking Cross-Lingual RDF Data Sets</article-title>
          .
          <source>In: Proc. 10th International Conference, ESWC 2013</source>
          , Vol.
          <volume>7882</volume>
          , pp.
          <fpage>671</fpage>
          -
          <lpage>675</lpage>
          . LNCS, Springer Berlin Heidelberg (
          <year>2013</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref21">
        <mixed-citation>
          21.
          <string-name>
            <surname>Rinser</surname>
            ,
            <given-names>D.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Lange</surname>
            ,
            <given-names>D.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Naumann</surname>
            ,
            <given-names>F.</given-names>
          </string-name>
          :
          <article-title>Cross-lingual entity matching and infobox alignment in Wikipedia</article-title>
          .
          <source>Information Systems</source>
          ,
          <volume>38</volume>
          (
          <issue>6</issue>
          ), pp.
          <fpage>887</fpage>
          -
          <lpage>907</lpage>
          (
          <year>2013</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref22">
        <mixed-citation>
          22.
          <string-name>
            <surname>Wang</surname>
            ,
            <given-names>Z.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Wang</surname>
            ,
            <given-names>Z.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Li</surname>
            <given-names>J.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Pan</surname>
            ,
            <given-names>J. Z.</given-names>
          </string-name>
          :
          <article-title>Knowledge extraction from Chinese wiki encyclopedias</article-title>
          .
          <source>Journal of Zhejiang</source>
          University - Science C
          <volume>13</volume>
          (
          <issue>4</issue>
          ):
          <fpage>268</fpage>
          -
          <lpage>280</lpage>
          (
          <year>2012</year>
          )
        </mixed-citation>
      </ref>
    </ref-list>
  </back>
</article>