<!DOCTYPE article PUBLIC "-//NLM//DTD JATS (Z39.96) Journal Archiving and Interchange DTD v1.0 20120330//EN" "JATS-archivearticle1.dtd">
<article xmlns:xlink="http://www.w3.org/1999/xlink">
  <front>
    <journal-meta />
    <article-meta>
      <title-group>
        <article-title>Linked Open Data Framework for Serendipity in History of Art Research</article-title>
      </title-group>
      <contrib-group>
        <aff id="aff0">
          <label>0</label>
          <institution>Dept. of Information Engineering, University of Padua</institution>
          ,
          <country country="IT">Italy</country>
        </aff>
      </contrib-group>
      <abstract>
        <p>In this paper we outline the main lines of research for de ning a framework based on Linked Open Data (LOD) for supporting knowledge creation in the Cultural Heritage (CH) eld with a particular focus on History of Art research. We delineate the main challenges we need to deal with and we explore the state-of-the-art in LOD publishing systems, LOD citation and authority management. Furthermore, we introduce the idea of computer-aided serendipity in History of Art research with the purpose of contributing to the advancement of the eld and to the de nition of new methodologies for entity linking and retrieval.</p>
      </abstract>
    </article-meta>
  </front>
  <body>
    <sec id="sec-1">
      <title>Motivation</title>
      <p>
        One of the most relevant socio-economical and scienti c changes in recent years
has been the recognition of data as a valuable asset. The Economist magazine
recently wrote that \data is the new raw material of business"; and the European
Commission stated that data-related \technology and services are expected to
grow from EUR 2.4 billion in 2010 to EUR 12.7 billion in 2015" [
        <xref ref-type="bibr" rid="ref9">9</xref>
        ]. The
principal driver of this evolution is the Web of Data, the size of which is estimated to
have exceeded 100 billion facts (i.e. semantically connected entities). The actual
paradigm realizing the Web of Data is LOD [
        <xref ref-type="bibr" rid="ref14">14</xref>
        ], which by exploiting Web
technologies, such as the Resource Description Framework (RDF) [20], allows public
data in machine-readable formats to be opened up ready for consumption and
re-use. LOD is becoming the de-facto standard for data publishing, accessing and
sharing because it allows for exible manipulation, enrichment and discovery of
data as well as for overcoming interoperability issues. The ground breaking
potential of this approach resides in the semantic connections among data enabling
new knowledge creation and discovery possibilities.
      </p>
      <p>The CH domain and in particular the History of Art research eld, provide a
fertile ground where the LOD potential can grow and bloom; indeed, in History
of Art the preponderant way to produce new knowledge is to reveal connections
between di erent items (illuminated manuscripts, pictures, frescos) that can cast
a new light on an artist, an artistic movement or an art-historical period.</p>
      <p>
        This potential is widely recognized by public and private agencies, which
keep investing in publishing CH resources as LOD. In the CH domain, which in
the EU alone \accounts for 3.3% of GDP and employs 6.7 million people (3% of
total employment)" [
        <xref ref-type="bibr" rid="ref9">9</xref>
        ]; as a consequence, in the last couple of years within the
Europeana digital library1, the European Commission started a major e ort for
publishing as LOD millions of multilingual and multimodal CH resources (e.g.
archival documents, illuminated manuscripts, pictures) gathered from more than
2300 institutions distributed across 36 countries.
      </p>
      <p>Nevertheless, publishing CH resources as LOD on the Web, is just the rst
step to enable new knowledge creation. To realize the full potential of the Web
of Data, we need to devise innovative methods for seeking and creating links
between datasets and to design user-oriented services for exploiting them. To
this purpose, History of Art is a fruitful domain that, not only can bene t,
but also can assist the development of these new methods because it provides:
rich and heterogeneous information needs, ample structured and unstructured
resource datasets and a proactive research community accustomed to seek new
connections between entities.</p>
      <p>The aim of this paper is to outline the basics to design a LOD framework
for supporting knowledge creation in the History of Art domain and to outline
some of the scienti c challenges we have to face. Overall, the framework we
envision pursues two main objectives: (i) to provide LOD-ready functionalities
to create, access and cite new knowledge and to represent and retain authority
information; (ii) to assist domain experts in the creation of semantic connections
and to empower them with (semi) automatic serendipity capabilities.</p>
      <p>Accomplishing the rst goal will overcome current LOD publishing practice
where the data are produced and stored by local systems and then mapped,
synchronized and published as LOD; this procedure requires a multiplication
of investments and produces scarcely connected datasets given that the expert
knowledge is not tackled in the publishing activity. Our idea is to con ate data
creation, connection, and publishing in one phase and let computer scientists
and domain experts to work back-to-back joining knowledge and e orts. In this
context there is the need also to provide concrete solutions for fundamental but
overlooked issues as data citation and authority management.</p>
      <p>Achieving the second goal will push History of Art research boundaries by
joining the experts' knowledge and intuition with automatic tools able to retrieve
relevant entities in the Web of Data. To this purpose, from the computer science
point-of-view, we need to devise query models from the experts information
needs, envisage and de ne methods for mapping entities to queries, use them
for retrieving target entities and then provide entity rankings. Current solutions
are biased toward Web search where user needs are expressed using very few
terms; as a consequence, they employ at models of user inputs where contextual
information, text, images, links, categories and feedbacks, when available, are
considered at the same level and equivalent one to the other.</p>
      <p>The History of Art domain provides us with rich and heterogeneous user
inputs that poses additional challenges, but that also provides the means for
advancing entity retrieval state-of-the-art. The framework we envision could also</p>
      <sec id="sec-1-1">
        <title>1 http://europeana.eu/</title>
        <p>push ahead the creation and distribution of CH resources o ering new ways for
addressing consumers' demand for data access and for greater participation in
the creativity process.</p>
        <p>The rest of the paper is organized as follows: Section 2 presents the
state-ofthe-art in LOD publishing in the CH context, serendipity oriented algorithm and
data citation methodologies. In Section 3 we present the main design lines, the
challenges we have to face for realising the envisioned framework and introduce
the art research use case on which the proposed framework will rely. In Section
4 we propose an architecture that could implement the envisioned framework.
Finally, in Section 5 we draw some nal remarks.
2</p>
      </sec>
    </sec>
    <sec id="sec-2">
      <title>State-of-the-art and some open questions</title>
      <p>The e orts for disclosing the LOD potential within the CH eld must go
towards the design of new methodologies for creating meaningful and possibly
unexpected semantic links between data and for managing the knowledge
created through these connections. This endeavor is needed to make the Web of
Data fully operational and so that its role can be compared to that played by
search engines over the last two decades in disclosing the full potential of the
Web.</p>
      <p>The CH domain and particularly the History of Art not only can be bene cial
of the full exploitation of the LOD potential, but they can also provide fertile
ground where new methods and technologies can be designed, developed and
applied. Indeed, in History of Art the preponderant way to produce new knowledge
is to reveal connections between di erent items (illuminations, pictures, frescos)
that can cast new light on an artist, an artistic movement or an art-historical
period. The most valuable connections are the unexpected ones linking elements
that may seem to have very few in common; indeed, in this domain, important
discoveries have often been done thanks to associations and connections of items
emerged also by chance in the research path identi ed by domain experts. This
process of discovery can be de ned as serendipity and it is especially encouraged
by LOD where meaningful links between entities allow us to move across diverse
and apparently unrelated knowledge domains. History of Art is a well suited
domain where algorithms fostering serendipity within LOD can be designed,
developed and evaluated because it provides: rich and heterogeneous
information needs, ample structured and unstructured resource datasets and a proactive
community accustomed to seek new semantic connections between entities.</p>
      <p>
        Given the scienti c and economic impact of LOD and the leading role of CH
for triggering its potential, a lot of research is being carried out in important elds
such as database systems, information retrieval and digital libraries focusing
especially on e cient ways for storing and querying datasets [
        <xref ref-type="bibr" rid="ref2 ref3">2,3</xref>
        ], extending Web
search to entity retrieval [
        <xref ref-type="bibr" rid="ref5">5</xref>
        ] and devising sustainable methods for publishing [
        <xref ref-type="bibr" rid="ref6">6</xref>
        ]
and linking LOD [
        <xref ref-type="bibr" rid="ref11 ref13">11, 13</xref>
        ].
      </p>
      <p>However, several crucial questions still remain unanswered or overlooked.
What merits to be linked? How can linking towards potentially unknown data
sources be eased? How can the serendipity process be enhanced and guided by
semantic methodologies? How can domain experts' knowledge be exploited for
creating meaningful semantic links? How the authoritative information about
entities, activities and people involved in data creation be established and
retained? How new created knowledge represented by a data subset can be cited?</p>
      <p>
        We can point out four main research areas concerning these questions that
must be taken into account:
{ LOD-based systems: The vast majority of systems in the CH domain
are not natively LOD-based and adopt a publishing paradigm involving
redundancy and multiplication of costs [15]. Moreover, in most cases (e.g.
Europeana), the links to external sources are established without involving
domain experts and are rather syntactic than semantic. The framework we
envision is aimed as developing user-oriented services for native LOD
creation and exploration. Moreover, it aims to provide services for connecting
data while experts are working, thus exploiting their knowledge for link
creation.
{ Serendipity capabilities: No system in the CH domain provides end-users
with this function. Related methodologies are proposed in the context of
entity linking and retrieval targeting Web search: they deal with sparse user
inputs [
        <xref ref-type="bibr" rid="ref1">1</xref>
        ] and general datasets [16]. These solutions employ at models of
user inputs mainly focusing on query expansion techniques [
        <xref ref-type="bibr" rid="ref4">4</xref>
        ].
{ LOD citation and authority management: Recently two EU projects
(i.e. PRELIDA2 and DIACHRON3) marginally considered these aspects
from the permanent preservation point-of-view, but there are as yet no
concrete solutions for LOD data. Other citation systems focused on
relational data [17] or hierarchical data such as eXtensible Markup Language
(XML) [
        <xref ref-type="bibr" rid="ref7">7</xref>
        ], but they are not applicable to the LOD case. [18] proposed
an initial methodology based on named graphs and RDF quad semantics
which enables persistent, dereferenceable, variable granularity and
humanand machine-readable citations of LOD subsets, but no ready-to-use solution
has been implemented yet. Several challenges are still open in data citation
of LOD such as the citation of evolving datasets which involves temporal
aspects of RDF, the de nition of equality between two citations and the
concept of citation closure [
        <xref ref-type="bibr" rid="ref8">8</xref>
        ] (something like closure of a paper references).
3
      </p>
    </sec>
    <sec id="sec-3">
      <title>Serendipity in art research: a proposal</title>
      <p>For what it is concerned with the serendipity algorithms the framework we
envision will employ (semi) automatic methods for suggesting entities to researchers;
the idea is that this methodology could help them in their daily work by
triggering the exploration of new research paths leading to new knowledge creation
in History of Art. Now, we present the concrete steps we are going to follow to</p>
      <sec id="sec-3-1">
        <title>2 http://prelida.eu/ 3 http://www.diachron-fp7.eu/</title>
        <p>realize the framework that we call Linked Open Data Enhanced Framework for
Serendipity in Art Research (LODESTAR).</p>
        <p>We will rely on a concrete use case rooted in history of art research which
aim is to identify, analyze and assess the in uence of the most ancient Divine
Comedy illuminated manuscript illustrations on the many representations of the
Hereafter and biblical and mythological characters that can be found in
lateXIV century and Renaissance art. Indeed, many inventions that can be found
in the rst Divine Comedy illuminated manuscripts were destined to have great
success and also to be used in later manuscripts, not only with Dante's poem,
but also with other texts and other works of art. In order to achieve large-scale
results, beside illuminations, we will also take into account products of the major
arts such as paintings, frescos and sculptures, thus bringing new knowledge on
di erent elds of the History of Art. The object of the research is highly original
and constitutes a eld of investigation that has never been explored before.
Given its ample scope, the large di usion of the Divine Comedy and its in uence
through the centuries on many elds, this use-case o ers substantial chances of
success especially for applying semi-automatic serendipity methodologies; it is
a valuable starting stage to design, apply and test methods that will be then
re-used in other use-cases and domains.</p>
        <p>Input streams
extracted from
the environment</p>
        <p>LODESTAR
work environment</p>
        <p>Rich and
heterogeneous inputs</p>
        <p>related images
multimodal related documents</p>
        <p>item connections
domain experts descriptions</p>
        <p>In Figure 1 we show the work ow realizing the objectives tackled by the
LODESTAR framework. In the center there is the work environment where
domain experts will carry out their daily research work. This environment provides
rich and heterogeneous input streams such as free and semi-structured text,
images, links, user pro le and contextual information; LODESTAR is required to
automatically catch these streams that are implicit expressions of user needs and
to extract entities from them.</p>
        <p>These entities will be sources for the entity retrieval algorithms realizing the
LODESTAR serendipity methodology shown on the right hand side of Figure
1. This would be a major shift from state-of-the art entity retrieval solutions,
which are typically employed in syntactic linking processes where input sources
are constituted by few keywords. In LODESTAR users are not asked to explicitly
issue queries, but the traces they leave behind while using the system will be
automatically gathered and exploited for devising queries. To tackle this goal
we have to employ new methods to model the extracted entities by associating
di erent groups of entities to di erent levels of relevance decided on the basis of
the user needs within the context under consideration; in other words we have
to go beyond the current \ at models" that consider all the user inputs at the
same level.</p>
        <p>This model will allow us to devise several queries corresponding to di erent
entities selections and compositions that will be used to retrieve \target entities"
from the LOD cloud. The purpose of issuing multiple queries is to cover a wide
spectrum of potential user needs as well as to ease the discovery of unknown
entities that may trigger the creation of unexpected connections. The rank module
gathers the target entities and orders them accordingly to the user need they
tackle.</p>
        <p>Lastly, the ranked entities are suggested to the users that will consider them
for establishing semantic connections fostering new knowledge creation.</p>
        <p>
          The left-hand side of Figure 1 reports data citation and authority
management components needed to give credit to data creators (e.g. produce a
diachronic reference to a data subset), retain context (e.g. who created a link,
when a link has been created) and control data con icts (e.g. choose the most
authoritative statement among contradicting ones). LODESTAR has to provide
general methodologies for tackling these issues for which, in literature, there are
no concrete solutions yet. We will employ the solution proposed in [18] by
extending it in order to consider also the temporal aspects of LOD datasets; indeed,
a citation should be resolved by referring to the precise version of the data that
was cited. To this end we plan to employ some RDF versioning methodology as
described in [
          <xref ref-type="bibr" rid="ref12">12, 19</xref>
          ].
4
        </p>
      </sec>
    </sec>
    <sec id="sec-4">
      <title>A possible implementing architecture</title>
      <p>Serendipity can be realized by following the work ow shown on the right of
Figure 1. Firstly, we need to trace the user information within the LODESTAR work
environment such as the item being studied, annotations, items being compared
to it, free text and semi-structured text (e.g. form elds), clickstream and user
pro les. Then, we apply mining algorithms, extending existing entity extraction
solutions, for extracting phrases, entities, categories from all these input streams;
afterwards, the enhancement module, exploiting suitable ontologies, will connect
the extracted entities one with the other via automatically established semantic
links. At this stage we obtain an RDF graph of connected entities representing
a very rich combination of user needs.</p>
      <p>There exists no ready-to-use solution for handling this \user needs graph" and
devise queries from it; LODESTAR represents the very rst e ort tackling this
problem. The methodology we envision is to hierarchically model the user needs
graph by associating di erent groups of entities to di erent levels of relevance
decided on the basis of contextual information.</p>
      <p>
        Machine learning methods have been proposed in literature for the task of
learn classes in ontologies and constructing knowledge from user-provided
examples; a possibility is to develop a similar approach to assign relevance weights
or probabilities (i.e. the estimated relatedness to the user needs) to the entities
in the graph. These weights will be exploited for modeling the user needs graph
with a innovative set-based model, i.e. the NESTOR model [
        <xref ref-type="bibr" rid="ref10">10</xref>
        ], organizing the
entities in sets with di erent levels relevance as shown in Figure 2.
      </p>
      <sec id="sec-4-1">
        <title>User needs graph modeled via NESTOR</title>
      </sec>
      <sec id="sec-4-2">
        <title>Adaptive Query Modeling</title>
        <p>I</p>
        <p>B</p>
        <p>A
C</p>
        <p>D</p>
        <p>S+(B) = {X 2 I | X ⇢ B} = {C, D}
E</p>
        <p>BE = B [ E
B = B \ S+(B)
B</p>
        <p>C</p>
        <p>D
B</p>
        <p>C</p>
        <p>D</p>
        <p>E</p>
      </sec>
      <sec id="sec-4-3">
        <title>Queries</title>
        <p>qS+(B)
qBE
qB</p>
        <p>NESTOR comes with a whole bunch of de ned set-operations that will be
exploited to design an algorithm for selecting alternative entity combinations
within the graph; from these selections a bunch of queries covering a wide
spectrum of potential user needs will be devised.</p>
        <p>The queries will be issued against the LOD cloud for retrieving groups of
target entities; each group will be ordered by a ranking model built on solutions
for structured search, e.g. statistical language models, and then aggregated by
exploiting meta-search engine techniques to be presented as link suggestions to
the users.</p>
        <p>This envisioned methodology can be embedded in a more comprehensive
architecture as shown in Figure 3 composed by 4-staked layers concerning
important aspects of a system architecture: scalability, robustness, reactivity and
suitability.</p>
        <p>The rst layer is designed by considering that LODESTAR will deal with
big datasets and large user and machine requests. Therefore, in order to provide
linear scalability and fault tolerance, several technologies should be taken into
L D E S t A R</p>
        <p>framework
y
t
i
l
i
b
a
t
i
u
S Analysis</p>
        <p>Service</p>
        <p>CREATE
y
itv SPARQL
itc end-point
a
e
R
s
s
e
tn Data
su Integration
b
o
R
y
t
i
l
i
b
a
l
a
c
S</p>
        <p>Domain Expert Community
Search
Service
FIND</p>
        <p>Visualization</p>
        <p>Service</p>
        <p>Linking
Service</p>
        <p>Citation</p>
        <p>Service
EXPLORE</p>
        <p>INNOVATE</p>
        <p>ATTRIBUTE</p>
        <p>Entity
Extraction and
Enhancement</p>
        <p>User's Inputs</p>
        <p>Modeling</p>
        <p>Adaptive</p>
        <p>Querying</p>
        <p>SERENDIPITY CAPABILITIES
Ontology</p>
        <p>Data Citation</p>
        <p>Authority
Management
Map-Reduce
Programming</p>
        <p>Model</p>
        <p>Hadoop
Distributed File</p>
        <p>System</p>
        <p>Mapping and</p>
        <p>Synchronization</p>
        <p>Distributed Multinode Cluster</p>
        <p>Fig. 3. A possible architecture implementing the LODESTAR framework.
account, in particular: Apache Cassandra4 as distributed column store solution
and Apache Hadoop5 in conjunction with the Map-Reduce functionalities to
provide reliable and distributed computing.</p>
        <p>Scalability is also concerned with the selection of datasets that will serve
the use-case we have in mind; in a real-world scenario we cannot assume that
all relevant data will be available as LOD because even if many relevant CH
institutions and systems are LOD-compliant { e.g. Europeana6, the Library of
Congress, the British Library, the French National Library, and Cultura Italia7
{ others still are not { e.g. Artstor, Dante On-line8, and the Web gallery of Art.</p>
        <p>To deal with all these sources, it is necessary to employ a mapping and
synchronizing module that will harvest, map to LOD and store these data. An
apposite data integration module will blend di erent resource kinds (as part of
the robustness layer). All the resources will be represented accordingly to
ontologies de ned within the domain of interest; computer scientist and art historians
will jointly carry out this activity by de ning an RDF Schema and reusing
existing vocabularies to establish bridges with third-party ontologies { e.g SKOS,
Europeana Data Model, CIDOC-CRM { that will also be exploited for
serendipity.</p>
        <p>In the reactivity layer, we could provide a SPARQL end-point, which will
work as a data provider allowing third-party services to discover, re-use and
connect to the LODESTAR data. This layer also accommodates the modules
implementing the serendipity algorithms.</p>
        <p>All the presented services and solutions must be made available to end-users
via pluggable user interface software components (i.e. Web portlets); the main
end-user interfaces need to deal with an environment for describing items,
annotating and visually comparing them, seeking and connecting entities, browsing
the data and making use of the data citation and serendipity modules.
5</p>
      </sec>
    </sec>
    <sec id="sec-5">
      <title>Final Remarks</title>
      <p>In this paper we discussed the role of LOD in the CH eld and we outlined some
of the cultural and computational challenges we have to face in order to de ne a
framework for supporting knowledge creation in CH. In particular, we selected a
stimulating use case based on History of Art research which provides us with rich
input data and a research community oriented to seek new connections between
di erent items to create new knowledge.</p>
      <p>The main challenges we outlined regard the necessity to de ne new
methodologies for entity linking and retrieval which take into account complex user
inputs, a work environment which on the one hand exploits existing LOD for
knowledge creation and on the other hand stimulates the production of new and
4 http://cassandra.apache.org/
5 https://hadoop.apache.org/
6 http://www.europeana.eu/
7 http://www.culturaitalia.it/
8 http://www.danteonline.it/
unexpected connections between CH resources and a data citation
methodology allowing us to cite evolving subset of data with variable granularity and to
produce both machine- and human-readable references.
15. Marden, J., Li-Madeo, C., Whysel, N., Edelstein, J.: Linked Open Data for Cultural
Heritage: Evolution of an Information Technology. In: Proc. of the 31st ACM
International Conference on Design of Communication. pp. 107{112. SIGDOC '13,
ACM, New York, NY, USA (2013)
16. Meij, E., Bron, M., Hollink, L., Huurnink, B., de Rijke, M.: Mapping Queries to
the Linking Open Data Cloud: A Case Study Using DBpedia. Web Semant. 9(4),
418{433 (2011)
17. Proll, S., Rauber, A.: Scalable Data Citation in Dynamic, Large Databases: Model
and Reference Implementation. In: Proc. of the 2013 IEEE International
Conference on Big Data, 6-9 October 2013, Santa Clara, CA, USA. pp. 307{312 (2013)
18. Silvello, G.: A Methodology for Citing Linked Open Data Subsets. D-Lib Magazine
21(1/2) (2015), http://dx.doi.org/10.1045/january2015-silvello
19. Stefanidis, K., Chrysakis, I., Flouris, G.: On Designing Archiving Policies for
Evolving RDF Datasets on the Web. Lecture Notes in Computer Science, vol. 8824, pp.
43{56. Springer (2014)
20. W3C: RDF 1.1 Concepts and Abstract Syntax { W3C Recommendation 25
February 2014 (February 2014)</p>
    </sec>
  </body>
  <back>
    <ref-list>
      <ref id="ref1">
        <mixed-citation>
          1.
          <string-name>
            <surname>Balog</surname>
            ,
            <given-names>K.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Bron</surname>
            ,
            <given-names>M.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>De Rijke</surname>
            ,
            <given-names>M.</given-names>
          </string-name>
          :
          <article-title>Query Modeling for Entity Search Based on Terms, Categories, and Examples</article-title>
          .
          <source>ACM Trans. Inf. Syst</source>
          .
          <volume>29</volume>
          (
          <issue>4</issue>
          ),
          <volume>22</volume>
          :1{
          <fpage>22</fpage>
          :
          <fpage>31</fpage>
          (
          <year>2011</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref2">
        <mixed-citation>
          2.
          <string-name>
            <surname>Blanco</surname>
            ,
            <given-names>R.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Mika</surname>
            ,
            <given-names>P.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Vigna</surname>
            ,
            <given-names>S.:</given-names>
          </string-name>
          <article-title>E ective and E cient Entity Search in RDF Data</article-title>
          .
          <source>In: Proc. of the 10th international conference on The semantic web - Volume Part I</source>
          . pp.
          <volume>83</volume>
          {
          <fpage>97</fpage>
          . Springer-Verlag, Berlin, Heidelberg (
          <year>2011</year>
          ), http://dl.acm.org/ citation.cfm?id=
          <volume>2063016</volume>
          .
          <fpage>2063023</fpage>
        </mixed-citation>
      </ref>
      <ref id="ref3">
        <mixed-citation>
          3.
          <string-name>
            <surname>Blanco</surname>
            ,
            <given-names>R.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Ottaviano</surname>
            ,
            <given-names>G.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Meij</surname>
          </string-name>
          , E.: Fast and
          <string-name>
            <surname>Space-E cient Entity</surname>
          </string-name>
          <article-title>Linking for Queries</article-title>
          .
          <source>In: Proc. of the Eighth ACM International Conference on Web Search and Data Mining - WSDM '15</source>
          . pp.
          <volume>179</volume>
          {
          <fpage>188</fpage>
          . ACM Press, New York, New York, USA (Feb
          <year>2015</year>
          ), http://dl.acm.org/citation.cfm?id=
          <volume>2684822</volume>
          .
          <fpage>2685317</fpage>
        </mixed-citation>
      </ref>
      <ref id="ref4">
        <mixed-citation>
          4. Boston,
          <string-name>
            <given-names>C.</given-names>
            ,
            <surname>Fang</surname>
          </string-name>
          ,
          <string-name>
            <given-names>H.</given-names>
            ,
            <surname>Carberry</surname>
          </string-name>
          ,
          <string-name>
            <given-names>S.</given-names>
            ,
            <surname>Wu</surname>
          </string-name>
          ,
          <string-name>
            <given-names>H.</given-names>
            ,
            <surname>Liu</surname>
          </string-name>
          ,
          <string-name>
            <surname>X.</surname>
          </string-name>
          :
          <article-title>Wikimantic: Toward E ective Disambiguation and Expansion of Queries</article-title>
          .
          <source>Data Knowl. Eng</source>
          .
          <volume>90</volume>
          ,
          <issue>22</issue>
          {
          <fpage>37</fpage>
          (
          <year>2014</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref5">
        <mixed-citation>
          5.
          <string-name>
            <surname>Bron</surname>
            ,
            <given-names>M.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Balog</surname>
            ,
            <given-names>K.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>de Rijke</surname>
          </string-name>
          , M.:
          <article-title>Example Based Entity Search in the Web of Data</article-title>
          .
          <source>In: Proc. of the 35th European conference on Advances in Information Retrieval</source>
          . pp.
          <volume>392</volume>
          {
          <fpage>403</fpage>
          . Springer-Verlag, Berlin, Heidelberg (
          <year>2013</year>
          ), http://dx. doi.org/10.1007/978-3-
          <fpage>642</fpage>
          -36973-5\_
          <fpage>33</fpage>
        </mixed-citation>
      </ref>
      <ref id="ref6">
        <mixed-citation>
          6.
          <string-name>
            <surname>Buccio</surname>
            ,
            <given-names>E.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Di Nunzio</surname>
            ,
            <given-names>G.M.D.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Silvello</surname>
          </string-name>
          , G.:
          <article-title>A Linked Open Data Approach for Geolinguistics Applications</article-title>
          .
          <source>Int. J. Metadata Semant. Ontologies</source>
          <volume>9</volume>
          (
          <issue>1</issue>
          ),
          <volume>29</volume>
          {
          <fpage>41</fpage>
          (
          <year>2014</year>
          ), http://dx.doi.org/10.1504/IJMSO.
          <year>2014</year>
          .059125
        </mixed-citation>
      </ref>
      <ref id="ref7">
        <mixed-citation>
          7.
          <string-name>
            <surname>Buneman</surname>
            ,
            <given-names>P.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Silvello</surname>
          </string-name>
          , G.:
          <article-title>A Rule-Based Citation System for Structured and Evolving Datasets</article-title>
          .
          <source>IEEE Data Eng. Bull</source>
          .
          <volume>33</volume>
          (
          <issue>3</issue>
          ),
          <volume>33</volume>
          {
          <fpage>41</fpage>
          (
          <year>2010</year>
          ), http://sites.computer. org/debull/A10sept/buneman.pdf
        </mixed-citation>
      </ref>
      <ref id="ref8">
        <mixed-citation>
          8.
          <string-name>
            <surname>Buneman</surname>
            ,
            <given-names>P.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Tannen</surname>
            ,
            <given-names>V.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Davidson</surname>
            ,
            <given-names>S.B.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Frew</surname>
            ,
            <given-names>J.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Cohen-Boulakia</surname>
            ,
            <given-names>S.</given-names>
          </string-name>
          :
          <source>Computational Challanges in Data Citation. Workshop report</source>
          , University of Pennsylvania (
          <year>2014</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref9">
        <mixed-citation>
          9.
          <article-title>Commission of the European Communities: Communication from the Commission to the European Parliament, the Council, the European Economic and Social Committee and the Committee of the Regions: Towards a Thriving Data-Driven Economy</article-title>
          .
          <source>COMM</source>
          (
          <year>2014</year>
          ) 442 Final (
          <year>2014</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref10">
        <mixed-citation>
          10.
          <string-name>
            <surname>Ferro</surname>
            ,
            <given-names>N.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Silvello</surname>
          </string-name>
          , G.:
          <article-title>NESTOR: A Formal Model for Digital Archives</article-title>
          .
          <source>Information Processing &amp; Management</source>
          <volume>49</volume>
          (
          <issue>6</issue>
          ),
          <volume>1206</volume>
          {1240 (November
          <year>2013</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref11">
        <mixed-citation>
          11.
          <string-name>
            <surname>Ferro</surname>
            ,
            <given-names>N.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Silvello</surname>
          </string-name>
          , G.:
          <article-title>Making it Easier to Discover, Re-Use and Understand Search Engine Experimental Evaluation Data</article-title>
          .
          <source>ERCIM News</source>
          <volume>96</volume>
          ,
          <issue>26</issue>
          {27 (
          <year>January 2014</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref12">
        <mixed-citation>
          12.
          <string-name>
            <surname>Flouris</surname>
            ,
            <given-names>G.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Konstantinidis</surname>
            ,
            <given-names>G.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Antoniou</surname>
            ,
            <given-names>G.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Christophides</surname>
            ,
            <given-names>V.</given-names>
          </string-name>
          :
          <article-title>Formal Foundations for RDF/S KB Evolution</article-title>
          . Knowl. Inf. Syst.
          <volume>35</volume>
          (
          <issue>1</issue>
          ),
          <volume>153</volume>
          {
          <fpage>191</fpage>
          (
          <year>2013</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref13">
        <mixed-citation>
          13.
          <string-name>
            <surname>Gottipati</surname>
            ,
            <given-names>S.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Jiang</surname>
          </string-name>
          , J.:
          <article-title>Linking Entities to a Knowledge Base with Query Expansion</article-title>
          .
          <source>In: Proc. of the 2011 Conference on Empirical Methods in Natural Language Processing</source>
          . pp.
          <volume>804</volume>
          {
          <fpage>813</fpage>
          . Association for Computational Linguistics (
          <year>Jul 2011</year>
          ), http://dl.acm.org/citation.cfm?id=
          <volume>2145432</volume>
          .
          <fpage>2145523</fpage>
        </mixed-citation>
      </ref>
      <ref id="ref14">
        <mixed-citation>
          14.
          <string-name>
            <surname>Heath</surname>
            ,
            <given-names>T.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Bizer</surname>
            ,
            <given-names>C.</given-names>
          </string-name>
          :
          <article-title>Linked Data: Evolving the Web into a Global Data Space</article-title>
          .
          <source>Synthesis Lectures on the Semantic Web: Theory and Technology</source>
          . Morgan &amp; Claypool Publishers, USA (
          <year>2011</year>
          )
        </mixed-citation>
      </ref>
    </ref-list>
  </back>
</article>