<!DOCTYPE article PUBLIC "-//NLM//DTD JATS (Z39.96) Journal Archiving and Interchange DTD v1.0 20120330//EN" "JATS-archivearticle1.dtd">
<article xmlns:xlink="http://www.w3.org/1999/xlink">
  <front>
    <journal-meta />
    <article-meta>
      <title-group>
        <article-title>Semantic Search Architecture for Retrieving Information in Biodiversity Repositories</article-title>
      </title-group>
      <contrib-group>
        <contrib contrib-type="author">
          <string-name>Flor K. Amanqui</string-name>
          <xref ref-type="aff" rid="aff1">1</xref>
        </contrib>
        <contrib contrib-type="author">
          <string-name>Kleberson J. Serique</string-name>
          <xref ref-type="aff" rid="aff1">1</xref>
        </contrib>
        <contrib contrib-type="author">
          <string-name>Franco Lamping</string-name>
          <email>lamping@grad.icmc.usp.br</email>
          <xref ref-type="aff" rid="aff1">1</xref>
        </contrib>
        <contrib contrib-type="author">
          <string-name>Andre´a C. F. Albuquerque</string-name>
        </contrib>
        <contrib contrib-type="author">
          <string-name>Jose´ L. C. Dos Santos</string-name>
          <xref ref-type="aff" rid="aff0">0</xref>
        </contrib>
        <contrib contrib-type="author">
          <string-name>Dilvan A. Moreira</string-name>
          <email>dilvang@icmc.usp.br</email>
          <xref ref-type="aff" rid="aff1">1</xref>
        </contrib>
        <aff id="aff0">
          <label>0</label>
          <institution>National Institute for Amazonian Research (INPA) CEP.</institution>
          <addr-line>: 69060-001 - Manaus - AM -</addr-line>
          <country country="BR">Brazil</country>
        </aff>
        <aff id="aff1">
          <label>1</label>
          <institution>University of Sa ̃o Paulo (USP) - CEP: 13566-590 - Sa ̃o Carlos - SP -</institution>
          <country country="BR">Brazil</country>
        </aff>
      </contrib-group>
      <fpage>83</fpage>
      <lpage>93</lpage>
      <abstract>
        <p>The amount of biological data available electronically is increasing at a rapid rate; for instance, over 16.500 specimens are available today in the National Institute for Amazonian Research (INPA) collections. However, this data is not semantically categorized and stored and thus is difficult to search. To tackle this problem, we present a semantic search architecture, implemented using state of the art semantic web tools, and test it on a set of representative data about biodiversity from INPA. This paper describes how the mechanism of mapping is designed so that the semantic search can find information, based on ontologies. We show a series of SPARQL queries and explain how the mapping mechanism works. Our experiments, using a prototype of the proposed architecture, showed that the prototype had better precision and recall then traditional keyword based search engines.</p>
      </abstract>
      <kwd-group>
        <kwd>Biodiversity</kwd>
        <kwd>Ontology</kwd>
        <kwd>Data Integration</kwd>
        <kwd>Semantic Search</kwd>
      </kwd-group>
    </article-meta>
  </front>
  <body>
    <sec id="sec-1">
      <title>1. Introduction</title>
      <p>Biological diversity, or biodiversity, is the term given to the variety of life on Earth.
Biodiversity is the combination of life forms and their interactions with one another, and with
the physical environment that has made Earth habitable.</p>
      <p>The biodiversity information that can be obtained via Internet continues to grow
significantly. Every day, new collections, databases, and applications are being added.
This information is stored in a variety of formats (spreadsheets, html, xml, pdf and
catalogues, amongst others). This proliferation of information from different sources means
that the search for information could be met by a variety of available resources, which
store data about the same domains but have different characteristics. For that reason,
much of this information is never found. The need for integration and analysis of
biodiversity information becomes evident.</p>
      <p>In this context, finding relevant and recent information is a hard task that is not
particularly well supported by current biodiversity software tools. Keyword-based search
have serious problems associated with its use: low or no recall; high recall, low precision;
initial keywords in search often do not get the wanted results.</p>
      <p>
        The semantic web (an extension of the current Web) tries to represent information
in such a way that it can be used by machines, not just for display purposes, but also for
automation, integration and reuse across applications [
        <xref ref-type="bibr" rid="ref4">Boley et al. 2001</xref>
        ].
      </p>
      <p>There are a number of important technologies related to the Semantic Web:
ontologies, languages for the Semantic Web, semantic search, semantic markup of pages and
services (that the Semantic Web is supposed to provide). Ontologies, one of the most
important ones, are implemented in the RDF(S) (Resource Description Framework/Schema)
and OWL (Web Ontology Languages) languages, two W3C recommended data
representation models.</p>
      <p>In this article, we propose a semantic search architecture that supports mapping
between biodiversity data, from INPA’s (National Institute for Amazonian Research)
collections, stored in relational databases and the ontologies describing it.</p>
      <p>The rest of this article is organized as follows: Section 2 describes related works.
Section 3 describes our biodiversity ontology. Section 4 presents our semantic search
architecture. Section 5 presents a synopsis of our experiments and Section 6 concludes
by summarizing our results and describing future works.</p>
    </sec>
    <sec id="sec-2">
      <title>2. Related Works</title>
      <p>Researchers have proposed various techniques and approaches designed to perform
semantic search. We studied a number of them that could be used in the area of biodiversity.</p>
      <p>
        In [
        <xref ref-type="bibr" rid="ref17">Xiong et al. 2009</xref>
        ], a method of search based on a smart query agent is
proposed (Geoonto). It retrieves information from data catalogs/databases using ontologies.
This method associates semantic information in the search process, and generates a
refined query string.
      </p>
      <p>
        In [
        <xref ref-type="bibr" rid="ref7">Latiri et al. 2012</xref>
        ], an automatic method of query expansion is proposed in
which user requests are expressed in natural language.
      </p>
      <p>
        In [
        <xref ref-type="bibr" rid="ref11">Mittal et al. 2010</xref>
        ], a method hybrid of personalized web information is
proposed in which ontology for retrieval of user context is used and a user profile is being
maintained.
      </p>
      <p>
        In [
        <xref ref-type="bibr" rid="ref8">Li and Yang 2008</xref>
        ], a method to construct a semantic search engine is
proposed. It provides a uniform platform to search, view and operate spatial on information.
      </p>
      <p>
        In [
        <xref ref-type="bibr" rid="ref15">Santos et al. 2011</xref>
        ], an architecture to support semantic search in a metadata
repository is proposed. This work discover similar concepts even when different terms
are used in their designation or description, since a domain ontology is used to
annotate information sources and to expand the user query with terms from the universe of
discourse.
      </p>
      <p>A number of techniques have been developed for using ontologies to retrieve
relevant documents in response to a query. However, none of them focused on the problem
of storage and retrieval of RDF triples. Most of these techniques require complex
analysis, involving natural language processing, to discover the context and semantics of query
terms. Also, an additional limitation, in many of the existing approaches, is the lack of a
quality evaluation of results.</p>
      <p>We have developed a semantic search application that uses key semantic web
concepts for information retrieval and also technologies such as mapping, triple store and
SPARQL queries.</p>
    </sec>
    <sec id="sec-3">
      <title>3. The Biodiversity ontology</title>
      <p>OntoBio is a biodiversity ontology developed by INPA and UFAM (Federal University
of Amazonas) and extended by USP (University of Sa˜o Paulo). Its main objective is to
provide a clear and precise conceptualization of the aspects considered in biodiversity
data collection, regardless of a specific application.</p>
      <p>
        The original version of OntoBio is presented in details at [
        <xref ref-type="bibr" rid="ref1">Albuquerque 2011</xref>
        ].
One of the advantages of having data annotated using OntoBio concepts is that it can be
reused as Linked Data. Linked Data describes a method of publishing structured data so
that it can be interlinked and become more useful [
        <xref ref-type="bibr" rid="ref6">Kauppinen and de Espindola 2011</xref>
        ].
      </p>
      <p>To better archive that, data annotated using OntoBio has to be easily interlinked
with other biodiversity data, already available on the web (as part of the wider Linked
Data community), through the use of as many shared concepts as possible. With that in
mind, we rewrote the first version of OntoBio to reuse, whenever possible, terms from
other public available ontologies to allow better ”linkability” with data already annotated
using them.</p>
      <sec id="sec-3-1">
        <title>We added terms from the following public ontologies:</title>
        <p>
          The Phenotypic Quality Ontology[
          <xref ref-type="bibr" rid="ref12">PATO 2010</xref>
          ], which is an ontology of
phenotypic qualities, intended for use in a number of applications, primarily defining
composite phenotypes and phenotype annotation;
Basic Geo Vocabulary [
          <xref ref-type="bibr" rid="ref16">WGS84 2003</xref>
          ] is a basic RDF vocabulary that
provides the semantic web community with a namespace for representing lat(itude),
long(itude) and other information about spatially-located things, using WGS84 as
a reference datum, and;
The Geoname Ontology[
          <xref ref-type="bibr" rid="ref5">GeoNames 2011</xref>
          ] makes it possible to add geospatial
semantic information to the Word Wide Web. All over 8.3 million geonames
toponyms now have a unique URL with a corresponding RDF web service.
        </p>
        <p>The OntoBio ontology is presented in the Figure 1. The Prote´ge´
4 ontology editor was used to write the OntoBio ontology in OWL 2 DL.
The new version of OntoBio is available through the NCBO’s Bioportal
http://bioportal.bioontology.org/ontologies/50517. There, users can download, browse
and suggest terms for the ontology.</p>
      </sec>
    </sec>
    <sec id="sec-4">
      <title>4. An Architecture for Semantic Search</title>
      <p>We have proposed the architecture of a semantic search that follows the mechanism of
mapping between OntoBio domain ontology, and Database from INPA the collections of
insects, fishes, and mammals.</p>
      <p>The system overall architecture is shown in Figure 2. It consists of four basic
modules: User Interface Layer, Query Reformulation, Mapping Component and Data
Access Layer.
1. User Interface Layer is responsible for the interaction between users and system.</p>
      <p>The search process begins with an initial keyword list, entered by the users, that
represents his/her search intentions.
2. The Query Reformulation component receives the input of search terms from
the user, selects and expands keyword lists by adding semantically related terms,
using techniques of expansion and semantic similarity. It uses the SPARQL Writer
component to take keyword lists and generate SPARQL queries from them. It uses
an algorithm that will be described on the following sections.
3. The Mapping Component loads the domain ontologies, taxonomic information
and the collection database and transforms them in a set of Resource
Description Framework (RDF) triples. We used Ontop, a platform to query databases as
Virtual RDF Graphs using SPARQL, to do the mapping between the relational
databases records and the OWL ontologies.</p>
      <p>
        Ontop is a platform to query databases as Virtual RDF Graphs using SPARQL. It
does the mapping between the relational databases records and the OWL
ontologies. Ontop has two tools: OntopPro, which is a Protege 4 plugin that implements
a graphic mapping editor; and Quest, which is a SPARQL query engine/reasoner
that supports RDFS and OWL 2 QL entailment regimes and SPARQL-to-SQL
query rewriting
        <xref ref-type="bibr" rid="ref10">(Mariano R and Calvanese, 2012)</xref>
        . The mapping process is
divided in three steps:
(a) Creation of Mapping Axioms: OntopPro mappings are done using
mapping axioms. A mapping axiom is defined by an SQL query and an ABox
assertion template (Figure 3). An Abox assertion template is a set of
RDF/OWL triples, written in a turtle-like syntax, in which the subject and
object of the triples allow for variables that reference columns of the SQL
query result [Mariano R and Calvanese 2012].
      </p>
      <p>In other words, a mapping axiom defines how the values in each row of
the results (of an SQL query) can be used to generate a set of ABox
assertions. The mapping axioms were created using information from the
OntoBio ontology and INPA experts. Each mapping must contain one or
more mapping axioms. Figure 3 shows a valid mapping.
(b) Generation of RDF Triples: Mapping axioms generate RDF triples. This
generation is done using the Quest tool from Ontop. The Quest reasoner
uses query-rewriting techniques to generate triples. The triples are created
by replacing the placeholders in the target with the values from the SQL
row.
(c) RDF Triples Loader: Using OntopPro, it is possible to export the RDF
triples generated by the Quest tool to a file. That file is then loaded into
the Virtuoso triple store, which is now ready to answer queries using them.
The Mapping Component can repeat the process described here, whenever
INPA releases updates to its collection records.
4. Data Access layer that is the architecture layer that provides access to the RDF
triples stored in the Virtuoso Triple Store, using SPARQL, both for the layer above
it and for other machines on the network. Triple Store is the common name given
to a database management system for RDF Data.</p>
    </sec>
    <sec id="sec-5">
      <title>4.1. Semantic Search Algorithm</title>
      <p>The basic idea of our algorithm is to compare input keywords with OntoBio resources
(subject, predicate and object) in the Virtuoso triple store. The Virtuoso platform was
chosen because it can store the triples generated from INPA data and work with multiple
graphs at the same time.
We implemented this algorithm in a prototype using: Java, Eclipse Indigo (as IDE),
Google Web Toolkit 2.5.1 to create a web client, Jena RDF framework to process
(simplified) SPARQL queries and Virtuoso Server as triple store.</p>
      <p>Figure 4 shows graphic interface to support user queries. We implemented a
SPARQL Endpoint for INPA http://143.107.231.220:8890/sparql and implemented a set
of queries described on experiments section.</p>
    </sec>
    <sec id="sec-6">
      <title>5. Experiments</title>
      <p>In order to validate our proposed architecture, researchers from our group and biodiversity
scientists were interviewed to categorize important information from the INPA data.</p>
      <p>We defined use cases (Table 1) with scenarios to identify the various user tasks
and built SPARQL queries related with these use cases.</p>
      <p>For each of the previous use cases, biodiversity experts identified the information
set each user needed for each task and examples of queries that should have returned this
information. After we tested each query, the same experts judged which results were
relevant and non relevant (relevance non relevance judgment).</p>
      <p>
        This process of information feedback is commonly referred to, in the literature, as
relevance feedback [
        <xref ref-type="bibr" rid="ref14">Salton 1971</xref>
        ] when experts explicitly provide information on relevant
documents to a query [
        <xref ref-type="bibr" rid="ref3">Baeza-Yates and Ribeiro-Neto 1999</xref>
        ]. In its original formulation,
expert users inspect the query results and indicate those that are really relevant to the
search. Table [tab:InfoNeeds] shows examples of users tasks and possible query strings
to get the relevant biodiversity information.
      </p>
      <p>Scientists can identify species using the taxonomic classification system no
matter what their language. The taxonomic classification system is composed by a
hierarchy (series of ranks) that shows the kinship of organisms and also, whenever possible,
ancestor-descendant relationships.</p>
      <p>The basic ranks of the taxonomic classification system are kingdom, phylum,
class, order, family, genus and species. The following SPARQL query (Listing 1) shows
taxonomic system of classification for the kingdom Animalia.</p>
      <p>Listing 1. SPARQL query returning the taxonomy of a specie
PREFIX oo : &lt;h t t p : / / www. owl o n t o l o g i e s . com /
B i o d i v e r s i t y O n t o l o g y F u l l . owl#&gt;
PREFIX owl : &lt;h t t p : / / www. w3 . o r g / 2 0 0 2 / 0 7 / owl#&gt;
PREFIX r d f : &lt;h t t p : / / www. w3 . o r g /1999/02/22 r d f s y n t a x ns#&gt;
PREFIX r d f s : &lt;h t t p : / / www. w3 . o r g / 2 0 0 0 / 0 1 / r d f schema#&gt;
s e l e c t ? phylum ? c l a s s ? o r d e r ? f a m i l y ? genus ? s p e c i e s
where f oo : kingdom A n i m a l i a oo : s u b k i n o f P h y K i n g ? phylum .
? phylum oo : s u b k i n o f C l a s s P h y ? c l a s s .
? c l a s s oo : s u b k i n o f O r d C l a s s ? o r d e r .
? o r d e r oo : subkinofFamOrd ? f a m i l y .
? f a m i l y oo : subkindOfGenFam ? genus .
? genus oo : subkindOfEspGen ? s p e c i e s .
g</p>
      <p>The following SPARQL query (Listing 2) shows important information from a
collect such as Collect, Research Institution, Method, Determinate Name.</p>
      <p>Listing 2. SPARQL query returning information of a collect
PREFIX : &lt;h t t p : / / www. owl o n t o l o g i e s . com /
B i o d i v e r s i t y O n t o l o g y F u l l . owl#&gt;
PREFIX owl : &lt;h t t p : / / www. w3 . o r g / 2 0 0 2 / 0 7 / owl#&gt;
PREFIX r d f : &lt;h t t p : / / www. w3 . o r g /1999/02/22 r d f s y n t a x ns#&gt;
PREFIX r d f s : &lt;h t t p : / / www. w3 . o r g / 2 0 0 0 / 0 1 / r d f schema#&gt;
s e l e c t ? c o l l e c t ? R e s e a r c h I n s t i t u t i o n ? M e t h o d C o l l e c t
? N a m e D e t e r m i n a t e C o l l e c t where f
? c o l l e c t : m e d i a t i o n I n s t i t u i c a o V i n c u l o ? R e s e a r c h I n s t i t u t i o n .
? c o l l e c t : i s C l a s s i f i e d A s C o l e t a T i p o C o l e t a ? M e t h o d C o l l e c t .
? c o l l e c t : m e d i a t i o n C o l e t a R e s p C o l e t a ? N a m e D e t e r m i n a t e C o l l e c t . g
The following SPARQL query (Listing 3) shows the geographical location of a specimen
collect and other data, such as collect local, geographic space, latitude and longitude.</p>
      <p>Listing 3. SPARQL query returning geographical location of a collect
PREFIX : &lt; h t t p : / / www. owl o n t o l o g i e s . com /
B i o d i v e r s i t y O n t o l o g y F u l l . owl#&gt;
PREFIX owl : &lt; h t t p : / / www. w3 . o r g / 2 0 0 2 / 0 7 / owl#&gt;
PREFIX r d f : &lt; h t t p : / / www. w3 . o r g / 1 9 9 9 / 0 2 / 2 2 r d f s y n t a x n s#&gt;
PREFIX r d f s : &lt; h t t p : / / www. w3 . o r g / 2 0 0 0 / 0 1 / r d f schema#&gt;
s e l e c t ? C o l l e c t L o c a l ? G e o g r a p h i c S p a c e ? l a t i t u d e ? l o n g i t u d e
where f
? C o l l e c t L o c a l : l o c a l i z a t i o n E s p a G e o C o o r d G e o ? G e o g r a p h i c S p a c e .
? G e o g r a p h i c S p a c e : l a t i t u d e ? l a t i t u d e .
? G e o g r a p h i c S p a c e : l o n g i t u d e ? l o n g i t u d e .
g
To evaluate our semantic search architecture, we measured precision and recall to assess
the performance of each approach dependent on input variable such as the user query. The
recall value measures whether a tool retrieves all possible items related to the search terms
contained in the data store, while precision measures to what extent only the relevant items
were actually returned.</p>
      <p>We compared the result in two search systems, our semantic search and keyword
based search from SpeciesLink with data from INPA. We used a total of 16 queries (8 for
each system).</p>
      <p>
        To compare the results of only two systems, we will employ the Students T-tests,
since they are designed for testing two data sets [
        <xref ref-type="bibr" rid="ref2">B. Rasch and Naumann 2004</xref>
        ]. When
checking two data sets, each characterized by its average, standard deviation, and number
of data points, it is possible to apply the T-test to identify, whether the means are in
fact distinct or not. A probability value (p-values) below 0.05 indicates a statistically
significant difference, whereas a p-value equal or exceeding 0.05 indicates no significant
evidence, that there exists no significant difference between the performance values of
two or more tools [
        <xref ref-type="bibr" rid="ref13">Sachs 2003</xref>
        ].
      </p>
      <p>In our experiments, Semantic Search resulted in is significant difference in recall
(p=0.0201 by t-test) and precision (p= 0.0006 by t-test) when compared to Keyword based
search. One reason might be that keyword based search is not enough to capture the
underlying semantics of user information needs, since it is content-oriented. This evaluation
is shown in Table 2.</p>
      <p>There is a significant difference in the mean of precision in Semantic Search
minus the mean precision in Keyword Search equals 0.50416. The confidence interval of
this difference from 0.302420283 to 0.705913042 is 95%. The mean of recall in
Semantic Search minus the mean precision in Keyword Search is equals 0.20624745663. The
confidence interval of this difference from 0.04330054489 to 0.36919436836 is 95%.
Group</p>
      <sec id="sec-6-1">
        <title>Mean Queries</title>
      </sec>
      <sec id="sec-6-2">
        <title>Semantic Search (Recall) 0.587638057 8</title>
      </sec>
    </sec>
    <sec id="sec-7">
      <title>6. Conclusions and Future Work</title>
      <p>
        The architecture presented in this work provides a new document retrieval process by
exploiting query terms to support scientists in the process of discovery and integrating
biodiversity data and domain knowledge. This architecture can be classified, according to
the categorization schema proposed by [
        <xref ref-type="bibr" rid="ref9">Mangold 2007</xref>
        ], as a Stand-Alone Search Engine.
The search process uses resources labels from classes, properties, mappings and instances
from domain ontologies represented in the OWL language.
      </p>
      <p>We defined a mapping mechanism between relational database data and OntoBio
ontology terms resulting in the generation of RDF triples (subject, object and predicate)
saved in a triple store (Virtuoso). The triple stores make it much easier to add new
predicates and write complicated queries or perform inferencing and rule processing.</p>
      <p>A comparative analysis showed a significant increase in recall and precision in
the semantic search. The possibility of creating queries that seek information based on
relationships between data offers many alternatives to semantic search systems, since the
results of these queries are not based only on specific information. Users can thus receive
data that, in traditional systems, would not be considered by the query, but by analyzing
their relations with other information, semantic search queries can consider them relevant.</p>
      <p>As future work, we intend to extend our current implementation with more
advanced structured searches in partnership with researches from INPA.</p>
    </sec>
    <sec id="sec-8">
      <title>7. Acknowledgment</title>
      <p>The authors would like to thank INPA for supporting this work. Thanks are also due to
researchers of INPA’s biological collections. This research was financed by the Brazilian
funding agency CNPq.</p>
    </sec>
  </body>
  <back>
    <ref-list>
      <ref id="ref1">
        <mixed-citation>
          <string-name>
            <surname>Albuquerque</surname>
            ,
            <given-names>A.</given-names>
          </string-name>
          (
          <year>2011</year>
          ). Desenvolvimento de uma Ontologia de DomA˜nio para Modelagem de Biodiversidade.
          <article-title>Disserta A˜x A˜£o de Mestrado. Universidade Federal do Amazonas</article-title>
          .
        </mixed-citation>
      </ref>
      <ref id="ref2">
        <mixed-citation>
          <string-name>
            <given-names>B.</given-names>
            <surname>Rasch</surname>
          </string-name>
          ,
          <string-name>
            <given-names>M.</given-names>
            <surname>Friese</surname>
          </string-name>
          ,
          <string-name>
            <surname>W. H.</surname>
          </string-name>
          <article-title>and</article-title>
          <string-name>
            <surname>Naumann</surname>
            ,
            <given-names>E.</given-names>
          </string-name>
          (
          <year>2004</year>
          ).
          <source>Quantitative Methoden Band. Springer, ISBN 978-3-540-33307-4.</source>
        </mixed-citation>
      </ref>
      <ref id="ref3">
        <mixed-citation>
          <string-name>
            <surname>Baeza-Yates</surname>
            ,
            <given-names>R. A.</given-names>
          </string-name>
          and
          <string-name>
            <surname>Ribeiro-Neto</surname>
            ,
            <given-names>B.</given-names>
          </string-name>
          (
          <year>1999</year>
          ).
          <article-title>Modern Information Retrieval</article-title>
          . AddisonWesley Longman Publishing Co., Inc., Boston, MA, USA.
        </mixed-citation>
      </ref>
      <ref id="ref4">
        <mixed-citation>
          <string-name>
            <surname>Boley</surname>
            ,
            <given-names>H.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Tabet</surname>
            ,
            <given-names>S.</given-names>
          </string-name>
          , and
          <string-name>
            <surname>Wagner</surname>
            ,
            <given-names>G.</given-names>
          </string-name>
          (
          <year>2001</year>
          ).
          <article-title>Design rationale of ruleml: A markup language for semantic web rules</article-title>
          . pages
          <fpage>381</fpage>
          -
          <lpage>401</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref5">
        <mixed-citation>
          <string-name>
            <surname>GeoNames</surname>
          </string-name>
          (
          <year>2011</year>
          ).
          <article-title>Geonames ontology</article-title>
          . http://www.geonames.org/ ontology/documentation.html. Accessed:
          <fpage>2013</fpage>
          -07-30.
        </mixed-citation>
      </ref>
      <ref id="ref6">
        <mixed-citation>
          <string-name>
            <surname>Kauppinen</surname>
          </string-name>
          , T. and
          <string-name>
            <surname>de Espindola</surname>
            ,
            <given-names>G. M.</given-names>
          </string-name>
          (
          <year>2011</year>
          ).
          <article-title>Linked Open Science-communicating, sharing and evaluating data, methods and results for executable papers</article-title>
          .
          <source>Proceedings of the International Conference on Computational Science (ICCS</source>
          <year>2011</year>
          ), Procedia Computer Science,
          <volume>4</volume>
          (
          <issue>0</issue>
          ):
          <fpage>726</fpage>
          -
          <lpage>731</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref7">
        <mixed-citation>
          <string-name>
            <surname>Latiri</surname>
            ,
            <given-names>C. C.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Haddad</surname>
            ,
            <given-names>H.</given-names>
          </string-name>
          , and
          <string-name>
            <surname>Hamrouni</surname>
            ,
            <given-names>T.</given-names>
          </string-name>
          (
          <year>2012</year>
          ).
          <article-title>Towards an effective automatic query expansion process using an association rule mining approach</article-title>
          . J.
          <string-name>
            <surname>Intell</surname>
          </string-name>
          . Inf. Syst.,
          <volume>39</volume>
          (
          <issue>1</issue>
          ):
          <fpage>209</fpage>
          -
          <lpage>247</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref8">
        <mixed-citation>
          <string-name>
            <surname>Li</surname>
            ,
            <given-names>W.</given-names>
          </string-name>
          and
          <string-name>
            <surname>Yang</surname>
            ,
            <given-names>C.</given-names>
          </string-name>
          (
          <year>2008</year>
          ).
          <article-title>A semantic search engine for spatial web portals</article-title>
          . volume
          <volume>2</volume>
          ,
          <string-name>
            <surname>pages</surname>
            <given-names>II</given-names>
          </string-name>
          -1278
          <string-name>
            <surname>-</surname>
          </string-name>
          II-1281.
        </mixed-citation>
      </ref>
      <ref id="ref9">
        <mixed-citation>
          <string-name>
            <surname>Mangold</surname>
            ,
            <given-names>C.</given-names>
          </string-name>
          (
          <year>2007</year>
          ).
          <article-title>A survey and classification of semantic search approaches</article-title>
          .
          <source>Int. J. Metadata Semant. Ontologies</source>
          ,
          <volume>2</volume>
          (
          <issue>1</issue>
          ):
          <fpage>23</fpage>
          -
          <lpage>34</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref10">
        <mixed-citation>
          <string-name>
            <surname>Mariano</surname>
            <given-names>R</given-names>
          </string-name>
          , M. and
          <string-name>
            <surname>Calvanese</surname>
            ,
            <given-names>D.</given-names>
          </string-name>
          (
          <year>2012</year>
          ).
          <article-title>Quest, an OWL 2 QL Reasoner for OntologyBased Data Access</article-title>
          .
          <source>KRDB Research Centre for Knowledge and Data</source>
          , Free University of Bozen-Bolzano, Bolzano, Italy.
        </mixed-citation>
      </ref>
      <ref id="ref11">
        <mixed-citation>
          <string-name>
            <surname>Mittal</surname>
            ,
            <given-names>N.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Nayak</surname>
            ,
            <given-names>R.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Govil</surname>
            ,
            <given-names>M. C.</given-names>
          </string-name>
          , and
          <string-name>
            <surname>Jain</surname>
            ,
            <given-names>K.</given-names>
          </string-name>
          (
          <year>2010</year>
          ).
          <article-title>A hybrid approach of personalized web information retrieval</article-title>
          .
          <source>In Web Intelligence and Intelligent Agent Technology (WI-IAT)</source>
          ,
          <year>2010</year>
          IEEE/WIC/ACM International Conference on, volume
          <volume>1</volume>
          , pages
          <fpage>308</fpage>
          -
          <lpage>313</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref12">
        <mixed-citation>
          <string-name>
            <surname>PATO</surname>
          </string-name>
          (
          <year>2010</year>
          ).
          <article-title>The phenotypic quality ontology</article-title>
          . http://bioportal. bioontology.org/ontologies/1069. Accessed:
          <fpage>2013</fpage>
          -07-30.
        </mixed-citation>
      </ref>
      <ref id="ref13">
        <mixed-citation>
          <string-name>
            <surname>Sachs</surname>
            ,
            <given-names>L.</given-names>
          </string-name>
          (
          <year>2003</year>
          ).
          <source>Angewandte Statistik: Anwendung statistischer Methoden</source>
          . Springer, November. ISBN 3540405550.
        </mixed-citation>
      </ref>
      <ref id="ref14">
        <mixed-citation>
          <string-name>
            <surname>Salton</surname>
          </string-name>
          , G., editor (
          <year>1971</year>
          ).
          <article-title>The SMART Retrieval System - Experiments in Automatic Document Processing</article-title>
          . Prentice Hall, Englewood, Cliffs, New Jersey.
        </mixed-citation>
      </ref>
      <ref id="ref15">
        <mixed-citation>
          <string-name>
            <surname>Santos</surname>
            ,
            <given-names>V. D.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Baiao</surname>
            ,
            <given-names>F. A.</given-names>
          </string-name>
          , and
          <string-name>
            <surname>Tanaka</surname>
            ,
            <given-names>A.</given-names>
          </string-name>
          (
          <year>2011</year>
          ).
          <article-title>An architecture to support information sources discovery through semantic search</article-title>
          .
          <source>In Information Reuse and Integration.</source>
        </mixed-citation>
      </ref>
      <ref id="ref16">
        <mixed-citation>
          <string-name>
            <surname>WGS84</surname>
          </string-name>
          (
          <year>2003</year>
          ). W3C Semantic Web Interest Group:
          <article-title>Basic Geo (WGS84 lat/long</article-title>
          ) Vocabulary.
        </mixed-citation>
      </ref>
      <ref id="ref17">
        <mixed-citation>
          <string-name>
            <surname>Xiong</surname>
            ,
            <given-names>J.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Huang</surname>
            ,
            <given-names>W.</given-names>
          </string-name>
          , and
          <string-name>
            <surname>Jin</surname>
            ,
            <given-names>C.</given-names>
          </string-name>
          (
          <year>2009</year>
          ).
          <article-title>An ontology-based semantic search approach for geosciences</article-title>
          . volume
          <volume>3</volume>
          , pages
          <fpage>87</fpage>
          -
          <lpage>90</lpage>
          .
        </mixed-citation>
      </ref>
    </ref-list>
  </back>
</article>