<!DOCTYPE article PUBLIC "-//NLM//DTD JATS (Z39.96) Journal Archiving and Interchange DTD v1.0 20120330//EN" "JATS-archivearticle1.dtd">
<article xmlns:xlink="http://www.w3.org/1999/xlink">
  <front>
    <journal-meta />
    <article-meta>
      <article-id pub-id-type="urn">nbn:de:0074-596-3</article-id>
      <title-group>
        <article-title>ORES-2010 Ontology Repositories and Editors for the Semantic Web</article-title>
      </title-group>
      <contrib-group>
        <contrib contrib-type="author">
          <string-name>Proceedings of the</string-name>
          <xref ref-type="aff" rid="aff0">0</xref>
          <xref ref-type="aff" rid="aff1">1</xref>
        </contrib>
        <contrib contrib-type="author">
          <string-name>st Workshop on Ontology Repositories</string-name>
          <xref ref-type="aff" rid="aff0">0</xref>
          <xref ref-type="aff" rid="aff1">1</xref>
        </contrib>
        <contrib contrib-type="author">
          <string-name>Editors for the Semantic Web</string-name>
          <xref ref-type="aff" rid="aff0">0</xref>
          <xref ref-type="aff" rid="aff1">1</xref>
        </contrib>
        <aff id="aff0">
          <label>0</label>
          <institution>Hersonissos</institution>
          ,
          <addr-line>Crete</addr-line>
          ,
          <country country="GR">Greece</country>
        </aff>
        <aff id="aff1">
          <label>1</label>
          <institution>Mathieu d'Aquin, The Open University, UK Alexander García Castro, Universität Bremen, Germany Christoph Lange, Jacobs University Bremen, Germany Kim Viljanen, Aalto University</institution>
          ,
          <addr-line>Helsinki</addr-line>
          ,
          <country country="FI">Finland</country>
        </aff>
      </contrib-group>
      <volume>596</volume>
      <abstract>
        <p>pCaoppeyrrsightby© 2th0e10 fpoarpethrse' inaduivthidoursa.l Copying permitted only for private and aepdcuaibtdloisershm.eidc paunrdposecos.pyTrihgihstedvolubmye itiss</p>
      </abstract>
    </article-meta>
  </front>
  <body>
    <sec id="sec-1">
      <title>-</title>
      <p>10-Jun-2010: submitted by Christoph Lange
11-Jun-2010: published on CEUR-WS.org</p>
      <p>Ontology Recommendation for the Data</p>
      <p>Publishers?</p>
      <p>Antoine Zimmermann1</p>
      <p>Digital Enterprise Research Institute
National University of Ireland, Galway, Ireland</p>
      <p>firstname.lastname@deri.org
1</p>
    </sec>
    <sec id="sec-2">
      <title>Introduction</title>
      <p>RDF data publishing on the Web has gathered momentum in the last few years,
thanks to a general e ort to link open data all over the Web. Yet, this trend
is slowed down by the di culty to nd appropriate terms for the data to be
described. Indeed, apart from a handful of well known ontologies, there are no
readily available, easily ndable vocabularies for most of the domains that would
be good candidates for publishing data in a standard, linkable way.</p>
      <p>The typical problems faced by would-be data publishers are: (1) ontologies
de ning the domain of interest do not exist; (2) they exist but are di cult to nd
because developed by small groups for experimentation, lacking advertisement;
(3) they exist and can be found but they are of poor quality, not complying with
standards or best practices; (4) they exist and can be found but there are too
many, of mixed quality, and it is di cult to assess which ones are appropriate
for a speci c use case.</p>
      <p>To address the rst issue, user-friendly ontology editors have been developed
Swoop1, Protege2, etc. But this mostly requires that ontology and domain
experts publish more and more data and terminologies. We assume that this will
naturally happen when Linked Data and Semantic Web technologies will reach a
critical mass which trigger a virtuous circle. We will not address this issue here.</p>
      <p>
        The second item is somehow addressed by Semantic Web search engines.
Several of them have been proposed, such as Swoogle [
        <xref ref-type="bibr" rid="ref1">1</xref>
        ], Sindice [
        <xref ref-type="bibr" rid="ref2">2</xref>
        ], SWSE [
        <xref ref-type="bibr" rid="ref3">3</xref>
        ],
? I would like to thank the Pedantic Web Group (http://www.pedantic-web.org/)
for their useful discussions, and more particularly Stephane Corlosquet, Richard
Cyganiak, Renaud Delbru, Alexandre Passant and Axel Polleres. This work is partly
funded by Science Foundation Ireland (SFI) project Lion-2 (SFI/08/CE/I1380).
1 http://code.google.com/p/swoop/
2 http://protege.stanford.edu/
FalconS [
        <xref ref-type="bibr" rid="ref4">4</xref>
        ], Watson [
        <xref ref-type="bibr" rid="ref5">5</xref>
        ], OntoSearch [
        <xref ref-type="bibr" rid="ref6">6</xref>
        ], Ontosearch 2 [
        <xref ref-type="bibr" rid="ref7">7</xref>
        ]. Besides, ontologies can
be gathered together in repositories [
        <xref ref-type="bibr" rid="ref8 ref9">8, 9</xref>
        ] that provide additional functionalities
for maintaining them. This can address to some extent the third and fourth items
because voluntarily submitted ontologies are more likely to be considered by their
authors as being of su cient quality rather than ontologies randomly retrieved
from the Web. Moreover, the implemented functionalities may help correcting
possible errors and eventually would only display formally valid terminologies. To
address the last issue, it has been proposed to add, e.g., Web 2.0-like rating and
voting functions to search engines and repositories (e.g., Revyu [
        <xref ref-type="bibr" rid="ref10">10</xref>
        ]). However,
reviewing ontologies needs advanced knowledge in elds that the ontology users
may not have, and the ontology experts are not necessarily inclined to judge
ontologies in the same way as social website assess conversations, products, etc.
      </p>
      <p>Therefore, we believe that to guarantee access to the ontologies that are
available, well supported and of quality, there is a need for promoting them more
actively. In this paper, we argue in favour of having a committee of experts
analyse Web vocabularies with respect to their suitability as reusable terminologies
for data publishers. As a result of this analysis and evaluation|which can be
partly automatised|the committee would advertise the ontology as a \quality
vocabulary" and recommend it for describing information in the eld applicable
to such terminology.</p>
      <p>To present this idea in more details, we rst discuss the criteria that a Web
terminology should ful l to be labelled as \quality vocabulary" (Section 2).
Then, we describe a possible approach to implement such an evaluation and
recommendation framework (Section 3). Finally, we show how this integrates
with ontology repositories (Section 4).
2</p>
    </sec>
    <sec id="sec-3">
      <title>Criteria for a recommended</title>
    </sec>
    <sec id="sec-4">
      <title>Web vocabulary</title>
      <p>Our objective in recommending Web vocabularies is focused on helping data
publishers to nd adequate terms for describing their data. For this reason our
proposed initiative distinguishes itself from other ontology evaluation activities
that focus more on engineering, design and logical issues. Moreover, we do not
pretend to assess the quality of the modelling of the domain of interest, which
could only be judged by a domain expert. Also, we encourage small, lightweight
ontologies, which are easier to assess, reuse and scale. In this section we discuss
possible criteria for quality vocabularies, recommended for reuse over the Web
of Data. More precisely, data publishers would expect vocabularies that are: (1)
justi ed by use cases; (2) easy to reuse and publish; (3) well interoperable with
published Linked Data. We analyse these requirements to determine the criteria
for quality vocabularies.</p>
      <p>Justifying the existence of the vocabulary. As a primary requirement, a
vocabulary should be accompanied by a statement about its utility. This
includes a general description of the vocabulary and its scope as well as, more
importantly, its related use cases. To avoid too much subjectivity in deciding
the relevance of a vocabulary in terms of usage, it can be required that a Web
vocabulary proposed for recommendation should be supported by at least some
data publishers. We consider this criteria a very important one and would not
recommend a vocabulary, be it very well designed, if nobody considers using it.
Usage should not be restricted to a unique dataset, not even to a big one by a
major player in the eld. At least two independent parties should be using the
terms, or there should be strong evidence that the terms will be used by
several distinct publishers in the near future. As an alternative proof of relevance,
the vocabulary publisher could claim potential adoption by showing precise
examples of possible usage. E.g., a geo-location vocabulary can be proven to be
useful if the authors show that translating existing geographic databases into
linked data can be done in an easy and straightforward way to create and
publish multiple datasets at a low cost. Finally, the utility of the vocabulary can be
demonstrated if existing applications are usefully exploiting the associated data.
This can justify the recommendation of a vocabulary since all data conforming
to it will interoperate with those existing applications.</p>
      <p>
        Ease of reuse and publication. Since one of the goal of recommending
vocabularies is to increase interoperability by reducing the number of heterogeneous
terminologies, it is important that the vocabulary be reusable as easily as
possible. To achieve this, the content should be made understandable by non-ontology
experts. Thus, one of the criteria is the presence of clear labels and textual
descriptions for each term in the ontology. Moreover, granularity and complexity
should be in line with the use cases. Highly expressive or too speci c
ontologies should be discouraged if they are not seriously justi ed. Finally, publication
of data conforming to the proposed vocabulary should be made easier, e.g., by
providing tools that automatise (partly or totally) the publishing process. For
instance, FOAF and SIOC exporters make the creation of online community RDF
metadata fully automatic when integrated in a content management system.
Interoperability. To ensure better interoperability, several guidelines have to
be followed. Obviously, vocabularies should be published in a standard format,
namely RDF(S) and OWL. Additionally, since vocabularies are themselves part
of the Linked Data, they should follow the best practices in the eld [
        <xref ref-type="bibr" rid="ref11">11</xref>
        ]. This
includes, e.g., URI dereferencability, entity naming schemes, or
authoritativeness of term de nition. A term de nition is considered \authoritative" if it
describes an entity which is in the namespace of the document describing it.
Most of these practices can be checked automatically, using, e.g., RDF:Alerts3.
Moreover, vocabularies are also used to reason about the data, so special care
must be taken with this respect. It is desirable to enable interoperability of
the vocabulary with both OWL tools and RDFS tools. On the one hand, an
OWL ontology can be made more RDFS-friendly by de ning all classes as both
owl:Class and rdfs:Class. Similarly, properties de ned as owl:ObjectProperty,
owl:DatatypeProperty and owl:AnnotationProperty should also be de ned as
3 http://swse.deri.org/RDFAlerts/
rdf:Property. On the other hand, RDFS terminologies should declare each term
as either of the aforementioned types, unless a strong justi cation comes from
the use cases. For instance, the Dublin Core vocabulary does not specify the type
of its properties in order to preserve exibility. Also, an OWL ontology should be
kept compatible with OWL DL as much as possible, and any exception should be
justi ed. More generally, vocabularies should use the least expressive fragment
of OWL that ful ls the desired requirements.
3
      </p>
    </sec>
    <sec id="sec-5">
      <title>Implementing the recommendation process</title>
      <p>This section shows how a framework for recommending Web vocabularies could
be implemented in practice.</p>
      <p>We notice that most ontologies and Web vocabularies, especially the most
popular ones, are developed by academic researchers (FOAF, SIOC, Good
Relations, Music Ontology, etc.) Thus, it would be possible to incite the ontology
builders to publish their creation through our quality assessment process by
establishing a regular submit/review/accept-reject process. An ontology or
terminology for data publishing o ers a solution to a problem or ful l a certain
need. It can thus be seen as a scienti c contribution. We propose to have a call
for vocabularies with a review process and editorial constraints.</p>
      <p>First, an automatic tool will verify the compliance of the submitted
vocabularies to the well established criteria mentioned in Section 2. If these are matched,
then a peer-reviewing process will be undertaken by the committee, considering
what has been discussed in the previous section. A submitted ontology must be
accompanied by a description that will be published together with the ontology.
The description|which can take the form of an article|must explains the
purpose of the ontology|not only its domain but also its scope and granularity as
well as possible or existing applications using it. It must show the utility and the
need for such a vocabulary. This is partly proven by the fact that existing
(independent) datasets are already using the terms or publishers have committed
to use it in the near future.</p>
      <p>The descriptions should be kept understandable by non-ontology specialists,
and technical details can be given if, and only if, it contributes to showing the
utility and interoperability of the vocabulary. As a result of this process, the
ontology is either deemed not suitable for recommendation or recommended as
a \quality vocabulary". Rejected ontologies can be improved and resubmitted
later, eventually leading to better quality of the vocabularies. A centralised Web
site would advertise these ontologies, provide documentation about them and
allow searching and browsing them. Such a central place could be an existing
ontology repository, as discussed in the next section.
4</p>
    </sec>
    <sec id="sec-6">
      <title>Integrating recommendations in an ontology repository</title>
      <p>The recommended ontologies should be easily searchable, browsable as well as
matchable. These are common tasks made possible by ontology repositories.
Therefore, the recommendation process that we described previously may be
integrated into a repository that could possibly include other non-recommended
ontologies. However, the recommended ones should be emphasised and all
operation should be applicable to the quality ontologies only. Moreover, non-recommended
vocabularies could be marked by speci c labels indicating that, although they
are not recommended, they validate some of the criteria for quality.
5</p>
    </sec>
    <sec id="sec-7">
      <title>Conclusion</title>
      <p>In this paper, we argued that data publishers should be guided in their choice
of vocabularies by actively recommending them ontologies that are considered
of quality. The recommendation should be based on criteria that span from
support by existing publishers and applications to the compliance to the best
practices in this eld. Evaluating those criteria could take the same form as an
academic publication process, with a review phase. As a rst step, we would like
to launch a workshop on this topic which would create, we hope, an incentive
for researchers to design and publish quality ontologies for the Web of Data.</p>
    </sec>
  </body>
  <back>
    <ref-list>
      <ref id="ref1">
        <mixed-citation>
          1.
          <string-name>
            <surname>Finin</surname>
            ,
            <given-names>T.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Ding</surname>
            ,
            <given-names>Z.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Pan</surname>
            ,
            <given-names>R.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Joshi</surname>
            ,
            <given-names>A.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Kolari</surname>
            ,
            <given-names>P.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Java</surname>
            ,
            <given-names>A.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Peng</surname>
            ,
            <given-names>Y.</given-names>
          </string-name>
          :
          <article-title>Swoogle: Searching for Knowledge on the Semantic Web</article-title>
          .
          <source>In: Proc. of AAAI</source>
          <year>2005</year>
          , AAAI Press / The MIT Press (
          <year>July 2005</year>
          )
          <volume>1682</volume>
          {
          <fpage>1683</fpage>
        </mixed-citation>
      </ref>
      <ref id="ref2">
        <mixed-citation>
          2.
          <string-name>
            <surname>Oren</surname>
            ,
            <given-names>E.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Delbru</surname>
            ,
            <given-names>R.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Catasta</surname>
            ,
            <given-names>M.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Cyganiak</surname>
            ,
            <given-names>R.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Stenzhorn</surname>
            ,
            <given-names>H.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Tummarello</surname>
          </string-name>
          , G.:
          <article-title>Sindice.com: a document-oriented lookup index for open linked data</article-title>
          .
          <source>International Journal of Metadata, Semantics and Ontologies</source>
          <volume>3</volume>
          (
          <issue>1</issue>
          ) (
          <year>2008</year>
          )
          <volume>37</volume>
          {
          <fpage>52</fpage>
        </mixed-citation>
      </ref>
      <ref id="ref3">
        <mixed-citation>
          3.
          <string-name>
            <surname>Harth</surname>
            ,
            <given-names>A.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Hogan</surname>
            ,
            <given-names>A.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Umbrich</surname>
            ,
            <given-names>J.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Decker</surname>
            ,
            <given-names>S.</given-names>
          </string-name>
          : SWSE:
          <article-title>Objects before documents!</article-title>
          <source>In: Proc. of the Billion Triple Semantic Web Challenge</source>
          . (
          <year>2008</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref4">
        <mixed-citation>
          4. Cheng, G.,
          <string-name>
            <surname>Ge</surname>
            ,
            <given-names>W.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Qu</surname>
          </string-name>
          , Y.:
          <article-title>FalconS: Searching and Browsing Entities on the Semantic Web</article-title>
          .
          <source>In: Proc. of WWW</source>
          <year>2008</year>
          , ACM Press (
          <year>April 2008</year>
          )
          <volume>1101</volume>
          {
          <fpage>1102</fpage>
        </mixed-citation>
      </ref>
      <ref id="ref5">
        <mixed-citation>
          5.
          <string-name>
            <surname>d'Aquin</surname>
            ,
            <given-names>M.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Sabou</surname>
            ,
            <given-names>M.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Dzbor</surname>
            ,
            <given-names>M.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Baldassare</surname>
            ,
            <given-names>C.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Gridinoc</surname>
            ,
            <given-names>L.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Angeletou</surname>
            ,
            <given-names>S.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Motta</surname>
          </string-name>
          , E.:
          <article-title>WATSON: A Gateway for the Semantic Web</article-title>
          . In:
          <article-title>Poster session of the European Semantic Web Conference</article-title>
          , ESWC. (
          <year>2007</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref6">
        <mixed-citation>
          6. Zhang,
          <string-name>
            <given-names>Y.</given-names>
            ,
            <surname>Vasconcelos</surname>
          </string-name>
          ,
          <string-name>
            <given-names>W.</given-names>
            ,
            <surname>Sleeman</surname>
          </string-name>
          ,
          <string-name>
            <surname>D.:</surname>
          </string-name>
          <article-title>OntoSearch: An Ontology Search Engine</article-title>
          .
          <source>In: Proc. of AI-2004. BCS Conference Series</source>
          , Springer (
          <year>2004</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref7">
        <mixed-citation>
          7.
          <string-name>
            <surname>Thomas</surname>
            ,
            <given-names>E.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Pan</surname>
            ,
            <given-names>J.Z.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Sleeman</surname>
            ,
            <given-names>D.:</given-names>
          </string-name>
          <article-title>ONTOSEARCH2: Searching Ontologies Semantically</article-title>
          .
          <source>In: Proc. of OWLED 2007. Volume 258 of CEUR Workshop Proceedings</source>
          .,
          <source>Sun SITE Central Europe (CEUR)</source>
          (
          <year>June 2007</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref8">
        <mixed-citation>
          8.
          <string-name>
            <surname>Pan</surname>
            ,
            <given-names>J.</given-names>
          </string-name>
          , Crane eld, S.,
          <string-name>
            <surname>Carter</surname>
            ,
            <given-names>D.:</given-names>
          </string-name>
          <article-title>A lightweight ontology repository</article-title>
          .
          <source>In: Proc. of AAMAS</source>
          <year>2003</year>
          , ACM Press (
          <year>July 2003</year>
          )
          <volume>632</volume>
          {
          <fpage>638</fpage>
        </mixed-citation>
      </ref>
      <ref id="ref9">
        <mixed-citation>
          9.
          <string-name>
            <surname>d'Aquin</surname>
            ,
            <given-names>M.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Lewen</surname>
          </string-name>
          , H.:
          <article-title>Cupboard - A Place to Expose Your Ontologies to Applications and the Community</article-title>
          .
          <source>In: Proc. of ESWC 2009</source>
          . Volume
          <volume>5554</volume>
          ., Springer (
          <year>June 2009</year>
          )
          <volume>913</volume>
          {
          <fpage>918</fpage>
        </mixed-citation>
      </ref>
      <ref id="ref10">
        <mixed-citation>
          10.
          <string-name>
            <surname>Heath</surname>
            ,
            <given-names>T.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Motta</surname>
          </string-name>
          , E.: Revyu:
          <article-title>Linking reviews and ratings into the Web of Data</article-title>
          .
          <source>Journal of Web Semantics</source>
          <volume>6</volume>
          (
          <issue>4</issue>
          ) (
          <year>2008</year>
          )
          <volume>266</volume>
          {
          <fpage>273</fpage>
        </mixed-citation>
      </ref>
      <ref id="ref11">
        <mixed-citation>
          11.
          <string-name>
            <surname>Bizer</surname>
            ,
            <given-names>C.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Cyganiak</surname>
            ,
            <given-names>R.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Heath</surname>
            ,
            <given-names>T.</given-names>
          </string-name>
          :
          <article-title>How to Publish Linked Data on the Web</article-title>
          .
          <source>web published (July</source>
          <year>2007</year>
          )
        </mixed-citation>
      </ref>
    </ref-list>
  </back>
</article>