<!DOCTYPE article PUBLIC "-//NLM//DTD JATS (Z39.96) Journal Archiving and Interchange DTD v1.0 20120330//EN" "JATS-archivearticle1.dtd">
<article xmlns:xlink="http://www.w3.org/1999/xlink">
  <front>
    <journal-meta />
    <article-meta>
      <title-group>
        <article-title>Feature Representation for Cross-Lingual, Cross- Media Semantic Web Applications</article-title>
      </title-group>
      <contrib-group>
        <contrib contrib-type="author">
          <string-name>Paul Buitelaar♦</string-name>
          <email>paulb@dfki.de</email>
        </contrib>
        <contrib contrib-type="author">
          <string-name>Michael Sintek</string-name>
          <email>sintek@dfki.de</email>
        </contrib>
        <contrib contrib-type="author">
          <string-name>Malte Kiesel♣</string-name>
          <email>kiesel@dfki.de</email>
        </contrib>
      </contrib-group>
      <fpage>89</fpage>
      <lpage>94</lpage>
      <abstract>
        <p>Currently, ontology development has been mostly directed at the representation of domain knowledge (i.e., classes, relations and instances) and much less at the representation of corresponding text and image features. To allow for cross-media knowledge markup, a richer representation of features is needed. At present, such information is mostly missing or represented only in a very impoverished way. In this paper we propose an RDF/S-based ontology for the integrated representation of domain knowledge and text/image features.</p>
      </abstract>
    </article-meta>
  </front>
  <body>
    <sec id="sec-1">
      <title>Introduction</title>
      <p>
        Ontologies define the semantics for a set of objects in the world using a set of
classes, each of which may be identified by a particular symbol (either
linguistic, as image, or otherwise). In this way, ontologies cover all three sides of
the “semiotic triangle” that includes object, referent and symbol, i.e., an object
in the world is defined by its referent and represented by a symbol
        <xref ref-type="bibr" rid="ref12">(Ogden and
Richards, 1923 – based on Peirce, de Saussure and others)</xref>
        .
      </p>
      <p>
        Currently, ontology development and the Semantic Web effort in general have
been mostly directed at the referent side of the triangle, and much less at the
symbol side. To allow for cross-media knowledge markup, a richer
representation is needed of these symbols, i.e. of the text and image features for the
object classes that are defined by the ontology. At present, such information is
mostly missing1 or represented only in a very impoverished way, leaving the
semantic information in an ontology without a grounding to the human
cognitive and linguistic domain.
1 According to the collection of ontologies available through OntoSelect
        <xref ref-type="bibr" rid="ref4">(see Buitelaar et al.,
2004)</xref>
        currently only about 9% of ontologies represent multilingual terms for classes and/or
properties (http://views.dfki.de/ontologies/index.php?mode=stats)).
      </p>
    </sec>
    <sec id="sec-2">
      <title>Cross-Lingual,</title>
      <p>Representation</p>
    </sec>
    <sec id="sec-3">
      <title>Cross-Media</title>
    </sec>
    <sec id="sec-4">
      <title>Feature</title>
    </sec>
    <sec id="sec-5">
      <title>Extraction and</title>
      <p>An ontology describes a knowledge model of a particular domain of discourse
at a particular point of time and is shared between two or more actors in the
domain. As the ontology defines the agreed semantics of the domain, all
relevant content will be marked-up with knowledge according to the ontology.
The definition of the ontology in turn depends primarily2 on the content that
has already been interpreted. Accordingly, content production and
interpretation will drive the adaptation of the ontology infrastructure, and ontology
adaptation will drive content interpretation and production. In order to arrive
at such a continuous ‘hermeneutic cycle’ of content and knowledge
production and interpretation, a rich representation of domain knowledge and content
features is needed. Here we propose an integrated approach that organizes
content and knowledge in several layers, as displayed below:</p>
      <sec id="sec-5-1">
        <title>Images</title>
        <p>informal
formal</p>
      </sec>
      <sec id="sec-5-2">
        <title>Other Media</title>
        <p>content
features
feature
associations
ontology
al al
rmrm
o o
f f
in
formal
informal
formal
inforfmormianlformal
a
l</p>
      </sec>
      <sec id="sec-5-3">
        <title>English</title>
      </sec>
      <sec id="sec-5-4">
        <title>Text</title>
        <p>…</p>
      </sec>
      <sec id="sec-5-5">
        <title>German</title>
      </sec>
      <sec id="sec-5-6">
        <title>Text</title>
        <p>The content layer (outermost layer) consists of cross-media data (images,
video and/or mixed image and text documents).</p>
        <p>The features layer (1st inner layer) consists of extracted features for the data in
the content layer. For multilingual data, this ranges from comparatively
informal feature vectors gathered by use of statistical methods to formalized
descriptions of the content of text documents, typically extracted by use of
natural language processing and information extraction methods. For
multimedia data, this will be mostly limited to informal features as used in colour
histograms and similar.
2 Aside from more generic knowledge of the physical world, time, space, etc. that will be
inherited from an upper-level ontology.
The feature association layer (2nd inner layer) consists of feature
representations occurring in the features layer. While in the features layer features are
associated with cross-media data, in the feature association layer the features
are associated with classes in the semantic model.</p>
        <p>The semantic model layer (central layer) consists of classes, with which the
data in the content layer is to be interpreted (i.e., annotated) by use of the
extracted and represented features in the features layer and the feature
association layer.</p>
        <p>This integrated approach allows for cross-lingual, cross-media feature
extraction and representation as follows:
image2text – For instance, if we know which terms express a class in English,
we will be able to build a classifier for the classification of images that occur
in the context of English terms for this class.
text2image – For instance, if we know which images represent instances for a
specific class, we will be able to extract German terms for this class from
surrounding German text.
text2text – For instance, if we know which terms express a class in English,
and we know the context features (i.e. words) for these terms and possible
translations for these words into German, we will be able to build a
crosslingual classifier for recognition of unseen German terms for this class.
image2class or text2class – For instance, if we know which terms express a
class in English, and we know the context words for these terms, we will be
able to detect a change in the semantic model for this class by monitoring any
change in the context words, and similar with image feature models.
3</p>
        <p>
          Towards Ontology-Based Feature Representation
The integrated ontology-based feature representation we propose is based on
ongoing work in the context of the SmartWeb project on mobile Semantic
Web access for intelligent information services in the football domain
(http://www.smartweb-project.de/). To represent terminology for concepts in
different languages we initiated an extension of RDF-based domain
knowledge representation with the meta-class ClassWithFeats.
Although there is some overlap with the SKOS
          <xref ref-type="bibr" rid="ref10 ref5 ref9">(Miles and Brickley, 2005)</xref>
          model for RDF-based thesauri, the proposed representation is richer as it will
include not only multilingual terms for classes (and properties) but also
context models for disambiguating these terms in knowledge markup (i.e., “world
cup” as EVENT or ARTIFACT). More specifically, there is a technical and a
conceptual reason why SKOS5 does not fulfill the needs of our scenario:
SKOS uses sub-properties of rdfs:label (skos:prefLabel,
skos:altLabel) together with xml:lang to attach multilingual terms to
concepts. Furthermore, the RDFS specification
          <xref ref-type="bibr" rid="ref3 ref7">(Brickley and Guha, 2004;
Hayes, 2004)</xref>
          defines the range of rdfs:label to be rdfs:Literal. From
the definition of rds:subPropertyOf follows that the range of
skos:prefLabel and skos:altLabel is also rdfs:Literal (or a
specialization of rdfs:Literal). This is not sufficient in our scenario since we
want to attach more information as linguistic information to classes than
simple multilingual strings. This led to our decision to use the meta-class
ClassWithFeats, which allows us to attach complex information to classes with
the properties lingFeat and imgFeat (in the future, more properties will be
defined for other media types like audio and video).
        </p>
        <p>
          The conceptual problem we see with SKOS for the use in our scenario is that
it mixes linguistic and semantic knowledge. SKOS uses skos:broader and
skos:narrower to express “semantic” relations without clearly stating the
semantics of these relations intentionally, and defines the sub-properties
skos:broaderGeneric and skos:narrowerGeneric to have class
subsumption semantics (i.e., they inherit the rdfs:subClassOf semantics from
RDFS). We clearly keep the linguistic and semantic, ontology-based
knowledge representations apart6: the ontology is represented using the semantic
relations defined in RDFS or OWL (Full)7
          <xref ref-type="bibr" rid="ref8">(McGuinnes and van Harmelen,
2004)</xref>
          , and attach linguistic knowledge to the classes (and properties).
We further propose to integrate image-related features in this representation,
which is beyond the scope of SKOS. Note that SKOS uses
foaf:depiction, skos:prefSymbol, and skos:altSymbol to attach
images to concepts, but not complex feature descriptions.
5 Our argumentation applies to all approaches based on rdfs:label and xml:lang to
attach multilingual labels to classes and relations.
6 Note that our approach in effect integrates a domain-specific multilingual Wordnet into the
ontology, although also the Wordnet model does not distinguish clearly between linguistic
and semantic information
          <xref ref-type="bibr" rid="ref11">(Miller et al., 1995)</xref>
          . Alternative lexicon models that are more
similar to our approach include
          <xref ref-type="bibr" rid="ref2">(Bateman et al., 1995)</xref>
          and
          <xref ref-type="bibr" rid="ref1">(Alexa et al., 2002)</xref>
          , but these
concentrate on the definition of a top ontology for lexicons instead of text/image features for
domain ontology classes and properties as in our case.
7 OWL Lite and OLW DL do not support meta-classes and meta-properties.
        </p>
      </sec>
    </sec>
    <sec id="sec-6">
      <title>Application</title>
      <p>The proposed feature representation is currently used in the SmartWeb
ontology on sport events and related issues. Figure 2 shows the ontology with
example classes and associated linguistic and image features: the ontology
contains the class o:FootballPlayer with subclasses o:Defender and
o:Midfielder. All these classes are instances of the meta-class
feat:ClassWithFeats which allows them to use the feature-association
properties feat:lingFeat and feat:imgFeat. The figure shows the
linguistic features of German terms for the class o:Defender (Abwehrspieler)
and o:Midfielder (Mittelfeldspieler). Note that the decomposition of
Abwehrspieler contains Spieler, therefore implicitly relating the two classes in
the ontology via linguistic features. Furthermore, the figure shows an image
feature representation associated with the class o:Midfielder, stating that
an instance of this class has a ‘human’ shape and certain color and texture
features.</p>
      <p>drdf:type
en URI
egproperty ...</p>
      <p>L
rdfs:Class</p>
      <p>rdfs:subClassOf
feat:ClassWithFeats
feat:ClassWithFeats
o:FootballPlayer</p>
      <p>rdfs:
feat:ClassWithFeats subClassOf feat:ClassWithFeats</p>
      <p>o:Defender o:Midfielder
feat:lingFeat feat:imgFeat
... feat:lingFeat
lf:LingFeat lf:LingFeat
lf:lang “de” lf:lang “de”
lf:term “Abwehrspieler” lf:term “Mittelfeldspieler”
lf:syntacticDecomp lf:syntacticDecomp
lf:Lexeme lf:Lexeme
lf:lexeme “Abwehr” lf:lexeme “Spieler”
lf:partOfSpeech “noun” lf:partOfSpeech “noun”
lf:morphDecomp lf:morphDecomp
lf:hasLemma lf:hasLemma
lf:Lemma lf:Lemma
lf:lemma “Abwehr” lf:lemma “Spieler”
meta-classes
...</p>
      <p>classes
rdfs:Class</p>
      <p>if:ImgFeat
rdfs:Class
lf:LingFeat
if:ImgFeat
if:color “#111111”
if:shape “human”
lf:texture “&amp;keypatchSet_223
lf:Lexeme
lf:lexeme “Mittelfeld”
lf:partOfSpeech “noun”
lf:morphDecomp
lf:hasLemma
lf:Lemma
lf:lemma “Mittelfeld”
instances
This research has been supported in part by the SmartWeb project, which is
funded by the German Ministry of Education and Research under grant 01
IMD01 A. We acknowledge input on the representation of image-related
features by Yannis Avrithis, Eric Gaussier and Yiannis Kompatsiaris.</p>
    </sec>
  </body>
  <back>
    <ref-list>
      <ref id="ref1">
        <mixed-citation>
          <string-name>
            <given-names>M.</given-names>
            <surname>Alexa</surname>
          </string-name>
          ,
          <string-name>
            <given-names>B.</given-names>
            <surname>Kreissig</surname>
          </string-name>
          ,
          <string-name>
            <given-names>M.</given-names>
            <surname>Liepert</surname>
          </string-name>
          ,
          <string-name>
            <given-names>K.</given-names>
            <surname>Reichenberger</surname>
          </string-name>
          ,
          <string-name>
            <given-names>L.</given-names>
            <surname>Rostek</surname>
          </string-name>
          ,
          <string-name>
            <given-names>K.</given-names>
            <surname>Rautmann</surname>
          </string-name>
          , W. ScholzeStubenrecht, S.
          <source>Stoye The Duden Ontology: an Integrated Representation of Lexical and Ontological Information In: Proc. of the OntoLex Workshop</source>
          at LREC, Spain, May
          <year>2002</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref2">
        <mixed-citation>
          <string-name>
            <given-names>J. A.</given-names>
            <surname>Bateman</surname>
          </string-name>
          ,
          <string-name>
            <given-names>R.</given-names>
            <surname>Henschel</surname>
          </string-name>
          and F.
          <source>Rinaldi Generalized Upper Model 2.0: documentation Report of GMD/Institut für Integrierte</source>
          Publikations- und
          <string-name>
            <surname>Informationssysteme</surname>
          </string-name>
          , Darmstadt, Germany,
          <year>1995</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref3">
        <mixed-citation>
          <string-name>
            <given-names>D.</given-names>
            <surname>Brickley</surname>
          </string-name>
          , R.V. Guha, editors.
          <source>RDF Vocabulary Description Language 1</source>
          .0:
          <string-name>
            <given-names>RDF</given-names>
            <surname>Schema</surname>
          </string-name>
          .
          <source>World Wide Web Consortium</source>
          ,
          <year>2004</year>
          . http://www.w3.org/TR/rdf-schema/
        </mixed-citation>
      </ref>
      <ref id="ref4">
        <mixed-citation>
          <string-name>
            <given-names>P.</given-names>
            <surname>Buitelaar</surname>
          </string-name>
          , Th. Eigner,
          <string-name>
            <surname>Th. Declerck OntoSelect: A Dynamic Ontology</surname>
          </string-name>
          <article-title>Library with Support for Ontology Selection In: Proc. of the Demo Session at the International Semantic Web Conference</article-title>
          , Hiroshima, Japan, Nov.
          <year>2004</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref5">
        <mixed-citation>
          <string-name>
            <given-names>P.</given-names>
            <surname>Buitelaar</surname>
          </string-name>
          and
          <string-name>
            <given-names>S.</given-names>
            <surname>Ramaka Unsupervised</surname>
          </string-name>
          Ontology-based
          <source>Semantic Tagging for Knowledge Markup In: Proc. of the Workshop on Learning in Web Search at the International Conference on Machine Learning</source>
          , Bonn, Germany,
          <year>August 2005</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref6">
        <mixed-citation>
          <string-name>
            <surname>Th. Declerck</surname>
            ,
            <given-names>O.</given-names>
          </string-name>
          <string-name>
            <surname>Vela</surname>
            ,
            <given-names>Z.</given-names>
          </string-name>
          <string-name>
            <surname>Gantner</surname>
          </string-name>
          and
          <string-name>
            <surname>D.</surname>
          </string-name>
          Manzano-Macho
          <source>Esperonto Deliverable</source>
          <volume>5</volume>
          .2: Multingualism and
          <string-name>
            <given-names>Ontologies</given-names>
            <surname>Dec</surname>
          </string-name>
          .
          <year>2004</year>
        </mixed-citation>
      </ref>
      <ref id="ref7">
        <mixed-citation>
          <string-name>
            <given-names>P.</given-names>
            <surname>Hayes</surname>
          </string-name>
          , editor.
          <source>RDF Semantics. World Wide Web Consortium</source>
          ,
          <year>2004</year>
          . http://www.w3.org/TR/rdf-mt/
        </mixed-citation>
      </ref>
      <ref id="ref8">
        <mixed-citation>
          <string-name>
            <surname>D.L. McGuinness</surname>
            ,
            <given-names>F.</given-names>
          </string-name>
          van Harmelen, editors.
          <source>OWL Web Ontology Language Overview. W3C Recommendation 10 February</source>
          <year>2004</year>
          . http://www.w3.org/TR/owl-features/
        </mixed-citation>
      </ref>
      <ref id="ref9">
        <mixed-citation>
          <string-name>
            <given-names>A.</given-names>
            <surname>Miles</surname>
          </string-name>
          , D. Brickley, editors.
          <source>SKOS Core Vocabulary Specification. W3C Working Draft 10 May</source>
          <year>2005</year>
          . http://www.w3.org/TR/swbp
          <article-title>-skos-core-spec/</article-title>
        </mixed-citation>
      </ref>
      <ref id="ref10">
        <mixed-citation>
          <string-name>
            <given-names>A.</given-names>
            <surname>Miles</surname>
          </string-name>
          , D. Brickley, editors.
          <source>SKOS Core Guide. W3C Working Draft 10 May</source>
          <year>2005</year>
          . http://www.w3.org/TR/swbp
          <article-title>-skos-core-guide/</article-title>
        </mixed-citation>
      </ref>
      <ref id="ref11">
        <mixed-citation>
          <string-name>
            <given-names>G. A.</given-names>
            <surname>Miller</surname>
          </string-name>
          <string-name>
            <surname>WORDNET</surname>
          </string-name>
          :
          <article-title>A Lexical Database for English</article-title>
          .
          <source>Communications of ACM (11)</source>
          :
          <fpage>39</fpage>
          -
          <lpage>41</lpage>
          ,
          <year>1995</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref12">
        <mixed-citation>
          <string-name>
            <given-names>Ch. K.</given-names>
            <surname>Ogden</surname>
          </string-name>
          and
          <string-name>
            <surname>I. A. Richards</surname>
          </string-name>
          <article-title>The meaning of meaning - A study of the influence of language upon thought and of the science of symbolism</article-title>
          . London: Kegan Paul, Trench, Trubner &amp; Co.,
          <year>1923</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref13">
        <mixed-citation>
          <string-name>
            <given-names>K.</given-names>
            <surname>Petridis</surname>
          </string-name>
          , I. Kompatsiaris,
          <string-name>
            <given-names>M. G.</given-names>
            <surname>Strintzis</surname>
          </string-name>
          ,
          <string-name>
            <given-names>S.</given-names>
            <surname>Bloehdorn</surname>
          </string-name>
          ,
          <string-name>
            <given-names>S.</given-names>
            <surname>Handschuh</surname>
          </string-name>
          ,
          <string-name>
            <given-names>S.</given-names>
            <surname>Staab</surname>
          </string-name>
          and
          <string-name>
            <surname>N.</surname>
          </string-name>
          <article-title>Simou Knowledge Representation for Semantic Multimedia Content Analysis</article-title>
          and
          <source>Reasoning In: Proc. of the European Workshop on the Integration of Knowledge, Semantics and Digital Media Technology, Royal Statistical Society</source>
          , London,
          <fpage>25</fpage>
          -
          <lpage>26</lpage>
          Nov.
          <year>2004</year>
          .
        </mixed-citation>
      </ref>
    </ref-list>
  </back>
</article>