<!DOCTYPE article PUBLIC "-//NLM//DTD JATS (Z39.96) Journal Archiving and Interchange DTD v1.0 20120330//EN" "JATS-archivearticle1.dtd">
<article xmlns:xlink="http://www.w3.org/1999/xlink">
  <front>
    <journal-meta />
    <article-meta>
      <title-group>
        <article-title>Mind the Gap - Requirements for the Combination of Content and Knowledge</article-title>
      </title-group>
      <contrib-group>
        <contrib contrib-type="author">
          <string-name>Tobias B u¨rger</string-name>
          <xref ref-type="aff" rid="aff0">0</xref>
        </contrib>
        <contrib contrib-type="author">
          <string-name>Rupert Westenthaler</string-name>
          <xref ref-type="aff" rid="aff0">0</xref>
        </contrib>
        <aff id="aff0">
          <label>0</label>
          <institution>Tobias Bu ̈rger and Rupert Westenthaler are with Salzburg Research Forschungsgesellschaft mbH</institution>
          ,
          <addr-line>Salzburg</addr-line>
          ,
          <country>Austria. Contact:</country>
        </aff>
      </contrib-group>
      <abstract>
        <p>- In this short paper we report on a semantic model for content and knowledge which distinguishes between three descriptive levels: information relating directly to the resource, to the meta data of the resource, and to the subject matter addressed by the content. This model addresses five fundamental requirements for automation: formality, interoperability, multiple interpretations, contextualisation, and independence of knowledge items from the resource's content.</p>
      </abstract>
    </article-meta>
  </front>
  <body>
    <sec id="sec-1">
      <title>-</title>
      <p>Index Terms— knowledge content objects, intelligent content
models, rich media content.</p>
    </sec>
    <sec id="sec-2">
      <title>I. INTRODUCTION</title>
      <p>
        Semantics – i.e. the interpretation of the content – is
important to make content machine-processable and to enable
the definition of tasks in workflow-environments for
knowledge workers in the content industries. Some of the recent
research projects in the area of semantic (or symbolic) video
annotation try to derive the semantics from the videos’ low
level features or from other available basic metadata. Most
of these approaches are – as also pointed out in [
        <xref ref-type="bibr" rid="ref1">1</xref>
        ] – not
capable of fully exploiting the semantics of multimedia content
because the meaning of the content is not localized just in the
media that is being analysed. The construction of meaning is
– for humans – an act of interpretation that has much more
to do with pre-existing knowledge and the context of the user
and/or the media than with recognition of the contents’
lowlevel-features. This is known as the semantic gap [
        <xref ref-type="bibr" rid="ref2">2</xref>
        ]. Popular
examples on the Web show that there are currently many
service-based platforms (like Flickr1 or LastFM2) that make
use of their users’ knowledge to understand the meaning of
multimedia content.
      </p>
      <p>We suggest that representing richer semantics for
multimedia content requires more expressive and more sophisticated
knowledge content models than those currently used. We
therefore introduce a model for representing knowledge and
content alongside each other, with clear separation of the
content and the knowledge items, so as to obtain optimal
conditions for content and knowledge reuse, and for subsequent
re-contextualization of content.</p>
    </sec>
    <sec id="sec-3">
      <title>II. RELATED WORK</title>
      <p>
        For a long time, combining knowledge and content did not
play a great role in the research communities: On the one hand
there were metadata models for content like MPEG-7 [
        <xref ref-type="bibr" rid="ref3">3</xref>
        ], and
on the other hand there were domain ontologies developed by
the Semantic Web community. However, in the past few years
much work has been done on the specification of ontologies
that aim to combine traditional multimedia description models
[
        <xref ref-type="bibr" rid="ref4">4</xref>
        ], [
        <xref ref-type="bibr" rid="ref1">1</xref>
        ]. Related approaches include work from two different
research communities: First there are traditional content
models like MPEG-7 or MPEG-21 [
        <xref ref-type="bibr" rid="ref5">5</xref>
        ] which are coming from the
multimedia community. Besides these traditional approaches
some efforts in modeling of intelligent content objects exist:
More recent efforts include amongst others the ACEMEDIA3
ACE-objects [
        <xref ref-type="bibr" rid="ref6">6</xref>
        ] or the knowledge content objects (KCOs) of
the METOKIS project4.
      </p>
    </sec>
    <sec id="sec-4">
      <title>III. REQUIREMENTS FOR KNOWLEDGE CONTENT</title>
      <p>
        Possible relations between knowledge and content are
manifold. Based on observed applications and requirements of
current projects (see [
        <xref ref-type="bibr" rid="ref7">7</xref>
        ] for details) we have derived the
following requirements for knowledge content:
1) Knowledge must be encoded using a formal language
2) Interoperability especially for cross domain aspects
3) Different interpretations of content objects
4) Link content with knowledge that cannot be directly
derived from the content
5) Make knowledge independent of content
      </p>
    </sec>
    <sec id="sec-5">
      <title>IV. KCO – A MODEL FOR KNOWLEDGE CONTENT</title>
      <p>KCOs are based on the DOLCE foundational ontology5 and
have so-called semantic facets that form modular entities to
describe the properties of KCOs, including the raw content
object or media file, metadata and knowledge specific to the
content object and knowledge about the topics of the content
(its meaning).</p>
      <p>Knowledge is represented by the structure of the KCO in
three different levels:
1) Resource Level: This level refers to the actual content
object (File, stream, image, etc).
2) Meta Level: This level refers to knowledge describing
features of the content object, eg. frame rate,
compression type or colour coding scheme.
3) Subject Matter Level: This level comprises knowledge
about the topic (subject) of the content as interpreted by
an actor. The content object realizes this interpretation.
3http://www.acemedia.org
4http://metokis.salzburgresearch.at
5http://www.loa-cnr.it/DOLCE.html</p>
      <p>
        In addition to this knowledge structure the KCO also defines
a structure based on the different domains of the knowledge
objects. This structure is divided into six so-called facets, each
of them optimized for a specific usage. Facets include for
example a content- or community description facet [
        <xref ref-type="bibr" rid="ref8">8</xref>
        ].
      </p>
      <p>Relating the above description to the requirements from
section III on knowledge content, we suggest that KCOs
provide a good foundation for modeling combinations of
content and knowledge.</p>
      <p>1) Knowledge must be encoded using a formal
language: KCOs are based on the information objects
design pattern, which is an extension of the DOLCE
foundational ontology. The main concepts and relations
used in the description of the KCO are well grounded
on this foundational framework. The current definition
of KCOs is based on OWL-DL6.
2) Interoperability especially for cross domain aspects:
The facet based structure of the KCO serves as a good
starting point for the alignment of standards to the KCO
structure. Some of the facets and elements there use parts
of different standards like NewsML7 or MPEG-7.
3) Different interpretations of content objects: The
possibility of different interpretations of content objects is
the main reason for distinguishing between the mesta
level and the subject matter level. Thus the KCO model
support multiple interpretations of one content object.
4) Clear definition of possible relations between content
and knowledge: The KCO defines two different
interrelations between content and knowledge: First,
knowledge objects can be about a content object, meaning
that the subject matter of the knowledge object is the
content object itself. Second, knowledge objects can be
realized by content objects. This relation is used for all
knowledge objects which are about the subject matter of
the content object.
5) Make knowledge independent of content: This
requirement is modeled by the fact that knowledge objects
which belong to the subject matter level are about an
arbitrary topic and only realized by the actual content.</p>
    </sec>
    <sec id="sec-6">
      <title>V. BRIDGING THE GAP</title>
      <p>In this section we will shortly sketch how the ideas of KCOs
can be applied to address the various conceptual relationships
between content and knowledge in media-rich systems.</p>
      <p>a) Search based on meta data: KCOs can be used to
model complex queries: In mental models that are representing
users intentions, subject matter level information about the
topic of the content is often mixed up with meta level
information about the content objects: By posing queries to
the system, users create descriptions about topics or subjects
that they are interested in.</p>
      <p>b) Collaborative filtering: A query in that setting
specifies the actual context of the actor by considering some of
the concepts and relations within the knowledge space of
the actor, which are typically stored in the form of user
profiles containing additional knowledge about preferences of
the users. This information can be used to further contextualize
queries by combining the context specified by the query, with
the characteristics of the user profile. Based on the active
concepts and relations of the contextualized query, the system
can find similar interpretations. Such a system is sensitive to
different interpretations of one and the same content object,
because it handles different interpretations of different users
separately.</p>
      <p>c) Context-based content classification to minimise the
semantic gap: This scenario refers to the problem of how
to overcome the gap between low level features and higher
level semantics. It assumes that new content objects typically
have to be analysed at the expense of some time and effort.
The complextity of this operation can be reduced based on
background information about the content or some predefined
knowledge of parts of the content as knowledge about one part
can help to understand other parts of the content.</p>
    </sec>
    <sec id="sec-7">
      <title>VI. CONCLUSIONS AND FUTURE WORK</title>
      <p>
        More details about the reported work can be found in [
        <xref ref-type="bibr" rid="ref7">7</xref>
        ].
We are currently trying to apply KCOs in several national and
international projects. Amongst them is the recently started
IST project LIVE8, in which we are responsible for the
definition of an intelligent media framework to support broadcasters
in the live staging of media events.
      </p>
    </sec>
    <sec id="sec-8">
      <title>VII. ACKNOWLEDGEMENTS The work reported here was part-funded by the EU projects METOKIS (contract number IST-FP6-507164) and LIVE (contract number IST-FP6-27312).</title>
    </sec>
  </body>
  <back>
    <ref-list>
      <ref id="ref1">
        <mixed-citation>
          [1]
          <string-name>
            <surname>Bloehdorn</surname>
            ,
            <given-names>S.</given-names>
          </string-name>
          et al.:
          <source>Semantic Annotation of Images and Videos for Multimedia Analysis In: Proceedings of the 2nd European Semantic Web Conference, ESWC</source>
          <year>2005</year>
          , Heraklion, Greece, May
          <year>2005</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref2">
        <mixed-citation>
          [2]
          <string-name>
            <surname>Smeulders</surname>
            ,
            <given-names>A. W. M.</given-names>
          </string-name>
          et al.:
          <article-title>Content-Based Image Retrieval at the End of the Early Years In: IEEE Transactions on Pattern Analysis</article-title>
          and
          <source>Machine Intelligence</source>
          Vol.
          <volume>22</volume>
          No.
          <issue>12</issue>
          ,
          <year>December 2000</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref3">
        <mixed-citation>
          [3]
          <string-name>
            <surname>Martinez-Sanchez</surname>
            ,
            <given-names>J. M.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Koenen</surname>
            ,
            <given-names>R.</given-names>
          </string-name>
          , and
          <string-name>
            <surname>Pereira</surname>
            ,
            <given-names>F.</given-names>
          </string-name>
          :
          <article-title>Mpeg-7: The generic multimedia content description standard, part 1</article-title>
          . IEEE MultiMedia,
          <volume>9</volume>
          (
          <issue>2</issue>
          ): pp.
          <fpage>78</fpage>
          -
          <lpage>87</lpage>
          ,
          <year>2002</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref4">
        <mixed-citation>
          [4]
          <string-name>
            <surname>Hunter</surname>
            ,
            <given-names>J.</given-names>
          </string-name>
          : Adding Multimedia to the
          <source>Semantic Web - Building an MPEG-7 Ontology, Proc. of the 1st International Semantic Web Working Symposium (SWWS</source>
          <year>2001</year>
          ), Stanford, USA,
          <year>2001</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref5">
        <mixed-citation>
          [5]
          <string-name>
            <surname>Bormans</surname>
            ,
            <given-names>J.</given-names>
          </string-name>
          and
          <string-name>
            <surname>Hill</surname>
            ,
            <given-names>K.</given-names>
          </string-name>
          : Mpeg-21 overview v.
          <volume>5</volume>
          . http://www.chiariglione.org/mpeg/standards/mpeg-21/mpeg-21.htm,
          <year>2002</year>
          . ISO/IEC JTC1/SC29/WG11. (last visited:
          <volume>06</volume>
          .
          <fpage>10</fpage>
          .
          <year>2006</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref6">
        <mixed-citation>
          [6]
          <string-name>
            <surname>Kompatsiaris</surname>
            ,
            <given-names>I.</given-names>
          </string-name>
          et al.:
          <article-title>Integrating knowledge, semantics and content for user-centred intelligent media services: the acemedia project</article-title>
          .
          <source>In: Proc. of WIAMIS</source>
          <year>2004</year>
          , Lisboa, Portugal, 23
          <string-name>
            <surname>April</surname>
          </string-name>
          <year>2004</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref7">
        <mixed-citation>
          [7] Bu¨rger, T. and
          <string-name>
            <surname>Westenthaler</surname>
          </string-name>
          , R.:
          <article-title>Why the combination of content and knowledge matters</article-title>
          ,
          <source>Salzburg Research Technical Report</source>
          ,
          <year>2006</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref8">
        <mixed-citation>
          [8]
          <string-name>
            <surname>Behrendt</surname>
            ,
            <given-names>W.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Gangemi</surname>
            ,
            <given-names>A.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Maass</surname>
            ,
            <given-names>W.</given-names>
          </string-name>
          , and
          <string-name>
            <surname>Westenthaler</surname>
          </string-name>
          , R.:
          <article-title>Towards an Ontology-Based Distributed Architecture for Paid Content In: A. Gomez-Perez and</article-title>
          J.
          <string-name>
            <surname>Euzenat</surname>
          </string-name>
          (Eds.) :
          <source>ESWC</source>
          <year>2005</year>
          , LNCS 3532, pp.
          <fpage>257</fpage>
          -
          <lpage>271</lpage>
          ,
          <year>2005</year>
          .
        </mixed-citation>
      </ref>
    </ref-list>
  </back>
</article>