<!DOCTYPE article PUBLIC "-//NLM//DTD JATS (Z39.96) Journal Archiving and Interchange DTD v1.0 20120330//EN" "JATS-archivearticle1.dtd">
<article xmlns:xlink="http://www.w3.org/1999/xlink">
  <front>
    <journal-meta />
    <article-meta>
      <title-group>
        <article-title>LinkedCulture: browsing related Europeana objects while watching a cultural heritage TV programme</article-title>
      </title-group>
      <contrib-group>
        <contrib contrib-type="author">
          <string-name>Lyndon Nixon</string-name>
          <email>lyndon.nixon@ modul.ac.at</email>
          <xref ref-type="aff" rid="aff0">0</xref>
        </contrib>
        <contrib contrib-type="author">
          <string-name>Lotte Belice Baltussen, Johan Oomen</string-name>
          <email>lbbaltussen,joomen@ beeldengeluid.nl</email>
          <xref ref-type="aff" rid="aff1">1</xref>
        </contrib>
        <aff id="aff0">
          <label>0</label>
          <institution>MODUL University</institution>
          ,
          <addr-line>Am Kahlenberg 1, 1190 Vienna</addr-line>
          ,
          <country country="AT">Austria</country>
        </aff>
        <aff id="aff1">
          <label>1</label>
          <institution>Sound and Vision</institution>
          ,
          <addr-line>Sumatralaan 45, Hilversum</addr-line>
          ,
          <country country="NL">Netherlands</country>
        </aff>
      </contrib-group>
      <abstract>
        <p>This short/demo paper describes LinkedCulture, a Web based application which complements the viewing of a well known Dutch cultural heritage TV program with the ability of viewers to explore art objects from Europeana related to those in the program.</p>
      </abstract>
    </article-meta>
  </front>
  <body>
    <sec id="sec-1">
      <title>-</title>
      <p>
        TRENDS AND RELATED WORK
Households have more and more connected devices and
consumers are increasingly using devices in parallel: this is
clearest when it comes to viewing audiovisual
programming on one device and browsing online content on
another. A 2013 published survey [
        <xref ref-type="bibr" rid="ref1">1</xref>
        ] backed up earlier
UK/US-focused surveys on “second screen” usage1 that
40% of continental Europeans were using a second device
to follow what they were watching on TV, with Google
Copyright held by the author(s).
being first choice to look up related information. As a
global trend, it has led to Forbes’ magazine to announce
“using a second screen while watching TV is the new
normal”2. LinkedTV’s own user trial [
        <xref ref-type="bibr" rid="ref2">2</xref>
        ] with viewers of
TKK confirmed that, even when the viewers were older as
is a typical demographic for such a program, there is a
significant interest in being able to explore further
information beyond what is provided by the program, as
long as it is easy enough to access and can be available also
after the program has been viewed.
      </p>
      <p>In the commercial domain, there is not yet a widely
successful application of second screen TV enrichment due
to the excessive cost of annotating TV programming and
manually preparing the enrichment. Shazam for TV
(http://www.shazam.com/music/web/productfeatures.html?i
d=1266), for example, focuses on a TV program as a whole,
e.g. actor information or episode trivia. Some demos have
been made with concept-level approaches, e.g. using
Mozilla Popcorn (https://popcorn.webmaker.org/), linking
terms in video subtitles to content shown in other frames
alongside the video. These demos suffered from the lack of
disambiguation of terms (such as “Paris”) and limited link
relevance (typically Wikipedia, a map etc.). LinkedTV
provides an automated workflow (cf. Sec 4) with better
disambiguation of concepts referred to within audiovisual
material as well as richer linking to sets of relevant content
from varied online sources. We believe this provides
significantly better automated results than any prior work
and thus also forms the basis for a usable cost-effective
solution by content owners. The rest of this paper will focus
on how this approach was used in enriching a cultural
heritage TV program.</p>
      <p>
        DEMONSTRATOR FOR CULTURAL HERITAGE
LinkedCulture3 [
        <xref ref-type="bibr" rid="ref3">3</xref>
        ] shows the provision of complementary
information related to art objects being discussed in the TV
program. Trials with viewers [
        <xref ref-type="bibr" rid="ref2">2</xref>
        ] validated their interest in
being suggested links to other program segments where
similar objects are discussed and links to information on
3 Demo video: http://vimeo.com/108891238
similar art objects in digital collections (Europeana), which
they can explore while pausing or completing their current
viewing, on the same screen or - casting the TV program to
another screen (e.g. from tablet to a TV) – alongside
viewing. The application allows also for bookmarking so
that viewing can continue but the viewer can easily refer
back to the content they were interested in afterwards.
After accessing the LinkedCulture application (Fig. 1) the
viewer can explore past episodes (vertically) and for each,
select which segment to start viewing (horizontally). Each
segment discusses an art object. An example from our demo
is a silver Frisian tea jar from the mid-18th century (Fig. 2).
Our viewer is very interested in Frisian heritage and the
expert in the TV program states that the tea jar is a typical
example of Frisian Silver. The viewer didn’t even know
there was such a thing, and would love to learn more about
what sort of other objects exist in this category.
During any segment, the viewer can switch to further
information about the art object being discussed such as for
the silver Frisian tea jar, the screenshot shows an
information card about the location Friesland (Fig. 3).
In the next screenshot, other examples of silver tea jars in
Dutch collections can be examined (Fig. 4). Providing the
most relevant related Europeana Cultural Heritage Objects
(CHOs) is the subject of the implementation of a dedicated
Europeana API wrapper, discussed in the next section.
TECHNICAL IMPLEMENTATION OF LINKEDCULTURE
LinkedCulture is built on top of the LinkedTV platform,
implementing a dedicated workflow which ingests video,
analyses and annotates it and generates links to related Web
content (enrichment) (Fig. 5). A Web-based Editor Tool
allows editors to view and curate the annotations and
enrichments of each episode prior to playout.
4 UI design in Figs. 3, 4 &amp; 5 courtesy Lilia Perez Romero (CWI),
cf. LinkedTV D3.5 “Requirements Document LinkedTV User
Interfaces (v2)”
http://slideshare.net/linkedtv/requirementsdocument-for-linkedtv-user-interfaces
      </p>
      <p>
        Immediately at the video analysis step at the beginning of
the LinkedTV workflow, the program is split into distinct
chapters each of which having a different art object as its
focus of discussion between the experts and the audience.
The segmentation approach uses the work of [
        <xref ref-type="bibr" rid="ref4">4</xref>
        ]. In the
TKK case, a set of visual cues have been identified for the
beginning of a new section of the TV episode where a new
art object is introduced and discussed, such as the repetition
of a textual overlay of the name of the expert discussing the
object. From a shot with this textual overlay our approach
searches for prior and subsequent gradual transitions
between shots to fix the chapter boundaries.
      </p>
      <p>
        Once the “art object” chapters are known, Named Entity
Recognition [
        <xref ref-type="bibr" rid="ref5">5</xref>
        ] is performed across the transcript of each
chapter so that entities (concepts) can be associated to them.
In the Editor Tool a specific interface to curators is
available to complete the description of the art object in
each chapter (Fig. 6). The task of the editor is aided by this
extraction of candidate entities from the transcript (i.e.
instances of art object characteristics) which then are
available as suggestions in the Editor Tool.
!
Art object descriptions are stored in the LinkedTV platform
as RDF metadata. The usage of semantics (uniquely
identifying concepts as URIs, following Linked Data
principles) allows us to disambiguate the intended meaning
and retrieve additional data for each concept. The
annotation model focuses on a small set of characteristics of
CHOs typically present in Europeana metadata: object type,
creator, creation location, time period, material. We
followed Dublin Core just as the Europeana Data Model
(EDM) but defined more specific properties for location
and time:
      </p>
    </sec>
    <sec id="sec-2">
      <title>Object type</title>
    </sec>
    <sec id="sec-3">
      <title>Creator</title>
    </sec>
    <sec id="sec-4">
      <title>Material</title>
    </sec>
    <sec id="sec-5">
      <title>Creation location</title>
      <p>Time period
dc:type
dc:creator
dc:medium
vra:locationCreationSite5
dct:temporal6
5 This property was the most specific we found in use in CH
vocabularies to capture the sense of location where a CHO was
created. Europeana does not seem to have a clear approach to this,
often only the current location of the CHO is indicated.
6 Compared to edm:year which takes a single integer for a
calendar year and dct:created which can take an unstructured data
The aggregated annotation is used for the subsequent
enrichment. Distinct Web services provide suggested links
to Web content based on those annotations. These
enrichment services are accessed by a single call to a Web
service called TVEnricher which integrates the individual
services listed above to a single request/response. A shared
service called EntityProxy has also been implemented to
provide information cards on all entities selected in the
video annotations similar to the Google Knowledge Graph
(seen in Fig. 3). We introduce here briefly only the
Europeana enrichment service. The goal is to provide, for
an art object annotated in the TV program chapter, a set of
related art objects from the Europeana collection.
Considering the annotation of our silver Frisian tea jar:</p>
    </sec>
    <sec id="sec-6">
      <title>Type</title>
    </sec>
    <sec id="sec-7">
      <title>Material</title>
    </sec>
    <sec id="sec-8">
      <title>Location</title>
    </sec>
    <sec id="sec-9">
      <title>Period</title>
      <p>db:Container
db:Silver
db:Friesland
Since these properties can occur in the Europeana metadata
under different fields, we tested different approaches to
determine the most effective query. Since we query largely
textual metadata, it was immediately clear that we needed
to consider two key aspects in a query:
Synonyms and similar terms. Metadata property values are
not typed formally in Europeana to a taxonomy so that
generalisations or specialisations can be included in results.
Thus our query must expand its search terms to relevant
synonyms and similar terms.</p>
    </sec>
    <sec id="sec-10">
      <title>Language. Metadata is generally textual and in the</title>
      <p>language of the providing organization. So queries need to
express terms in the local language.</p>
      <p>To address the above and temporal queries, SPARQL
would be complex to model and execute. Thus we focused
our experiment on using the Europeana API. Firstly we
found the best fields to query for each art object property
and then the best approach to expand each field-based
query to aim for the best precision in search results. With
all queries restricted to “COUNTRY:netherlands”7:
Type. “what:container” finds 240 CHOs. In contrast,
"skos_concept= http://vocab.getty.edu/aat/300045611” is
empty, and finds only 2 CHOs in the whole Europeana.
Material. “proxy_dc_format:zilver” finds 448 CHOs. Note
the use of the Dutch zilver, the string silver returns no
matches in Dutch collections. In contrast, “what:zilver”
finds 10256 CHOs as often the material is referenced in the
object type or in subject categories.
string, this property takes as value a dct:PeriodOfTime which is
modelled with a distinct start and end date-time.
7 API query results from 28 January 2015
Location. No metadata property clearly refers to creation
location, not even dc:source. “proxy_dc_source:Friesland”
is empty. “location:Friesland” returns 1271 CHOs and
“where:Friesland” returns 49 556 CHOs, yet both seem to
draw from the value of ‘geographical coverage’
Period. Temporal period queries are well supported, e.g.
“YEAR:[1690+TO+1742]” (60 440 CHOs)
Creator. While not annotated for this object, the API offers
the “who:” field over the metadata property of creator.
Our approach was to expand the query to capture all
possible variations of characteristic values in the metadata
and focus on finding the Europeana CHOs closest in
relevance by matching on at least 3 characteristics.
Expansion is importance since, e.g. type lacks a formal type
system and thus requires us to consider synonyms and
related types in the query, or the spelling and formatting of
creator names can vary in the metadata while the API tries
to make an exact string match. Combining the queries, no
CHO matches on all four characteristics. We do find 4
CHOs for the combination of Frisian+silver+”from 1690 to
1742” and 4 CHOs for the combination of
silver+container+”from 1690 to 1742” (if container is
expanded to also query on synonyms). Combining just two
characteristics in the query, result sets varied from 6 to
1 287 CHOs, and thus often provided too many options on
CHOs which were only weakly related to the original.
To handle this for any arbitrary art object annotated in TKK
episodes we implemented a Europeana API wrapper which
follows the above approach. Since English DBPedia is used
in annotation, we follow owl:sameAs property links to the
Dutch DBPedia to get a Dutch label, while the
dbo:wikiPageRedirects property can indicate alternative
spellings and synonyms. For creators, the foaf:name
property usually provides alternative forms. For locations,
this is the values of the dbo:alternativeName property,
while dbo:isPartOf and dbo:part link to locations which
supersume or subsume this location. The value of this
approach is clear by simple introspection, e.g. if the editor
annotates the object’s creation location as the city of
Franeker, which is quite correct, then the query on
Franeker+silver+”from 1690 to 1742” is empty, but since in
DBPedia the concept Franeker dbo:isPartOf Friesland, the
query can be expanded and the 4 relevant CHOs found.
FUTURE WORK AND CONCLUSION
The implementation of the described Europeana API
wrapper and the semantic annotation of art objects it draws
on has enabled the LinkedCulture application to
semiautomatically enrich TKK episodes with related art objects
from Europeana collections. Simple introspection can show
the benefits of the approach. Even for a single art object
example, we show how a numerically restricted set of most
related art objects can be found via Europeana metadata
overcoming problems of multilingualism and inconsistency
in the terms used. The final LinkedCulture application will
use the annotations of ca. 35 art objects in 6 TKK episodes
and be trialed with TKK viewers to measure satisfaction
with the suggested related Europeana CHOs.</p>
      <p>Europeana as an online portal to digitized collections of
Europe’s cultural heritage is confronted with the problem
that the general public do not search in europeana.eu when
interested in exploring CHOs. Rather, applications are
needed which can push relevant content from the portal to
people when appropriate. Additionally, exploration is made
difficult by the complexity of the domain – even when
turning to Google while watching TV, how does the viewer
seek other examples of silver Frisian tea jars? What
happens when they are interested in a painting they see but
missed the artist’s name? While not so apparent to the
viewer, we must add to this list the inconsistency in the
Europeana metadata. Until there is a more significant
amount of semantic annotation of CHOs, our approach
successfully uses domain knowledge to expand queries and
address issues of term ambiguity, alternative spelling,
synonyms and multilingualism. LinkedCulture is a step
towards eased entry for the public into Europeana’s rich
and deep collection of digital objects tied to the trending
activity of “second screen” usage with television.
ACKNOWLEDGMENTS
The author(s) wish to acknowledge the hard work of the
entire project consortium
(http://www.linkedtv.eu/aboutthe-project/consortium/). We are indebted to AVROTROS
for their permission to use TKK video. LinkedTV has only
been feasible thanks to an EU research grant (FP7-287911).</p>
    </sec>
  </body>
  <back>
    <ref-list>
      <ref id="ref1">
        <mixed-citation>
          1.
          <string-name>
            <given-names>J.</given-names>
            <surname>Abreu</surname>
          </string-name>
          ,
          <string-name>
            <given-names>P.</given-names>
            <surname>Almeida</surname>
          </string-name>
          ,
          <string-name>
            <given-names>B.</given-names>
            <surname>Teles</surname>
          </string-name>
          , and
          <string-name>
            <given-names>M.</given-names>
            <surname>Reis</surname>
          </string-name>
          . „
          <article-title>Viewer behaviors and practices in the (new) television environment”</article-title>
          .
          <source>In Proceedings of the 11th European Conference on Interactive TV and Video</source>
          ,
          <source>EuroITV '13</source>
          , pages
          <fpage>5</fpage>
          -
          <lpage>12</lpage>
          , New York, NY, USA,
          <year>2013</year>
          . ACM.
        </mixed-citation>
      </ref>
      <ref id="ref2">
        <mixed-citation>
          2.
          <string-name>
            <surname>Stanoevska</surname>
            ,
            <given-names>K.</given-names>
          </string-name>
          et al.,
          <article-title>“User Trial results”</article-title>
          ,
          <source>LinkedTV deliverable 6</source>
          .3,
          <string-name>
            <surname>March</surname>
          </string-name>
          <year>2014</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref3">
        <mixed-citation>
          3.
          <string-name>
            <surname>Nixon</surname>
            ,
            <given-names>L.</given-names>
          </string-name>
          and
          <string-name>
            <surname>Baltussen</surname>
            ,
            <given-names>L.</given-names>
          </string-name>
          ,
          <article-title>“Engaging TV viewers with AudioVisual heritage on second screens”</article-title>
          , EUScreenXL conference, Rome, Italy,
          <year>October 2014</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref4">
        <mixed-citation>
          4.
          <string-name>
            <given-names>E.</given-names>
            <surname>Apostolidis</surname>
          </string-name>
          ,
          <string-name>
            <given-names>V.</given-names>
            <surname>Mezaris</surname>
          </string-name>
          ,
          <article-title>"Fast Shot Segmentation Combining Global and Local Visual Descriptors"</article-title>
          ,
          <source>Proc. IEEE Int. Conf. on Acoustics, Speech and Signal Processing (ICASSP)</source>
          , Florence, Italy, May
          <year>2014</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref5">
        <mixed-citation>
          5.
          <string-name>
            <given-names>Y.</given-names>
            <surname>Li</surname>
          </string-name>
          ,
          <string-name>
            <given-names>G.</given-names>
            <surname>Rizzo</surname>
          </string-name>
          ,
          <string-name>
            <given-names>J. L. R.</given-names>
            <surname>García</surname>
          </string-name>
          ,
          <string-name>
            <given-names>R.</given-names>
            <surname>Troncy</surname>
          </string-name>
          ,
          <string-name>
            <given-names>M.</given-names>
            <surname>Wald</surname>
          </string-name>
          , and
          <string-name>
            <surname>G. Wills.</surname>
          </string-name>
          “
          <article-title>Enriching media fragments with named entities for video classification”</article-title>
          .
          <source>In 1st Worldwide Web Workshop on Linked Media (LiME'13)</source>
          , pages
          <fpage>469</fpage>
          -
          <lpage>476</lpage>
          , Rio de Janeiro, Brazil,
          <year>2013</year>
          .
        </mixed-citation>
      </ref>
    </ref-list>
  </back>
</article>