<!DOCTYPE article PUBLIC "-//NLM//DTD JATS (Z39.96) Journal Archiving and Interchange DTD v1.0 20120330//EN" "JATS-archivearticle1.dtd">
<article xmlns:xlink="http://www.w3.org/1999/xlink">
  <front>
    <journal-meta />
    <article-meta>
      <title-group>
        <article-title>Research-Centres Centred Living Labs</article-title>
      </title-group>
      <contrib-group>
        <contrib contrib-type="author">
          <string-name>Liadh Kelly</string-name>
          <email>Liadh.Kelly@tcd.ie</email>
          <xref ref-type="aff" rid="aff0">0</xref>
        </contrib>
        <aff id="aff0">
          <label>0</label>
          <institution>ADAPT Centre, Trinity College</institution>
          ,
          <addr-line>Dublin</addr-line>
          ,
          <country country="IE">Ireland</country>
        </aff>
      </contrib-group>
      <abstract>
        <p>Developing means to conduct shared evaluation in the user modelling, adaptation and personalization (UMAP) space is inherently difficult. Not least because of privacy concerns and individual differences in behaviours between users of systems. In this paper we propose a research-centre centred living labs approach as one potential way to overcome these difficulties and to allow for shared task generation in the UMAP domain.</p>
      </abstract>
    </article-meta>
  </front>
  <body>
    <sec id="sec-1">
      <title>Introduction</title>
      <p>
        A tool for creating a living lab that centres on research centres providing data and users
for shared evaluation is presented in [
        <xref ref-type="bibr" rid="ref3">3</xref>
        ] (see Figure 1(b)). The paper focuses on a
living lab for evaluation of retrieval techniques for personal desktop collections. In this
approach, researchers wishing to evaluate their technologies would participate in a
collaborative evaluation effort. Whereby required protocols and technology to gather data
for, and to conduct the evaluation, would be distributed to the participating research
centres. The retrieval algorithms/techniques developed by each participating research
centre would also be distributed to the research centres for evaluation. Individual research
centres would then recruit experiment subjects locally, who install and run the provided
tool on their personal computer (PC). The tool indexes the items on the PCs. Using
the provided protocols and tool, experiment subjects issue personal queries and
conduct relevance assessment. Participating research centres’ IR algorithms are evaluated
locally on subjects’ PCs using the generated index, queries and relevance assessments,
with only performance measures returned to investigators thus preserving privacy.
2.1 Proposal
The high-level concept of this research-centres centred living labs approach could be
generalized to allow for shared task evaluation in the UMAP space. Whereby
evaluation goal specific tools and protocols, and challenge participants algorithms/software
are distributed to individual research centres. These are then either used (as shown in
Figure 1(b)) to: generate static collections for evaluations as described above; run
controlled experiments with the participants’ software locally in each research centre; or
ideally run the participants’ software live, for evaluation purposes, in place of
individuals’ typical software as they go about their normal activities. Or indeed a hybrid of this
research-centres centred and the earlier API-centred approach might prove most useful
in the UMAP space, depending on the precise scenario to be evaluated. This general
evaluation paradigm has potential for evaluation of any tool or algorithm supporting
individuals interacting with digital data on both mobile and stationary devices.
      </p>
      <p>
        Realising such living labs requires addressing several challenges associated with
living labs architecture and design, hosting, maintenance, security, privacy, participant
recruiting, and scenarios and tasks for use development. Lessons can be learned here
from the experiences of the IR and recommender systems living labs shared tasks [
        <xref ref-type="bibr" rid="ref1 ref2">1, 2</xref>
        ].
3
      </p>
    </sec>
    <sec id="sec-2">
      <title>Conclusions</title>
      <p>Current instantiations of living labs in IR and recommender system shared challenges
focus on an API-centred living labs methodology. Other interpretations of living labs are
also possible. In this paper we put forward the use of a research-centres centred living
labs methodology for shared UMAP challenges. This research-centres centred living
labs methodology could also have application in evaluation of other research spaces.</p>
    </sec>
  </body>
  <back>
    <ref-list>
      <ref id="ref1">
        <mixed-citation>
          [1]
          <string-name>
            <given-names>K.</given-names>
            <surname>Balog</surname>
          </string-name>
          ,
          <string-name>
            <given-names>L.</given-names>
            <surname>Kelly</surname>
          </string-name>
          ,
          <article-title>and</article-title>
          <string-name>
            <given-names>A.</given-names>
            <surname>Schuth</surname>
          </string-name>
          . Head first:
          <article-title>Living labs for ad-hoc search evaluation</article-title>
          .
          <source>In CIKM'14</source>
          ,
          <year>2014</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref2">
        <mixed-citation>
          [2]
          <string-name>
            <given-names>F.</given-names>
            <surname>Hopfgartner</surname>
          </string-name>
          ,
          <string-name>
            <given-names>B.</given-names>
            <surname>Kille</surname>
          </string-name>
          ,
          <string-name>
            <given-names>A.</given-names>
            <surname>Lommatzsch</surname>
          </string-name>
          ,
          <string-name>
            <given-names>T.</given-names>
            <surname>Plumbaum</surname>
          </string-name>
          ,
          <string-name>
            <given-names>T.</given-names>
            <surname>Brodt</surname>
          </string-name>
          , and
          <string-name>
            <given-names>T.</given-names>
            <surname>Heintz</surname>
          </string-name>
          .
          <article-title>Benchmarking news recommendations in a living lab</article-title>
          .
          <source>In CLEF'14</source>
          ,
          <year>2014</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref3">
        <mixed-citation>
          [3]
          <string-name>
            <given-names>L.</given-names>
            <surname>Kelly</surname>
          </string-name>
          ,
          <string-name>
            <given-names>P.</given-names>
            <surname>Bunbury</surname>
          </string-name>
          , and
          <string-name>
            <given-names>G. J. F.</given-names>
            <surname>Jones</surname>
          </string-name>
          .
          <article-title>Evaluating personal information retrieval</article-title>
          .
          <source>In Proc. of ECIR '12</source>
          ,
          <year>2012</year>
          .
        </mixed-citation>
      </ref>
    </ref-list>
  </back>
</article>