<!DOCTYPE article PUBLIC "-//NLM//DTD JATS (Z39.96) Journal Archiving and Interchange DTD v1.0 20120330//EN" "JATS-archivearticle1.dtd">
<article xmlns:xlink="http://www.w3.org/1999/xlink">
  <front>
    <journal-meta />
    <article-meta>
      <title-group>
        <article-title>E ective Metadata for Social Book Search from a User Perspective</article-title>
      </title-group>
      <contrib-group>
        <contrib contrib-type="author">
          <string-name>Hugo Huurdeman</string-name>
          <xref ref-type="aff" rid="aff0">0</xref>
          <xref ref-type="aff" rid="aff2">2</xref>
        </contrib>
        <contrib contrib-type="author">
          <string-name>Jaap Kamps</string-name>
          <xref ref-type="aff" rid="aff0">0</xref>
          <xref ref-type="aff" rid="aff1">1</xref>
          <xref ref-type="aff" rid="aff2">2</xref>
        </contrib>
        <contrib contrib-type="author">
          <string-name>Marijn Koolen</string-name>
          <xref ref-type="aff" rid="aff2">2</xref>
        </contrib>
        <aff id="aff0">
          <label>0</label>
          <institution>Archives and Information Studies, Faculty of Humanities, University of Amsterdam</institution>
        </aff>
        <aff id="aff1">
          <label>1</label>
          <institution>ISLA, Faculty of Science, University of Amsterdam</institution>
        </aff>
        <aff id="aff2">
          <label>2</label>
          <institution>Institute for Logic, Language and Computation, University of Amsterdam</institution>
        </aff>
      </contrib-group>
      <fpage>543</fpage>
      <lpage>548</lpage>
      <abstract>
        <p>In this extended abstract we describe our participation in the INEX 2014 Interactive Social Book Search Track. In previous work, we have looked at the impact of professional and user-generated metadata in the context of book search, and compared these di erent categories of metadata in terms of retrieval e ectiveness. Here, we take a di erent approach and study the use of professional and user-generated metadata of books in an interactive setting, and the e ectivity of this metadata from a user perspective. We compare the perceived usefulness of general descriptions, publication metadata, user reviews and tags in focused and open-ended search tasks, based on data gathered in the INEX Interactive Social Book Search Track. Furthermore, we take a tentative look at the actual use of di erent types of metadata over time in the aggregated search tasks. Our preliminary ndings in the surveyed tasks indicate that user reviews are generally perceived to be more useful than other types of metadata, and they are frequently mentioned in users' rationales for selecting books. Furthermore, we observe a varying usage frequency of traditional and user-generated metadata across time in the aggregated search tasks, providing initial indications that these types of metadata might be useful at di erent stages of a search task.</p>
      </abstract>
    </article-meta>
  </front>
  <body>
    <sec id="sec-1">
      <title>-</title>
      <p>In 2014, the Interactive Social Book Search (iSBS) task has been introduced.
The goal of this task is \to investigate how book searchers deal with professional
and user-generated content at di erent stages of the of the search process" 1.</p>
      <p>The iSBS track uses the Amazon/LibraryThing collection with book
descriptions for 1.5 million books, as crawled by the University of Duisburg-Essen in
early 2009. This dataset contains both professional metadata and user-generated
descriptions from Amazon and LibraryThing.</p>
      <p>The task has introduced two interfaces for book search: a baseline interface,
including search elements, results lists, and item details, and a `multistage'
interface, in which di erent features for di erent stages of a search have been
introduced. Both interfaces feature a \book-bag", to which users can add books
when completing the assigned tasks2. Two kinds of tasks are employed: a
goaloriented task, and a non-goal oriented task.</p>
      <p>The goal-oriented task was the following: Imagine you are looking for some
interesting physics and mathematics books for a layperson. You have heard about
the Feynman books but you have never really read anything in this area. You
would also like to nd an \interesting facts" sort of book on mathematics.</p>
      <p>The non-goal oriented task, without a prede ned information need, was the
following: Imagine you are waiting to meet a friend in a co ee shop or pub or the
airport or your o ce. While waiting, you come across this website and explore
it looking for any book that you nd interesting, or engaging or relevant...</p>
      <p>The participants in the iSBS experiment carried out one goal, and one
nongoal oriented task, and were randomly assigned the baseline or multistage
interface. A total of 41 users participated in the experiment, at Aalborg University
Copenhagen, Edge Hill University, Humboldt University, and the University of
Amsterdam. Of these participants, 19 used the baseline interface, and 22 users
utilized the multistage interface. The resulting data, in the form of usage logs
and questionnaires gathered in the experiment, was shared among the
partipating teams.</p>
      <p>In this extended abstract, we take an initial look at the results, and
specifically focus on the usefulness and e ectivity of professional and user-generated
metadata for books in an interactive setting, and compare the perceived
usefulness of general descriptions, publication metadata, user reviews and tags across
tasks.
2</p>
    </sec>
    <sec id="sec-2">
      <title>Previous work</title>
      <p>
        Previous work related to the iSBS track has been carried out in the INEX
Interactive Retrieval Experiments (2004-2010) [
        <xref ref-type="bibr" rid="ref7">7</xref>
        ] and the Cultural Heritage in
CLEF (CHiC) Interactive Track 2013 [
        <xref ref-type="bibr" rid="ref8">8</xref>
        ]. In these tracks, a standard procedure
for collecting data was being used by participating research groups, including
common topics and tasks, standardized search systems, document corpora and
procedures. The system used for the iSBS track is a modi ed version of the one
used for CHiC and is based on the Interactive IR evaluation platform
developed by Hall and Toms [
        <xref ref-type="bibr" rid="ref1">1</xref>
        ], where di erent search engines and interfaces can be
plugged into fully developed IIR framework that runs the entire user study [
        <xref ref-type="bibr" rid="ref2">2</xref>
        ].
      </p>
      <p>
        The Interactive SBS Track complements the system-oriented Social Book
Search Track [
        <xref ref-type="bibr" rid="ref6">6</xref>
        ]. Both tracks use the Amazon/LibraryThing collection and
investigate the value of professional metadata and user-generated content. The
system-oriented evaluation of the SBS Track has shown that retrieval e
ectiveness increases when systems include user-generated content for a broad range of
tasks [
        <xref ref-type="bibr" rid="ref4 ref5">4, 5</xref>
        ]. These ndings prompted the desire to study how book searchers use
professional metadata and user-generated content.
2 More information about the experiment and di erent interfaces is available in [
        <xref ref-type="bibr" rid="ref3">3</xref>
        ]
      </p>
    </sec>
    <sec id="sec-3">
      <title>Results</title>
      <p>The results described in this extended abstract are preliminary, and based on a
relatively small dataset. Nevertheless, they can provide basic insights into book
search from a user perspective.</p>
      <p>We focus on four di erent types of displayed metadata in the baseline and
multistage interface:
{ general book descriptions which contain publisher-provided product
descriptions and editorial reviews (available for 78% of all books).
{ publication metadata, including information on publisher, price, number of
pages and binding (available for all books)
{ reviews metadata, containing up to 10 Amazon user reviews and include
ratings (available for 45% of all books)
{ user-provided tags from LibraryThing (LT), the displayed size based on how
many LT users assigned each tag to the book (available for 83% of all books).
usefulness First of all, we look at the perceived usefulness of the general
descriptions, publication metadata, user reviews and user tags.</p>
      <p>In the post-questionnaire following the focused and open task of the
experiment, users had to indicate how useful the metadata items about a book were
on a rating scale of 1 (not at all useful) to 5 (extremely useful). Table 1
contains the results, di erentiated by interface. It shows that for both interfaces,
the general descriptions were deemed most useful, with an average rating of 4.2
out of 5. This is followed by the reviews, that were almost as useful, with an
average rating of 3.9. Finally, both formal publication metadata, and tags were
not considered highly useful by the participants.</p>
      <p>Secondly, the usage of each feature is of importance, since in the same survey
question participants could also indicate that they did not use an element. Table
2 indicates that only a small minority (7.5%) of participants did not use the
general book descriptions. Again, it is followed by the reviews, that are not used
by 25% of all participants. Finally, the tags and publication metadata are unused
by more than 30% of all users.</p>
      <p>So despite the fact that user reviews and ratings are only available for 45%
of all books, they are actually considered almost as useful as general descriptive
metadata by the participants in the experiment, and used more often than
publication metadata and user tags. To contextualize this nding, we now look at
the rationales of users for adding books to their bookbag selections.
contextualizing perceived usefulness In the post-task questionnaire, users could
review their bookbag immediately after nishing their task, and were asked
to describe why they selected these books. Users describe various reasons for
adding books to their bookbags. A commonly mentioned selection criterion is
related to the book's title, subject and description. One user, for example, points
out that \the rst book, due to the title, seemed like a simpler read then the
Feynman volumes". Other users, surprisingly, mention that they chose a result
because of it's high position in the results list (being \closer to the top of the
results"). Another reason for choosing particular books, in line with the results
discussed in the previous section, is based on their ratings and reviews. For
the open task, 7 out of 41 participants mention ratings and reviews as being
important in their nal book selections, while for the focused task another 7
participants participants explicitly mention ratings and reviews. For example,
one user reports that based on the number of reviews, (s)he decided that a
book is engaging: \I tried to look at that in the comments, if books had no
comments I didn't select them". Reviews are also frequently used to choose
among alternatives: \the rst book has also a very good review". Hence, the
usergenerated ratings and reviews seem to play an important role in the selection
process in the context of the employed tasks.
usage statistics of metadata elements over time Based on the available usage
logs, we can take a tentative look at the usage of the metadata available in the
multistage experimental user interface. Here, we do not yet have statistics on
the number of times the general book descriptions were seen, since they appear
by default, but we can get insights on the usage of the publication, review and
tag metadata.</p>
      <p>Table 3 shows the number of times publication metadata, reviews and tags
were explicitly selected in the focused and open-ended search task. While these
numbers, upon a total number of 22 users across two tasks are quite low, they
do indicate a higher utilization of user-generated reviews in the combined tasks,
used in total 52 times by 22 users.</p>
      <p>Furthermore, the use of these metadata features over time can be surveyed.
Therefore we divided the total task time into three segments, begin, middle and
end, each taking up one third of the task process. From this data, we see that
the number of times these data elements are used varies over time: the
publication metadata is used more in the middle, and the reviews and tags are used
more frequently in the end. Due to the low number of data points, we cannot
derive strong conclusions, but the use of metadata features over time could be
interesting to assess in future studies.
4</p>
    </sec>
    <sec id="sec-4">
      <title>Conclusion</title>
      <p>The preliminary results documented in this extended abstract provide
indications that besides professional metadata, also user-generated metadata might
be useful for book search, from the perspective of users. In terms of perceived
usefulness, and usage, especially user reviews seem to aid users in determining
their book selections for di erent types of tasks, even though they are not
available for all books. While the ndings here are based on the rst iteration of the
Interactive Social Book Search task, they provide handles for future research,
and we intend to extend these ndings in future work in the context of Social
Book Search.</p>
      <p>Acknowledgments This research was supported by the Netherlands
Organization for Scienti c Research (NWO projects # 612.066.513, 639.072.601, and
640.005.001) and by the European Community's Seventh Framework Program
(FP7 2007/2013, Grant Agreement 270404).</p>
    </sec>
  </body>
  <back>
    <ref-list>
      <ref id="ref1">
        <mixed-citation>
          1.
          <string-name>
            <given-names>M.</given-names>
            <surname>Hall</surname>
          </string-name>
          and
          <string-name>
            <given-names>E.</given-names>
            <surname>Toms</surname>
          </string-name>
          .
          <article-title>Building a common framework for iir evaluation</article-title>
          . In P. Forner, H. Muller, R. Paredes,
          <string-name>
            <given-names>P.</given-names>
            <surname>Rosso</surname>
          </string-name>
          , and B. Stein, editors,
          <source>Information Access Evaluation</source>
          . Multilinguality, Multimodality, and Visualization, volume
          <volume>8138</volume>
          of Lecture Notes in Computer Science, pages
          <volume>17</volume>
          {
          <fpage>28</fpage>
          . Springer Berlin Heidelberg,
          <year>2013</year>
          .
          <source>ISBN 978-3-642-40801-4. doi: 10.1007/978-3-642-40802-1 3.</source>
        </mixed-citation>
      </ref>
      <ref id="ref2">
        <mixed-citation>
          2.
          <string-name>
            <given-names>M.</given-names>
            <surname>Hall</surname>
          </string-name>
          ,
          <string-name>
            <given-names>S.</given-names>
            <surname>Katsaris</surname>
          </string-name>
          , and
          <string-name>
            <given-names>E.</given-names>
            <surname>Toms</surname>
          </string-name>
          .
          <article-title>A pluggable interactive ir evaluation work-bench</article-title>
          .
          <source>In European Workshop on Human-Computer Interaction and Information Retrieval</source>
          , pages
          <volume>35</volume>
          {
          <fpage>38</fpage>
          ,
          <year>2013</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref3">
        <mixed-citation>
          3.
          <string-name>
            <given-names>M.</given-names>
            <surname>Hall</surname>
          </string-name>
          ,
          <string-name>
            <given-names>H.</given-names>
            <surname>Huurdeman</surname>
          </string-name>
          ,
          <string-name>
            <given-names>M.</given-names>
            <surname>Koolen</surname>
          </string-name>
          ,
          <string-name>
            <given-names>M.</given-names>
            <surname>Skov</surname>
          </string-name>
          , and
          <string-name>
            <given-names>D.</given-names>
            <surname>Walsh</surname>
          </string-name>
          .
          <article-title>Overview of the INEX 2014 interactive social book search track</article-title>
          . In L. Cappellato,
          <string-name>
            <given-names>N.</given-names>
            <surname>Ferro</surname>
          </string-name>
          ,
          <string-name>
            <given-names>M.</given-names>
            <surname>Halvey</surname>
          </string-name>
          , and W. Kraaij, editors,
          <source>CLEF 2014 Labs and Workshops</source>
          , Notebook Papers,
          <source>CEUR Workshop Proceedings (CEUR-WS.org)</source>
          ,
          <year>2014</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref4">
        <mixed-citation>
          4.
          <string-name>
            <given-names>M.</given-names>
            <surname>Koolen</surname>
          </string-name>
          .
          <article-title>"user reviews in the search index? that'll never work!"</article-title>
          . In M. de Rijke,
          <string-name>
            <given-names>T.</given-names>
            <surname>Kenter</surname>
          </string-name>
          ,
          <string-name>
            <surname>A. P. de Vries</surname>
            ,
            <given-names>C.</given-names>
          </string-name>
          <string-name>
            <surname>Zhai</surname>
          </string-name>
          , F. de Jong, K. Radinsky,
          <article-title>and</article-title>
          K. Hofmann, editors,
          <source>ECIR</source>
          , volume
          <volume>8416</volume>
          of Lecture Notes in Computer Science, pages
          <volume>323</volume>
          {
          <fpage>334</fpage>
          . Springer,
          <year>2014</year>
          . ISBN 978-3-
          <fpage>319</fpage>
          -06027-9.
        </mixed-citation>
      </ref>
      <ref id="ref5">
        <mixed-citation>
          5.
          <string-name>
            <given-names>M.</given-names>
            <surname>Koolen</surname>
          </string-name>
          ,
          <string-name>
            <given-names>J.</given-names>
            <surname>Kamps</surname>
          </string-name>
          , and
          <string-name>
            <given-names>G.</given-names>
            <surname>Kazai. Social Book</surname>
          </string-name>
          <article-title>Search: The Impact of Professional and User-Generated Content on Book Suggestions</article-title>
          .
          <source>In CIKM 2012. ACM</source>
          ,
          <year>2012</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref6">
        <mixed-citation>
          6.
          <string-name>
            <given-names>M.</given-names>
            <surname>Koolen</surname>
          </string-name>
          , G. Kazai,
          <string-name>
            <given-names>M.</given-names>
            <surname>Preminger</surname>
          </string-name>
          ,
          <article-title>and</article-title>
          <string-name>
            <given-names>A.</given-names>
            <surname>Doucet</surname>
          </string-name>
          .
          <article-title>Overview of the INEX 2013 social book search track</article-title>
          .
          <source>In CLEF 2013 Evaluation Labs and Workshop</source>
          , Online Working Notes,
          <year>2013</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref7">
        <mixed-citation>
          7.
          <string-name>
            <given-names>R.</given-names>
            <surname>Nordlie</surname>
          </string-name>
          and
          <string-name>
            <given-names>N.</given-names>
            <surname>Pharo</surname>
          </string-name>
          .
          <article-title>Seven years of inex interactive retrieval experiments - lessons and challenges</article-title>
          . In T. Catarci,
          <string-name>
            <given-names>P.</given-names>
            <surname>Forner</surname>
          </string-name>
          ,
          <string-name>
            <given-names>D.</given-names>
            <surname>Hiemstra</surname>
          </string-name>
          ,
          <string-name>
            <surname>A</surname>
          </string-name>
          . Pen~as, and G. Santucci, editors,
          <source>CLEF</source>
          , volume
          <volume>7488</volume>
          of Lecture Notes in Computer Science, pages
          <volume>13</volume>
          {
          <fpage>23</fpage>
          . Springer,
          <year>2012</year>
          . ISBN 978-3-
          <fpage>642</fpage>
          -33246-3.
        </mixed-citation>
      </ref>
      <ref id="ref8">
        <mixed-citation>
          8.
          <string-name>
            <given-names>V.</given-names>
            <surname>Petras</surname>
          </string-name>
          ,
          <string-name>
            <given-names>T.</given-names>
            <surname>Bogers</surname>
          </string-name>
          , E. Toms,
          <string-name>
            <given-names>M.</given-names>
            <surname>Hall</surname>
          </string-name>
          ,
          <string-name>
            <given-names>J.</given-names>
            <surname>Savoy</surname>
          </string-name>
          ,
          <string-name>
            <given-names>P.</given-names>
            <surname>Malak</surname>
          </string-name>
          ,
          <string-name>
            <given-names>A.</given-names>
            <surname>Pawowski</surname>
          </string-name>
          ,
          <string-name>
            <given-names>N.</given-names>
            <surname>Ferro</surname>
          </string-name>
          ,
          <string-name>
            <surname>and I. Masiero.</surname>
          </string-name>
          <article-title>Cultural heritage in clef (chic) 2013</article-title>
          . In P. Forner, H. Muller, R. Paredes,
          <string-name>
            <given-names>P.</given-names>
            <surname>Rosso</surname>
          </string-name>
          , and B. Stein, editors,
          <source>Information Access Evaluation</source>
          . Multilinguality, Multimodality, and Visualization, volume
          <volume>8138</volume>
          of Lecture Notes in Computer Science, pages
          <volume>192</volume>
          {
          <fpage>211</fpage>
          . Springer Berlin Heidelberg,
          <year>2013</year>
          . ISBN 978-3-
          <fpage>642</fpage>
          -40801-4. doi:
          <volume>10</volume>
          .1007/978-3-
          <fpage>642</fpage>
          -40802-1
          <fpage>23</fpage>
          .
        </mixed-citation>
      </ref>
    </ref-list>
  </back>
</article>