<!DOCTYPE article PUBLIC "-//NLM//DTD JATS (Z39.96) Journal Archiving and Interchange DTD v1.0 20120330//EN" "JATS-archivearticle1.dtd">
<article xmlns:xlink="http://www.w3.org/1999/xlink">
  <front>
    <journal-meta />
    <article-meta>
      <title-group>
        <article-title>A Study of Users' Image Seeking Behaviour in Flickling</article-title>
      </title-group>
      <contrib-group>
        <contrib contrib-type="author">
          <string-name>Evgenia Vassilakaki</string-name>
          <email>evgenia.vassilakaki@student.mmu.ac.uk</email>
          <xref ref-type="aff" rid="aff0">0</xref>
          <xref ref-type="aff" rid="aff1">1</xref>
        </contrib>
        <contrib contrib-type="author">
          <string-name>Frances Johnson</string-name>
          <xref ref-type="aff" rid="aff0">0</xref>
          <xref ref-type="aff" rid="aff1">1</xref>
        </contrib>
        <contrib contrib-type="author">
          <string-name>R.J. Hartley</string-name>
          <xref ref-type="aff" rid="aff0">0</xref>
          <xref ref-type="aff" rid="aff1">1</xref>
        </contrib>
        <contrib contrib-type="author">
          <string-name>David Randall</string-name>
          <email>d.randall@mmu.ac.uk</email>
          <xref ref-type="aff" rid="aff0">0</xref>
          <xref ref-type="aff" rid="aff1">1</xref>
        </contrib>
        <contrib contrib-type="author">
          <string-name>General Terms</string-name>
          <xref ref-type="aff" rid="aff0">0</xref>
          <xref ref-type="aff" rid="aff1">1</xref>
        </contrib>
        <contrib contrib-type="editor">
          <string-name>Languages, Human Factors, Experimentation</string-name>
        </contrib>
        <aff id="aff0">
          <label>0</label>
          <institution>Dept. Information &amp; Communications</institution>
        </aff>
        <aff id="aff1">
          <label>1</label>
          <institution>Manchester Metropolitan University</institution>
        </aff>
      </contrib-group>
      <abstract>
        <p>This study aims to explore users' image seeking behaviour when searching for a known, non-annotated image in Flickling provided by iCLEF2008 track. The task assigned to users was to search for the three rst images given after rst login. Users did not know in advance in which of the six languages (English, German, Dutch, Spanish, French, Italian) the images were described, forcing them to search across languages. The main focus of our study was threefold: a) to identify the reasons that determined users' choice over a speci c interface, b) to examine whether users were thinking about languages when searching for images and to what extent and c) to examine if used, how helpful the translations proved to be for nding the images. This study used four di erent, both quantitative and qualitative methods (questionnaires, retrospective thinking aloud, observation and interviews) to meet its research questions. Results show that two out of ten users were using only the monolingual interface because they did not feel con dent with languages and the rest were switching between interfaces for a variety of reasons in which languages played a small part. Only four out of ten users were actually thinking about languages when searching for the images, while the rest were more preoccupied with nding the images and completing the task successfully. As a consequence, only four users paid attention to translations and only judged the translations in languages known to them. Overall, the translations were not considered to be helpful due to their inconsistency in coverage and their tendency to lead to irrelevant results.</p>
      </abstract>
      <kwd-group>
        <kwd>H</kwd>
        <kwd>3 [Information Storage and Retrieval]</kwd>
        <kwd>H</kwd>
        <kwd>3</kwd>
        <kwd>3 Information Search and Retrieval</kwd>
        <kwd>H</kwd>
        <kwd>2</kwd>
        <kwd>3 [Database Management]</kwd>
        <kwd>Languages|Query Languages</kwd>
      </kwd-group>
    </article-meta>
  </front>
  <body>
    <sec id="sec-1">
      <title>Introduction</title>
      <p>
        Cross Language Evaluation Forum (CLEF) is an annual evaluation campaign that aims to
promote the development of monolingual and multilingual information retrieval systems for European
Languages. The main research interest of CLEF has gradually moved over the years, from only
textual document retrieval to question answering (QA) and Geographic information retrieval [
        <xref ref-type="bibr" rid="ref2">2</xref>
        ].
In this context, a pilot interactive track, known as interactive CLEF (iCLEF), was introduced in
2001 focusing initially on document selection questions. Since then, iCLEF has been included as
a regular event in CLEF.
      </p>
      <p>
        The iCLEF tracks in CLEF2002 and 2003 focused on examining support mechanisms for query
formulation and re nement, as well as user-assisted translations experiments. In the 2004 track,
the participants, all using a common evaluation design, tried to assess the ability of their own
interactive systems to nd speci c answers to speci c questions in a language other than the one
of the initial query. In addition, the 2005 track studied the problem of QA from a user-inclusive
perspective [
        <xref ref-type="bibr" rid="ref8">8</xref>
        ]. Finally in 2006, the iCLEF track moved to Flickr, a photo-sharing multilingual
database in order to promote interactive experimentation on multilingual search tasks and study
users' behaviour [
        <xref ref-type="bibr" rid="ref6">6</xref>
        ]. In particular, three studies were submitted: a) UNED [
        <xref ref-type="bibr" rid="ref1">1</xref>
        ] examined the
attitude of users towards cross-language searching when the search system allows three search
modes (no translation, automatic translation, assisted translation), b) U. She eld (and IBM)
[
        <xref ref-type="bibr" rid="ref3">3</xref>
        ] used bilingual Arabic-English students to test the Arabic interface of Flickr that they have
developed and nally c) Swedish Institute of Computer Science (SICS) [
        <xref ref-type="bibr" rid="ref7">7</xref>
        ] focused on evaluating
information access based on user satisfaction and user con dence.
      </p>
      <p>
        In this context, the 2008 iCLEF track focuses both on acquiring a large set of search session
logs for the participants to mine and on allowing participants to perform their own interactive
experiments with the Flickling interface provided and adopting the task prede ned by the
Organizers [
        <xref ref-type="bibr" rid="ref4">4</xref>
        ]. The aim of this study is to explore users' image seeking behaviour when searching and
retrieving known, non-annotated images across languages in Flickling. In particular, the research
questions that will be addressed in this paper are:
      </p>
      <p>Identify the reasons that determined our users' choice of a speci c interface
(monolingual/multilingual).</p>
      <p>Examine if and/or to what extent users were thinking about languages when searching and
retrieving images.</p>
      <p>Examine if and/or to what extent users were paying attention to translations when searching
and retrieving images.</p>
      <p>The remainder of this paper is structured as follows: a description of the Flickling interface,
of our user sample, of the task given, of the four di erent methods that we used in assembling the
data, of the way that the study was carried out and the data processed are illustrated in section
2. We provide an analysis of our ndings and an extended discussion of them in sections 3 and
4 respectively. Finally, we conclude summarizing the di erent image seeking behaviours that our
users developed while using Flickling in section 5.
2</p>
    </sec>
    <sec id="sec-2">
      <title>Method</title>
      <p>In this section, further details about the test object of the study, the users, the task given, the
speci c methods used such as retrospective thinking aloud, observation and interviews, and the
data processing are presented.
2.1</p>
      <sec id="sec-2-1">
        <title>Test Object</title>
        <p>
          The test object of this study was the Flickling interface, a basic cross-language search front-end to
the well-known web application Flickr. Flickr was adopted as the target collection by the iCLEF
organizers mainly for two reasons [
          <xref ref-type="bibr" rid="ref4">4</xref>
          ]: a) it is a multilingual database enabling the users to tag
and comment in di erent languages the uploaded images and b) it can provide the baseline for a
series of both realistic and challenging multilingual search tasks.
        </p>
        <p>
          The Flickling interface is a multilingual information retrieval interface that encompasses the
following functionalities [
          <xref ref-type="bibr" rid="ref4">4</xref>
          ]: a) multilingual search across six languages (English, Spanish, German,
French, Dutch, Italian), b) term-to-term translations between six languages using freely available
dictionaries, c) select of best target translations, d) pick/remove translations and adding of new
ones by the users, f) provision of search suggestions and g) control over the game-like features of
the task.
        </p>
        <p>The Flickling interface was intended for two user groups: a) CLEF participants, researchers
who expressed interest in conducting experiments based on the provided multilingual interface
and b) Flickr/Web users, ordinary users of Flickr application that would like to participate in this
game o ered by CLEF organizers.</p>
        <p>Our main reason of interest in participating in the iCLEF2008 Flickr challenge was to
investigate the behaviour of users when asked to search and retrieve a known, non-annotated image
across languages.
2.2</p>
      </sec>
      <sec id="sec-2-2">
        <title>Users</title>
        <p>The study was carried out with a sample of 10 users, three male and seven female, ranging in age
from 20 to 40. They were all related in one way or another to Manchester Metropolitan University
(MMU). In particular, from the sample of ten users, seven were research postgraduate students,
one taught postgraduate student, one lecturer and one MMU sta member. In addition, four of
the users were English native speakers, two Greek, one German, one Spanish, one Arabic and one
Luganda (see Table 1). Moreover, one of the users was monolingual, four stated knowledge of a
language other than their native and ve were multilingual.</p>
        <p>Native Language</p>
        <p>Arabic
English
German</p>
        <p>Greek
Luganda
Spanish</p>
        <p>In addition, the users were asked to state their level of comprehension for the languages used
in Flickling but also any other additional language. In particular, from the six non English native
speakers, two stated an Excellent knowledge of English, three Very Good knowledge and one Basic
knowledge. In regards of German, three of the nine non-German native speakers stated a Basic
knowledge of German. Four out of ten users stated knowledge of French, three of whom Basic and
one Good. Concerning Italian two out of ten users stated knowledge of Italian language, Basic
and Good respectively. Two out of ten stated a Basic knowledge of Dutch and nally, three out
of nine non Spanish native speakers stated a Basic knowledge of Spanish (see Table 2).</p>
        <p>All ten users have searched in the past for an image on the web. In particular, four stated that
they \rarely" have, three \sometimes", two \very often" and one \often". In addition, nine out of
ten stated that they have searched for an image on the web in a language other than their native
and only one had not. In particular, the nine users identi ed the following reasons for having done
so: \university research, searching for holiday info", \to increase the numbers of results because I
could not nd any relevant image by using keywords in my native language", \because there are
only few web resources that I am interested in the Greek language", \for my course arguments,
assignments", \looking for shoes and clothing in Portuguese and French", \because the images I
wanted were provided in English [Luganda native speaker]", \I was looking for an image of a region
of Poland" and last \because language was not an issue". The user who who had not searched for
images on the web in other languages justi ed it as \not necessary".</p>
        <p>In addition to the users' previous knowledge and experience with Flickr, only nine out of ten
users answered this question. Three of whom stated that \Yes" they have used Flickr in the past.
When asked the reason why they have used it, one stated for searching images, one \just to see
what it is" and the third user stated for uploading, sharing and searching images. The other six
users gave the following justi cation for not having used Flickr before: \because I was not aware
of it", \not interested, security issues", \never needed to", \I don't Know what Flickr is". The
participants were evenly assigned to the conditions in the experiment with no di erence in gender,
age, and prior knowledge of Flickling interface.
2.3</p>
      </sec>
      <sec id="sec-2-3">
        <title>Task</title>
        <p>Our users were asked to nd a given image which it was not annotated from Flickr using the
Flicking interface. The users did not know in advance in which of the six languages (English,
German, Dutch, Spanish, French, Italian) the image was described enforcing them to use both
monolingual and multilingual features to nd the given image. Each of our 10 users was asked
to search and retrieve the rst three given images after login by using all the features and the
help instructions of the Flickling interface. The images presented to users were not controlled but
given randomly from a set of 100 stored in the Flickling database.
2.4</p>
      </sec>
      <sec id="sec-2-4">
        <title>Retrospective Thinking Aloud</title>
        <p>
          Retrospective thinking aloud is a widely used method for usability testing of software and
interfaces. Its basic principle is to ask from potential users to complete a certain task with the testing
object in question and to describe their thoughts and actions afterwards on the basis of a video
recording their task performance [
          <xref ref-type="bibr" rid="ref5">5</xref>
          ]. This method focuses on peoples' cognitive processes after
having completed a speci c task. It is a method that enables the users and not the experts to
point out the problems concerning the test object in the usability test.
        </p>
        <p>The use of retrospective thinking aloud method to carry out our study, like any other method,
entails both drawbacks and bene ts. In particular, it allows users to complete the task in their own
way and pace, spending as much time as they wish and are therefore not likely to perform better or
worse than usual. In addition, it enables the exact recording of the time spent for each part of the
task, as it re ects the real time that a user has spent on completing the task. Moreover, it provides
the possibility to users to re ect on their own actions while using the test object and highlight
particular causes or milestones that they have personally encountered. In addition, retrospective
thinking aloud is considered to be an appealing method for conducting tests across languages as it
is less di cult for users to disclose their thoughts in a foreign language after completing the task.</p>
        <p>Apart from bene ts, the use of retrospective thinking aloud has also some drawbacks the most
important of which are: a) the duration of every user session varies according to the time that
user will spend on completing the task plus the time that user will need to describe the video in
retrospect and b) there is also risk that users may forget what they were thinking during speci c
phases of the task or withhold others for reasons of social desirability.</p>
        <p>In this context, we have used the Camtasia Studio (v.5.1), a premiere screen recorder, in order
to capture the users' search sessions in individual videos and a digital recorder for the retrospective
thinking aloud and for the individual interviews that followed.
2.5</p>
      </sec>
      <sec id="sec-2-5">
        <title>Observation</title>
        <p>In addition to retrospective thinking aloud, observation was also adopted. The observation method
was used to form speci c questions regarding preselected research areas of the test object
(translations, layout of the interface, etc) in an attempt to shed light on speci c behaviours of the users on
speci c occasions. A form was created to assist the work of the facilitator at focusing on speci c
areas of interest and at the same time re ecting on users' behaviour. This form was categorized
according to the areas that they were to be tested. Every category had a set of prede ned
questions/ remarks that the facilitator had to ll in according to user's behaviour each time and write
additional comments for the questions to be asked to users during individual interviews. The
facilitator, one of the organizers of our study, had to ll in a set of three forms, one for each image
of the task, for each user. These forms were coded according to the number assigned to each user
and the order of the images (eg. 01/01, 01/02, 01/03).
2.6</p>
      </sec>
      <sec id="sec-2-6">
        <title>Interviews</title>
        <p>The last part of the study consisted of small scale individual interviews with every user after the
completion of the retrospective thinking aloud. The interviews lasted no more than 10 minutes
for every user. The questions asked varied according to user's answers to the questionnaire, search
session, retrospective thinking aloud and the notes gathered throughout the experiment. The main
goal of these questions was to clarify speci c actions of the user's image seeking behaviour during
the search session and expressions that the user used to describe what he/she was doing.
2.7</p>
      </sec>
      <sec id="sec-2-7">
        <title>Experimental Procedure</title>
        <p>The study was carried out in 10 individual sessions, which they were all held in the same lab
and each lasted from one to two hours approximately. During each session, users were given
general instructions about the way that the study will be carried out and about the task that they
had to complete. These instructions were read to each user explaining that there were no more
instructions to be given throughout the session and the facilitator was there to observe only. After
that, users were asked to ll in the questionnaire on personal details and prior experience. Users
were then instructed to register, login and start completing the task while screen recording software
was taping the computer screen. Having done that, users were asked to watch the recorded session
that was played back to them and describe what they were doing and what they were thinking
in retrospect. Finally, a no more than 10 minutes semi-structured interview was carried out with
each user.
2.8</p>
      </sec>
      <sec id="sec-2-8">
        <title>Processing of the data</title>
        <p>Once the 10 sessions were completed, transcripts were created based on users' retrospective
thinking aloud and interviews, as well as analysis of the questionnaires and comments on video
recordings of the users' search sessions. The analysis of the data gathered focused on the way that users
were interacting and using the Flickling interface and its features to complete the given task.</p>
        <p>The transcripts of the retrospective thinking aloud and interviews, as well as the video sessions
and observation notes were examined speci cally to identify the users decision to choose a speci c
interface (monolingual/ multilingual) and identify the extent of the role played by languages and
translations in forming their image seeking strategy. These parts were then grouped, when possible,
to enable better presentation and discussion of the ndings.</p>
        <p>In addition, users also occasionally experienced technology problems, such as trouble with
the function of the interface (problem messages coming up, interrupting the search and thought
process of the users), the ickering of the cursor due to the software used to record users' search
sessions. These problems were excluded from the study.
3</p>
      </sec>
    </sec>
    <sec id="sec-3">
      <title>Findings</title>
      <p>The analysis of data gathered by retrospective thinking aloud, video recordings and interviews
focused on three distinct areas, such as: a) to identify the reasons that determined users' choice
over a speci c interface, b) if and/or to what extent users were thinking about languages when
searching for images and c) if and/or to what extent they were paying attention to translations
and how helpful they proved to be for nding the images. The ndings will be presented according
to the three research questions in the subsections 3.1, 3.2 and 3.3 consecutively.
3.1</p>
      <sec id="sec-3-1">
        <title>Reasons</title>
        <p>The Flickling was providing to users with two di erent interfaces, monolingual and multilingual
in order to cope with the problem of searching across languages for the target three images. This
study's rst research question aims at identifying the reasons that determined each time users'
choice of a speci c interface (monolingual/ multilingual).</p>
        <p>Out of the ten users, two used only the monolingual interface and the rest switched between
interfaces. The reasons our ten users gave for their behaviour in the thinking aloud process and
interviews are stated below.</p>
        <sec id="sec-3-1-1">
          <title>1. Only Monolingual Interface</title>
          <p>Two out of ten users did not use at all the multilingual interface, even though they were
given images to search in a language unknown to them. The rst user, an English native
speaker with basic knowledge in French when informed by the system that the image was
annotated in French, stated while still on the monolingual interface: \I kindly instantly gave
up, because I am not good in French...I realized that I am never going to nd it...So, I
decided to give up". When asked why the subject didn't use the multilingual interface, the
subject answered: \Because I did not trust my abilities with other languages, to be able
to put the decent search words in...Because I did not know the keywords to search in other
languages". As a nal remark, the subject added: \I was not con dent with the languages".
The second user, a Luganda native speaker with no knowledge of French, stated: \I went for
the hint and it said that the image is described in French. So, I felt there is no need to...Well,
I thought, I do not speak French, I can't understand that". When asked why the subject
did not use the multilingual interface, answered: \If I knew how to use another language,
then I could use the multilingual and access the same image in another language. But
because my rst language is not accessible in there [Flickling], then I thought I should keep
to monolingual to where I know what I am looking for". When asked how the subject was
planning to cope with the problem of searching a French annotated image on monolingual
interface by using English keywords, the subject stated: \I thought that the image was not
available and all images should be described in English [as well]. So, I thought that it was
inaccessible...that I could not get it".</p>
        </sec>
        <sec id="sec-3-1-2">
          <title>2. Switching between Monolingual &amp; Multilingual Interface</title>
          <p>The remaining eight users switched between monolingual and multilingual interfaces in order
to complete the given task. A variety of reasons to justify these actions were reported by the
users during the retrospective thinking aloud process and interviews. In particular, users
identi ed the following reasons why: \In order to increase or decrease the number of results,
depending on the results that I had on the beginning of my search", \I have chosen to use the
multilingual interface because I assumed that it would give me the highest possible number
of relevant results in relation to my query", \I am trying to nd the right combination
of keywords", \Because of the setting of the image...I believe that this system, if you know
where the picture was from, or if you know the place then you can like recognize the language
in which you can type in", \I was looking to isolate words and translate them", \Simple
because I wasn't getting any of the results that I wanted", \I tried to increase my chances of
getting the image...I am widening my possibilities", \I am just trying out the system", \So,
it was not there [monolingual English], I guess it was in other language" and lastly: \For
me the problem was more kind of how to nd where the image was from".
Also, hints played a signi cant part in users' choice over an interface. As stated by the
users: \I switched to monolingual because the hint told me that the image was described in
English", \ok, I have learned about the hints, so I gave up and asked for a hint...O I went
to multilingual to ask to translate...and then I went back to monolingual and searched for
it" and \I went to ask for a hint on language just in case because that seamed to save me
lots of time".</p>
          <p>Two users, both English native speakers, stayed on multilingual interface though after taking
the hint, they both knew that their image was described in English. When asked why
they haven't switched to monolingual, they gave the following explanations for their choice
consecutively: \Because I did not think that would make any di erence, because I was
assuming that it is in English as well" and \Well, because I was there. I did not realize
that...I thought, to be honest, I thought, it's not going to make that much di erence really.
It is set to do a search in English, so if it does search in other languages that does not make
any di erence...is not going to increase my chances in monolingual English...maybe, it would
but I don't know that".</p>
          <p>There were also some cases that although users were seemingly using a speci c interface
(monolingual or multilingual) they stated during retrospective thinking aloud process and
con rmed afterward with the interviews that: \I did not really, even think about it. I
was just...I was at that point...I was thinking about getting this title", \I was not paying
attention to the fact that it was multilingual. Maybe, I forgot about that and left it as it
was" and \I was so focused on trying to see how to describe the image that I was not paying
attention to the interface".
3.2</p>
        </sec>
      </sec>
      <sec id="sec-3-2">
        <title>Role of Languages</title>
        <p>The second research question that our study was set out to explore was if and/or to what extent
languages were forming the image seeking behaviour of our users. As already stated, our users
did not know in advance in which of the six languages (English, German, Dutch, Spanish, French,
Italian) supported by the system the given images were described. The task was set in that way
so users had to include the element of di erent languages. As a consequence, a set of di erent
behaviours were identi ed which can be grouped in the following, again through the analysis of
the data gathered from retrospective thinking aloud and interviews:
1. Two users out of ten used only the monolingual interface searching in English, although they
knew that the images may not be described in English (see subsection 3.1.1). In particular,
the English native speaker stated: \My French are not good, so I decided to give up because
I could not nd the appropriate translations for my keywords, so I was never to nd it". The
subject also added: \I was not con dent with the languages". The Luganda native speaker
admitted that: \I would not search an image in any other language; I would only search
images in English. If I would search images from my home country, from my background,
then I would use my rst language. But any other image, I would search in English". When
asked if the subject was thinking about languages while searching, the user answered: \No.
It did not...because when I am searching for images on the Internet, I normally get them in
English because I imagine that...I guess it's a little bit of arrogance, I speak English and I
imagine that images...That if you put them in Internet, they should have English tags".
2. The other eight users who were switching between monolingual and multilingual interfaces,
can be divided in two groups: a) those who were thinking about languages and b) those for
whom languages were not a variable when performing the given task. In particular, four of
them stated clearly during retrospective thinking aloud that: \Now, I made the relationship
of country, Florida...I write them [keywords] in English", \I had the feeling that the building
which I recognized, was described in German", \It was not within my results, so, I guessed
that it is in other language", \Because by looking at the tortoise had written on it...it was
written in English. So, I assumed that it would be in English...And I was also thinking at
this time, I wonder if it is English or not...because the child got a little blue and red hat and
I was thinking, maybe the child is French...Yes, I changed into multilingual because I think
that maybe it is French, with the outside possibilities that it might be Italian" and lastly
\Well, that's probable a bit Anglo-centrism. You know, well, it is a picture in England".
On the other hand, the remaining four users when asked if they were thinking about
languages during the task, they said that: \To be honest, I was not thinking about languages...I
did not consider it a variable that in uences my results", \I did not bother about languages...I
did not really think about them...I did not focus on languages while performing my searches.
Maybe, because I am not used to, is not widely used or maybe I am not using languages
when retrieving information on the web", \I was not taking languages under consideration
when searching for the images" and \For me it was not a question of language...In my mind
language was a very small factor in there [Flickling]. It did not really play any important
role".
3.3</p>
      </sec>
      <sec id="sec-3-3">
        <title>Role of Translations</title>
        <p>All users were given a minimum one image out of three which was described in a language unknown
to them. The multilingual interface was provided to cope with this problem and help retrieve the
image. The third and last research question of this study was to examine the use of translations
and the in uence of translations on the users' information seeking behaviour.</p>
        <p>We are obliged at this point to exclude the two users who used only the monolingual interface
and the four users who used the multilingual interface but with no thought to the translations.
The remaining four users tried both the monolingual and multilingual interfaces driven by the need
to identify the language of the image and the appropriate keywords to retrieve the given images.
In particular the four users, when asked if they were paying attention to the translations, stated:
\Yes, but it did not translate anything. I thought like, it did not give me anything, because it did
not translate anything", \Yes, I did use them", \Yes, at this point I am trying to gure out how
this translation thing works" and \Yes, I was paying attention to the translations".</p>
        <p>In addition, when users asked if they could judge the translations that were given to them,
users answered: \Overall, I had the feeling that the translations of the system were not that
good...I switched to monolingual because the translations were not doing anything", \I would
trust the system to give me the right translations...I would have to for languages unknown to me",
\Because I went for the languages that I had a vague idea about and it did not tell me something
that I did not really know" and last \I was not satis ed with the German translations because I
can understand German...it's not the right word in German for a man. So, it should have been
something else...in Dutch I don't know what the translation is, so, I had to accept it, whatever
it is...Yes, I was satis ed [Dutch translations] because the computer knows the Dutch language
better than I do... maybe that's not the best translation, so, I just had to accept it. There was
anything that I could do about it really".</p>
        <p>Finally, when the users were asked if the translations were helpful in terms of actually
contributing to the retrieval of the image, users stated: \Ok, I have got the translations but they
are not doing anything to me...at the end, I totally disregard the translations", \I think that I
stop searching for translations, when I stop having much con dence that it was bringing me the
right translations...So, I got used to asking for a hint...I thought I am gonna ask for another hint
now and then it would tell me in what it was translated it, not necessarily what translation [the
system] has given me", \The words that I was trying to isolate like particular words like London,
the di erent translations there were not coming up...and what it was saying, like gigante in Italian
for giant, it told something that I already knew. So it was not isolating the words in the way that
I wanted it to. It was just telling me the adjectives were, which I did not really need" and nally
\At the end I was not paying attention to the translations, I was purely interested in nding the
image as quickly as possible because once more I did not think the translations would necessarily
help me".
4</p>
      </sec>
    </sec>
    <sec id="sec-4">
      <title>Discussion</title>
      <p>
        The evaluation of CLIR e ectiveness often does not involve the end user. On the one hand, it
may be assumed that since the translation is automated the user has no role to play or possibly
that the user has no interest in the translation, providing the system is e ective. On the other
hand, the non trivial challenges posed in the e ort in designing realistic task scenarios, recruiting
participants, analyzing large amounts of data to obtain user assessments or to observe search
behaviour can be prohibitive. However, we take the view of Petrelli [
        <xref ref-type="bibr" rid="ref9">9</xref>
        ] that e ective system
design must be in accordance with the end users' needs and to best assist users involved in
crosslanguage information retrieval we need to understand their behaviours and the search problems
they face. Petrelli's study of users involved in CLIR presented a number of interesting ndings. In
particular, the users preferred the interface which hid the translation and that language knowledge
and sight of the translation a ected search behaviour.
      </p>
      <p>The present study shows that users form image seeking behaviours according to a series of
reasons from which thinking about languages is the least important one. In addition, users were
so focused on completing the task that were not paying attention to interfaces, languages and
translations though these were factors that a ected and prede ned the results. In particular,
regarding the rst research question, the reasons why two users used only the monolingual, we
conclude they were not feeling con dent of their language skills and they were not used to searching
for images in languages other than English. These users would use the multilingual interface only
if they could speak the language in which they were searching and that one would be other than
English.</p>
      <p>The other eight users described a variety of reasons from which languages played a small part.
Users' need to minimize or maximize the number of results was the driving forces of switching
between interfaces. Moreover, they were more concerned with identifying the images in terms
of the keywords and in the right combination. In the whole process, languages were playing an
insigni cant role. Only four users were trying to identify the language of the images from its context
and use it to their bene t. Some were treating the multilingual interface as a translator, trying
to isolate speci c words, translate them on the multilingual interface and use the translations to
retrieve the images on the monolingual interface. Another user stated that: \I think I just saw it
as a translation tool and not as an integrated translation thing that already was retrieving images.
I did not really use it in this way because in my mind, it was only translating my keywords". The
Hint feature was also a factor forming users' image seeking behaviour to a large extent. Users
became accustomed to using the hints after a few minutes of unsuccessfully searching for the image.
On the whole, it would appear that users were so focused on completing the task, \obsessed" (as a
user stated) of nding the images that even from the beginning of the task, they were not thinking
really which interface they are going to use and for what reason. Even users who were concerned
about languages, at the end of the task, also admitted that they were not paying much attention
to the interfaces because they thought that it was not making any di erence.</p>
      <p>Going back to Petrelli's ndings, our users did not express any preference in hiding or not the
translations, though they were not satis ed with the position of the translations. They stated
that the translations were not obvious because they were on the right side and not on the left,
close to the search box. They commented even the way that the translations were presented. It
was not clear to them how the translations were working or how they could write their query and
get translations. They also stated that it was taking some time to gure out how the multilingual
interface was built and how they could use the translations to their bene t. Since it was not
clear that the system was retrieving both their search terms and the translations, users were
typing both their search terms and the translations provided in the search box and running the
search again. In short, they used it as a translator tool and not as an integrated feature of the
multilingual interface. However, at no point, did the users state that the presence of translations
was distracting or that they would prefer them to be hidden. On the contrary, they said that they
used them in order to retrieve relevant results and disregarded them when they were not getting
the results that they were expecting or hoping for.</p>
      <p>Both Petrelli's and our study aimed to explore user information seeking behaviour in
crosslanguage information retrieval but from a di erent perspective and for di erent reasons. As a
consequence, the methodologies adopted in both studies di er. Petrelli used mainly observation
and interviews whereas we used retrospective thinking aloud and interviews as the main methods
to record user behaviour, supplemented with questionnaires and observation to further verify
our ndings. Our aim was to derive ndings entirely on users' thoughts, comments and search
behaviour rather than depending on the facilitator's observations, interpretations and questions
asked. Although, in retrospective thinking aloud there was some risk of users not remembering
what exactly they were thinking, our users were always able to recall what caused them search
for an image in such a way. Not only did they justify their actions, they also revealed details
about their thought process and the reasons that motivated them or made them feel distracted,
uncomfortable or even bored while using Flickling. Our study based on retrospective thinking
aloud revealed a complex picture of the in uence (or not) of language skills and con dence therein
and of perceptions of the role of the multilingual interface, language and translations in image
retrieval. Most revealing and of potential interest to future study of users of CLIR is the nding
that less than half of our users appeared to consider identi cation of the language to be essential
in retrieving the image. The majority either lacked con dence in using di erent languages or
were so focused on nding the given images and completing the task that were not thinking at all
about languages. Indicative of this was the comment \...completing the task successfully. What
was success for me? That you nd the image. In any way I possible could. I was not focusing on
translations...I thought my task is to nd that image and I will do whatever I could to nd it". Of
those for which languages played a signi cant role in the process of identifying keywords to search
for the images, the translations were judged to be poor as either the translations were not coming
up, were not corresponding to the users' keywords or were judged to be resulting in the retrieval of
irrelevant results. As a consequence, users were losing interest and trust in translations resulting
in no usage of them or not paying attention to them.</p>
      <p>One of the initial aims of this study was to try and look in greater detail at how working
with the translations a ected search behaviour with regards to the actual search terms entered by
users. Unfortunately, this study could not reach a conclusion because only four out of ten users
used the translations and this was in a way not anticipated. The reasons why this happened are:
the Experiment Design: the fact that users were given a speci c task to complete, in a
context of a game, was so overwhelming that they were not thinking about anything else
other than carrying out the task. The task to nd clues for unambiguous query terms from
the images was su ciently challenging in the Flickling interface that perhaps our users chose
to ignore the language and translations. Our users were even feeling that they were failing
the task or disappointing the organizers in some way when they had to give up. As a result,
completing the task was regarded as such a challenge that they were not paying attention
to anything else.
the Flickling Interface: the way that the Flickling interface was developed, providing the two
di erent interfaces, monolingual and multilingual, to search across languages created some
issues in studying users' information seeking behaviour. The option of the two interfaces
created a confusion to most of the users because they are accustomed to using only one
search box on one interface. In addition, users had to interpret what monolingual and
multilingual meant, something that again users are not accustomed in doing when using a
search engine. Others driven by this fact were not paying attention to interfaces, they were
just using what was in front of them without making any informed choice. The users who
used the multilingual interface did so for various reasons but, as stated, languages played
an insigni cant role. The four users who used the translations to enhance their chances in
nding the target image rather than viewing them as an integrated feature of the multilingual
interface could not regarded to be forming their search strategy based on the translations.
Although, it would have been interesting to look at how translations in uenced or not users'
information seeking behaviour because of the reasons stated above we were not able to collect
data as we might have hoped.
5</p>
    </sec>
    <sec id="sec-5">
      <title>Conclusion</title>
      <p>This study aimed at investigating the users' image seeking behaviour when retrieving a known,
non-annotated image in Flickling. The task assigned to ten users was to retrieve the rst three
images given after rst login. The images could have been described in any of the six languages
(English, German, Dutch, French, Spanish, Italian) supported by the Flickling. As a consequence,
users had to use monolingual and multilingual interface to search across languages and retrieve
the images.</p>
      <p>This study used a combination of four di erent methods: a) questionnaires, b) retrospective
thinking aloud, c) observation and d) interviews. These both quantitative and qualitative methods
contributed to meeting the research questions of our study. In particular, we identi ed the reasons
why two of our users were choosing to search only on the monolingual interface and the eight
switching between interfaces. We demonstrated that only four users were thinking about languages
when trying to retrieve the given images while the rest of our users were more preoccupied with
nding the images and completing \successfully" the task. Consequently, we showed that only
these four users were paying attention to translations provided by the system. These stated that
translations were not helpful or they were not making much di erence in nding the given images
since the results were irrelevant to what they were looking for.</p>
      <p>This small study has also shown that if we are to ask whether a CLIR system should display
query translations or not, then the answer is no. Our users were either not interested in the
translations or found them to be poor. However taking the ndings to such conclusion would be
foolhardy given the complexity of the activity highlighted in the users' comments that they were
so engaged in nding the image that language or translations played little or no part. Rather than
reaching rm conclusions, this small study has suggested the need for more research into users'
search behaviour with translations (and in image retrieval) if we are to design CLIR systems which
will not place additional or unnecessary cognitive demands on the user and will support e ective
search behaviour and performance.</p>
    </sec>
  </body>
  <back>
    <ref-list>
      <ref id="ref1">
        <mixed-citation>
          [1]
          <string-name>
            <given-names>J.</given-names>
            <surname>Artiles</surname>
          </string-name>
          ,
          <string-name>
            <given-names>J.</given-names>
            <surname>Gonzalo</surname>
          </string-name>
          ,
          <string-name>
            <given-names>F.</given-names>
            <surname>Lopez-Ostenero</surname>
          </string-name>
          , and
          <string-name>
            <given-names>V.</given-names>
            <surname>Peinado</surname>
          </string-name>
          .
          <article-title>Are users willing to search cross-language? an experiment with the ickr image sharing repository</article-title>
          .
          <source>In CLEF</source>
          , pages
          <volume>195</volume>
          {
          <fpage>204</fpage>
          ,
          <year>2006</year>
          . http://www.clef-campaign.org/2006/working_notes/workingnotes2006/ artilesCLEF2006.pdf.
        </mixed-citation>
      </ref>
      <ref id="ref2">
        <mixed-citation>
          [2]
          <string-name>
            <given-names>M.</given-names>
            <surname>Braschler</surname>
          </string-name>
          ,
          <string-name>
            <given-names>G. Di</given-names>
            <surname>Nunzio</surname>
          </string-name>
          ,
          <string-name>
            <given-names>N.</given-names>
            <surname>Ferro</surname>
          </string-name>
          ,
          <string-name>
            <given-names>J.</given-names>
            <surname>Gonzalo</surname>
          </string-name>
          ,
          <string-name>
            <given-names>C.</given-names>
            <surname>Peters</surname>
          </string-name>
          , and
          <string-name>
            <given-names>M.</given-names>
            <surname>Sanderson</surname>
          </string-name>
          .
          <article-title>From clef to trebleclef: promoting technology transfer for multilingual information retrieval</article-title>
          .
          <source>In Second DELOS Conference on Digital Libraries, 5-7 December</source>
          <year>2007</year>
          , Tirrenia,
          <source>Pisa (Italy)</source>
          , pages
          <fpage>1</fpage>
          <issue>{7</issue>
          ,
          <year>2007</year>
          . http://www.trebleclef.eu/getfile.php?id=
          <fpage>38</fpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref3">
        <mixed-citation>
          [3]
          <string-name>
            <given-names>P.</given-names>
            <surname>Clough</surname>
          </string-name>
          ,
          <string-name>
            <given-names>A.</given-names>
            <surname>Al-Maskari</surname>
          </string-name>
          , and
          <string-name>
            <given-names>K.</given-names>
            <surname>Darwish</surname>
          </string-name>
          .
          <article-title>Providing multilingual access to ickr for arabic users</article-title>
          .
          <source>In CLEF</source>
          , pages
          <volume>205</volume>
          {
          <fpage>216</fpage>
          ,
          <year>2006</year>
          . http://www.clef-campaign.org/2006/working_ notes/workingnotes2006/cloughCLEF2006.pdf.
        </mixed-citation>
      </ref>
      <ref id="ref4">
        <mixed-citation>
          [4]
          <string-name>
            <given-names>P.</given-names>
            <surname>Clough</surname>
          </string-name>
          ,
          <string-name>
            <given-names>J.</given-names>
            <surname>Gonzalo</surname>
          </string-name>
          ,
          <string-name>
            <given-names>J.</given-names>
            <surname>Karlgren</surname>
          </string-name>
          ,
          <string-name>
            <given-names>E.</given-names>
            <surname>Barker</surname>
          </string-name>
          ,
          <string-name>
            <given-names>J.</given-names>
            <surname>Artiles</surname>
          </string-name>
          , and
          <string-name>
            <given-names>V.</given-names>
            <surname>Peinado</surname>
          </string-name>
          .
          <article-title>Large-scale interactive evaluation of multilingual access systems: the iclef ickr challenge</article-title>
          .
          <source>In Workshop on Novel Methodologies for Evaluation in Information Retrieval, 30 March</source>
          <year>2008</year>
          , Glasgow, Scotland,
          <year>2008</year>
          . http://nlp.uned.es/iCLEF/ECIR-evaluation-workshop.pdf.
        </mixed-citation>
      </ref>
      <ref id="ref5">
        <mixed-citation>
          [5]
          <string-name>
            <given-names>M.</given-names>
            <surname>Haak</surname>
          </string-name>
          , M. de Jong, and
          <string-name>
            <given-names>P. J.</given-names>
            <surname>Schellens</surname>
          </string-name>
          .
          <article-title>Retrospective vs. concurrent think-aloud protocols: testing the usability of an online library catalogue</article-title>
          .
          <source>Behaviour &amp; Information Technology</source>
          ,
          <volume>22</volume>
          (
          <issue>5</issue>
          ):
          <volume>339</volume>
          {
          <fpage>351</fpage>
          ,
          <year>2003</year>
          . http://www.students.cs.uu.nl/people/jpwester/WO1/Artikelen/ p339.pdf.
        </mixed-citation>
      </ref>
      <ref id="ref6">
        <mixed-citation>
          [6]
          <string-name>
            <given-names>J.</given-names>
            <surname>Karlgren</surname>
          </string-name>
          ,
          <string-name>
            <given-names>J.</given-names>
            <surname>Gonzalo</surname>
          </string-name>
          , and
          <string-name>
            <given-names>P.</given-names>
            <surname>Clough</surname>
          </string-name>
          .
          <article-title>iclef 2006 overview: Searching the ickr www photo-sharing repository</article-title>
          .
          <source>In CLEF</source>
          , pages
          <volume>186</volume>
          {
          <fpage>194</fpage>
          ,
          <year>2006</year>
          . http://eprints.sics.se/321/ 01/iclef-2006
          <string-name>
            <surname>-</surname>
          </string-name>
          overview-v4.
          <fpage>pdf</fpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref7">
        <mixed-citation>
          [7]
          <string-name>
            <given-names>Fredrik</given-names>
            <surname>Olsson</surname>
          </string-name>
          and
          <string-name>
            <given-names>Jussi</given-names>
            <surname>Karlgren</surname>
          </string-name>
          .
          <article-title>Trusting the results in crosslingual keyword-based image retrieval</article-title>
          .
          <source>In Proceedings of Evaluation of Multilingual and Multi-modal Information Retrieval, 7th Workshop of the Cross-Language Evaluation Forum</source>
          ,
          <string-name>
            <surname>CLEF</surname>
          </string-name>
          <year>2006</year>
          , Alicante, Spain,
          <source>September 20-22</source>
          ,
          <year>2006</year>
          , Revised Selected Papers, page 3,
          <string-name>
            <surname>Alicante</surname>
          </string-name>
          , Spain,
          <year>2007</year>
          . http://eprints.sics.se/319/01/iclef-2006
          <source>-SICS.pdf.</source>
        </mixed-citation>
      </ref>
      <ref id="ref8">
        <mixed-citation>
          [8]
          <string-name>
            <given-names>C.</given-names>
            <surname>Peters</surname>
          </string-name>
          .
          <article-title>Comparative evaluation of cross-language information retrieval systems</article-title>
          .
          <source>In From Integrated Publication and Information Systems to Virtual Information and Knowledge Environments</source>
          , pages
          <volume>152</volume>
          {
          <fpage>161</fpage>
          ,
          <year>2005</year>
          . http://dienst.isti.cnr.it/Dienst/Repository/2.0/ Body/ercim.cnr.isti/2004-TR-43/pdf?tiposearch=cnr&amp;langver=.
        </mixed-citation>
      </ref>
      <ref id="ref9">
        <mixed-citation>
          [9]
          <string-name>
            <given-names>Daniela</given-names>
            <surname>Petrelli</surname>
          </string-name>
          , Micheline Beaulieu, Mark Sanderson, George Demetriou, Patrick Herring, and
          <string-name>
            <given-names>Preben</given-names>
            <surname>Hansen</surname>
          </string-name>
          .
          <article-title>Observing users, designing clarity: A case study on the user-centered design of a cross-language information retrieval system</article-title>
          .
          <source>JASIST</source>
          ,
          <volume>55</volume>
          (
          <issue>10</issue>
          ):
          <volume>923</volume>
          {
          <fpage>934</fpage>
          ,
          <year>2004</year>
          . http: //dblp.uni-trier.de/db/journals/jasis/jasis55.html#PetrelliBSDHH04.
        </mixed-citation>
      </ref>
    </ref-list>
  </back>
</article>