<!DOCTYPE article PUBLIC "-//NLM//DTD JATS (Z39.96) Journal Archiving and Interchange DTD v1.0 20120330//EN" "JATS-archivearticle1.dtd">
<article xmlns:xlink="http://www.w3.org/1999/xlink">
  <front>
    <journal-meta />
    <article-meta>
      <title-group>
        <article-title>Extending search facilities via bibliometric-enhanced stratagems</article-title>
      </title-group>
      <contrib-group>
        <contrib contrib-type="author">
          <string-name>Zeljko Carevic</string-name>
          <xref ref-type="aff" rid="aff0">0</xref>
        </contrib>
        <contrib contrib-type="author">
          <string-name>Philipp Mayr</string-name>
          <xref ref-type="aff" rid="aff0">0</xref>
        </contrib>
        <aff id="aff0">
          <label>0</label>
          <institution>GESIS - Leibniz Institute for the Social Sciences Unter Sachsenhausen</institution>
          <addr-line>6-8 50667 Cologne</addr-line>
          ,
          <country country="DE">Germany</country>
        </aff>
      </contrib-group>
      <abstract>
        <p>The paper introduces simple bibliometric-enhanced search facilities which are derived from the famous stratagems by Bates. Moves, tactics and stratagems are revisited from a Digital Library perspective. The potential of extended versions of "journal run" or "citation search" for interactive information retrieval is outlined. The authors elaborate on the future implementation and evaluation of new bibliometric-enhanced search services.</p>
      </abstract>
    </article-meta>
  </front>
  <body>
    <sec id="sec-1">
      <title>Introduction</title>
      <p>
        In information retrieval a change away from a mainly system-oriented
perspective towards a more user-oriented perspective can be observed [
        <xref ref-type="bibr" rid="ref7">7</xref>
        ]. A main
challenge is to gain insights into the way users perform searches in state-of-the-art
Digital Libraries (DLs). During the past numerous models have been proposed
that aim at modelling the information searching and seeking behaviour, e.g. the
information seeking behaviour model proposed by Bates. According to Bates [
        <xref ref-type="bibr" rid="ref1">1</xref>
        ]
search activities can be separated into four categories: moves, tactics, stratagems
and strategies. A move is a simple search activity like entering a query term or
selecting a speci c document. Tactics are a combination of many moves like the
selection of a broader search term or breaking complex search queries into
subproblems. Bates de nes a stratagem as follows: "a stratagem is a complex of a
number of moves and/or tactics, and generally involves both a particular
identied information search domain anticipated to be productive by the searcher, and
a mode of tackling the particular le organization of that domain." A stratagem
could be for instance a "journal run" or a "citation search". Finally, a strategy
is a combination of moves, tactics and stratagems that satis es an information
need like for instance searching for related work in a speci c research area. It is
di cult to implement strategic support in an information system as strategies
involve numerous moves, tactics and stratagems as well as experience gathered
in the entire search process. Although some work has been invested in
developing strategic support [
        <xref ref-type="bibr" rid="ref4">4</xref>
        ] for our position paper we focus on moves, tactics and
stratagems. In state-of-the-art DLs moves and tactics are widely supported. For
our real life example, the DL sowiport1, users are enabled to enter search terms
1 sowiport.gesis.org
selected from a search term recommender and to select broader or narrower
terms from a thesaurus etc. [
        <xref ref-type="bibr" rid="ref6">6</xref>
        ]. Support for advanced search activities on the
stratagem or strategy level does not exist. Though the "systems le
organization" covers information needed for stratagem support like for instance journal
and citation data, we are missing prede ned stratagems on the search interface
(see Section 3). Users therefore need to perform search stratagems manually.
This requires deep knowledge of the information structure and the formulation
of complex queries [
        <xref ref-type="bibr" rid="ref3">3</xref>
        ]. This may be adequate for expert or power users but can
hardly be accomplished by novices. We think that stratagems are an essential
part of complex search tasks that need to be supported on the user interface.
We believe that DLs like sowiport can largely bene t from prede ned stratagems
like footnote chasing, journal runs and citation search (see Section 4).
2
      </p>
    </sec>
    <sec id="sec-2">
      <title>Related work</title>
      <p>The following related work section brings together some papers which have been
identi ed as key ideas of the approach outlined later in this paper.</p>
      <p>
        First of all we want to mention the concepts developed by Bates [
        <xref ref-type="bibr" rid="ref1 ref2">1, 2</xref>
        ] which
received a lot of attention in information and computer science. Her concepts
describe the mechanisms of search activities and tasks in a very generalized way,
as an information seeking model. These concepts of speci c search tactics in an
evolving search have been implemented in an academic Web environment by Fuhr
et al. [
        <xref ref-type="bibr" rid="ref4 ref5">5, 4</xref>
        ], in the project Da odil. Today many other state-of-the-art DLs like
Web of Science or PubMed support the search tactics outlined in Bates. Another
important aspect is the ongoing popularity of the cognitive approach in IR and
the inclusion of di erent forms of context and interaction (e.g. in Interactive
IR) which has been combined by Ingwersen and Jarvelin [
        <xref ref-type="bibr" rid="ref7">7</xref>
        ]. As a consequence
of their framework the authors postulate a shift away from laboratory IR with
controlled settings and without direct user engagement towards a holistic
useroriented perspective on the search process.
      </p>
      <p>
        Bibliometric techniques are not yet widely applied to enhance the retrieval
processes in DLs, although they o er value-added e ects for users [
        <xref ref-type="bibr" rid="ref10 ref9">10, 9</xref>
        ]. The
objective of the IRM project2 was to introduce and evaluate bibliometric
valueadded services for information retrieval within a heterogeneous DL environment.
The authors [
        <xref ref-type="bibr" rid="ref10">10</xref>
        ] have investigated the use of informetrics as a ranking feature in
a retrieval system. They found that using informetrics can improve the retrieval
quality. Though the results are promising this has only been evaluated using
the classic Cran eld setting. Although a prototype3 has been implemented that
supports prede ned bibliometric stratagems the larger scenario with real users
and interactions has never been evaluated.
2 http://www.gesis.org/en/research/external-funding-projects/archive/irm/
3 http://multiweb.gesis.org/irsa/IRMPrototype
      </p>
    </sec>
    <sec id="sec-3">
      <title>Motivation</title>
      <p>In the following we develop two basic positions which try to support the
approach outlined in section 4. Position A aims at the missing support for
predened stratagems on the user interface. In position B we brie y explain
contextpreserving and context free moves and how these could be supported in
information systems.</p>
      <p>Position A: Missing prede ned stratagem support on the user
interface
Like most DLs sowiport supports basic search features like facets, ltering,
different types of ranking etc. When comparing the available features with Bates's
search activities it can be seen that all these functions belong to moves and
tactics. Domain speci c search activities that could be considered as a stratagem are
not supported on the interface. Thus, they need to be composed and executed
manually by the user. This requires knowledge of the domain and the
underlying data structure. In DLs for example users need to be aware that papers are
often published in journals or that co-cited documents are often related to each
other. One problem with stratagems is that they are only used by experts that
have experience in the given domain. A novice searcher may not be aware of
certain stratagems which may result in limiting his/her search activities to moves
and tactics. We think that stratagems are an essential part of complex search
tasks that need to be supported. For this purpose we de ne a set of prede ned
stratagems that we consider useful for each level of experience (see section 4).
Open Question: Which stratagems can be supported?
Position B: Supporting context-preserving and context free moves
When browsing DLs we can distinguish between a) context-preserving and b)
context free moves. A context is de ned e.g. by a query or lter criteria. In
context-preserved browsing the context remains and is transferred into the next
move. One example for a context-preserved move is faceted browsing. A user
enters a search term and is then provided with a ranked result set. He/She could
now reduce the result set by selecting a facet item. In this case, both the initial
query (context) and the facet item are combined to a new query. Browsing
without context (b) is typically a simple move that performs a certain action in a
retrieval system without transferring the context into the next step. An example
of a context-free move may be the selection of an author appearing in the result
set. This results in a new query where the retrieval system performs a new search
looking for the selected author name. We believe that both context-preserving
and context free moves should be supported by the system. A system should
o er both functions to the user and let him/her decide based on the current
task what is best for him/her. Users should be able to decide whether they want
to perform a context free or a context-preserved move.</p>
      <p>Open Question: How can context-preserving and context free moves be
supported on the interface? How can these functions be arranged for the user?</p>
    </sec>
    <sec id="sec-4">
      <title>Bibliometric-enhanced stratagems</title>
      <p>Bibliometric-enhanced stratagems aim at facilitating domain speci c search
activities by applying bibliometric measures for re-ranking and/or rearranging
DLentities like documents, journals or authors. Stratagems as described in the
previous section can be implemented in various ways. A journal run for example
can be implemented in the most simple way by o ering a list of issues ranked by
the publication date. We think that the implementation of stratagems depends
on the current task thus, making it necessary to o er various ways of
performing a stratagem search. To this end we describe a preliminary approach for
novel bibliometric-enhanced search facilities as an extension to basic stratagem
support. In the following we discuss two implementations for stratagem
support in a DL like sowiport: an extended journal run and an extended citation
search. Both examples are described using a mockup showing how to implement
bibliometric-enhanced search facilities and how to deal with context free and
context-preserving stratagems and moves.
4.1</p>
      <p>
        Journal run
A journal can be considered as a single specialized source for nding relevant
documents from a manageable number of potential documents. On the other hand
the focus on one journal results in the exclusion of other journals that might
contain relevant documents. One way to overcome this issue is to o er di erent
modes of a journal run on the user interface. For this section two stratagems are
described (other modes are possible as well).
1) Extended journal run: starting from a ranked result set (see Figure 1) the user
can perform an extended journal run that rearranges the articles based on the
journal they were published in (see Figure 2). It can be seen that an extended
journal run changes the ranking from a document-based to a journal-based
ranking. This journal-based ranking in our example is organized according to the
impact factor measure of the journal. Using the impact factor is only one possible
way of applying bibliometric journal metrics to re-rank the results. Other
possible journal metrics are for example: h-index, g-index, etc. Each journal shows
at least all documents that were available in the previous step. The list of
documents can be expanded to all documents that are available for the particular
journal by selecting "More from journal X".
2) Context-preserving journal run: A context-preserving journal run is performed
by selecting the name of a journal from a document appearing in the result set
(see Figure 1). In this example the previous moves and tactics that were
performed before the journal run form the context of the stratagem. A context can
be for instance a combination of the query term, a lter criterion and a single
document attribute. Instead of ranking the documents in that journal by issue or
by date we perform a ranking that is based on the context. Doing so we create
a ranking that is orientated on the current search task. In comparison to the
extended journal run this stratagem is limited to one journal.
Another example for a bibliometric-enhanced stratagem is displayed in Figure
3. In this example the user performs a citation search. We de ne a context menu
from which the user can select di erent citation analysis modes. For this example
three modes are proposed.
1) A simple citation overview where the user can see a list of documents that
cite the seed document. This is a simple move that performs a new search using
the seed document as a query term looking for citations.
2) The second mode allows the user to rank the citations based on bibliographic
coupling [
        <xref ref-type="bibr" rid="ref8">8</xref>
        ]. Bibliographic coupling aims at nding related documents under the
assumption that scienti c papers are related to each other when they have one
or more references in common. We now rank the citing documents based on
the similarity in their reference lists. This way we expect documents that are
strongly related to be ranked at the top of the list.
3) The third mode is a context-preserving list of citations ranked by their
relatedness to the seed. There are numerous methods of measuring the relatedness of
the seed to the citing documents. The relatedness could for example be measured
by comparing titles or by looking for keyword overlaps between the seed and the
citing documents. This way citing documents that are related to the seed are
ranked at the top of the list.
5
      </p>
    </sec>
    <sec id="sec-5">
      <title>Open Questions</title>
      <p>In this position paper we have described a preliminary approach for two novel
bibliometric-enhanced search facilities as an extension to basic stratagem
support. We strongly believe that bibliometric-enhanced search facilities can be a
substantial part of DLs. Crucial points are: the choice of stratagems that could
be supported and how these stratagems can be arranged on the interface.
Furthermore, we need to investigate which bibliometric metrics can be integrated.
One of the main challenges will be the evaluation of the stratagems. Our ideas for
an evaluation go into two directions. We suggest a log- le based evaluation and a
user evaluation. In the former we will measure the acceptance of the stratagems
based on di erent indicators like session duration and positiv follow-up search
activities (e.g. bookmarking or printing a document). Additionally, we will
create di erent A/B tests where a number of users are given some of the prede ned
stratagems instead of the current implementation. For our user evaluation we
plan to conduct several user studies with experts. This way we want to gain
an insight into which features could be helpful and which stratagems a system
should support.</p>
    </sec>
  </body>
  <back>
    <ref-list>
      <ref id="ref1">
        <mixed-citation>
          1.
          <string-name>
            <surname>Bates</surname>
            ,
            <given-names>M.J.:</given-names>
          </string-name>
          <article-title>The design of browsing and berrypicking techniques for the online search interface</article-title>
          .
          <source>Online Review</source>
          <volume>13</volume>
          (
          <issue>5</issue>
          ),
          <volume>407</volume>
          {424 (May
          <year>1989</year>
          ), http://www.emeraldinsight.com/doi/abs/10.1108/eb024320
        </mixed-citation>
      </ref>
      <ref id="ref2">
        <mixed-citation>
          2.
          <string-name>
            <surname>Bates</surname>
            ,
            <given-names>M.J.:</given-names>
          </string-name>
          <article-title>Where should the person stop and the information search interface start?</article-title>
          <source>Information Processing &amp; Management</source>
          <volume>26</volume>
          (
          <issue>5</issue>
          ),
          <volume>575</volume>
          {591 (Jan
          <year>1990</year>
          ), http://linkinghub.elsevier.com/retrieve/pii/0306457390901039
        </mixed-citation>
      </ref>
      <ref id="ref3">
        <mixed-citation>
          3.
          <string-name>
            <surname>Booth</surname>
            ,
            <given-names>A.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Harris</surname>
            ,
            <given-names>J.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Croot</surname>
            ,
            <given-names>E.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Springett</surname>
          </string-name>
          , J.,
          <string-name>
            <surname>Campbell</surname>
            ,
            <given-names>F.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Wilkins</surname>
          </string-name>
          , E.:
          <article-title>Towards a methodology for cluster searching to provide conceptual and contextual "richness" for systematic reviews of complex interventions: case study (CLUSTER)</article-title>
          .
          <source>BMC medical research methodology 13</source>
          ,
          <issue>118</issue>
          (
          <year>2013</year>
          ), http://www.biomedcentral.com/1471-2288/13/118
        </mixed-citation>
      </ref>
      <ref id="ref4">
        <mixed-citation>
          4.
          <string-name>
            <surname>Fuhr</surname>
          </string-name>
          , N.:
          <article-title>Information Retrieval | From Information Access to Contextual Retrieval</article-title>
          .
          <source>In: Designing Information Systems. Festschrift fur Jurgen Krause</source>
          , pp.
          <volume>47</volume>
          {
          <fpage>57</fpage>
          .
          <string-name>
            <given-names>UVK</given-names>
            <surname>Verlagsgesellschaft</surname>
          </string-name>
          (
          <year>2005</year>
          ), http://www.is.informatik.uniduisburg.de/bib/pdf/ir/Fuhr 05a.pdf
        </mixed-citation>
      </ref>
      <ref id="ref5">
        <mixed-citation>
          5.
          <string-name>
            <surname>Fuhr</surname>
            ,
            <given-names>N.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Klas</surname>
            ,
            <given-names>C.P.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Schaefer</surname>
            ,
            <given-names>A.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Mutschke</surname>
            ,
            <given-names>P.</given-names>
          </string-name>
          :
          <article-title>Da odil: An Integrated Desktop for Supporting High-Level Search Activities in Federated Digital Libraries</article-title>
          .
          <source>In: 6th European Conference on Digital Libraries</source>
          , vol.
          <volume>2458</volume>
          , pp.
          <volume>157</volume>
          {
          <issue>166</issue>
          (
          <year>2002</year>
          ), http://dx.doi.org/10.1007/3-540-45747-X 45
        </mixed-citation>
      </ref>
      <ref id="ref6">
        <mixed-citation>
          6.
          <string-name>
            <surname>Hienert</surname>
            ,
            <given-names>D.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Sawitzki</surname>
            ,
            <given-names>F.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Mayr</surname>
            ,
            <given-names>P.</given-names>
          </string-name>
          :
          <article-title>Digital library research in action supporting information retrieval in sowiport</article-title>
          . D-Lib
          <string-name>
            <surname>Magazine</surname>
          </string-name>
          (
          <year>2015</year>
          ), http://dx.doi.org/doi:10.1045/march2015-hienert
        </mixed-citation>
      </ref>
      <ref id="ref7">
        <mixed-citation>
          7.
          <string-name>
            <surname>Ingwersen</surname>
            ,
            <given-names>P.</given-names>
          </string-name>
          , Jarvelin,
          <string-name>
            <surname>K.</surname>
          </string-name>
          :
          <source>The Turn, The Information Retrieval Series</source>
          , vol.
          <volume>18</volume>
          . Springer-Verlag, Berlin/Heidelberg (
          <year>2005</year>
          ), http://link.springer.com/10.1007/1- 4020-3851-8
        </mixed-citation>
      </ref>
      <ref id="ref8">
        <mixed-citation>
          8.
          <string-name>
            <surname>Kessler</surname>
            ,
            <given-names>M.M.:</given-names>
          </string-name>
          <article-title>Bibliographic coupling between scienti c papers</article-title>
          .
          <source>American Documentation</source>
          <volume>14</volume>
          (
          <issue>1</issue>
          ),
          <volume>10</volume>
          {
          <fpage>25</fpage>
          (
          <year>1963</year>
          ), http://dx.doi.org/10.1002/asi.5090140103
        </mixed-citation>
      </ref>
      <ref id="ref9">
        <mixed-citation>
          9.
          <string-name>
            <surname>Mayr</surname>
            ,
            <given-names>P.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Scharnhorst</surname>
            ,
            <given-names>A.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Larsen</surname>
            ,
            <given-names>B.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Schaer</surname>
            ,
            <given-names>P.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Mutschke</surname>
            ,
            <given-names>P.</given-names>
          </string-name>
          :
          <article-title>Bibliometricenhanced Information Retrieval</article-title>
          . In: et al. de Rijke, M. (ed.)
          <source>36th European Conference on IR Research</source>
          , ECIR
          <year>2014</year>
          , Amsterdam, The Netherlands,
          <source>April 13-16</source>
          ,
          <year>2014</year>
          . Proceedings. pp.
          <volume>798</volume>
          {
          <fpage>801</fpage>
          . Springer International Publishing (
          <year>2014</year>
          ), http://arxiv.org/abs/1310.8226
        </mixed-citation>
      </ref>
      <ref id="ref10">
        <mixed-citation>
          10.
          <string-name>
            <surname>Mutschke</surname>
            ,
            <given-names>P.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Mayr</surname>
            ,
            <given-names>P.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Schaer</surname>
            ,
            <given-names>P.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Sure</surname>
            ,
            <given-names>Y.</given-names>
          </string-name>
          :
          <article-title>Science models as value-added services for scholarly information systems</article-title>
          .
          <source>Scientometrics</source>
          <volume>89</volume>
          (
          <issue>1</issue>
          ),
          <volume>349</volume>
          {364 (Jun
          <year>2011</year>
          ), http://arxiv.org/abs/1105.2441 http://www.springerlink.com/index/10.1007/s11192-011-0430-x
        </mixed-citation>
      </ref>
    </ref-list>
  </back>
</article>