<!DOCTYPE article PUBLIC "-//NLM//DTD JATS (Z39.96) Journal Archiving and Interchange DTD v1.0 20120330//EN" "JATS-archivearticle1.dtd">
<article xmlns:xlink="http://www.w3.org/1999/xlink">
  <front>
    <journal-meta />
    <article-meta>
      <title-group>
        <article-title>Editorial for the 3rd Bibliometric-Enhanced Information Retrieval Workshop at ECIR 2016</article-title>
      </title-group>
      <contrib-group>
        <contrib contrib-type="author">
          <string-name>Philipp Mayr</string-name>
          <email>philipp.mayr@gesis.org</email>
          <xref ref-type="aff" rid="aff0">0</xref>
        </contrib>
        <contrib contrib-type="author">
          <string-name>Ingo Frommholz</string-name>
          <email>ingo.frommholz@beds.ac.uk</email>
          <xref ref-type="aff" rid="aff1">1</xref>
        </contrib>
        <contrib contrib-type="author">
          <string-name>Guillaume Cabanac</string-name>
          <email>guillaume.cabanac@univ-tlse3.fr</email>
          <xref ref-type="aff" rid="aff2">2</xref>
        </contrib>
        <aff id="aff0">
          <label>0</label>
          <institution>GESIS - Leibniz-Institute for the Social Sciences</institution>
          ,
          <addr-line>Cologne</addr-line>
          ,
          <country country="DE">Germany</country>
        </aff>
        <aff id="aff1">
          <label>1</label>
          <institution>Institute for Research in Applicable Computing, University of Bedfordshire</institution>
          ,
          <addr-line>Luton</addr-line>
          ,
          <country country="UK">UK</country>
        </aff>
        <aff id="aff2">
          <label>2</label>
          <institution>University of Toulouse, Computer Science Department</institution>
          ,
          <addr-line>IRIT UMR 5505</addr-line>
          ,
          <country country="FR">France</country>
        </aff>
      </contrib-group>
      <pub-date>
        <year>2016</year>
      </pub-date>
      <abstract>
        <p>Overview of the papers This year 15 papers were submitted to the workshop, 7 of which were finally accepted for presentation and inclusion in the proceedings. The workshop featured one keynote talk and three paper sessions. The first session discussed text and reference mining approaches while the second session focused on bibliometric and IR tools. The final position paper session gave an outlook on further research. The following briefly describes the keynote and sessions.</p>
      </abstract>
    </article-meta>
  </front>
  <body>
    <sec id="sec-1">
      <title>1 Introduction</title>
      <p>
        Following the successful workshops at ECIR 20144 and 20155, respectively, this
workshop was the third in a series of events that brought together experts of
communities which often have been perceived as different ones: bibliometrics /
scientometrics / informetrics on the one hand and information retrieval on the
other. Our motivation as organizers of the workshop started from the
observation that main discourses in both fields are different, that communities are only
partly overlapping and from the belief that a knowledge transfer would be
profitable for both sides [
        <xref ref-type="bibr" rid="ref1">1</xref>
        ]. The first BIR workshop in 2014 set the research agenda
by introducing each group to the other, illustrating state-of-the-art methods,
reporting on current research problems, and brainstorming about common
interests. The second workshop in 2015 further elaborated these themes. This third
full-day BIR workshop6 at ECIR 2016 aimed to foster a common ground for the
incorporation of bibliometric-enhanced services into scholarly search engine
interfaces. In particular we addressed specific communities, as well as studies on
large, cross-domain collections like Mendeley and ResearchGate. This third BIR
workshop addressed explicitly both scholarly and industrial researchers.
2.1
      </p>
      <sec id="sec-1-1">
        <title>Keynote</title>
        <p>
          The keynote “Bibliometrics in online book discussions: Lessons for complex
search tasks” [
          <xref ref-type="bibr" rid="ref2">2</xref>
          ] was given by Marijn Koolen from the University of
Amsterdam. Koolen explores the potential relationships between book search
information needs and bibliometric analysis. The Social Book Search Lab is introduced,
which utilizes data from Amazon, LibraryThing (LT), the Library of Congress
and the British Library. LT discussions indicate some complex search tasks. Users
catalogue, tag, and relate books to each other. The hypothesis is that reviews,
catalogues, and discussion threads could be interpreted as (implicit) co-citation
and citation structures. Analyzing comments and reviews, several information
need patterns can be identified. Koolen also discusses how the data at hand can
be utilized for information retrieval.
2.2
        </p>
      </sec>
      <sec id="sec-1-2">
        <title>Text and Reference Mining</title>
        <p>
          In their paper “Weak links and strong meaning: The complex phenomenon of
negational citations” [
          <xref ref-type="bibr" rid="ref3">3</xref>
          ], Marc Bertin and Iana Atanassova designed a method
to extract negational citations from full-text publications. They revealed the
frequency distribution of such citations appearing throughout the regular IMRaD
structure of about 80,000 PLOS papers. Qualifying the polarity of citations has
many practical applications. This valuable knowledge might inform the scientific
community about papers attracting negative feedback that should be
reconsidered and potentially retracted.
        </p>
        <p>
          In their paper “Towards a more fine grained analysis of sceintific authorship:
Predicting the number of authors using stylometric features” [
          <xref ref-type="bibr" rid="ref4">4</xref>
          ] Andi Rexha,
Stefan Klampfl, Mark Kröll, and Roman Kern aimed to chunk papers
according to stylometric features. The resulting segments were then attributed to the
corresponding author(s) listed in the byline of the paper (i.e., the individuals
who co-signed the paper). This contribution is likely to enhance paper/passage
retrieval by author name.
        </p>
        <p>
          In their paper “The references of references: Enriching library catalogs via
domain-specific reference mining” [
          <xref ref-type="bibr" rid="ref5">5</xref>
          ], Giovanni Colavizza, Matteo Romanello,
and Frédéric Kaplan enhanced a digital library by collecting references from
domain-specific reference monographs in the Humanities. Their experiment on a
corpus dedicated to the history of Venice stresses the necessity of including such
overlooked references to improve search effectiveness in such corpora.
2.3
        </p>
      </sec>
      <sec id="sec-1-3">
        <title>Tools for Bibliometric IR</title>
        <p>
          In the paper “ Bibliometrics : a publication analysis tool” [
          <xref ref-type="bibr" rid="ref6">6</xref>
          ] by Rosa
PadrósCuxart, Clara Riera-Quintero, and Francesc March-Mir, the authors present a
bibliometric data management and consultation tool that can be utilized to
study and analyze an institution’s scientific activity. The tool is able to
generate bibliometric reports on scientific outputs at different analysis levels like
author, journal, and institution. The tool includes data from various sources
like WOS/Scopus and provides different indicators like productivity, visibility,
impact, and collaboration.
        </p>
        <p>
          In the paper “Engineering a tool to detect automatically generated papers” [
          <xref ref-type="bibr" rid="ref7">7</xref>
          ]
by Nguyen Minh Tien and Cyril Labbé, the authors are focussing on detecting
fake academic papers that are automatically created. The authors work on
detection approaches based on distance/similarity measurement and introduce a
tool which is able to detect automatically generated papers, the SciDetect
system. The authors evaluate the SciDetect system against pattern matching and
Kullback-Leibler Divergence on three different text corpora.
2.4
        </p>
      </sec>
      <sec id="sec-1-4">
        <title>IR Position Papers</title>
        <p>
          In his article “Bag of works retrieval: TF*IDF weighting of co-cited works”,
Howard D. White proposes an alternative to the well-known bag of words model
called bag of works [
          <xref ref-type="bibr" rid="ref8">8</xref>
          ]. This model can in particular be used for finding similar
documents to a given seed one. In the proposed bag of works model, the tf and
idf measures are re-defined based on (co-)citation counts. The properties of the
retrieved documents are discussed and an example is provided.
        </p>
        <p>
          In their article “On the need for and provision for an ‘IDEAL’ scientific
information retrieval test collection” [
          <xref ref-type="bibr" rid="ref9">9</xref>
          ], Birger Larsen and Christina Lioma argue
there is a need for test collections tailored to bibliometric IR. They discuss
several challenges coming along with creating such a collection (e.g., regarding
size, domain-specific dissemination and retrieval, realistic queries and relevance
judgements, pooling strategies as well as format). Furthermore, procedures to
create an ideal test collection are examined.
3
        </p>
      </sec>
    </sec>
    <sec id="sec-2">
      <title>Outlook</title>
      <p>With this continuing workshop series we have built up a sequence of explorations,
visions, results documented in scholarly discourse, and created a sustainable
bridge between bibliometrics and IR.</p>
      <p>As a next iteration we will organize a Joint Workshop on
Bibliometricenhanced Information Retrieval and Natural Language Processing for Digital
Libraries (BIRNDL 2016)7 at the JCDL conference 2016. The BIRNDL
workshop will be co-organized together with the natural language processing group
of Min-Yen Kan, National University of Singapore, which includes a shared task
(the CL-SciSumm Shared Task8). The shared task tackles automatic paper
summarization in the Computational Linguistics (CL) domain.
7 http://wing.comp.nus.edu.sg/birndl-jcdl2016/
8 http://wing.comp.nus.edu.sg/cl-scisumm2016/</p>
    </sec>
  </body>
  <back>
    <ref-list>
      <ref id="ref1">
        <mixed-citation>
          1.
          <string-name>
            <surname>Mayr</surname>
            ,
            <given-names>P.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Scharnhorst</surname>
            ,
            <given-names>A.</given-names>
          </string-name>
          :
          <article-title>Scientometrics and Information Retrieval: weak-links revitalized</article-title>
          .
          <source>Scientometrics</source>
          <volume>102</volume>
          (
          <issue>3</issue>
          ) (
          <year>2015</year>
          )
          <fpage>2193</fpage>
          -
          <lpage>2199</lpage>
        </mixed-citation>
      </ref>
      <ref id="ref2">
        <mixed-citation>
          2.
          <string-name>
            <surname>Koolen</surname>
            ,
            <given-names>M.:</given-names>
          </string-name>
          <article-title>Bibliometrics in online book discussions: Lessons for complex search tasks</article-title>
          .
          <source>In: Proc. of the 3rd Workshop on Bibliometric-enhanced Information Retrieval (BIR2016)</source>
          .
          <fpage>5</fpage>
          -
          <lpage>13</lpage>
        </mixed-citation>
      </ref>
      <ref id="ref3">
        <mixed-citation>
          3.
          <string-name>
            <surname>Bertin</surname>
            ,
            <given-names>M.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Atanassova</surname>
            ,
            <given-names>I.</given-names>
          </string-name>
          :
          <article-title>Weak links and strong meaning: The complex phenomenon of negational citations</article-title>
          .
          <source>In: Proc. of the 3rd Workshop on Bibliometricenhanced Information Retrieval (BIR2016)</source>
          .
          <fpage>14</fpage>
          -
          <lpage>25</lpage>
        </mixed-citation>
      </ref>
      <ref id="ref4">
        <mixed-citation>
          4.
          <string-name>
            <surname>Rexha</surname>
            ,
            <given-names>A.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Klampfl</surname>
            ,
            <given-names>S.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Kröll</surname>
            ,
            <given-names>M.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Kern</surname>
          </string-name>
          , R.:
          <article-title>Towards a more fine grained analysis of scientific authorship: Predicting the number of authors using stylometric features</article-title>
          .
          <source>In: Proc. of the 3rd Workshop on Bibliometric-enhanced Information Retrieval (BIR2016)</source>
          .
          <fpage>26</fpage>
          -
          <lpage>31</lpage>
        </mixed-citation>
      </ref>
      <ref id="ref5">
        <mixed-citation>
          5.
          <string-name>
            <surname>Colavizza</surname>
            ,
            <given-names>G.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Romanello</surname>
            ,
            <given-names>M.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Kaplan</surname>
            ,
            <given-names>F.</given-names>
          </string-name>
          :
          <article-title>The references of references: Enriching library catalogs via domain-specific reference mining</article-title>
          .
          <source>In: Proc. of the 3rd Workshop on Bibliometric-enhanced Information Retrieval (BIR2016)</source>
          .
          <fpage>32</fpage>
          -
          <lpage>43</lpage>
        </mixed-citation>
      </ref>
      <ref id="ref6">
        <mixed-citation>
          6.
          <string-name>
            <surname>Padrós-Cuxart</surname>
            ,
            <given-names>R.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Riera-Quintero</surname>
            ,
            <given-names>C.</given-names>
          </string-name>
          :
          <article-title>Bibliometrics: a publication analysis tool</article-title>
          .
          <source>In: Proc. of the 3rd Workshop on Bibliometric-enhanced Information Retrieval (BIR2016)</source>
          .
          <fpage>44</fpage>
          -
          <lpage>53</lpage>
        </mixed-citation>
      </ref>
      <ref id="ref7">
        <mixed-citation>
          7.
          <string-name>
            <surname>Tien</surname>
            ,
            <given-names>N.M.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Labbé</surname>
            ,
            <given-names>C.</given-names>
          </string-name>
          :
          <article-title>Engineering a tool to detect automatically generated papers</article-title>
          .
          <source>In: Proc. of the 3rd Workshop on Bibliometric-enhanced Information Retrieval (BIR2016)</source>
          .
          <fpage>54</fpage>
          -
          <lpage>62</lpage>
        </mixed-citation>
      </ref>
      <ref id="ref8">
        <mixed-citation>
          8.
          <string-name>
            <surname>White</surname>
          </string-name>
          , H.D.:
          <article-title>Bag of works retrieval: TF*IDF weighting of co-cited works</article-title>
          .
          <source>In: Proc. of the 3rd Workshop on Bibliometric-enhanced Information Retrieval (BIR2016)</source>
          .
          <fpage>63</fpage>
          -
          <lpage>72</lpage>
        </mixed-citation>
      </ref>
      <ref id="ref9">
        <mixed-citation>
          9.
          <string-name>
            <surname>Larsen</surname>
            ,
            <given-names>B.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Lioma</surname>
            ,
            <given-names>C.</given-names>
          </string-name>
          :
          <article-title>On the need for and provision for an 'IDEAL' scholarly information retrieval test collection</article-title>
          .
          <source>In: Proc. of the 3rd Workshop on Bibliometricenhanced Information Retrieval (BIR2016)</source>
          .
          <fpage>73</fpage>
          -
          <lpage>81</lpage>
        </mixed-citation>
      </ref>
    </ref-list>
  </back>
</article>