<!DOCTYPE article PUBLIC "-//NLM//DTD JATS (Z39.96) Journal Archiving and Interchange DTD v1.0 20120330//EN" "JATS-archivearticle1.dtd">
<article xmlns:xlink="http://www.w3.org/1999/xlink">
  <front>
    <journal-meta />
    <article-meta>
      <title-group>
        <article-title>Identifying Citation Contexts: a Review of Strategies and Goals.</article-title>
      </title-group>
      <contrib-group>
        <contrib contrib-type="author">
          <string-name>Agata Rotondi</string-name>
          <email>agata.rotondi@unibo.it</email>
          <xref ref-type="aff" rid="aff0">0</xref>
        </contrib>
        <contrib contrib-type="author">
          <string-name>Angelo Di Iorio</string-name>
          <email>angelo.diiorio@unibo.it</email>
          <xref ref-type="aff" rid="aff0">0</xref>
        </contrib>
        <contrib contrib-type="author">
          <string-name>Freddy Limpens</string-name>
          <email>freddy.limpens@unibo.it</email>
          <xref ref-type="aff" rid="aff0">0</xref>
        </contrib>
        <aff id="aff0">
          <label>0</label>
          <institution>Department of Computer Science and Engineering University of Bologna</institution>
          ,
          <country country="IT">Italy</country>
        </aff>
      </contrib-group>
      <abstract>
        <p />
      </abstract>
    </article-meta>
  </front>
  <body>
    <sec id="sec-1">
      <title>Italiano. Possiamo pensare ai Contesti</title>
      <p>Citazionali come tante tessere che, unite,
possono essere sfruttate per seguire
l’opinione della comunita` scientifica
riguardo ad un determinato lavoro o per
riassumerne i contenuti piu` importanti.
Questo mosaico di informazioni puo`
essere utilizzato per identificare
sinonimi specifici e Index Terms nonche` per
individuare i motivi degli autori dietro
le citazioni. Identificare il Contesto</p>
    </sec>
    <sec id="sec-2">
      <title>Citazionale ottimale e` il primo passo per</title>
      <p>numerose analisi e ricerche. Il Contesto
Citazionale e` stato definito in diversi modi
in letteratura, in relazione a differenti
scopi, domini e applicazioni. In questo
paper presentiamo le principali
dimensioni testuali di Contesto Citazionale
investigate dai ricercatori nel corso degli
anni.</p>
      <sec id="sec-2-1">
        <title>1 Introduction and Background</title>
        <p>
          Researchers consider as Citation Context (CC)
different snippets of text around a citation marker.
These differences of width influence the
applications that exploit CC as source of
information. For example,
          <xref ref-type="bibr" rid="ref30">Qazvinian and Radev (2010)</xref>
          showed that using also implicit citations (i.e.
sentences that contain information about a specific
secondary source but do not explicitly cite it) for
generating surveys, rather than citing sentences
alone, improve the results.
          <xref ref-type="bibr" rid="ref31">Ritchie et al. (2008)</xref>
          compared different widths of CC in order to find
the most appropriate window for identifying
Index Terms. They proved that varying the context
from which the Index Terms are gathered has a
significant effect on retrieval effectiveness.
          <xref ref-type="bibr" rid="ref3">Aljaber et al. (2010)</xref>
          tested different sizes of CC for
a document clustering experiment. They claimed
that a window size of 50 words from either side
of the citation marker works better than taking 10
or 30 terms or the citing sentence alone, whatever
its size is. From their analysis, relevant
synonymous and related vocabulary extracted from this
window of text, in combination with an original
full-text representation of the cited document, are
effective for document clustering. We can claim
that the issue of finding the optimal CC for a
specific application is a challenging task that interests
researchers and which is at the base of every study
that exploits the CC as a source of information.
1 With the purpose of providing a useful
background to anyone approaching this question, in the
following sections we give an overview of
different dimensions of textual CC investigated in
literature. We classified them in 3 main categories:
a) fixed number of characters b) citing sentence
c) extended context (fixed and adaptive), and we
summarized our analysis in Figure1. We focus
on the strategies to identify the correct textual CC
of a citation, nevertheless other CC related topics
have been investigated in literature as for example
citation recommendations (see
          <xref ref-type="bibr" rid="ref13">Farber (2018)</xref>
          and
          <xref ref-type="bibr" rid="ref11">Ebesu (2017)</xref>
          )
The belief of the need of a clear introductory
survey about how CC has been differently shaped in
literature came to our mind when we faced the
problem of defining the optimal CC for the
Semantic Coloring of Academic References (SCAR)
project1
          <xref ref-type="bibr" rid="ref10">(Di Iorio et al., 2018)</xref>
          . The goal of the
SCAR project is to enrich bibliographies of
scientific articles by adding explicit meta data about
individual bibliographic entries and to characterize
these entries according to multiple criteria. With
this purpose, we are studying a set of properties
to support the automatic characterization of
bibliographic entries and one of our primary source of
information is the textual content around citation
markers, i.e. the CC. We are currently
investigating on finding the best span of text for our needs.
By reviewing the literature, we realized that
different approaches correspond to different tasks and
are also related to the linguistic domain of
application. The SCAR project as well as this review are
focused on the English language but it would be
interesting to extend this study to other languages.
1http://dasplab.cs.unibo.it/index.php/scar/
2
        </p>
      </sec>
      <sec id="sec-2-2">
        <title>Fixed Number of Characters</title>
        <p>
          A good way to start exploring how the CC can be
diversely defined is to look for well known
examples. One of these is the public search engine and
digital library for scientific and academic papers
CiteSeerX2. This web platform allows users to
browse papers’ references and to read the context
in which a reference is cited. The function enables
the reading of 200 characters before and after the
citation marker. Here the choice of the CC width
is not directly related to further analysis and
applications as the purpose is the mere reading of text
by users. As
          <xref ref-type="bibr" rid="ref19">Ii et al. (2014)</xref>
          describe, CiteSeerX
uses ParsCit
          <xref ref-type="bibr" rid="ref9">(Councill et al., 2008)</xref>
          for citation
extraction. ParsCit is a freely available, open-source
implementation of a reference string parsing
package which performs reference string segmentation
and CC extraction. The size of the context is
configurable, but by default extends to 200 characters
on either side of the match. ParsCit is a well know
software and is used in different projects. For
example, the Association Of Computational
Linguistics (ACL) Anthology Network3 uses ParsCit
for curation.
          <xref ref-type="bibr" rid="ref24">Doslu and Bingol (2016)</xref>
          also used
ParsCit in their work regarding how to rank
articles for a given topic. The authors exploited the
information contained in the CC of a certain
paper for detecting important articles and providing
focused directions to access the literature about a
topic. They stated that the words that are used to
describe a cited paper stand close to the citation
marker, and this is their motivation for choosing a
fixed window size context. Before Doslu and
Bingol, also
          <xref ref-type="bibr" rid="ref7">Bradshaw (2003)</xref>
          used CC to index cited
2http://citeseerx.ist.psu.edu/index
3http://aan.how/index.php/home/about
paper for specific topics. He designed the
Reference Direct Indexing in which measures of
relevance and impact are joined in a single retrieval
metric based on the comparison of the terms
authors use in multiple CC of a document. The CC
Bradshaw used to index the documents are directly
gathered from CiteSeerX. Also the tool presented
by
          <xref ref-type="bibr" rid="ref22">Knoth et al. (2017)</xref>
          , who address the problem
of automatically retrieving and collecting CC for
a given unstructured research paper, extract a CC
window of fixed length corresponding to 300
characters before and after a citation marker. The
approach of considering as CC a fixed length
snippet around the citation marker is a naive baseline
method. It can be used to retrieve terms related to
a cited entity and the accuracy of applications that
employ it might be improved for example by
considering sentence or paragraph boundaries
          <xref ref-type="bibr" rid="ref3">(Aljaber
et al., 2010)</xref>
          . This kind of context is unsuitable if
the CC needs to be further analyzed, for example
by using syntactic parsers, or if its content have
to be represented in a coherent formal way where
the meaning and structure of sentences have to be
preserved.
3
        </p>
      </sec>
      <sec id="sec-2-3">
        <title>Citing Sentence</title>
        <p>
          Another famous platform among scholars is
Semantic Scholar4. This subjective search service
for journal articles provides several functions for
browsing papers among which the possibility of
quickly read the CC of each citation. This service
allows reading more than one excerpt of text for
each entity (when available). Each CC shown
corresponds exactly to a citing sentence, i.e. the
sentence that contains the targeted reference marker.
Implicit citations5 are also investigated by
exploiting lexical hooks and also in these cases the CC
excerpts shown are in the form of a full sentence.
The same CC window has been adopted in
several projects.
          <xref ref-type="bibr" rid="ref26">Nakov et al. (2004)</xref>
          investigated
the use of CC for semantic interpretation of
bioscience articles. Starting from the collection of the
citing sentences related to a specific cited entity
(that they call citances), they used the output of a
4https://www.semanticscholar.org
5More in details, with implicit citations we refer to those
mentions of a work where the relation cited entity-citing
entity is not provided by a citation marker but rather by a lexical
object related to the cited entity. E.g.: The heuristics based on
WordNet and Wikipedia ontologies are very sensitive to
preprocessing is an implicit citation of George A. Miller (1995).
WordNet: A Lexical Database for English. Communications
of the ACM Vol. 38, No. 11: 39-41.
dependency parser to build paraphrase expressing
relations between two named entities. As
commented before, parsers need to be fed with full
sentences in order to provide proper
representations and this work is a clear example where a
fixed length CC would not have been an
appropriate input. Also
          <xref ref-type="bibr" rid="ref12">Elkiss et al. (2008)</xref>
          focused
their research on the set of citing sentences of a
given article (named by the authors citation
summaries) testing the biomedical domain. Despite
Elkiss study did not rely on any strictly sentence
based technique (they employed cosine
similarity and tf-idf), both their hypothesis are grounded
on the importance of citing sentences boundaries.
          <xref ref-type="bibr" rid="ref32">Sula and Miller (2014)</xref>
          presented an experimental
tool for extracting and classifying citation contexts
in humanities. Their approach is based on
citing sentences from which they extracted features
(e.g. location in document) and polarity
(evaluating n-grams with a naive Bayes classifier).
          <xref ref-type="bibr" rid="ref6">Bertin
et al. (2016)</xref>
          followed a similar approach to
identify n-grams and sentiment in CC. They chose to
work on a sentence basis stating that sentences are
the natural building blocks of text and likely to
include the context of a specific reference. Starting
from citing sentences they extracted 3-grams
containing verbs, together with position in the paper
and type of section according to the IMRaD
structure in order to analyze the combination and
distribution of these features in the biomedical domain.
Citing sentence as a base unit for CC is mostly
chosen in hard sciences domains. In fact,
scientific communities have particular ways of
using language and specific conventions that reveal
clear disciplinary differences.
          <xref ref-type="bibr" rid="ref16">Hyland (2009)</xref>
          describes some of these language variations that go
from terminology differences to different citations
practices and rhetorical preferences. Writers use
different sets of reporting verbs to refer to others
work (engineers show, philosophers argue,
biologists find and linguists suggest); frequencies of
hedges and self citations, directives and n-grams
also diverge across fields. In the humanities
writers tend to include extensive referencing and build
a background for the heterogeneous readership
while in hard sciences most of the readers share a
common context with writers. This attitude
clarifies citers’ behaviors in different domains and
makes us presume that CC in humanities might
be more complex than in hard sciences.
Following these considerations, it is reasonable to
conclude that for choosing the appropriate CC width
one needs to take into account not only the task
he is going to face but also the domain of
applications and the specificity of the language. In this
sense, CC as citing sentence might not always
correspond to the entire fragment of text referring to
a targeted citation marker.
4
        </p>
      </sec>
      <sec id="sec-2-4">
        <title>Extended Context</title>
        <p>
          Extending CC beyond the citing sentence can
prove useful in many cases as illustrated by
the social networking site for researchers
ResearchGate6. Every document in this platform’s
database can be inspected according to different
prospectives. Among them, readers can browse
documents citations lists and access CC (when
available) displayed in the form of: 1 sentence
before the citing sentence + citing sentence + 1
sentence after the citing sentence. This window
size allows users to better understand the full
context of a citation without loosing any possible
informations contained in the nearby sentences.
This is particularly relevant for the task of polarity
identification of citations.
          <xref ref-type="bibr" rid="ref5">Athar and Teufel (2012)</xref>
          have shown that authors’ sentiments are most
likely expressed outside the citing sentences.
Sentiment in citations is often hidden and especially
criticism might be hedged both for politeness
and for political reasons
          <xref ref-type="bibr" rid="ref25">(MacRoberts and
MacRoberts, 1984)</xref>
          . Citing sentences are typically
neutral and in particular negative polarity occurs
in the following sentences
          <xref ref-type="bibr" rid="ref33">(Teufel et al., 2006)</xref>
          ,
see for example (from
          <xref ref-type="bibr" rid="ref29">(Platt, 1990)</xref>
          ):
        </p>
        <p>In [19, sec. 11.11], Vapnik suggests a method
for mapping the output of SVM to probabilities by
decomposing the feature space []. Preliminary
results for this method, are promising.However,
there are some limitations that are overcome by
the method of this chapter.</p>
        <p>Particularly for, but not limited to, polarity
identification tasks, a context extended to the nearby
sentences can supply the complete set of
information about a citation to applications and readers.
Sentences nearby a citing sentence can be add as
part of the CC according to a fixed schema or by
following an adaptive approach.</p>
        <p>
          6https://www.researchgate.net
4.1
Besides ResearchGate and the aforementioned
Ritchie’s work, who studied different window
sizes of CC for identifying Index Terms, also
          <xref ref-type="bibr" rid="ref23">Mei
and Zhai (2008)</xref>
          implemented a fixed extended
context for their study of summarizing articles
influence. For their impact-based summarization
task they used a 5 sentences window size, with
2 sentences before and after the citing sentence.
This technique allows to include more info in the
CC but at the same time the risk of adding noise is
high. This is why most of the literature concerning
extended CC rather provides adaptive methods.
A mention is needed to the work of
          <xref ref-type="bibr" rid="ref14">Fujiwara and
Yamamoto (2015)</xref>
          , mostly for their overall project
than for the CC retrieval approach which relies on
a very basic technique (they include the sentence
after the citing one if the reference marker is at
the end of the citing sentence and limit long citing
sentences to 240 characters before and after
citation markers). The authors built the Colil database
where CC of the life sciences domain are stored,
and made it available to users through a web-based
search service. For each resource stored in the
database, a list of CC in which the resource has
been cited is returned to the user who can easily
read how a work is perceived and used by
different authors.
4.2
        </p>
        <sec id="sec-2-4-1">
          <title>Adaptive Extended Context</title>
          <p>
            <xref ref-type="bibr" rid="ref28">O’Connor (1982</xref>
            ) was the first who investigated
the CC as a sequence of sentences - a
multisentence citing statement. His purpose was to
study the words of CC as possible improvement
for the retrieval of the related cited entities. He
wrote 16 complex and detailed computer rules (not
completely computer procedures at that time) with
linguistic, structural and more general features for
the selection of citing statements.
            <xref ref-type="bibr" rid="ref27">Nanba and
Okumura (1999)</xref>
            presented a system to support
writing surveys of a specific domain. They see the
CC as a succession of sentences where the
possible connections are indicated by 6 kinds of cue
words (anaphora, negative expression, 1st and 3rd
person pronoun, adverb, other) that they use for
retrieving the suitable CC for their system. To
identify the full span of CC,
            <xref ref-type="bibr" rid="ref20">Kaplan et al. (2009)</xref>
            presented a different method based on co-reference
chains. They built a SVM
            <xref ref-type="bibr" rid="ref8">(Cortes and Vapnik,
1995)</xref>
            classifier with 13 features (among which:
cosine similarity, gender and number agreement,
semantic class agreement etc.) that are tested in
order to find the best configuration. Results of the
classifier alone and in combination with cue-based
techniques are promising. Despite the little data
analyzed for the project, Kaplan raised some
interesting remarks about CC. Particularly, they stated
that sentences of CC are not necessarily
contiguous.
            <xref ref-type="bibr" rid="ref30">Qazvinian and Radev (2010)</xref>
            explored the
task of retrieving background information close to
explicit citations by implementing a probabilistic
inference model (Markov Random Field). Like
previous authors, they observed that the majority
of sentences related to a citation directly occur
after or before the citation or another context
sentence; however they also confirmed Kaplan’s
intuition about possible gaps between sentences
describing a cited paper.
            <xref ref-type="bibr" rid="ref5">Athar and Teufel (2012)</xref>
            tried to go further by attempting to retrieve all the
mentions of a cited entity within the full text of the
citing paper. As claimed by the authors, mentions
to a cited entity can occur in the full article and are
necessary to identify the real sentiment toward the
cited work. Their first experiment of manual
annotation proved the insight that retrieving all the
mentions of a cited entity increases citation
sentiment coverage. Also the SVM framework
implemented by the authors, despite limited to a 4
sentence window, outperformed a single sentence
baseline system.
            <xref ref-type="bibr" rid="ref2">Abu-Jbara et al. (2013)</xref>
            , with
the purpose of adding qualitative aspects to
standard quantitative bibliometrics (H-Index, G-Index,
etc.), analyzed the text surrounding a citation in
order to define the citer’s purposes and polarity. This
piece of text (CC), is retrieved with a sequence
labeling method. Starting from the citing sentence,
Abu-Jbara’s team used CRF
            <xref ref-type="bibr" rid="ref18">(Lafferty et al., 2001)</xref>
            to determine if the sentence before and the two
sentences after the citing sentences have to be
included in the CC. The features for the CRF model
are both structural (e.g. position of the current
sentence with respect to the citing sentence) and
lexical (e.g. presence of demonstrative determiners).
            <xref ref-type="bibr" rid="ref21">Kaplan et al. (2016)</xref>
            named Citation Block
Determination(CBD) the task of detecting non-explicit
citing sentences and faced it by testing various
features representing different aspects of textual
coherence. Non local mentions are excluded from
what they formalized as a binary classification task
of sentences from the citing one. They tested
different relational and entity coherence features and
their combinations. Experiments showed that the
CRF method fits better the task than the SVM
approach.
          </p>
          <p>The different works briefly described so far give
an overview of the most interesting techniques
explored by researchers. From rule-based
approaches to probability methods, the implemented
features are most of the time domain-specific
relying on particular vocabulary and on stylistic and
rhetorical habits.
4.2.1</p>
        </sec>
        <sec id="sec-2-4-2">
          <title>Citation Scope</title>
          <p>Related to the Adaptive Extended Context topic is
the identification of the Scope of a citation. So far
we have discussed different ways of including in
the CC what is outside the citing sentence but at
the same time related to it. The idea is to extend
the context. However, there are cases in which the
citing sentence does not completely refer to the
targeted citation or where the context of multiple
citations overlap. In these cases the
aforementioned approaches of CC extraction would include
noise and affect applications results. See for
instance the following example where the whole
citing sentence might produce a negative polarity
despite the neutral value of the citation:</p>
          <p>
            The negative results produced by the BoW
approach led our team to change direction and
we tested a SVM
            <xref ref-type="bibr" rid="ref8">(CORTES, 1995)</xref>
            classifier.
          </p>
          <p>
            Finding a procedure to cut out the precise scope
of a citation is a tricky and challenging task for
which little experiments have been done.
            <xref ref-type="bibr" rid="ref4">Athar (2011)</xref>
            suggested to trim the parse tree of
each citing sentence and to keep only the deepest
clause in the subtree of which the citation is a part.
            <xref ref-type="bibr" rid="ref1">Abu-Jbara and Radev (2012)</xref>
            explored 3 different
methods for identifying the scope: word
classification, sequence labeling and segment
classification. Results showed that the scope of a given
reference consists of units of higher granularity than
words. In fact, the segment classification
technique achieved the best performance. Despite the
interesting results, we agree with
            <xref ref-type="bibr" rid="ref15">Hernandez and
Gomez (2016)</xref>
            who stated that additional work is
required to improve the citation scope
identification task. The need of further research in this
field is also encouraged by the analysis of
            <xref ref-type="bibr" rid="ref17">Jha et
al. (2017)</xref>
            who performed an annotation
experiment on a sample of the ACL Anthology Network
revealing that, on average, the reference scope for
a given target reference contains only 57.63 per
cent of the original citing sentence.
5
          </p>
        </sec>
      </sec>
      <sec id="sec-2-5">
        <title>Conclusion</title>
        <p>We have reviewed what we consider the most
interesting works about CC identification in order to
provide a solid background to anyone interested in
the topic and especially to those researchers who
are facing the task of identifying the best approach
for their studies. We did not compare the
different strategies with the purpose of ranking them,
but we rather showed that there exists various
relations between a methodology and the usage,
domain, and language specificity of its possible
applications.</p>
      </sec>
    </sec>
  </body>
  <back>
    <ref-list>
      <ref id="ref1">
        <mixed-citation>
          <string-name>
            <surname>Abu-Jbara</surname>
            ,
            <given-names>A.</given-names>
          </string-name>
          , and
          <string-name>
            <surname>Radev</surname>
            ,
            <given-names>D.</given-names>
          </string-name>
          <year>2012</year>
          .
          <article-title>Reference Scope Identification in Citing Sentences</article-title>
          .
          <source>In Proc. of NAACL HLT</source>
          , (p.
          <fpage>80</fpage>
          -
          <lpage>90</lpage>
          ).
        </mixed-citation>
      </ref>
      <ref id="ref2">
        <mixed-citation>
          <string-name>
            <surname>Abu-Jbara</surname>
            ,
            <given-names>A.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Ezra</surname>
            ,
            <given-names>J.</given-names>
          </string-name>
          , and
          <string-name>
            <surname>Radev</surname>
            ,
            <given-names>D.</given-names>
          </string-name>
          <year>2013</year>
          .
          <article-title>Purpose and Polarity of Citation: Towards NLP-based Bibliometrics</article-title>
          .
          <source>In Proc. of NAACL HLT</source>
          , (p.
          <fpage>596</fpage>
          -
          <lpage>606</lpage>
          ).
        </mixed-citation>
      </ref>
      <ref id="ref3">
        <mixed-citation>
          <string-name>
            <surname>Aljaber</surname>
            ,
            <given-names>B.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Stokes</surname>
            ,
            <given-names>N.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Bailey</surname>
            ,
            <given-names>J.</given-names>
          </string-name>
          , and
          <string-name>
            <surname>Pei</surname>
            ,
            <given-names>J.</given-names>
          </string-name>
          <year>2010</year>
          .
          <article-title>Document Clustering of Scientific Texts Using Citation Contexts</article-title>
          .
          <source>Information Retrieval</source>
          ,
          <volume>13</volume>
          , (p.
          <fpage>101</fpage>
          -
          <lpage>131</lpage>
          ).
        </mixed-citation>
      </ref>
      <ref id="ref4">
        <mixed-citation>
          <string-name>
            <surname>Athar</surname>
            ,
            <given-names>A.</given-names>
          </string-name>
          <year>2011</year>
          .
          <article-title>Sentiment Analysis of Citations using Sentence Structure-Based Features</article-title>
          .
          <source>Proceedings of NAACL-HLT</source>
          , (p.
          <fpage>81</fpage>
          -
          <lpage>87</lpage>
          ).
        </mixed-citation>
      </ref>
      <ref id="ref5">
        <mixed-citation>
          <string-name>
            <surname>Athar</surname>
            ,
            <given-names>A.</given-names>
          </string-name>
          , and
          <string-name>
            <surname>Teufel</surname>
            ,
            <given-names>S.</given-names>
          </string-name>
          <year>2012</year>
          .
          <article-title>Detection of Implicit Citations for Sentiment Detection.</article-title>
          .
          <source>In Proc. of DSSD</source>
          , (p.
          <fpage>18</fpage>
          -
          <lpage>26</lpage>
          ).
        </mixed-citation>
      </ref>
      <ref id="ref6">
        <mixed-citation>
          <string-name>
            <surname>Bertin</surname>
            ,
            <given-names>M.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Atanassova</surname>
            ,
            <given-names>I.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Sugimoto</surname>
            ,
            <given-names>C.</given-names>
          </string-name>
          , and
          <string-name>
            <surname>Lariviere</surname>
            ,
            <given-names>V.</given-names>
          </string-name>
          <year>2016</year>
          .
          <article-title>The Linguistic Patterns and Rhetorical Structure of Citation Context: an Approach Using N-Grams</article-title>
          . Scientometrics,
          <volume>109</volume>
          (
          <issue>3</issue>
          ).
        </mixed-citation>
      </ref>
      <ref id="ref7">
        <mixed-citation>
          <string-name>
            <surname>Bradshaw</surname>
            ,
            <given-names>S.</given-names>
          </string-name>
          <year>2003</year>
          .
          <article-title>Reference Directed Indexing: Redeeming Relevance for Subject Search in Citation Indexes</article-title>
          .
          <source>In Proc. of ECDL</source>
          , (p.
          <fpage>499</fpage>
          -
          <lpage>510</lpage>
          ).
        </mixed-citation>
      </ref>
      <ref id="ref8">
        <mixed-citation>
          <string-name>
            <surname>Cortes</surname>
            ,
            <given-names>C.</given-names>
          </string-name>
          and
          <string-name>
            <surname>Vapnik</surname>
            <given-names>V.</given-names>
          </string-name>
          <year>1995</year>
          .
          <article-title>Support-Vector Networks</article-title>
          .
          <source>Machine Learning</source>
          ,
          <volume>20</volume>
          (
          <issue>3</issue>
          ), (p.
          <fpage>273</fpage>
          -
          <lpage>297</lpage>
          ).
        </mixed-citation>
      </ref>
      <ref id="ref9">
        <mixed-citation>
          <string-name>
            <surname>Councill</surname>
            ,
            <given-names>I.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Giles</surname>
            ,
            <given-names>C.</given-names>
          </string-name>
          , and
          <string-name>
            <surname>Kan</surname>
            ,
            <given-names>M.</given-names>
          </string-name>
          <year>2008</year>
          .
          <article-title>ParsCit : An Open-Source CRF Reference String Parsing Package</article-title>
          .
          <source>In Proc. of LREC</source>
          , (p.
          <fpage>661</fpage>
          -
          <lpage>667</lpage>
          ).
        </mixed-citation>
      </ref>
      <ref id="ref10">
        <mixed-citation>
          <string-name>
            <given-names>Di</given-names>
            <surname>Iorio</surname>
          </string-name>
          ,
          <string-name>
            <given-names>A.</given-names>
            ,
            <surname>Limpens</surname>
          </string-name>
          ,
          <string-name>
            <given-names>F.</given-names>
            ,
            <surname>Peroni</surname>
          </string-name>
          ,
          <string-name>
            <given-names>S.</given-names>
            ,
            <surname>Rotondi</surname>
          </string-name>
          ,
          <string-name>
            <given-names>A.</given-names>
            , and
            <surname>Tsatsaronis</surname>
          </string-name>
          ,
          <string-name>
            <surname>G.</surname>
          </string-name>
          <year>2018</year>
          .
          <article-title>Investigating Facets to Characterise Citations for Scholars</article-title>
          .
          <source>In Proc. of SAVESD Workshop.</source>
        </mixed-citation>
      </ref>
      <ref id="ref11">
        <mixed-citation>
          <string-name>
            <surname>Ebesu</surname>
            ,
            <given-names>T.</given-names>
          </string-name>
          , and
          <string-name>
            <surname>Fang</surname>
            ,
            <given-names>Y.</given-names>
          </string-name>
          <year>2017</year>
          .
          <article-title>Neural Citation Network for Context-Aware Citation Recommendation</article-title>
          .
          <source>In Proc. of SIGIR</source>
          , (p.
          <fpage>10931096</fpage>
          ).
        </mixed-citation>
      </ref>
      <ref id="ref12">
        <mixed-citation>
          <string-name>
            <given-names>Elkiss A.</given-names>
            ,
            <surname>Shen</surname>
          </string-name>
          <string-name>
            <given-names>S.</given-names>
            ,
            <surname>Fader</surname>
          </string-name>
          <string-name>
            <given-names>A.</given-names>
            ,
            <surname>Erkan</surname>
          </string-name>
          <string-name>
            <given-names>G.</given-names>
            ,
            <surname>States</surname>
          </string-name>
          <string-name>
            <given-names>D.</given-names>
            , and
            <surname>Radev</surname>
          </string-name>
          <string-name>
            <surname>D.</surname>
          </string-name>
          <year>2008</year>
          .
          <article-title>Blind Men and Elephants: What Do Citation Summaries Tell Us About a Research Article?</article-title>
          .
          <source>American Society for Information Science and Technology</source>
          ,
          <volume>59</volume>
          (
          <issue>1</issue>
          ), (p.
          <fpage>51</fpage>
          -
          <lpage>62</lpage>
          ).
        </mixed-citation>
      </ref>
      <ref id="ref13">
        <mixed-citation>
          <string-name>
            <surname>Farber</surname>
            <given-names>M.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Thiemann</surname>
            <given-names>A.</given-names>
          </string-name>
          , and
          <string-name>
            <surname>Jatowt</surname>
            <given-names>A.</given-names>
          </string-name>
          <year>2018</year>
          . To Cite, or Not to Cite?
          <article-title>Detecting Citation Contexts in Text</article-title>
          .
          <source>In Proc. of ECIR: Advances in Information Retrieval</source>
          , (p.
          <fpage>598</fpage>
          -
          <lpage>603</lpage>
          ).
        </mixed-citation>
      </ref>
      <ref id="ref14">
        <mixed-citation>
          <string-name>
            <surname>Fujiwara</surname>
            ,
            <given-names>T.</given-names>
          </string-name>
          , and
          <string-name>
            <surname>Yamamoto</surname>
            ,
            <given-names>Y.</given-names>
          </string-name>
          <year>2015</year>
          .
          <article-title>Colil: a Database and Search Service for Citation Contexts in the Life Sciences Domain</article-title>
          .
          <source>Biomedical Semantics</source>
          ,
          <volume>6</volume>
          (
          <issue>38</issue>
          ).
        </mixed-citation>
      </ref>
      <ref id="ref15">
        <mixed-citation>
          <string-name>
            <surname>Hernandez-Alvarez</surname>
            ,
            <given-names>M.</given-names>
          </string-name>
          , and
          <string-name>
            <surname>Gomez</surname>
            ,
            <given-names>J.</given-names>
          </string-name>
          <year>2016</year>
          .
          <article-title>Survey About Citation Context Analysis: Tasks, Techniques, and Resources</article-title>
          .
          <source>Natural Language Engineering</source>
          ,
          <volume>22</volume>
          (
          <issue>3</issue>
          ), (p.
          <fpage>327</fpage>
          -
          <lpage>349</lpage>
          ).
        </mixed-citation>
      </ref>
      <ref id="ref16">
        <mixed-citation>
          <string-name>
            <surname>Hyland</surname>
            ,
            <given-names>K.</given-names>
          </string-name>
          <year>2009</year>
          .
          <article-title>Writing in the Disciplines: Research Evidence for Specificity</article-title>
          .
          <source>Taiwan International ESP Journal</source>
          ,
          <volume>1</volume>
          (
          <issue>1</issue>
          ), (p.
          <fpage>5</fpage>
          -
          <lpage>22</lpage>
          ).
        </mixed-citation>
      </ref>
      <ref id="ref17">
        <mixed-citation>
          <string-name>
            <surname>Jha</surname>
            ,
            <given-names>R.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Jbara</surname>
            ,
            <given-names>A.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Qazvinian</surname>
            ,
            <given-names>V.</given-names>
          </string-name>
          , and
          <string-name>
            <surname>Radev</surname>
            <given-names>D.</given-names>
          </string-name>
          <year>2017</year>
          .
          <article-title>NLP-driven Citation Analysis for Scientometrics</article-title>
          .
          <source>Natural Language Engineering</source>
          ,
          <volume>23</volume>
          (
          <issue>1</issue>
          ), (p.
          <fpage>93</fpage>
          -
          <lpage>130</lpage>
          ).
        </mixed-citation>
      </ref>
      <ref id="ref18">
        <mixed-citation>
          <string-name>
            <surname>Lafferty</surname>
            ,
            <given-names>J.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>McCallum</surname>
            ,
            <given-names>A.</given-names>
          </string-name>
          , and
          <string-name>
            <given-names>C.N.</given-names>
            <surname>Pereira</surname>
          </string-name>
          ,
          <string-name>
            <surname>F.</surname>
          </string-name>
          <year>2001</year>
          .
          <article-title>Conditional Random Fields: Probabilistic Models for Segmenting and Labeling Sequence Data</article-title>
          .
          <source>In Proc. of ICML</source>
          , (p.
          <fpage>282</fpage>
          -
          <lpage>289</lpage>
          ).
        </mixed-citation>
      </ref>
      <ref id="ref19">
        <mixed-citation>
          <string-name>
            <surname>Ii</surname>
            ,
            <given-names>A.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Wu</surname>
            ,
            <given-names>J.</given-names>
          </string-name>
          , and
          <string-name>
            <surname>Giles</surname>
            ,
            <given-names>C.</given-names>
          </string-name>
          <year>2014</year>
          .
          <article-title>CiteSeerX : Intelligent Information Extraction and Knowledge Creation from Web-Based Data</article-title>
          .
          <source>In Proc. of AKBC</source>
          , (p.
          <fpage>1</fpage>
          -
          <lpage>7</lpage>
          ).
        </mixed-citation>
      </ref>
      <ref id="ref20">
        <mixed-citation>
          <string-name>
            <surname>Kaplan</surname>
            ,
            <given-names>D.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Iida</surname>
            ,
            <given-names>R.</given-names>
          </string-name>
          , and
          <string-name>
            <surname>Tokunaga</surname>
            ,
            <given-names>T.</given-names>
          </string-name>
          <year>2009</year>
          .
          <article-title>Automatic Extraction of Citation Contexts for Research Paper Summarization: A Coreference-Chain Based Approach</article-title>
          .
          <source>In Proc. of NLPIR4DL</source>
          , (p.
          <fpage>88</fpage>
          -
          <lpage>95</lpage>
          ).
        </mixed-citation>
      </ref>
      <ref id="ref21">
        <mixed-citation>
          <string-name>
            <surname>Kaplan</surname>
            ,
            <given-names>D.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Tokunaga</surname>
            ,
            <given-names>T.</given-names>
          </string-name>
          , and
          <string-name>
            <surname>Teufel</surname>
            ,
            <given-names>S.</given-names>
          </string-name>
          <year>2016</year>
          .
          <article-title>Citation Block Determination Using Textual Coherence</article-title>
          .
          <source>Information Processing</source>
          ,
          <volume>24</volume>
          (
          <issue>3</issue>
          ), (p.
          <fpage>540</fpage>
          -
          <lpage>553</lpage>
          ).
        </mixed-citation>
      </ref>
      <ref id="ref22">
        <mixed-citation>
          <string-name>
            <surname>Knoth</surname>
            ,
            <given-names>P.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Gooch</surname>
            ,
            <given-names>P.</given-names>
          </string-name>
          and
          <string-name>
            <surname>Jack</surname>
            ,
            <given-names>K.</given-names>
          </string-name>
          <year>2017</year>
          .
          <article-title>What Others Say About This Work? Scalable Extraction of Citation Contexts from Research Papers</article-title>
          .
          <source>In Proc. of TPDL</source>
          , (p.
          <fpage>287299</fpage>
          ).
        </mixed-citation>
      </ref>
      <ref id="ref23">
        <mixed-citation>
          <string-name>
            <surname>Mei</surname>
            ,
            <given-names>Q.</given-names>
          </string-name>
          , and
          <string-name>
            <surname>Zhai</surname>
            ,
            <given-names>C.</given-names>
          </string-name>
          <year>2008</year>
          .
          <article-title>Generating ImpactBased Summaries for Scientific Literature</article-title>
          .
          <source>In Proc. of ACL-HLT</source>
          , (p.
          <fpage>816</fpage>
          -
          <lpage>824</lpage>
          ).
        </mixed-citation>
      </ref>
      <ref id="ref24">
        <mixed-citation>
          <string-name>
            <surname>Doslu</surname>
            ,
            <given-names>M.</given-names>
          </string-name>
          , and
          <string-name>
            <surname>Bingol</surname>
            ,
            <given-names>H.</given-names>
          </string-name>
          <year>2016</year>
          .
          <article-title>Context Sensitive Article Ranking with Citation Context Analysis</article-title>
          .
          <source>Scientometrics</source>
          ,
          <volume>108</volume>
          (
          <issue>2</issue>
          ), (p.
          <fpage>653671</fpage>
          ).
        </mixed-citation>
      </ref>
      <ref id="ref25">
        <mixed-citation>
          <string-name>
            <surname>MacRoberts</surname>
            ,
            <given-names>M.H.</given-names>
          </string-name>
          ,
          <article-title>and</article-title>
          <string-name>
            <surname>MacRoberts</surname>
            ,
            <given-names>B.R.</given-names>
          </string-name>
          <year>1984</year>
          .
          <article-title>The Negational Reference: or the Art of Dissembling</article-title>
          .
          <source>Social Studies of Science</source>
          ,
          <volume>14</volume>
          , (p.
          <fpage>91</fpage>
          -
          <lpage>94</lpage>
          ).
        </mixed-citation>
      </ref>
      <ref id="ref26">
        <mixed-citation>
          <string-name>
            <surname>Nakov</surname>
            ,
            <given-names>P.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Schwartz</surname>
            ,
            <given-names>A.</given-names>
          </string-name>
          , and
          <string-name>
            <surname>Hearst</surname>
            ,
            <given-names>M.</given-names>
          </string-name>
          <year>2004</year>
          .
          <article-title>Citances: Citation Sentences for Semantic Analysis of Bioscience Text</article-title>
          .
          <source>In Proc. of SIGIR.</source>
        </mixed-citation>
      </ref>
      <ref id="ref27">
        <mixed-citation>
          <string-name>
            <surname>Nanba</surname>
            ,
            <given-names>H.</given-names>
          </string-name>
          , and
          <string-name>
            <surname>Okumura</surname>
            ,
            <given-names>M.</given-names>
          </string-name>
          <year>1999</year>
          .
          <article-title>Towards Multipaper Summarization Using Reference Information</article-title>
          .
          <source>In Proc. of IJCAI</source>
          , (p.
          <fpage>926</fpage>
          -
          <lpage>931</lpage>
          ).
        </mixed-citation>
      </ref>
      <ref id="ref28">
        <mixed-citation>
          <string-name>
            <surname>O'Connor</surname>
            ,
            <given-names>J.</given-names>
          </string-name>
          <year>1982</year>
          .
          <article-title>Citing Statements: Computer Recognition</article-title>
          and
          <article-title>Use to Improve Retrieval</article-title>
          .
          <source>Information Processing and Management</source>
          ,
          <volume>18</volume>
          (
          <issue>3</issue>
          ), (p.
          <fpage>125</fpage>
          -
          <lpage>131</lpage>
          ).
        </mixed-citation>
      </ref>
      <ref id="ref29">
        <mixed-citation>
          <string-name>
            <surname>Platt J.C.</surname>
          </string-name>
          <year>1990</year>
          .
          <article-title>Probabilistic Outputs for Support Vector Machines and Comparisons to Regularized Likelihood Methods</article-title>
          .
          <source>Advances in Large Margin Classifiers</source>
          , (p.
          <fpage>61</fpage>
          -
          <lpage>74</lpage>
          ).
        </mixed-citation>
      </ref>
      <ref id="ref30">
        <mixed-citation>
          <string-name>
            <surname>Qazvinian</surname>
            ,
            <given-names>V.</given-names>
          </string-name>
          , and
          <string-name>
            <surname>Radev</surname>
            ,
            <given-names>D.</given-names>
          </string-name>
          <year>2010</year>
          .
          <article-title>Identifying Nonexplicit Citing Sentences for Citation-based Summarization</article-title>
          .
          <source>In Proc. of ACL</source>
          , (p.
          <fpage>555</fpage>
          -
          <lpage>564</lpage>
          ).
        </mixed-citation>
      </ref>
      <ref id="ref31">
        <mixed-citation>
          <string-name>
            <surname>Ritchie</surname>
            ,
            <given-names>A.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Robertson</surname>
            ,
            <given-names>S.</given-names>
          </string-name>
          , and
          <string-name>
            <surname>Teufel</surname>
            ,
            <given-names>S.</given-names>
          </string-name>
          <year>2008</year>
          .
          <article-title>Comparing Citation Contexts for Information Retrieval</article-title>
          .
          <source>In Proc. of ACM-CIKM</source>
          , (p.
          <fpage>213</fpage>
          -
          <lpage>222</lpage>
          ).
        </mixed-citation>
      </ref>
      <ref id="ref32">
        <mixed-citation>
          <string-name>
            <surname>Sula</surname>
            ,
            <given-names>C.</given-names>
          </string-name>
          , and
          <string-name>
            <surname>Miller</surname>
            ,
            <given-names>M.</given-names>
          </string-name>
          <year>2014</year>
          . Citations, Contexts, and Humanistic Discourse:
          <article-title>Toward Automatic Extraction and Classification</article-title>
          .
          <source>Literary and Linguistic Computing</source>
          ,
          <volume>29</volume>
          , (p.
          <fpage>453</fpage>
          -
          <lpage>464</lpage>
          ).
        </mixed-citation>
      </ref>
      <ref id="ref33">
        <mixed-citation>
          <string-name>
            <surname>Teufel</surname>
            ,
            <given-names>S.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Siddharthan</surname>
            ,
            <given-names>A.</given-names>
          </string-name>
          , and
          <string-name>
            <surname>Tidhar</surname>
            ,
            <given-names>D.</given-names>
          </string-name>
          <year>2006</year>
          .
          <article-title>Automatic Classification of Citation Function</article-title>
          .
          <source>In Proc. of EMNLP</source>
          , (p.
          <fpage>103</fpage>
          -
          <lpage>110</lpage>
          ).
        </mixed-citation>
      </ref>
    </ref-list>
  </back>
</article>