<!DOCTYPE article PUBLIC "-//NLM//DTD JATS (Z39.96) Journal Archiving and Interchange DTD v1.0 20120330//EN" "JATS-archivearticle1.dtd">
<article xmlns:xlink="http://www.w3.org/1999/xlink">
  <front>
    <journal-meta />
    <article-meta>
      <title-group>
        <article-title>Style Change Detection with Feed-forward Neural Networks</article-title>
      </title-group>
      <contrib-group>
        <contrib contrib-type="author">
          <string-name>Department of Computer Science</string-name>
          <xref ref-type="aff" rid="aff0">0</xref>
          <xref ref-type="aff" rid="aff1">1</xref>
        </contrib>
        <contrib contrib-type="author">
          <string-name>chzuo</string-name>
          <xref ref-type="aff" rid="aff0">0</xref>
          <xref ref-type="aff" rid="aff1">1</xref>
        </contrib>
        <contrib contrib-type="author">
          <string-name>yuzhao</string-name>
          <xref ref-type="aff" rid="aff0">0</xref>
          <xref ref-type="aff" rid="aff1">1</xref>
        </contrib>
        <contrib contrib-type="author">
          <string-name>rbanerjee}@cs.stonybrook.edu</string-name>
          <xref ref-type="aff" rid="aff0">0</xref>
          <xref ref-type="aff" rid="aff1">1</xref>
        </contrib>
        <aff id="aff0">
          <label>0</label>
          <institution>Chaoyuan Zuo</institution>
          ,
          <addr-line>Yu Zhao, and Ritwik Banerjee</addr-line>
        </aff>
        <aff id="aff1">
          <label>1</label>
          <institution>Stony Brook University</institution>
          ,
          <addr-line>Stony Brook New York 11794</addr-line>
          ,
          <country country="US">USA</country>
        </aff>
      </contrib-group>
      <pub-date>
        <year>2019</year>
      </pub-date>
      <abstract>
        <p>The majority of previous authorship attribution studies mainly focus on a dataset of documents (or parts of documents) with labeled authorship. This scenario, however, is not applicable to documents written by more than one author. Detecting the authorship switches within multi-author documents has been shown to be a challenging task in previous PAN tasks. A simplified version of the style change task is thus organized by PAN 2019, which aims at identifying the number of authors in a given document. To this end, we present a system consisting of two modules, one for distinguishing the single-author documents from the multiauthor documents and the other for determining the exact number of authors in the multi-author documents.</p>
      </abstract>
    </article-meta>
  </front>
  <body>
    <sec id="sec-1">
      <title>Introduction</title>
      <p>
        Authorship attribution is a difficult task with a long history [
        <xref ref-type="bibr" rid="ref10">10</xref>
        ]. With the advent of
the web and social media, however, the nature of authorship is changing. On one hand,
collaborative writing is becoming more commonplace, while on the other, the easy
availability of vast amounts of source material makes plagiarism easy to carry out but
hard to detect. There have been several settings of the authorship attribution task, with
the most common scenario focusing on a closed set of documents and candidate authors
with the assumption that each document is written by a single author from among the
candidates [
        <xref ref-type="bibr" rid="ref18 ref4">4,18</xref>
        ]. These models are not applicable if a document is written by more
than one author, however. The style breach detection task at PAN 2017 was designed
to bridge this gap. The task was to find the border positions where authorship changes,
but the results showed it to be an extremely challenging task [
        <xref ref-type="bibr" rid="ref20">20</xref>
        ]. A simplified version
was presented in the following year at PAN, where the task was to detect whether or not
there was any stylistic change, i.e., whether or not the document had multiple authors.
The results of this task were quite promising, attaining accuracy up to 0:893. The current
task [
        <xref ref-type="bibr" rid="ref22">22</xref>
        ] goes further and aims at detecting the exact number of authors in a document.
      </p>
      <p>This paper reports on the PAN 2019 shared task on style change detection. A
twostep pipeline is presented to solve this problem. The first step of the system aims at
distinguishing multi-author documents from the documents with a single author, and the
second identifies the exact number of authors within the multi-author document. The
evaluation results of this task suggest that automated detection of writing style change
remains a challenging task.
2</p>
    </sec>
    <sec id="sec-2">
      <title>Related Work</title>
      <p>
        Earlier research in author profiling worked under the assumption that each article has
exactly one author [
        <xref ref-type="bibr" rid="ref18">18</xref>
        ] where authorship attribution could largely be done on the basis
of lexical, syntactic, and semantic features [
        <xref ref-type="bibr" rid="ref14 ref2">2,14</xref>
        ]. Documents may have multiple authors,
however. For instance, in collaborative work, different sections may be written by
different authors. The stylistic cues linking authors to text can be used to partition a
document into stylistic clusters, thus providing insight into the number of authors of
a document, and who those authors might be. This is known as the author diarization
problem, an important component of which is to identify the number of authors within a
document.
      </p>
      <p>
        Prior research on the author diarization task at PAN-2016 [
        <xref ref-type="bibr" rid="ref19">19</xref>
        ] indicates that capturing
stylometric changes is perhaps the most promising approach on author clustering within
documents [
        <xref ref-type="bibr" rid="ref16 ref9">9,16</xref>
        ]. The text is split into sentences, and features including word frequency,
selected part-of-speech (POS) tag counts, and average word length are calculated for
each sentence. The K-means clustering algorithm is applied to generate clusters based
on the distances computed with these features [
        <xref ref-type="bibr" rid="ref16">16</xref>
        ]. The style breach detection task at
PAN-2017 [
        <xref ref-type="bibr" rid="ref20">20</xref>
        ] is to find the exact position of authorship changing with a document.
Here, too, the three submissions extract shallow stylometric features like character
ngrams, word frequency, function words and punctuation. Such features are explored
at the level of sentences [
        <xref ref-type="bibr" rid="ref12 ref6">6,12</xref>
        ] as well as paragraphs [
        <xref ref-type="bibr" rid="ref5">5</xref>
        ]. The similarity between the
consecutive objects is then evaluated to detect a change of authorship.
      </p>
      <p>
        Given the evident difficulty of author diarization, PAN-2018 presented a simpler
binary classification task of identifying whether or not a document was written by a single
author. Going beyond typical stylometric features [
        <xref ref-type="bibr" rid="ref13 ref7">7,13</xref>
        ], several other approaches were
explored. The winning submission by Zlatkova et al. [
        <xref ref-type="bibr" rid="ref23">23</xref>
        ] split each document into three
segments of equal length and use several classifiers to obtain the final results. Hosseinia
and Mukherjee [
        <xref ref-type="bibr" rid="ref3">3</xref>
        ] relied solely on grammatical structures and lexical features where
each sentence was represented by a collection of features extracted from its parse tree.
This representation was provided as input to two recurrent neural networks in the original
or reversed order of the sentences of the document, respectively. Multiple similarity
measures were then computed the difference between the two network representations,
yielding the final binary classification. Another approach was taken by Schaetti [
        <xref ref-type="bibr" rid="ref15">15</xref>
        ],
based on a character embedding layer used in a convolutional neural network.
3
      </p>
    </sec>
    <sec id="sec-3">
      <title>Style Change Detection</title>
      <p>
        This paper is the result of our participation in the ‘Style Change Detection’ taks as part
of PAN at CLEF 2019. The task is defined as follows: Given a document, determine
whether it contains style changes or not, i.e., if it was written by a single or multiple
authors. If it is written by more than one author, determine the number of involved
collaborators. As is evident from the task definition, an individual author is implicitly
associated with a particular style of writing. Earlier PAN tasks have shown that complete
author diarization is a particularly difficult problem [
        <xref ref-type="bibr" rid="ref19 ref20">19,20</xref>
        ]. As such, this task sets the
relatively modest goal of detecting the number of authors in each document, omitting
author attribution. The task definition lends itself naturally to a two-step pipeline process:
(i) binary classification to determine whether a document has a single author or multiple
authors, and (ii) if a document has multiple authors (as determined by the first step),
identify how many. Next, we present the details of the data used in this work and the
evaluation framework.
      </p>
      <p>Data: Each document in the dataset for Table 1. Distribution of the number of authors of
this task is a post (or a concatenation the documents in the dataset
of multiple posts) from StackExchange1. # authors 1 2 3 4 5
The training set comprises 2,546
documents, and a separate validation set of ## ddooccss iinn tvraaliidnaintigon 1623763 312759 311532 312680 310475
1,272 documents is also provided. For
each document, the gold-standard labels
are provided in the form of (i) the number of authors, and (ii) annotations marking who
authored exactly which portions of the document. Exactly half of the training documents
have a single author. The remaining half have a nearly uniform distribution over the
number of authors, ranging from 2 to 5 authors. This is shown in Table 1.The test set
has the same size as the validation set, but has been provided without any labels, i.e., no
information about where within a document authorship switched, or how many authors
are there for a given document.</p>
      <p>
        Evaluation: In this task, the performance of a model is evaluated by combining (i) the
accuracy, which serves as a measurement for the binary classification of single against
multiple authorship, with (ii) the ordinal classification index (OCI) [
        <xref ref-type="bibr" rid="ref1">1</xref>
        ], which measures
the error of predicting the number of authors for documents with multiple authors. Thus,
the final rank r that captures both is the arithmetic mean of the accuracy and the inverted
OCI is given by
      </p>
      <p>1
r = (accuracy + (1 OCI)) (1)</p>
      <p>
        2
To evaluate the submission, the participants are asked to submit the created software for
this task through a virtual machine in TIRA [
        <xref ref-type="bibr" rid="ref11">11</xref>
        ], a web platform that supports software
submission for shared tasks.
4
      </p>
    </sec>
    <sec id="sec-4">
      <title>Methodology</title>
      <p>Given the evaluation framework, which consists of combining two independent
measurement, we build a system that treats them separately by applying a two-step pipeline.
Th first step separartes single-author from multi-author documents, and the second step</p>
      <sec id="sec-4-1">
        <title>1 https://stackexchange.com/</title>
        <p>identifies the number of authors in multi-author documents, In our approach, we divide
the whole document into several segments, and cluster the segments based on writing
style. The objective is to have the number of clusters be equal to the the number of
authors in the document. Figure 1 illustrates this process2.</p>
        <p>Data preprocessing: About 1% of the documents in the dataset are in Spanish instead
of English. Since it is difficult to extract features and build models for them separately
within the scope of this task, we randomly assign the number of authors (ranging from 1
to 5) for these documents. For the rest, we filter out some frequent phrases that carry
little or no linguistic relevance. These are typically URL or technical specifications like
“OSX 10.11.2”.</p>
        <p>Binary classification of documents (single vs multiple authors): Here we describe
how we identify whether a document is authored by one person or not. We treat all
documents with multiple authors as one category, and build binary classifiers for
separating them from documents with a single author. We use Keras3 to implement this. Each
documents is first tokenized and then converted into a term-document matrix where
each word is encoded with its term-frequency inverse-document-frequency (TF-IDF)
score. Further, we set the maximum number of words to keep as 40,000. Only the most
common 40,000 words on the dataset are used, based on word frequency.</p>
        <p>
          This classification is along the lines of the PAN-2018 task. Given the success of
non-linear neural networks in that task (e.g., Zlatkova et al. [
          <xref ref-type="bibr" rid="ref23">23</xref>
          ]), we decide to adopt a
neural network as well. We use a simple feedforward neural network as the classifier,
with the term-document matrix being used as the input. The network has only one
hidden layer with 128 units and a softmax output layer with two nodes for the binary
classification. For activation, we use the sigmoid function since it achieved better results
when compared to rectified linear units (ReLU) on validation, and for optimization we
        </p>
      </sec>
      <sec id="sec-4-2">
        <title>2 The implementation is available at https://github.com/chzuo/PAN_2019. 3 https://keras.io/</title>
        <p>
          use Adam [
          <xref ref-type="bibr" rid="ref8">8</xref>
          ]. To avoid overfitting, we use dropout [
          <xref ref-type="bibr" rid="ref17">17</xref>
          ] with the probability value set to
p = 0:5.
        </p>
        <p>Detecting the number of authors in multi-author documents: Our key idea for this
step is to divide the whole document into several parts and then cluster them to groups.
Since some documents are poorly structured (e.g.), upon tokenizing using the NLTK
tokenizer, some documents yield more than 200 sentences, we divide each document
into paragraph-level segments instead of sentences. Further, upon studying the test and
validation sets, we notice that almost all authorship changes happen near an empty line or
newline symbol, based on the observation that half of the switches happen after an empty
line while 80% of the writing by new authors starts after the newline symbols. Thus, to
divide a document into segments, firstly the empty line is used to separate the segments,
and if we can get more than 15 segments in the document after splitting, we then use
these segments for the next clustering step, otherwise, we use the newline symbol for
splitting segments in the text and use that results for the next step. In general, using the
newline symbol results in more segments than using an empty line. Finally, segments
with less than 20 tokens are discarded.</p>
        <p>
          Feature Extraction: Then the feature extraction module is conducted on each segment
after splitting. Following the feature design of the winning submission of PAN-2018 [
          <xref ref-type="bibr" rid="ref23">23</xref>
          ],
we use the following features:
– Token and POS distribution features. This includes the distribution of token length
in the segment, the distribution of POS-tags in the segment, the number of sentences
in the segment, the number of each special character and punctuation in the segment.
The special characters included are ‘#’, ‘$’, ‘%’, ‘&amp;’, ‘*’, ‘@’, parentheses and the
forward and backward slashes.
– Contracted word forms. Writers may have their own preferences about the usage
of contractions like “I’ll” instead of “I will” or “I’m” instead of “I am”. Thus, we
maintain two lists, one containing the original words and the other containing their
corresponding contracted forms. We count the total number of occurrences of the
words in each list and use these two number as the feature.
– British/American English spelling. We use a list of spelling variations4, and encode
this feature as the number of occurrence of variant spellings in a single segment.
– Function word frequencies. We combine the list of function words from NLTK and
the list used by Zlatkova et al. [
          <xref ref-type="bibr" rid="ref23">23</xref>
          ], and encode the frequency of each function words
as additional features.
– Readability. We use Textstat5 to obtain the Flesch reading ease score, SMOG grade,
Flesch-Kincaid grade, Coleman-Liau index, automated readability index, Dale-Chall
readability score, Linsear Write readability metric, Gunning-Fog index, and the
number of difficult words in the text. We use all the above measures of readability
and keep them as separate features.
        </p>
        <p>Further, we also use the TF-IDF scores of the words in each segment, as are calculated
for the creation of the term-document matrix earlier.
4 https://en.wikipedia.org/wiki/Wikipedia:List_of_spelling_variants (Accessed May 2019).
5 https://github.com/shivam5992/textstat
Segment Clustering: To cluster the segments into groups, we build an ensemble system
of different algorithms on various combinations of features. It consists of three models: (i)
the K-means clustering where each segment is represented as a bag-of-words vector with
TF-IDF encoding, (ii) hierarchical clustering with all the rest of the features mentioned in
the above section, and (iii) a feedforward neural network classifier to detect the similarity
between segments and create the similarity matrix for all the segments using the output
of the NN. Then we use K-means clustering with silhouette analysis to determine the
number of clusters in this similarity matrix. The three models share the same weight
when determining the final results of the number of authors.</p>
        <p>K-means clustering: For each document, all the segments are represented as a
bagof-words vector with TF-IDF encoding. We use the scikit-learn 6 tool to implement the
K-means clustering for these segments. The silhouette analysis is then used to select the
number (ranging from 2 to 5) of best clusters.</p>
        <p>
          Hierarchical clustering: The hierarchical clustering algorithm is used to group the
segments in each document. We use the extracted features for clustering except for the
TF-IDF encoding. The tool we use for the implementation of hierarchical clustering is the
SciPy7. We use the Ward’s minimum variance method [
          <xref ref-type="bibr" rid="ref21">21</xref>
          ] for computing the distance
between the nodes . Once the distance between nodes has been calculated, the linkage
function is used for paring the objects that are close to the binary clusters (clusters
consisting of two objects), and it links the newly formed clusters to each other to create
bigger clusters based on the distance information until all the nodes are linked in a
hierarchical tree. The distance information of each clustering step in this tree can be used
as the cutoff argument to determine the number of clusters. Moreover, the inconsistent
coefficient for each link in the hierarchical clustering tree can be used for determining
the number of clusters as well. It is calculated by comparing the height of a link with the
average height of other links at the same cluster hierarchy. The lower the coefficient, the
smaller the difference between the object with those around it.
        </p>
        <p>The number of clusters for each document, however, is hard to determine, as there
are no methods to set a universal cut-off value for all documents. Thus, we build a
feedforward neural network with one hidden layer consisting of 20 units. The input
vector with the dimension as 50 is created for each document, using the distance and
inconsistency value for the last 25 clustering step on a linkage matrix. If the number
of segments is less than 25 in the document, we add zeros to the start of the vectors to
increase its length to 50. The output target of the NN classifier is the number of authors
in the document. A softmax layer with 4 output nodes is added. We use the ReLU for
activation function and Adam for stochastic optimization, as well as the dropout with
p = 0:5.</p>
        <p>We train the network on all the documents with more than two authors in the training
and validation set. During the test phase, only the documents that get categorized as
multi-author documents after the first step of our process (i.e., the binary classification)
is sent as the input to the NN model for prediction.</p>
      </sec>
      <sec id="sec-4-3">
        <title>6 https://scikit-learn.org/stable/ 7 https://www.scipy.org/</title>
        <p>zuo19
nath19
Random Baseline</p>
        <p>Accuracy</p>
        <p>K-means with similarity matrix We create a dataset by splitting the documents into
several parts using the provided authors switches information from the gold-standard
label and pairing the obtained segments. Over 40k segment pairs are selected from the
documents. For half of the pairs, the two segments are written by the same author and
we treat them as one category. We build a binary NN classifier for separating them with
pairs written by different authors. We use all the features mentioned above except for
the TF-IDF representation. The network has 2 hidden layers with 50 units and 8 units in
them. We use ReLU for activation function. For each pair of segments, we define the
similarity of this pair is the probability that the two segments are written by the same
author. Then for a document containing n segments, we generate a similarity matrix
M of size n n where Mi;j is the similarity of segments i; j, using the output the NN
classifier. Finally, we employ the K-means clustering method with silhouette analysis
for determining the best number of authors in this similarity matrix.
5</p>
      </sec>
    </sec>
    <sec id="sec-5">
      <title>Results</title>
      <p>The results of our system are shown in Table 2. Two teams participated in this task, and
our submission outperforms the other in the OCI evaluation measure, where the lower
the value, the better the performance. For the first metric - accuracy, the performance of
our binary classifier reaches the value as 0.6, 20 percent increase of 0.5, which serves
as the baseline for a random guess. And for the second metric - OCI for multi-authors
detection, our system achieves 0.808, slightly better than the random guess. The results
of our system suggest that style change detection for multi-author documents is remains
a challenging task and requires significant further research to be adequately resolved.
6</p>
    </sec>
    <sec id="sec-6">
      <title>Conclusion</title>
      <p>We develope a two-step pipeline system to detect the number of authors in the given
document. The purpose of the first step is to identify whether or not the document is
written by more than one author. This is achieved by using a feedforward neural network
as a binary classifier. The second step is an ensemble model of different clustering
methods. This identifies the number of authors for the multi-author documents. This
task, however, is quite challenging, and there is scope for significant improvement in
this direction.</p>
    </sec>
  </body>
  <back>
    <ref-list>
      <ref id="ref1">
        <mixed-citation>
          1.
          <string-name>
            <surname>Cardoso</surname>
            ,
            <given-names>J.S.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Sousa</surname>
            ,
            <given-names>R.</given-names>
          </string-name>
          :
          <article-title>Measuring the performance of ordinal classification</article-title>
          .
          <source>International Journal of Pattern Recognition and Artificial Intelligence</source>
          <volume>25</volume>
          (
          <issue>08</issue>
          ),
          <fpage>1173</fpage>
          -
          <lpage>1195</lpage>
          (
          <year>2011</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref2">
        <mixed-citation>
          2.
          <string-name>
            <surname>Feng</surname>
            ,
            <given-names>S.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Banerjee</surname>
            ,
            <given-names>R.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Choi</surname>
            ,
            <given-names>Y.</given-names>
          </string-name>
          :
          <article-title>Characterizing Stylistic Elements in Syntactic Structure</article-title>
          .
          <source>In: Proceedings of the 2012 Joint Conference on Empirical Methods on Natural Language Processing and Computational Natural Language Learning</source>
          . pp.
          <fpage>1522</fpage>
          -
          <lpage>1533</lpage>
          . Association for Computational Linguistics (
          <year>2012</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref3">
        <mixed-citation>
          3.
          <string-name>
            <surname>Hosseinia</surname>
            ,
            <given-names>M.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Mukherjee</surname>
            ,
            <given-names>A.</given-names>
          </string-name>
          :
          <article-title>A Parallel Hierarchical Attention Network for Style Change Detection - Notebook for PAN at CLEF 2018</article-title>
          . In: Cappellato,
          <string-name>
            <given-names>L.</given-names>
            ,
            <surname>Ferro</surname>
          </string-name>
          ,
          <string-name>
            <given-names>N.</given-names>
            ,
            <surname>Nie</surname>
          </string-name>
          ,
          <string-name>
            <given-names>J.Y.</given-names>
            ,
            <surname>Soulier</surname>
          </string-name>
          ,
          <string-name>
            <surname>L</surname>
          </string-name>
          . (eds.)
          <article-title>CLEF 2018 Evaluation Labs</article-title>
          and Workshop - Working Notes Papers,
          <volume>10</volume>
          -
          <fpage>14</fpage>
          September, Avignon, France. CEUR-WS.org (
          <year>2018</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref4">
        <mixed-citation>
          4.
          <string-name>
            <surname>Iqbal</surname>
            ,
            <given-names>F.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Binsalleeh</surname>
            ,
            <given-names>H.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Fung</surname>
            ,
            <given-names>B.C.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Debbabi</surname>
            ,
            <given-names>M.</given-names>
          </string-name>
          :
          <article-title>Mining writeprints from anonymous e-mails for forensic investigation</article-title>
          .
          <source>digital investigation 7(1-2)</source>
          ,
          <fpage>56</fpage>
          -
          <lpage>64</lpage>
          (
          <year>2010</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref5">
        <mixed-citation>
          5. Karas´,
          <string-name>
            <surname>D.</surname>
            , S´ piewak,
            <given-names>M.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Sobecki</surname>
            ,
            <given-names>P.</given-names>
          </string-name>
          :
          <article-title>OPI-JSA at CLEF 2017: Author Clustering and Style Breach Detection-Notebook for PAN at CLEF 2017</article-title>
          . In: Cappellato,
          <string-name>
            <given-names>L.</given-names>
            ,
            <surname>Ferro</surname>
          </string-name>
          ,
          <string-name>
            <given-names>N.</given-names>
            ,
            <surname>Goeuriot</surname>
          </string-name>
          ,
          <string-name>
            <given-names>L.</given-names>
            ,
            <surname>Mandl</surname>
          </string-name>
          , T. (eds.)
          <article-title>CLEF 2017 Evaluation Labs</article-title>
          and Workshop - Working Notes Papers,
          <volume>11</volume>
          -
          <fpage>14</fpage>
          September, Dublin, Ireland.
          <source>CEUR-WS.org (Sep</source>
          <year>2017</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref6">
        <mixed-citation>
          6.
          <string-name>
            <surname>Khan</surname>
          </string-name>
          , J.:
          <article-title>Style Breach Detection: An Unsupervised Detection Model-Notebook for PAN at CLEF 2017</article-title>
          . In: Cappellato,
          <string-name>
            <given-names>L.</given-names>
            ,
            <surname>Ferro</surname>
          </string-name>
          ,
          <string-name>
            <given-names>N.</given-names>
            ,
            <surname>Goeuriot</surname>
          </string-name>
          ,
          <string-name>
            <given-names>L.</given-names>
            ,
            <surname>Mandl</surname>
          </string-name>
          , T. (eds.)
          <article-title>CLEF 2017 Evaluation Labs</article-title>
          and Workshop - Working Notes Papers,
          <volume>11</volume>
          -
          <fpage>14</fpage>
          September, Dublin, Ireland.
          <source>CEUR-WS.org (Sep</source>
          <year>2017</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref7">
        <mixed-citation>
          7.
          <string-name>
            <surname>Khan</surname>
            ,
            <given-names>J.A.</given-names>
          </string-name>
          :
          <article-title>A model for style breach detection at a glance: Notebook for PAN at CLEF 2018</article-title>
          . In: Working Notes of CLEF 2018 -
          <article-title>Conference and Labs of the Evaluation Forum</article-title>
          , Avignon, France,
          <source>September 10-14</source>
          ,
          <year>2018</year>
          . (
          <year>2018</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref8">
        <mixed-citation>
          8.
          <string-name>
            <surname>Kingma</surname>
            ,
            <given-names>D.P.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Ba</surname>
          </string-name>
          , J.:
          <article-title>Adam: A method for stochastic optimization</article-title>
          .
          <source>arXiv preprint arXiv:1412.6980</source>
          (
          <year>2014</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref9">
        <mixed-citation>
          9.
          <string-name>
            <surname>Kuznetsov</surname>
            ,
            <given-names>M.P.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Motrenko</surname>
            ,
            <given-names>A.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Kuznetsova</surname>
            ,
            <given-names>R.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Strijov</surname>
            ,
            <given-names>V.V.</given-names>
          </string-name>
          :
          <article-title>Methods for intrinsic plagiarism detection and author diarization</article-title>
          . In: Working Notes of CLEF 2016 -
          <article-title>Conference and Labs of the Evaluation forum</article-title>
          , Évora, Portugal,
          <fpage>5</fpage>
          -
          <lpage>8</lpage>
          September,
          <year>2016</year>
          . pp.
          <fpage>912</fpage>
          -
          <lpage>919</lpage>
          (
          <year>2016</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref10">
        <mixed-citation>
          10.
          <string-name>
            <surname>Mendenhall</surname>
          </string-name>
          , T.C.:
          <article-title>The characteristic curves of composition</article-title>
          .
          <source>Science</source>
          <volume>9</volume>
          (
          <issue>214</issue>
          ),
          <fpage>237</fpage>
          -
          <lpage>249</lpage>
          (
          <year>1887</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref11">
        <mixed-citation>
          11.
          <string-name>
            <surname>Potthast</surname>
            ,
            <given-names>M.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Gollub</surname>
            ,
            <given-names>T.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Wiegmann</surname>
            ,
            <given-names>M.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Stein</surname>
            ,
            <given-names>B.</given-names>
          </string-name>
          :
          <article-title>TIRA Integrated Research Architecture</article-title>
          . In: Ferro,
          <string-name>
            <given-names>N.</given-names>
            ,
            <surname>Peters</surname>
          </string-name>
          ,
          <string-name>
            <surname>C</surname>
          </string-name>
          . (eds.)
          <article-title>Information Retrieval Evaluation in a Changing World - Lessons Learned from 20 Years of</article-title>
          CLEF. Springer (
          <year>2019</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref12">
        <mixed-citation>
          12.
          <string-name>
            <surname>Safin</surname>
            ,
            <given-names>K.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Kuznetsova</surname>
          </string-name>
          , R.:
          <article-title>Style Breach Detection with Neural Sentence Embeddings-Notebook for PAN at CLEF 2017</article-title>
          . In: Cappellato,
          <string-name>
            <given-names>L.</given-names>
            ,
            <surname>Ferro</surname>
          </string-name>
          ,
          <string-name>
            <given-names>N.</given-names>
            ,
            <surname>Goeuriot</surname>
          </string-name>
          ,
          <string-name>
            <given-names>L.</given-names>
            ,
            <surname>Mandl</surname>
          </string-name>
          , T. (eds.)
          <article-title>CLEF 2017 Evaluation Labs</article-title>
          and Workshop - Working Notes Papers,
          <volume>11</volume>
          -
          <fpage>14</fpage>
          September, Dublin, Ireland.
          <source>CEUR-WS.org (Sep</source>
          <year>2017</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref13">
        <mixed-citation>
          13.
          <string-name>
            <surname>Safin</surname>
            ,
            <given-names>K.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Ogaltsov</surname>
            ,
            <given-names>A.</given-names>
          </string-name>
          :
          <article-title>Detecting a change of style using text statistics: Notebook for PAN at CLEF 2018</article-title>
          . In: Working Notes of CLEF 2018 -
          <article-title>Conference and Labs of the Evaluation Forum</article-title>
          , Avignon, France,
          <source>September 10-14</source>
          ,
          <year>2018</year>
          . CEUR-WS.org (
          <year>2018</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref14">
        <mixed-citation>
          14.
          <string-name>
            <surname>Sapkota</surname>
            ,
            <given-names>U.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Bethard</surname>
            ,
            <given-names>S.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Montes</surname>
            ,
            <given-names>M.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Solorio</surname>
            ,
            <given-names>T.</given-names>
          </string-name>
          :
          <article-title>Not All Character N-grams Are Created Equal: A Study in Authorship Attribution</article-title>
          .
          <source>In: Proceedings of the</source>
          <year>2015</year>
          <article-title>Conference of the North American chapter of the Association for Computational Linguistics: Human Language Technologies</article-title>
          . pp.
          <fpage>93</fpage>
          -
          <lpage>102</lpage>
          (
          <year>2015</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref15">
        <mixed-citation>
          15.
          <string-name>
            <surname>Schaetti</surname>
          </string-name>
          , N.:
          <article-title>Character-based Convolutional Neural Network and ResNet18 for Twitter Author Profiling - Notebook for PAN at CLEF 2018</article-title>
          . In: Cappellato,
          <string-name>
            <given-names>L.</given-names>
            ,
            <surname>Ferro</surname>
          </string-name>
          ,
          <string-name>
            <given-names>N.</given-names>
            ,
            <surname>Nie</surname>
          </string-name>
          ,
          <string-name>
            <given-names>J.Y.</given-names>
            ,
            <surname>Soulier</surname>
          </string-name>
          ,
          <string-name>
            <surname>L</surname>
          </string-name>
          . (eds.)
          <article-title>CLEF 2018 Evaluation Labs</article-title>
          and Workshop - Working Notes Papers,
          <volume>10</volume>
          -
          <fpage>14</fpage>
          September, Avignon, France. CEUR-WS.org (
          <year>2018</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref16">
        <mixed-citation>
          16.
          <string-name>
            <surname>Sittar</surname>
            ,
            <given-names>A.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Iqbal</surname>
            ,
            <given-names>H.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Nawab</surname>
          </string-name>
          , R.:
          <article-title>Author Diarization Using Cluster-Distance Approach-Notebook for PAN at CLEF 2016</article-title>
          . In: Balog,
          <string-name>
            <given-names>K.</given-names>
            ,
            <surname>Cappellato</surname>
          </string-name>
          ,
          <string-name>
            <given-names>L.</given-names>
            ,
            <surname>Ferro</surname>
          </string-name>
          ,
          <string-name>
            <given-names>N.</given-names>
            ,
            <surname>Macdonald</surname>
          </string-name>
          , C. (eds.)
          <article-title>CLEF 2016 Evaluation Labs</article-title>
          and Workshop - Working Notes Papers,
          <fpage>5</fpage>
          -
          <lpage>8</lpage>
          September, Évora, Portugal.
          <source>CEUR-WS.org (Sep</source>
          <year>2016</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref17">
        <mixed-citation>
          17.
          <string-name>
            <surname>Srivastava</surname>
            ,
            <given-names>N.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Hinton</surname>
            ,
            <given-names>G.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Krizhevsky</surname>
            ,
            <given-names>A.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Sutskever</surname>
            ,
            <given-names>I.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Salakhutdinov</surname>
          </string-name>
          , R.:
          <article-title>Dropout: a simple way to prevent neural networks from overfitting</article-title>
          .
          <source>The Journal of Machine Learning Research</source>
          <volume>15</volume>
          (
          <issue>1</issue>
          ),
          <fpage>1929</fpage>
          -
          <lpage>1958</lpage>
          (
          <year>2014</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref18">
        <mixed-citation>
          18.
          <string-name>
            <surname>Stamatatos</surname>
          </string-name>
          , E.:
          <article-title>A Survey of Modern Authorship Attribution Methods</article-title>
          .
          <source>Journal of the American Society for Information Science and Technology</source>
          <volume>60</volume>
          (
          <issue>3</issue>
          ),
          <fpage>538</fpage>
          -
          <lpage>556</lpage>
          (
          <year>2009</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref19">
        <mixed-citation>
          19.
          <string-name>
            <surname>Stamatatos</surname>
            ,
            <given-names>E.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Tschnuggnall</surname>
            ,
            <given-names>M.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Verhoeven</surname>
            ,
            <given-names>B.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Daelemans</surname>
            ,
            <given-names>W.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Specht</surname>
            ,
            <given-names>G.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Stein</surname>
            ,
            <given-names>B.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Potthast</surname>
            ,
            <given-names>M.</given-names>
          </string-name>
          :
          <article-title>Clustering by Authorship Within and Across Documents</article-title>
          .
          <source>In: Working Notes Papers of the CLEF</source>
          <year>2016</year>
          <article-title>Evaluation Labs</article-title>
          . CEUR Workshop Proceedings/Balog, Krisztian [edit.]; et al. pp.
          <fpage>691</fpage>
          -
          <lpage>715</lpage>
          (
          <year>2016</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref20">
        <mixed-citation>
          20.
          <string-name>
            <surname>Tschuggnall</surname>
            ,
            <given-names>M.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Stamatatos</surname>
            ,
            <given-names>E.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Verhoeven</surname>
            ,
            <given-names>B.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Daelemans</surname>
            ,
            <given-names>W.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Specht</surname>
            ,
            <given-names>G.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Stein</surname>
            ,
            <given-names>B.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Potthast</surname>
            ,
            <given-names>M.</given-names>
          </string-name>
          :
          <article-title>Overview of the author identification task at pan-2017: style breach detection and author clustering</article-title>
          .
          <source>In: Working Notes Papers of the CLEF 2017 Evaluation Labs/Cappellato</source>
          , Linda [edit.]; et al. pp.
          <fpage>1</fpage>
          -
          <lpage>22</lpage>
          (
          <year>2017</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref21">
        <mixed-citation>
          21.
          <string-name>
            <surname>Ward</surname>
            <given-names>Jr</given-names>
          </string-name>
          ,
          <string-name>
            <surname>J.H.</surname>
          </string-name>
          :
          <article-title>Hierarchical grouping to optimize an objective function</article-title>
          .
          <source>Journal of the American statistical association</source>
          <volume>58</volume>
          (
          <issue>301</issue>
          ),
          <fpage>236</fpage>
          -
          <lpage>244</lpage>
          (
          <year>1963</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref22">
        <mixed-citation>
          22.
          <string-name>
            <surname>Zangerle</surname>
            ,
            <given-names>E.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Tschuggnall</surname>
            ,
            <given-names>M.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Specht</surname>
            ,
            <given-names>G.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Potthast</surname>
            ,
            <given-names>M.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Stein</surname>
            ,
            <given-names>B.</given-names>
          </string-name>
          :
          <article-title>Overview of the Style Change Detection Task at PAN 2019</article-title>
          . In: Cappellato,
          <string-name>
            <given-names>L.</given-names>
            ,
            <surname>Ferro</surname>
          </string-name>
          ,
          <string-name>
            <given-names>N.</given-names>
            ,
            <surname>Losada</surname>
          </string-name>
          ,
          <string-name>
            <given-names>D.</given-names>
            ,
            <surname>Müller</surname>
          </string-name>
          , H. (eds.)
          <article-title>CLEF 2019 Labs and Workshops, Notebook Papers</article-title>
          .
          <source>CEUR-WS.org (Sep</source>
          <year>2019</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref23">
        <mixed-citation>
          23.
          <string-name>
            <surname>Zlatkova</surname>
            ,
            <given-names>D.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Kopev</surname>
            ,
            <given-names>D.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Mitov</surname>
            ,
            <given-names>K.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Atanasov</surname>
            ,
            <given-names>A.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Hardalov</surname>
            ,
            <given-names>M.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Koychev</surname>
            ,
            <given-names>I.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Nakov</surname>
            ,
            <given-names>P.</given-names>
          </string-name>
          :
          <article-title>An Ensemble-Rich Multi-Aspect Approach for Robust Style Change Detection - Notebook for PAN at CLEF 2018</article-title>
          . In: Cappellato,
          <string-name>
            <given-names>L.</given-names>
            ,
            <surname>Ferro</surname>
          </string-name>
          ,
          <string-name>
            <given-names>N.</given-names>
            ,
            <surname>Nie</surname>
          </string-name>
          ,
          <string-name>
            <given-names>J.Y.</given-names>
            ,
            <surname>Soulier</surname>
          </string-name>
          ,
          <string-name>
            <surname>L</surname>
          </string-name>
          . (eds.)
          <article-title>CLEF 2018 Evaluation Labs</article-title>
          and Workshop - Working Notes Papers,
          <volume>10</volume>
          -
          <fpage>14</fpage>
          September, Avignon, France. CEUR-WS.org (
          <year>2018</year>
          )
        </mixed-citation>
      </ref>
    </ref-list>
  </back>
</article>