<!DOCTYPE article PUBLIC "-//NLM//DTD JATS (Z39.96) Journal Archiving and Interchange DTD v1.0 20120330//EN" "JATS-archivearticle1.dtd">
<article xmlns:xlink="http://www.w3.org/1999/xlink">
  <front>
    <journal-meta />
    <article-meta>
      <title-group>
        <article-title>USDB at eRisk 2020: Deep learning models to measure the Severity of the Signs of Depression using Reddit Posts</article-title>
      </title-group>
      <contrib-group>
        <contrib contrib-type="author">
          <string-name>Amina MADANI</string-name>
          <email>a_madani@esi.dz</email>
          <xref ref-type="aff" rid="aff1">1</xref>
        </contrib>
        <contrib contrib-type="author">
          <string-name>Fatima BOUMAHDI</string-name>
          <email>f_boumahdi@esi.dz</email>
          <xref ref-type="aff" rid="aff1">1</xref>
        </contrib>
        <contrib contrib-type="author">
          <string-name>Anfel BOUKENAOUI</string-name>
          <email>anfelboukenaoui@gmail.com</email>
          <xref ref-type="aff" rid="aff1">1</xref>
        </contrib>
        <contrib contrib-type="author">
          <string-name>Mohamed Chaouki KRITLI</string-name>
          <email>kritlichawki56@gmail.com</email>
          <xref ref-type="aff" rid="aff1">1</xref>
        </contrib>
        <contrib contrib-type="author">
          <string-name>Hamza HENTABLI</string-name>
          <email>hentabli_hamza@yahoo.fr</email>
          <xref ref-type="aff" rid="aff0">0</xref>
        </contrib>
        <aff id="aff0">
          <label>0</label>
          <institution>Information Assurance and Security Research Group, Faculty of Computing</institution>
          ,
          <addr-line>Universiti Teknologi Malaysia</addr-line>
          ,
          <country country="MY">Malaysia</country>
        </aff>
        <aff id="aff1">
          <label>1</label>
          <institution>LRDSI Laboratory, BLIDA 1 University</institution>
          ,
          <addr-line>Blida</addr-line>
          ,
          <country country="DZ">Algeria</country>
        </aff>
      </contrib-group>
      <abstract>
        <p>In this paper, we describe the participation of our USDB group (University of Saad Dahleb Blida) in the shared task T2 of the eRisk Lab at the CLEF 2020 workshop. This task focused on measuring the severity of the signs of depression from a thread of user posts. In response to this task, we study the performance of two different deep learning models (CNN and BiLSTM) in order to provide more perspectives for depression researches.</p>
      </abstract>
      <kwd-group>
        <kwd>Depression severity processing</kwd>
        <kwd>Deep learning</kwd>
        <kwd>CNN</kwd>
        <kwd>Social networks Natural language BiLSTM Sentiment analysis</kwd>
      </kwd-group>
    </article-meta>
  </front>
  <body>
    <sec id="sec-1">
      <title>-</title>
      <p>
        Depression identification has been the subject of research of many fields,
psychiatry, psychology, medicine and even sociolinguistics fields. Depression comes
in different degrees and the examinations are usually done through one of the
popular questionnaires used by psychologists, such as the Center of
Epidemiologic Studies Depression Scale (CES-D) [
        <xref ref-type="bibr" rid="ref26">26</xref>
        ], Beck’s Depression Inventory (BDI)
[
        <xref ref-type="bibr" rid="ref4">4</xref>
        ] and Zung’s Self-Rating Depression Scale (SDS) [
        <xref ref-type="bibr" rid="ref38">38</xref>
        ]. But, these examinations
lack empirical data as they use the patient’s observations or a third-party’s ones
which puts the results under the risk of flawed subjective human testing that
can be manipulated easily, often with the purpose of gaining antidepressants or
just to hide one’s own depression from peers [
        <xref ref-type="bibr" rid="ref21">21</xref>
        ].
      </p>
      <p>Twitter, Facebook and Reddit are different social media platforms that allow
people to share their opinions and their personal thoughts. It has been proven
that such data can be used to study clinical matters, especially when it comes
to mental illnesses like depression.</p>
      <p>Many research works have analyzed the prowess of these data for determining
indications of depression. Furthermore, the scientifc community has set forth
different shared tasks like eRisk (Early Risk Prediction on the Internet) of CLEF
(Conference and Labs of the Evaluation Forum).</p>
      <p>In this paper, we describe the participation of our team USDB to the CLEF
eRisk 2020 task 2. The goal of the task was to measure the severity of the signs
of depression, considering a set of user posts on Reddit. Participants had to
answer the standard BDI depression questionnaires using the text of postings
for 70 users.</p>
      <p>
        Our team proposed an approach based on deep learning models to
automatically fill the BDI questionnaire that are Convolutional Neural Networks
model (CNN) [
        <xref ref-type="bibr" rid="ref15">15</xref>
        ] and Bidirectional Long Short Term Memory model (BiLSTM)
[
        <xref ref-type="bibr" rid="ref3 ref35 ref36">3,35,36</xref>
        ], which is a type of recurrent neural network. Subsequently, to generate
different runs, we use two statistical methods. For the first time, neural networks
are used for the task of measuring the severity of the signs of depression.
      </p>
      <p>The rest of this paper is organized as follows: section 2 presents related works
that tackled the same problem as ours. In section 3, we explore our proposed
approach. The following three sections are dedicated to the description of the
dataset, the metrics evaluation used and the results obtained. Finally, we
conclude the paper and discuss future perspectives for our proposed approach.
2</p>
    </sec>
    <sec id="sec-2">
      <title>Related work</title>
      <p>
        Since 2017 until now a shared task on eRisk has been organised. In 2017 and 2018,
the challenge consists in performing a task on early risk detection of depression.
Several researchers were focused on detecting depression. [
        <xref ref-type="bibr" rid="ref19 ref22 ref27 ref28 ref30 ref31 ref33 ref7 ref9">9,28,30,19,22,27,33,31,7</xref>
        ]
are different approaches that propose interesting models and evaluate them using
eRisk dataset.
      </p>
      <p>
        In 2019, a new task was added. Measuring the severity of the signs of
depression consists of estimating the level of depression from a thread of user
submissions. The results of all participating teams can be found in [
        <xref ref-type="bibr" rid="ref17">17</xref>
        ]. [
        <xref ref-type="bibr" rid="ref2 ref29 ref32 ref6">29,2,6,32</xref>
        ] are
several works of different teams.
      </p>
      <p>
        In paper [
        <xref ref-type="bibr" rid="ref29">29</xref>
        ], authors proposed a rule-based method that combines machine
learning with psycholinguistics and behavioural patterns. They divided the 21
questions into 6 groups that are : Depression, Guilt, Appetite, Anxiety, Fatigue
and Sleep. Based on the presence of occurrences of the features considered for
each group, they produced responses for each user.
      </p>
      <p>
        In their work, [
        <xref ref-type="bibr" rid="ref2">2</xref>
        ] extracted features from users using the GPT-1 (Generative
Pre-trained Transformer version 1) language model [
        <xref ref-type="bibr" rid="ref25 ref37">25,37</xref>
        ] and Linguistic Inquiry
Word Count tool (LIWC) [
        <xref ref-type="bibr" rid="ref23">23</xref>
        ]. They predict responses in two ways, an
unsupervised and supervised one. For the unsupervised manner, they used an approach
based on vectorial representation of the user and vectorial representation of
possible response using GPT-1. Cosine Similarity between vectors was calculated
to choose the user’s response. For the supervised way, they used data of a PhD
study, where Psychology students answer the BDI and other questionnaires and
also completed parts of writing about a negative personal problem. Next, they
trained support vector machines using the training data. They also submitted
another supervised approach that used GPT-1 features and AutoSklearn [
        <xref ref-type="bibr" rid="ref10">10</xref>
        ].
Paper [
        <xref ref-type="bibr" rid="ref2">2</xref>
        ] concluded that without training dataset, tasks were not easy and are
unsupervised and data must be annotated to improve the quality of depression
prediction.
      </p>
      <p>
        The Paper of [
        <xref ref-type="bibr" rid="ref6">6</xref>
        ] used SS3 [
        <xref ref-type="bibr" rid="ref5">5</xref>
        ] which is a word-based classifier that estimates
risk based on term statistics. To train SS3, they adapted the dataset of eRisk
2018 depression detection task. They transformed the output of SS3 from a
2dimensional vector into a BDI depression rank. All the questions were answered
with 0 for persons whose depression rank was less or equal to 0. For the others,
they used different methods (based on textual hint, word matching. . . ). SS3
obtained the best AHR and ACR values, and the second-best ADODL and
DCHR.
      </p>
      <p>
        To automatically fill in the BDI questionnaire, the authors of [
        <xref ref-type="bibr" rid="ref32">32</xref>
        ] developed
four models. They submitted only the results of the fourth model. The models
are:
– Word polarity model using the Multi Perspective Question Answering (MPQA)
subjectivity lexicon [
        <xref ref-type="bibr" rid="ref1 ref34">34,1</xref>
        ].
– Mutual information model by creating a training dataset from Reddit and
using the mutual information measure to extract important tokens from
depressive messages [
        <xref ref-type="bibr" rid="ref14">14</xref>
        ].
– Semantic similarity model which is based on post-level representation. The
pre-trained GloVe word embeddings [
        <xref ref-type="bibr" rid="ref24">24</xref>
        ] is used to represent the words.
– In the fourth model, the results of the three models are combined using
voting.
      </p>
      <p>
        [
        <xref ref-type="bibr" rid="ref16">16</xref>
        ] say that no team was able to reach best results for each of the evaluation
measures because of the difficulty of the challenge and probably the similarity
of the approaches.
3
      </p>
    </sec>
    <sec id="sec-3">
      <title>Method</title>
      <p>In this section, we will introduce the architecture of our proposed approach (see
Fig. 1). First, we do preprocessing of posts. We extract keywords that contain
most important information by removing special characters, punctuations, URL
and stop-words. Words would all be stemmed and lemmatized to remove noise
from posts.</p>
      <p>Some of the publications are longer or shorter. Padding is then necessary,
because we need to have the inputs with the same size. We fixed the sequence
length of posts to 250 and shorter input sequences are padded with zeros.</p>
      <p>
        Next, we transform distributed representations of words in a vector space
using the Skip-gram model [
        <xref ref-type="bibr" rid="ref20">20</xref>
        ] which is used to predict the context word for a
given target word. For a given sequence of words, the objective is to find word
representations that are useful for predicting the surrounding words in a post.
      </p>
      <p>After that, the sentences are encoded by means of a CNN or a Bi-LSTM
model. Our first method is based on a CNN model, one of the most popular
deep neural networks. Our model consists of :
– two one-dimensional convolutional layers,
– max-pooling layer : maintains only the most important words in each feature
obtained from the previous layers,
– two others one-dimensional convolutional layers,
– fully connected input layer: flattens the outputs of the convolutional layer
to map them into a single vector,
– fully connected layer:applies weights to predict the correct label,
– fully connected output layer:gives the final probabilities.</p>
      <p>In these layers, each linear activation is run through ReLU (Rectified Linear
Unit). The rectified linear activation function will output the input directly if is
positive, otherwise, it will output zero.</p>
      <p>
        Our second method involves the usage of Recurrent Neural Networks (RNN)
[
        <xref ref-type="bibr" rid="ref12 ref8">8,12</xref>
        ] with Long BiLSTM [
        <xref ref-type="bibr" rid="ref11">11</xref>
        ] to make predictions on sequences of texts. LSTM
[
        <xref ref-type="bibr" rid="ref13">13</xref>
        ] is used in different problems due to their ability to remember information
over long periods of time. LSTM use 3 gates to capture Long Term
Dependencies: Forget Gate for adding information to the cell state, Input Gate to choose
what component of previous cell state must be forgotten and Output Gate to
ascertain information to output at current cell state. The BiLSTM model brings
the advantage of maintaining two separate states for inputs using two different
LSTMs. The first LSTM is going forward from the beginning of the sentence,
while in the second LSTM, the input sequence are fed in backward. BiLSTM
allows capturing information of surrounding inputs and learning faster than LSTM
model.
      </p>
      <p>Last, for the two deep learning models we have a final layer, in which the
representation is fed to the final fully connected softmax layer as an output
feature vector. The results will be decimal probabilities for each answer. Each
publication has only one answer for each question.</p>
      <p>For each post, our two models generate 21 outputs that are answers to the
BDI’s questions. Finally, in order to know the level of depression for a user, we use
two statistical methods to generate 4 runs. The two first runs concern the CNN
model when the BiLSTM model is applied for the others. For run1 (CNN_max)
and run3 (BiLSTM_max), we calculate for a question the frequency of each
generated answer to choose the most recurrent answer, which makes it as relevant
as possible. For run2 (CNN_suite) and run4 (BiLSTM_suite), we calculate for a
question, the higher length sequence of a same answer for all generated answers
from which we choose the answer of the higher value to be a response of this
question.</p>
      <p>Fig. 1. The architecture of our proposed approach.</p>
    </sec>
    <sec id="sec-4">
      <title>Dataset description</title>
      <p>eRisk 2020 Task 2 is a continuation of 2019’s task 3. This year, a dataset with
70 files was provided. Each file contains a set of posts of one user on Reddit. In
Table 1, we present the characteristics of the dataset.</p>
      <p>Based on a user’s history of posts, task 2 was aimed to estimate the depression
level of a user by automatically answering each individual question derived from
the BDI questionnaire. The questionnaire has 21 questions in different classes of
feelings like sadness, pessimism, crying, loss of energy, etc.</p>
      <p>The possible responses are 0, 1a, 1b, 2a, 2b, 3a, 3b for questions 16 and 18
and 0, 1, 2, 3 for the rest of the questions.</p>
      <p>The 2019’s questionnaires and their golden truth responses were provided.
Thus, they can be used as a training dataset.
5</p>
    </sec>
    <sec id="sec-5">
      <title>Metrics evaluation</title>
      <p>Four metrics are used for results evaluation:
– Average Hit Rate (AHR): where HR calculates the percentage of cases
where our answers are the same as the golden truth responses.
– Average Closeness Rate (ACR): computes the absolute difference
between our answer and the true one.
– Average Difference between Overall Depression Levels (ADODL):
for a user, the absolute difference between the generated overall depression
score (sum of all the answers) and the real score is calculated.
– Depression Category Hit Rate (DCHR): Four depression categories
based on the sum of all answers of the 21 questions can be found. The
categories are minimal depression, mild depression, moderate depression and
severe depression. DCHR verify the depression category of the real
questionnaire with the category of the automated obtained answers.</p>
      <p>
        More details about these used measures and examples are given in [
        <xref ref-type="bibr" rid="ref17">17</xref>
        ].
      </p>
    </sec>
    <sec id="sec-6">
      <title>Results</title>
      <p>
        In this year, 5 teams submitted 17 different runs to the eRisk task 2. Our team
submitted 4 runs to this task. Table 2 shows our results comparing to the best
results for each metric. The results of all participants can be found in [
        <xref ref-type="bibr" rid="ref18">18</xref>
        ]. We
observe that no single team was able to achieve the best results for each of the
four metrics evaluation.
      </p>
      <p>Comparing only our runs, run 2 and run 4 did not perform well on the
dataset. Although, run 1 performed the best in the AHR with 34.97% of the
answers right and the best in the DCHR which is able to predict the correct
depression severity category for 25.71% of the users. In contrast to the 3 runs,
run 3 had higher ACR (67.78%) and ADODL (79.30%) scores. Therefore, we
notice that using the first statistical method based on the frequency of answers
is better than using the second method based on the higher length sequence of
answers.</p>
      <p>Most importantly, we believe combining the CNN model with the BiLSTM
one could improve the feature extraction process and enhance the model’s
performance to predict better results.
The aim of this article is to exploit artificial intelligence’s deep learning models
in order to measure automatically the severity of the signs of depression from an
individual’s posts. We described the participation of our research group at task
2 of the CLEF eRisk 2020 using using the CNN model, the BiLSTM model and
statistical methods to generate runs.</p>
      <p>We conclude that no run was able to predict overall depression better than
other because this task is not easy and a training dataset with 20 users is not
sufficient.</p>
      <p>For future work, we plan to principally combine the CNN model with the
BiLSTM one and we will analyze in more details the obtained results.</p>
    </sec>
  </body>
  <back>
    <ref-list>
      <ref id="ref1">
        <mixed-citation>1. Mpqa resources. http://mpqa.cs.pitt.edu/#subj_lexicon</mixed-citation>
      </ref>
      <ref id="ref2">
        <mixed-citation>
          2.
          <string-name>
            <surname>Abed-Esfahani</surname>
            ,
            <given-names>P.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Howard</surname>
            ,
            <given-names>D.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Maslej</surname>
            ,
            <given-names>M.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Patel</surname>
            ,
            <given-names>S.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Mann</surname>
            ,
            <given-names>V.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Goegan</surname>
            ,
            <given-names>S.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>French</surname>
            ,
            <given-names>L.</given-names>
          </string-name>
          :
          <article-title>Transfer learning for depression: Early detection and severity prediction from social media postings</article-title>
          .
          <source>In: CLEF (Working Notes)</source>
          (
          <year>2019</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref3">
        <mixed-citation>
          3.
          <string-name>
            <surname>Baziotis</surname>
            ,
            <given-names>C.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Pelekis</surname>
            ,
            <given-names>N.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Doulkeridis</surname>
            ,
            <given-names>C.</given-names>
          </string-name>
          : Datastories at semeval
          <article-title>-2017 task 4: Deep lstm with attention for message-level and topic-based sentiment analysis</article-title>
          .
          <source>In: Proceedings of the 11th international workshop on semantic evaluation (SemEval2017)</source>
          . pp.
          <fpage>747</fpage>
          -
          <lpage>754</lpage>
          (
          <year>2017</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref4">
        <mixed-citation>
          4.
          <string-name>
            <surname>Beck</surname>
            ,
            <given-names>A.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Ward</surname>
            ,
            <given-names>C.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Mendelson</surname>
            ,
            <given-names>M.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Mock</surname>
            ,
            <given-names>J.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Erbaugh</surname>
            ,
            <given-names>J.:</given-names>
          </string-name>
          <article-title>An inventory for measuring depression</article-title>
          .
          <source>archives of general psychiatry</source>
          , vol.
          <volume>4</volume>
          (
          <year>1961</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref5">
        <mixed-citation>
          5.
          <string-name>
            <surname>Burdisso</surname>
            ,
            <given-names>S.G.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Errecalde</surname>
            ,
            <given-names>M.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Montes-y Gómez</surname>
            ,
            <given-names>M.:</given-names>
          </string-name>
          <article-title>A text classification framework for simple and effective early depression detection over social media streams</article-title>
          .
          <source>Expert Systems with Applications</source>
          <volume>133</volume>
          ,
          <fpage>182</fpage>
          -
          <lpage>197</lpage>
          (
          <year>2019</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref6">
        <mixed-citation>
          6.
          <string-name>
            <surname>Burdisso</surname>
            ,
            <given-names>S.G.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Errecalde</surname>
            ,
            <given-names>M.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Montes-y Gómez</surname>
            ,
            <given-names>M.</given-names>
          </string-name>
          : Unsl at erisk
          <year>2019</year>
          :
          <article-title>a unified approach for anorexia, self-harm and depression detection in social media</article-title>
          .
          <source>In: CLEF (Working Notes)</source>
          (
          <year>2019</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref7">
        <mixed-citation>
          7.
          <string-name>
            <surname>Cacheda</surname>
            ,
            <given-names>F.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Iglesias</surname>
            ,
            <given-names>D.F.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Nóvoa</surname>
            ,
            <given-names>F.J.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Carneiro</surname>
            ,
            <given-names>V.</given-names>
          </string-name>
          :
          <article-title>Analysis and experiments on early detection of depression</article-title>
          .
          <source>CLEF (Working Notes)</source>
          <volume>2125</volume>
          (
          <year>2018</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref8">
        <mixed-citation>
          8.
          <string-name>
            <surname>Elman</surname>
            ,
            <given-names>J.L.</given-names>
          </string-name>
          :
          <article-title>Finding structure in time</article-title>
          .
          <source>Cognitive science 14</source>
          (
          <issue>2</issue>
          ),
          <fpage>179</fpage>
          -
          <lpage>211</lpage>
          (
          <year>1990</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref9">
        <mixed-citation>
          9.
          <string-name>
            <surname>Fatima</surname>
            ,
            <given-names>B.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Amina</surname>
            ,
            <given-names>M.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Nachida</surname>
            ,
            <given-names>R.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Hamza</surname>
          </string-name>
          , H.:
          <article-title>A mixed deep learning based model to early detection of depression</article-title>
          .
          <source>Journal of Web</source>
          Engineering pp.
          <fpage>429</fpage>
          -
          <lpage>456</lpage>
          (
          <year>2020</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref10">
        <mixed-citation>
          10.
          <string-name>
            <surname>Feurer</surname>
            ,
            <given-names>M.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Klein</surname>
            ,
            <given-names>A.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Eggensperger</surname>
            ,
            <given-names>K.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Springenberg</surname>
            ,
            <given-names>J.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Blum</surname>
            ,
            <given-names>M.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Hutter</surname>
            ,
            <given-names>F.</given-names>
          </string-name>
          :
          <article-title>Efficient and robust automated machine learning</article-title>
          .
          <source>In: Advances in neural information processing systems</source>
          . pp.
          <fpage>2962</fpage>
          -
          <lpage>2970</lpage>
          (
          <year>2015</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref11">
        <mixed-citation>
          11.
          <string-name>
            <surname>Graves</surname>
            ,
            <given-names>A.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Schmidhuber</surname>
            ,
            <given-names>J.</given-names>
          </string-name>
          :
          <article-title>Framewise phoneme classification with bidirectional lstm and other neural network architectures</article-title>
          .
          <source>Neural networks 18(5-6)</source>
          ,
          <fpage>602</fpage>
          -
          <lpage>610</lpage>
          (
          <year>2005</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref12">
        <mixed-citation>
          12.
          <string-name>
            <surname>Graves</surname>
            ,
            <given-names>A.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Schmidhuber</surname>
            ,
            <given-names>J.</given-names>
          </string-name>
          :
          <article-title>Offline handwriting recognition with multidimensional recurrent neural networks</article-title>
          .
          <source>In: Advances in neural information processing systems</source>
          . pp.
          <fpage>545</fpage>
          -
          <lpage>552</lpage>
          (
          <year>2009</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref13">
        <mixed-citation>
          13.
          <string-name>
            <surname>Hochreiter</surname>
            ,
            <given-names>S.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Schmidhuber</surname>
            ,
            <given-names>J.:</given-names>
          </string-name>
          <article-title>Long short-term memory</article-title>
          .
          <source>Neural computation 9(8)</source>
          ,
          <fpage>1735</fpage>
          -
          <lpage>1780</lpage>
          (
          <year>1997</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref14">
        <mixed-citation>
          14.
          <string-name>
            <surname>Kraskov</surname>
            ,
            <given-names>A.S.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Stogbauer</surname>
          </string-name>
          , H.: H. &amp;
          <article-title>grassberger. Estimating mutual information</article-title>
          .
          <source>Phys. Rev. E</source>
          <volume>69</volume>
          (
          <issue>6</issue>
          ) (
          <year>2004</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref15">
        <mixed-citation>
          15.
          <string-name>
            <surname>LeCun</surname>
          </string-name>
          , Y.,
          <string-name>
            <surname>Bengio</surname>
            ,
            <given-names>Y.</given-names>
          </string-name>
          , et al.:
          <article-title>Convolutional networks for images, speech, and time series</article-title>
          .
          <source>The handbook of brain theory and neural networks</source>
          <volume>3361</volume>
          (
          <issue>10</issue>
          ),
          <year>1995</year>
          (
          <year>1995</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref16">
        <mixed-citation>
          16.
          <string-name>
            <surname>Losada</surname>
            ,
            <given-names>D.E.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Crestani</surname>
            ,
            <given-names>F.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Parapar</surname>
            ,
            <given-names>J.</given-names>
          </string-name>
          :
          <article-title>Early detection of risks on the internet: an exploratory campaign</article-title>
          .
          <source>In: European Conference on Information Retrieval</source>
          . pp.
          <fpage>259</fpage>
          -
          <lpage>266</lpage>
          . Springer (
          <year>2019</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref17">
        <mixed-citation>
          17.
          <string-name>
            <surname>Losada</surname>
            ,
            <given-names>D.E.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Crestani</surname>
            ,
            <given-names>F.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Parapar</surname>
          </string-name>
          , J.:
          <article-title>Overview of erisk 2019 early risk prediction on the internet</article-title>
          .
          <source>In: International Conference of the Cross-Language Evaluation Forum for European Languages</source>
          . pp.
          <fpage>340</fpage>
          -
          <lpage>357</lpage>
          . Springer (
          <year>2019</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref18">
        <mixed-citation>
          18.
          <string-name>
            <surname>Losada</surname>
            ,
            <given-names>D.E.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Crestani</surname>
            ,
            <given-names>F.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Parapar</surname>
          </string-name>
          , J.:
          <article-title>Overview of erisk 2019 early risk prediction on the internet</article-title>
          .
          <source>In: International Conference of the Cross-Language Evaluation Forum for European Languages</source>
          . Springer (
          <year>2020</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref19">
        <mixed-citation>
          19.
          <string-name>
            <surname>Maupomé</surname>
            ,
            <given-names>D.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Meurs</surname>
            ,
            <given-names>M.J.:</given-names>
          </string-name>
          <article-title>Using topic extraction on social media content for the early detection of depression</article-title>
          .
          <source>CLEF (Working Notes)</source>
          <volume>2125</volume>
          (
          <year>2018</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref20">
        <mixed-citation>
          20.
          <string-name>
            <surname>Mikolov</surname>
            ,
            <given-names>T.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Sutskever</surname>
            ,
            <given-names>I.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Chen</surname>
            ,
            <given-names>K.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Corrado</surname>
            ,
            <given-names>G.S.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Dean</surname>
            ,
            <given-names>J.</given-names>
          </string-name>
          :
          <article-title>Distributed representations of words and phrases and their compositionality</article-title>
          .
          <source>In: Advances in neural information processing systems</source>
          . pp.
          <fpage>3111</fpage>
          -
          <lpage>3119</lpage>
          (
          <year>2013</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref21">
        <mixed-citation>
          21.
          <string-name>
            <surname>Nadeem</surname>
            ,
            <given-names>M.</given-names>
          </string-name>
          :
          <article-title>Identifying depression on twitter</article-title>
          .
          <source>arXiv preprint arXiv:1607.07384</source>
          (
          <year>2016</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref22">
        <mixed-citation>
          22.
          <string-name>
            <surname>Paul</surname>
            ,
            <given-names>S.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Jandhyala</surname>
            ,
            <given-names>S.K.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Basu</surname>
            ,
            <given-names>T.</given-names>
          </string-name>
          :
          <article-title>Early detection of signs of anorexia and depression over social media using effective machine learning frameworks</article-title>
          .
          <source>In: CLEF (Working Notes)</source>
          (
          <year>2018</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref23">
        <mixed-citation>
          23.
          <string-name>
            <surname>Pennebaker</surname>
            ,
            <given-names>J.W.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Chung</surname>
            ,
            <given-names>C.K.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Ireland</surname>
            ,
            <given-names>M.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Gonzales</surname>
            ,
            <given-names>A.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Booth</surname>
            ,
            <given-names>R.J.:</given-names>
          </string-name>
          <article-title>The development and psychometric properties of liwc2007: Liwc. net</article-title>
          . Google
          <string-name>
            <surname>Scholar</surname>
          </string-name>
          (
          <year>2007</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref24">
        <mixed-citation>
          24.
          <string-name>
            <surname>Pennington</surname>
            ,
            <given-names>J.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Socher</surname>
            ,
            <given-names>R.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Manning</surname>
          </string-name>
          , C.D.: Glove:
          <article-title>Global vectors for word representation</article-title>
          .
          <source>In: Proceedings of the 2014 conference on empirical methods in natural language processing (EMNLP)</source>
          . pp.
          <fpage>1532</fpage>
          -
          <lpage>1543</lpage>
          (
          <year>2014</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref25">
        <mixed-citation>
          25.
          <string-name>
            <surname>Radford</surname>
            ,
            <given-names>A.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Narasimhan</surname>
            ,
            <given-names>K.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Salimans</surname>
            ,
            <given-names>T.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Sutskever</surname>
            ,
            <given-names>I.</given-names>
          </string-name>
          :
          <article-title>Improving language understanding by generative pre-training (</article-title>
          <year>2018</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref26">
        <mixed-citation>
          26.
          <string-name>
            <surname>Radloff</surname>
            ,
            <given-names>L.S.:</given-names>
          </string-name>
          <article-title>The ces-d scale: A self-report depression scale for research in the general population</article-title>
          .
          <source>Applied psychological measurement 1</source>
          (
          <issue>3</issue>
          ),
          <fpage>385</fpage>
          -
          <lpage>401</lpage>
          (
          <year>1977</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref27">
        <mixed-citation>
          27.
          <string-name>
            <surname>Ragheb</surname>
            ,
            <given-names>W.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Moulahi</surname>
            ,
            <given-names>B.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Azé</surname>
            ,
            <given-names>J.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Bringay</surname>
            ,
            <given-names>S.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Servajean</surname>
            ,
            <given-names>M.</given-names>
          </string-name>
          :
          <article-title>Temporal mood variation: at the clef erisk-2018 tasks for early risk detection on the internet (</article-title>
          <year>2018</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref28">
        <mixed-citation>
          28.
          <string-name>
            <surname>Stankevich</surname>
            ,
            <given-names>M.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Isakov</surname>
            ,
            <given-names>V.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Devyatkin</surname>
            ,
            <given-names>D.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Smirnov</surname>
            ,
            <given-names>I.</given-names>
          </string-name>
          :
          <article-title>Feature engineering for depression detection in social media</article-title>
          .
          <source>In: ICPRAM</source>
          . pp.
          <fpage>426</fpage>
          -
          <lpage>431</lpage>
          (
          <year>2018</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref29">
        <mixed-citation>
          29.
          <string-name>
            <surname>Trifan</surname>
            ,
            <given-names>A.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Oliveira</surname>
            ,
            <given-names>J.L.</given-names>
          </string-name>
          :
          <article-title>Bioinfo@ uavr at erisk 2019: delving into social media texts for the early detection of mental and food disorders</article-title>
          .
          <source>In: CLEF (Working Notes)</source>
          (
          <year>2019</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref30">
        <mixed-citation>
          30.
          <string-name>
            <surname>Trotzek</surname>
            ,
            <given-names>M.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Koitka</surname>
            ,
            <given-names>S.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Friedrich</surname>
            ,
            <given-names>C.M.</given-names>
          </string-name>
          :
          <article-title>Utilizing neural networks and linguistic metadata for early detection of depression indications in text sequences</article-title>
          .
          <source>IEEE Transactions on Knowledge and Data Engineering</source>
          (
          <year>2018</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref31">
        <mixed-citation>
          31.
          <string-name>
            <surname>Trotzek</surname>
            ,
            <given-names>M.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Koitka</surname>
            ,
            <given-names>S.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Friedrich</surname>
            ,
            <given-names>C.M.:</given-names>
          </string-name>
          <article-title>Word embeddings and linguistic metadata at the clef 2018 tasks for early detection of depression and anorexia</article-title>
          .
          <source>In: CLEF (Working Notes)</source>
          (
          <year>2018</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref32">
        <mixed-citation>
          32. Van Rijen,
          <string-name>
            <given-names>P.</given-names>
            ,
            <surname>Teodoro</surname>
          </string-name>
          ,
          <string-name>
            <given-names>D.</given-names>
            ,
            <surname>Naderi</surname>
          </string-name>
          ,
          <string-name>
            <given-names>N.</given-names>
            ,
            <surname>Mottin</surname>
          </string-name>
          ,
          <string-name>
            <given-names>L.</given-names>
            ,
            <surname>Knafou</surname>
          </string-name>
          ,
          <string-name>
            <given-names>J.</given-names>
            ,
            <surname>Jeffryes</surname>
          </string-name>
          ,
          <string-name>
            <given-names>M.</given-names>
            ,
            <surname>Ruch</surname>
          </string-name>
          ,
          <string-name>
            <surname>P.</surname>
          </string-name>
          :
          <article-title>A data-driven approach for measuring the severity of the signs of depression using reddit posts</article-title>
          .
          <source>In: CLEF (Working Notes)</source>
          (
          <year>2019</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref33">
        <mixed-citation>
          33.
          <string-name>
            <surname>Wang</surname>
            ,
            <given-names>Y.T.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Huang</surname>
            ,
            <given-names>H.H.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Chen</surname>
            ,
            <given-names>H.H.:</given-names>
          </string-name>
          <article-title>A neural network approach to early risk detection of depression and anorexia on social media text</article-title>
          .
          <source>In: CLEF (Working Notes)</source>
          (
          <year>2018</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref34">
        <mixed-citation>
          34. Wilson,
          <string-name>
            <given-names>T.</given-names>
            ,
            <surname>Wiebe</surname>
          </string-name>
          ,
          <string-name>
            <given-names>J.</given-names>
            ,
            <surname>Hoffmann</surname>
          </string-name>
          ,
          <string-name>
            <surname>P.</surname>
          </string-name>
          :
          <article-title>Recognizing contextual polarity in phraselevel sentiment analysis</article-title>
          .
          <source>In: Proceedings of human language technology conference and conference on empirical methods in natural language processing</source>
          . pp.
          <fpage>347</fpage>
          -
          <lpage>354</lpage>
          (
          <year>2005</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref35">
        <mixed-citation>
          35.
          <string-name>
            <surname>Zhang</surname>
            ,
            <given-names>Y.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Wang</surname>
            ,
            <given-names>J.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Zhang</surname>
            ,
            <given-names>X.</given-names>
          </string-name>
          :
          <article-title>Ynu-hpcc at semeval-2018 task 1: Bilstm with attention based sentiment analysis for affect in tweets</article-title>
          .
          <source>In: Proceedings of The 12th International Workshop on Semantic Evaluation</source>
          . pp.
          <fpage>273</fpage>
          -
          <lpage>278</lpage>
          (
          <year>2018</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref36">
        <mixed-citation>
          36.
          <string-name>
            <surname>Zhou</surname>
            ,
            <given-names>P.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Shi</surname>
            ,
            <given-names>W.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Tian</surname>
            ,
            <given-names>J.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Qi</surname>
            ,
            <given-names>Z.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Li</surname>
            ,
            <given-names>B.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Hao</surname>
            ,
            <given-names>H.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Xu</surname>
            ,
            <given-names>B.</given-names>
          </string-name>
          :
          <article-title>Attention-based bidirectional long short-term memory networks for relation classification</article-title>
          .
          <source>In: Proceedings of the 54th Annual Meeting of the Association for Computational Linguistics (Volume 2: Short Papers)</source>
          . pp.
          <fpage>207</fpage>
          -
          <lpage>212</lpage>
          (
          <year>2016</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref37">
        <mixed-citation>
          37.
          <string-name>
            <surname>Zhu</surname>
            ,
            <given-names>Y.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Kiros</surname>
            ,
            <given-names>R.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Zemel</surname>
            ,
            <given-names>R.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Salakhutdinov</surname>
            ,
            <given-names>R.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Urtasun</surname>
            ,
            <given-names>R.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Torralba</surname>
            ,
            <given-names>A.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Fidler</surname>
            ,
            <given-names>S.</given-names>
          </string-name>
          :
          <article-title>Aligning books and movies: Towards story-like visual explanations by watching movies and reading books</article-title>
          .
          <source>In: Proceedings of the IEEE international conference on computer vision</source>
          . pp.
          <fpage>19</fpage>
          -
          <lpage>27</lpage>
          (
          <year>2015</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref38">
        <mixed-citation>
          38.
          <string-name>
            <surname>Zung</surname>
            ,
            <given-names>W.W.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Richards</surname>
            ,
            <given-names>C.B.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Short</surname>
            ,
            <given-names>M.J.</given-names>
          </string-name>
          :
          <article-title>Self-rating depression scale in an outpatient clinic: further validation of the sds</article-title>
          .
          <source>Archives of general psychiatry 13(6)</source>
          ,
          <fpage>508</fpage>
          -
          <lpage>515</lpage>
          (
          <year>1965</year>
          )
        </mixed-citation>
      </ref>
    </ref-list>
  </back>
</article>