<!DOCTYPE article PUBLIC "-//NLM//DTD JATS (Z39.96) Journal Archiving and Interchange DTD v1.0 20120330//EN" "JATS-archivearticle1.dtd">
<article xmlns:xlink="http://www.w3.org/1999/xlink">
  <front>
    <journal-meta />
    <article-meta>
      <title-group>
        <article-title>Language Independent Linguistic Features and Transformers in a Multi-label Emotion Detection Challenge in Urdu using Nastalīq Script</article-title>
      </title-group>
      <contrib-group>
        <contrib contrib-type="author">
          <string-name>José Antonio García-Díaz</string-name>
          <email>joseantonio.garcia8@um.es</email>
          <xref ref-type="aff" rid="aff1">1</xref>
        </contrib>
        <contrib contrib-type="author">
          <string-name>Manuel Valencia-García</string-name>
          <email>manuelv@um.es</email>
          <email>valencia@um.es</email>
          <xref ref-type="aff" rid="aff1">1</xref>
        </contrib>
        <contrib contrib-type="author">
          <string-name>Gema Alcaraz Mármol</string-name>
          <email>Gema.Alcaraz@uclm.es</email>
          <xref ref-type="aff" rid="aff0">0</xref>
        </contrib>
        <contrib contrib-type="author">
          <string-name>Rafael Valencia-García</string-name>
          <xref ref-type="aff" rid="aff1">1</xref>
        </contrib>
        <contrib contrib-type="author">
          <string-name>of BERT</string-name>
        </contrib>
        <contrib contrib-type="author">
          <string-name>RoBERTA.</string-name>
        </contrib>
        <aff id="aff0">
          <label>0</label>
          <institution>Departamento de Filología Moderna, Universidad de Castilla-La Mancha</institution>
          ,
          <country country="ES">Spain</country>
        </aff>
        <aff id="aff1">
          <label>1</label>
          <institution>Facultad de Informática, Universidad de Murcia, Campus de Espinardo</institution>
          ,
          <addr-line>30100</addr-line>
          ,
          <country country="ES">Spain</country>
        </aff>
      </contrib-group>
      <pub-date>
        <year>2022</year>
      </pub-date>
      <abstract>
        <p>Emotion Analysis is a Natural Language Processing task whose objective is to obtain fine-grained emotions from a text. The understanding of emotions in written communication has applications in marketing, e-commerce and infodemiology among others. Besides, Emotion Analysis can be applied to identify threats that could represent a threat to citizens, from a Smart City perspective. In this working notes we describe the participation of the UMUTeam in the EmoThreat shared task, proposed at FIRE'2022 workshop. Out of the subtasks proposed, our team only participated in the main subtask, which consisted in a multi-label emotion classification based on Ekman's six basic emotions in documents written in Urdu using Nastalīq script. We achieved the second best result, from a total of 8 participants, achieving 66.9% of macro average F1-score. Our proposal combines in the same neural network four feature sets that include a subset of language independent linguistic features extracted from UMUTextStats, a non-contextual sentence embeddings from fastText and two contextual sentence embeddings from multilingual versions 0000-0002-3651-2660 (J. A. García-Díaz); 0000-0001-7703-3829 (G. A. Mármol); 0000-0003-2457-1791</p>
      </abstract>
      <kwd-group>
        <kwd>Detection</kwd>
        <kwd>Nastalīq</kwd>
      </kwd-group>
    </article-meta>
  </front>
  <body>
    <sec id="sec-1">
      <title>1. Introduction</title>
      <p>The proliferation of social media platforms has made it easier for people all over the world
to communicate and share experiences. It has also provided benefits in international trade
and improved public health policies, as social media posts can refer threats that potentially
endanger citizens. Natural Language Processing (NLP) tools provide a useful way to process
and to understand what the users want to express in an automatic manner. However, NLP
tools have some challenges, highlight that some of the methods and state-of-the-art tools and
datasets are based on English, and the fact that natural language is complex to understand, as it
is highly subjective.
(R. Valencia-García)</p>
      <p>
        The organisers of the EmoThreat 2022 shared task [
        <xref ref-type="bibr" rid="ref1">1</xref>
        ] release a dataset for conducting
multilabel emotion detection in written Urdu using Nastalīq script (written from right to left). Urdu
language is spoken by more than 170 million people worldwide, highlighting India, Pakistan
and Nepal. The underlying objective of this shared task is to understand public emotions from
social networks applicable in NLP tools that can help to monitor events such as disasters, or to
improve public policies in e-commerce and public health.
      </p>
      <p>
        These working notes describe the participation of the UMUTeam at EmoThreat 2022 shared
task [
        <xref ref-type="bibr" rid="ref2">2</xref>
        ], proposed in FIRE 2022 [
        <xref ref-type="bibr" rid="ref3">3</xref>
        ]. The main challenge in this shared task is a multi-label
emotion classification task in Urdu [
        <xref ref-type="bibr" rid="ref4">4</xref>
        ]. The participants of the shared task are required to
classify each document with one, or more of the Ekman’s six basic emotions (plus one neutral
emotion).
      </p>
      <p>
        It is worth noting that our team has previous experience dealing with emotion classification.
Specifically, we participate in the EmoEvalEs 2021 shared task [
        <xref ref-type="bibr" rid="ref5">5</xref>
        ], achieving the sixth position
[
        <xref ref-type="bibr" rid="ref6">6</xref>
        ]. This shared task consisted in a multi-classification emotion detection with texts written in
Spanish. However, this shared task allowed our team to continue validating subsets of language
independent linguistic features tools that we have already applied in other languages such as
Tamil [
        <xref ref-type="bibr" rid="ref7">7</xref>
        ].
      </p>
      <p>The remainder of these working notes is organised as follows. First, Section 2 provides
a review of related work focused in Urdu and emotion detection. Second, in Section 3, the
developed pipeline for solving this task is described. Third, Section 4 includes the results
achieved in the challenge and a comparison with the rest of runs submitted by our team and by
the rest of the participants. Forth, Section 5 presents the findings obtained and it also includes
some promising research lines.</p>
    </sec>
    <sec id="sec-2">
      <title>2. Related work</title>
      <p>
        The EmoThreat 2022 shared task is a continuation of a previously shared task [
        <xref ref-type="bibr" rid="ref8">8</xref>
        ]. In the previous
edition of this shared task, the organisers proposed two challenges of abusive and threatening
language detection in Urdu. The past shared task consisted only in binary classification tasks,
one for detecting abusive documents (2400 documents for training, 1100 for testing) and another
shared task for detecting threatening messages (6000 documents for training, 3950 documents
for testing). The datasets were extracted from the micro-blogging platform Twitter. A total of
10 teams submitted their proposals for the abusive classification task and 9 for the threatening
classification task. The best result was achieved with a F1-score value of 0.880 for Subtask A
and 0.545 for Subtask B, using both run architectures based on Transformers.
      </p>
      <p>
        Sentiment Analysis is another NLP field that has been explored in Urdu. Recent works such
as the one described at [
        <xref ref-type="bibr" rid="ref9">9</xref>
        ] performed a multi-classification task from an Urdu dataset of 9312
reviews manually annotated compiled from user reviews about food, movies, apps, politics and
sports. The dataset was annotated with three labels (positive, negative and neutral). The authors
explored diferent baselines based on traditional machine-learning, deep-learning and models
based on multilingual Transformers. Their experiments confirmed that multilingual BERT
outperforms traditional models for Urdu, reaching an F1 score of 81.49%. In [
        <xref ref-type="bibr" rid="ref10">10</xref>
        ], the authors
compile a dataset in Urdu for sentiment analysis and evaluate several traditional machine
learning classifiers. The features were extracted using count-based techniques and pre-trained
word embeddings from fastText. The authors found that the combination of features of these
features outperformed the results achieved separately and compared with existing
state-of-theart approaches, reaching a F1-score of 82.05%. Another relevant work focused on Sentiment
Analysis in Urdu is [
        <xref ref-type="bibr" rid="ref11">11</xref>
        ], in which the authors [
        <xref ref-type="bibr" rid="ref11">11</xref>
        ] evaluated several word embeddings using an
architecture based on convolutional and recurrent neural networks, combined with traditional
machine learning classifiers for the final classification. The authors evaluate their proposal
with four corpora. Among the diferent architectures evaluated, the authors achieve their best
results using a classifier based on Support Vector Machines and features based on Word2Vec
based on Continuous Bag of Words.
      </p>
    </sec>
    <sec id="sec-3">
      <title>3. Methodology</title>
      <p>The first step of our pipeline is to explore the dataset and to create a custom validation split.
The validation split is created using a stratified sample in a ratio of 80-20. In Table 1 it can
be observed that there is an important imbalance among the emotions, being fear, anger, and
disgust underrepresented.</p>
      <p>As we deal with a multi-label classification challenge, we analyse the co-occurrence of
emotions (see Figure 1). As it can be observed, anger and disgust are the two sentiments that
usually appear together in the same tweet. It is also noticed that the neutral class is not used
combined with other sentiments.</p>
      <p>The next step in our pipeline is the feature extraction. Four diferent feature sets are involved
in our participation. The first feature set is a subset of language independent linguistic features
extracted with UMUTextStats (LF) [12, 13, 14, 15]. The second feature set are non-contextual
sentence embeddings from FastText (SE) [16]. The third and forth feature sets correspond to
multilingual contextual embeddings from BERT (BF) [17] and RoBERTa (RF) [18].</p>
      <p>To obtain the contextual sentence embeddings from BF and RF we do hyperparameter tuning.
A total of 20 transformers models (10 for BF, 10 for RF) were trained using the EmoThreat
training split, and deciding which the best model is by using our custom validation split. From
the best model, we extracted [CLS] token [19]. The hyperparameters involved in this process
are 1) the weight decay, 2) the batch size, 3) the warm-up speed, 4) the number of epochs, and
the 5) learning rate. The combination of these hyperparameters is performed using Tree of
Parzen Estimators (TPE) [20].</p>
      <p>Once all feature sets had been extracted, we evaluated two strategies for combining the
strengths of each one. The first strategy is called Knowledge Integration (KI), and consists in
training a multi-input deep-learning model that combines all feature sets at once. The second
strategy involves ensemble learning (EL), which combines the predictions of models focused on
one specific feature set. For this, we obtain a model for each feature set using hyperparameter
tuning (described below) and then, we evaluate two ways to use ensemble learning. The first
strategy is soft voting, which consists in calculating the mode of the predictions. The second
strategy corresponds to average probabilities, which consists in averaging the probabilities
predicted of each individual model to generate the final prediction.</p>
      <p>Regardless the training of the KI or the ensemble learning, we perform a hyperparameters
tuning stage. The hyperparameters involved are the shape of the network (that is, the number
of neurons and the number of hidden layers), the dropout mechanism, the learning rate and
several activation functions. The results of the hyperparameter tuning stage can be found in
Table 2.</p>
      <p>In all cases, the best result is achieved with shallow neural networks (that is, neural networks
with only one or two hidden layers, and the same number of neurons in all layers). Besides,
except for RF, all experiments achieved better results with a small dropout rate of .1. The learning
rate varies, being 0.001 for SE, RF and KI. In case of the activation function, all experiments
achieved better results with non-linear activation functions except LF.</p>
      <p>anger
disgust
fear
happiness
neutral
sadness
surprise
1
0.63
0.081
0.036
0
0.29
0.21
anger
0.68
0.075
0.021
1
0
0.39
0.25
disgust
0.11
0.094
1
0.1
0
0.51
0.079
fear
0.028
0.015
0.058
1
0
0.069
0.16
happiness
0
0
0
0
1
0
0
neutral
0.11
0.14
0.14
0.033
0
1
0.31
sadness
0.11
0.12
0.031
0.11
0.44
0
1
surprise</p>
    </sec>
    <sec id="sec-4">
      <title>4. Results and analysis</title>
      <p>First, we report in Table 3 the macro average results achieved by each feature set (for the
EL strategy) and the KI strategy. It can be observed that the best results achieved separately
are obtained with RF. However, when combined with the rest of the feature sets, the recall is
higher and the precision is lower. As it is expected, the results achieved by LF are limited, as
they are based on stylometry and PoS features. It draw out attention the limited recall of the
embeddings based on BERT compared with the embeddings based on RoBERTa. As we deal
with classification tasks, we consider that this diference is not related to the tasks in which
these models has been trained (Next Sentence Prediction and Masked Language Model), but
with the tokenizer and the dataset used to learn the embeddings.</p>
      <p>Next, we report in Table 4 the results per emotion achieved with the KI strategy using
the custom validation split. It can be observed that the model reaches almost a perfect score
concerning documents without attached emotions. In non-neutral documents, all emotions
achieve similar scores. Sadness reaches the best f1-score and happiness gets very good precision
but limited recall.</p>
      <p>The results of the oficial leader board are reported in Table 5. The results are ranked using
the Macro F1 score. The rest of the evaluated metrics are the multi-label accuracy, the Hamming
loss and the micro and weighted versions of the F1-score. We achieve the second best position
with our run based on KI.</p>
      <p>The results of our three runs are depicted in Table 6. As it can be observed, the results
achieved with ensemble learning are more limited in all metrics. The reason for this is that
the majority of correct predictions are performed by the RF feature set and the contribution of
the rest of the feature sets dismisses the performance of the model applying ensemble learning
strategies.
4.1. Error Analyses
For the error analysis we get our best run and collect the wrong predictions with the test split.
Next, we sort the multi-label output by euclidean distance in order to get the predictions with
the higher number of wrong labels. We obtained that the wrong classifications represent the
39.18% of total test split. 129 documents get one wrong label, 529 two wrong labels, 92 three
wrong labels and 14 four wrong labels.</p>
      <p>Next, we present the most notable failures. It is worth noting that the texts presented here
are translated using Google Translator. The most distant classifications made by our system are
those in which our system could not be able to identify any emotion. That is the case of: 1)
You interpret me very well. You hate Maulana Tariq Jameel Sahib. Fear Allah. Those holy persons
should respect him., and 2) Even if you express the pain in a happy way, the pain will still hurt..
For the first sentence, we consider that the problem is that the application does not have enough
context to understand the sentence. For the second sentence, we consider that the text does not
express any emotion but a refrain. The case of 3) The pain and sadness of Imran Niazi on the
death of such a close friend of Imran Niazi is not seen, or even Imran Niazi is just the opposite.
This document was rated as anger, sadness, and surprise, but the annotators did not find any
emotion in the document. Besides, we identified other documents related to SARS-Covid 2019
diseases. That is the case of 4) Those who ask for permission to open shops are not afraid of Corona.
Watch the program Live with Nasrullah Malik only New, and 5) Sami Ibrahim sir, what are you
most afraid of Corona till now? Imran Ahmed Khan Niazi still scares me the most. Besides, there
are other errors with short texts. That is the case of I hate this game. In this case, our system
correctly predicted the anger emotion, but missclassified disgust with sadness, which can be
considered a minor mistake.</p>
    </sec>
    <sec id="sec-5">
      <title>5. Conclusions</title>
      <p>We achieved the second position in a multi-label classification task in Urdu (66.9% of macro
F1-score), in which our pipeline is based on the combination of a subset of independent linguistic
features and transformers. Our best result combines the features using a knowledge integration
strategy; however, the runs submitted with ensemble learning achieved limited results, losing
several positions in the oficial ranking. Although we are very happy with our participation,
as we have evaluated our tools with non-Latin languages, we could not participate in the
second subtask of the competition due to lack of time. The source code is available at: https:
//github.com/Smolky/umuteam-emothreat-2022</p>
      <p>As promising future research lines, we would include nested cross validation to prevent
the hyperparameter tinning stages from being biased to the custom validation split and we
will apply data augmentation to increase the number of instances and reduce the efects of
class imbalance. We also explore the reliability of using transformers focused on Urdu rather
than multilingual. Besides, we will include features concerning figurative language [ 21], as its
identification may increase the generalisation of emotion analysis detectors. Another research
line is to apply emotion analysis to authors profiling tasks. In this sense, we are planning to
extend the PoliticES 2022 shared task [22] to compile tweets from politicians and journalist and
to extract emotions per author profile. For this, we will use the UMUCorpusClassifier tool [ 23].</p>
    </sec>
    <sec id="sec-6">
      <title>Acknowledgments</title>
      <p>This work is part of the research project LaTe4PSP (PID2019-107652RB-I00) funded by MCIN/
AEI/10.13039/501100011033. This work is also part of the research projects AIInFunds
(PDC2021121112-I00) funded by MCIN/AEI/10.13039/501100011033, by the European Union
NextGenerationEU/PRTR, LT-SWM (TED2021-131167B-I00) funded by MCIN/AEI/10.13039/501100011033
and by the European Union NextGenerationEU/PRTR, and by “Programa para la Recualificación
del Sistema Universitario Español 2021-2023”. In addition, José Antonio García-Díaz is supported
by Banco Santander and the University of Murcia through the Doctorado Industrial programme.
[12] J. A. García-Díaz, P. J. Vivancos-Vicente, Á. Almela, R. Valencia-García, Umutextstats: A
linguistic feature extraction tool for spanish, in: Proceedings of the Language Resources
and Evaluation Conference, European Language Resources Association, Marseille, France,
2022, pp. 6035–6044. URL: https://aclanthology.org/2022.lrec-1.649.
[13] J. A. García-Díaz, R. Colomo-Palacios, R. Valencia-García, Psychographic traits
identification based on political ideology: An author analysis study on spanish politicians’ tweets
posted in 2020, Future Generation Computer Systems 130 (2022) 59–74.
[14] J. A. García-Díaz, S. M. Jiménez-Zafra, M. A. García-Cumbreras, R. Valencia-García,
Evaluating feature combination strategies for hate-speech detection in spanish using linguistic
features and transformers, Complex &amp; Intelligent Systems (2022) 1–22.
[15] J. A. García-Díaz, R. Valencia-García, Compilation and evaluation of the spanish saticorpus
2021 for satire identification using linguistic features and transformers, Complex &amp;
Intelligent Systems 8 (2022) 1723–1736.
[16] E. Grave, P. Bojanowski, P. Gupta, A. Joulin, T. Mikolov, Learning word vectors
for 157 languages, CoRR abs/1802.06893 (2018). URL: http://arxiv.org/abs/1802.06893.
arXiv:1802.06893.
[17] J. Devlin, M. Chang, K. Lee, K. Toutanova, BERT: pre-training of deep bidirectional
transformers for language understanding, CoRR abs/1810.04805 (2018). URL: http://arxiv.
org/abs/1810.04805. arXiv:1810.04805.
[18] A. Conneau, K. Khandelwal, N. Goyal, V. Chaudhary, G. Wenzek, F. Guzmán, E. Grave,
M. Ott, L. Zettlemoyer, V. Stoyanov, Unsupervised cross-lingual representation
learning at scale, CoRR abs/1911.02116 (2019). URL: http://arxiv.org/abs/1911.02116.
arXiv:1911.02116.
[19] N. Reimers, I. Gurevych, Sentence-bert: Sentence embeddings using siamese
bertnetworks, in: Proceedings of the 2019 Conference on Empirical Methods in Natural
Language Processing, Association for Computational Linguistics, 2019, pp. 3982–3992.</p>
      <p>URL: https://arxiv.org/abs/1908.10084.
[20] J. Bergstra, D. Yamins, D. Cox, Making a science of model search: Hyperparameter
optimization in hundreds of dimensions for vision architectures, in: International conference
on machine learning, PMLR, 2013, pp. 115–123.
[21] M. del Pilar Salas-Zárate, G. Alor-Hernández, J. L. Sánchez-Cervantes, M. A.
ParedesValverde, J. L. García-Alcaraz, R. Valencia-García, Review of english literature on figurative
language applied to social networks, Knowledge Information Systems 62 (2020) 2105–2137.</p>
      <p>URL: https://doi.org/10.1007/s10115-019-01425-3. doi:10.1007/s10115- 019- 01425- 3.
[22] J. A. García-Díaz, S. M. J. Zafra, M. T. M. Valdivia, F. García-Sánchez, L. A. U. López,
R. Valencia-García, Overview of politices 2022: Spanish author profiling for political
ideology, Proces. del Leng. Natural 69 (2022) 265–272. URL: http://journal.sepln.org/sepln/
ojs/ojs/index.php/pln/article/view/6446.
[23] J. A. García-Díaz, Á. Almela, G. Alcaraz-Mármol, R. Valencia-García, Umucorpusclassifier:
Compilation and evaluation of linguistic corpus for natural language processing tasks,
Proces. del Leng. Natural 65 (2020) 139–142. URL: http://journal.sepln.org/sepln/ojs/ojs/
index.php/pln/article/view/6292.</p>
    </sec>
  </body>
  <back>
    <ref-list>
      <ref id="ref1">
        <mixed-citation>
          [1]
          <string-name>
            <given-names>S.</given-names>
            <surname>Butt</surname>
          </string-name>
          ,
          <string-name>
            <given-names>M.</given-names>
            <surname>Amjad</surname>
          </string-name>
          ,
          <string-name>
            <given-names>F.</given-names>
            <surname>Balouchzahi</surname>
          </string-name>
          ,
          <string-name>
            <given-names>N.</given-names>
            <surname>Ashraf</surname>
          </string-name>
          ,
          <string-name>
            <given-names>R.</given-names>
            <surname>Sharma</surname>
          </string-name>
          ,
          <string-name>
            <given-names>G.</given-names>
            <surname>Sidorov</surname>
          </string-name>
          ,
          <string-name>
            <given-names>A.</given-names>
            <surname>Gelbukh</surname>
          </string-name>
          ,
          <source>Overview of EmoThreat: Emotions and Threat Detection in Urdu at FIRE</source>
          <year>2022</year>
          , in: CEUR Workshop Proceedings,
          <year>2022</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref2">
        <mixed-citation>
          [2]
          <string-name>
            <given-names>N.</given-names>
            <surname>Ashraf</surname>
          </string-name>
          ,
          <string-name>
            <given-names>L.</given-names>
            <surname>Khan</surname>
          </string-name>
          ,
          <string-name>
            <given-names>S.</given-names>
            <surname>Butt</surname>
          </string-name>
          , H.-T. Chang,
          <string-name>
            <given-names>G.</given-names>
            <surname>Sidorov</surname>
          </string-name>
          ,
          <string-name>
            <given-names>A.</given-names>
            <surname>Gelbukh</surname>
          </string-name>
          <article-title>, Multi-label emotion classification of urdu tweets</article-title>
          ,
          <source>PeerJ Computer Science</source>
          <volume>8</volume>
          (
          <year>2022</year>
          )
          <article-title>e896</article-title>
          .
        </mixed-citation>
      </ref>
      <ref id="ref3">
        <mixed-citation>
          [3]
          <string-name>
            <given-names>S.</given-names>
            <surname>Butt</surname>
          </string-name>
          ,
          <string-name>
            <given-names>M.</given-names>
            <surname>Amjad</surname>
          </string-name>
          ,
          <string-name>
            <given-names>F.</given-names>
            <surname>Balouchzahi</surname>
          </string-name>
          ,
          <string-name>
            <given-names>N.</given-names>
            <surname>Ashraf</surname>
          </string-name>
          ,
          <string-name>
            <given-names>R.</given-names>
            <surname>Sharma</surname>
          </string-name>
          ,
          <string-name>
            <given-names>G.</given-names>
            <surname>Sidorov</surname>
          </string-name>
          ,
          <string-name>
            <surname>A</surname>
          </string-name>
          . Gelbukh, EmoThreat@FIRE2022:
          <article-title>Shared Track on Emotions and Threat Detection in Urdu, in: Forum for Information Retrieval Evaluation</article-title>
          ,
          <string-name>
            <surname>FIRE</surname>
          </string-name>
          <year>2022</year>
          ,
          <article-title>Association for Computing Machinery</article-title>
          , New York, NY, USA,
          <year>2022</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref4">
        <mixed-citation>
          [4]
          <string-name>
            <given-names>I.</given-names>
            <surname>Ameer</surname>
          </string-name>
          ,
          <string-name>
            <given-names>N.</given-names>
            <surname>Ashraf</surname>
          </string-name>
          , G. Sidorov,
          <string-name>
            <surname>H.</surname>
          </string-name>
          <article-title>Gómez Adorno, Multi-label emotion classification using content-based features in twitter</article-title>
          ,
          <source>Computación y Sistemas</source>
          <volume>24</volume>
          (
          <year>2020</year>
          )
          <fpage>1159</fpage>
          -
          <lpage>1164</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref5">
        <mixed-citation>
          [5]
          <string-name>
            <given-names>F. M.</given-names>
            <surname>Plaza-del Arco</surname>
          </string-name>
          ,
          <string-name>
            <given-names>S. M.</given-names>
            <surname>Jiménez-Zafra</surname>
          </string-name>
          ,
          <string-name>
            <given-names>A.</given-names>
            <surname>Montejo-Ráez</surname>
          </string-name>
          ,
          <string-name>
            <given-names>M. D.</given-names>
            <surname>Molina-González</surname>
          </string-name>
          ,
          <string-name>
            <given-names>L. A.</given-names>
            <surname>Ureña-López</surname>
          </string-name>
          ,
          <string-name>
            <given-names>M. T.</given-names>
            <surname>Martín-Valdivia</surname>
          </string-name>
          ,
          <article-title>Overview of the emoevales task on emotion detection for spanish at iberlef 2021</article-title>
          ,
          <source>Procesamiento del Lenguaje Natural</source>
          <volume>67</volume>
          (
          <year>2021</year>
          )
          <fpage>155</fpage>
          -
          <lpage>161</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref6">
        <mixed-citation>
          [6]
          <string-name>
            <given-names>J. A.</given-names>
            <surname>García-Díaz</surname>
          </string-name>
          ,
          <string-name>
            <given-names>R. C.</given-names>
            <surname>Palacios</surname>
          </string-name>
          ,
          <string-name>
            <given-names>R.</given-names>
            <surname>Valencia-García</surname>
          </string-name>
          , Umuteam at emoevales 2021:
          <article-title>Emotion analysis for spanish based on explainable linguistic features and transformers</article-title>
          ,
          <source>in: Proceedings of the Iberian Languages Evaluation Forum (IberLEF</source>
          <year>2021</year>
          )
          <article-title>co-located with the Conference of the Spanish Society for Natural Language Processing (SEPLN 2021), XXXVII International Conference of the Spanish Society for Natural Language Processing</article-title>
          .,
          <string-name>
            <surname>Málaga</surname>
          </string-name>
          , Spain, September,
          <year>2021</year>
          , volume
          <volume>2943</volume>
          <source>of CEUR Workshop Proceedings, CEUR-WS.org</source>
          ,
          <year>2021</year>
          , pp.
          <fpage>59</fpage>
          -
          <lpage>71</lpage>
          . URL: http://ceur-ws.
          <source>org/</source>
          Vol-
          <volume>2943</volume>
          /emoeval_paper6.pdf.
        </mixed-citation>
      </ref>
      <ref id="ref7">
        <mixed-citation>
          [7]
          <string-name>
            <given-names>J.</given-names>
            <surname>García-Díaz</surname>
          </string-name>
          ,
          <string-name>
            <given-names>M.</given-names>
            <surname>Á. R. García</surname>
          </string-name>
          ,
          <string-name>
            <given-names>R.</given-names>
            <surname>Valencia-García</surname>
          </string-name>
          , Umuteam@
          <fpage>tamilnlp</fpage>
          -acl2022:
          <article-title>Emotional analysis in tamil</article-title>
          ,
          <source>in: Proceedings of the Second Workshop on Speech and Language Technologies for Dravidian Languages</source>
          ,
          <year>2022</year>
          , pp.
          <fpage>39</fpage>
          -
          <lpage>44</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref8">
        <mixed-citation>
          [8]
          <string-name>
            <given-names>M.</given-names>
            <surname>Amjad</surname>
          </string-name>
          ,
          <string-name>
            <given-names>A.</given-names>
            <surname>Zhila</surname>
          </string-name>
          ,
          <string-name>
            <given-names>G.</given-names>
            <surname>Sidorov</surname>
          </string-name>
          ,
          <string-name>
            <given-names>A.</given-names>
            <surname>Labunets</surname>
          </string-name>
          ,
          <string-name>
            <given-names>S.</given-names>
            <surname>Butt</surname>
          </string-name>
          ,
          <string-name>
            <given-names>H. I.</given-names>
            <surname>Amjad</surname>
          </string-name>
          ,
          <string-name>
            <given-names>O.</given-names>
            <surname>Vitman</surname>
          </string-name>
          ,
          <string-name>
            <given-names>A.</given-names>
            <surname>Gelbukh</surname>
          </string-name>
          , Urduthreat@ fire2021:
          <article-title>Shared track on abusive threat identification in urdu</article-title>
          ,
          <source>in: Forum for Information Retrieval Evaluation</source>
          ,
          <year>2021</year>
          , pp.
          <fpage>9</fpage>
          -
          <lpage>11</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref9">
        <mixed-citation>
          [9]
          <string-name>
            <given-names>L.</given-names>
            <surname>Khan</surname>
          </string-name>
          ,
          <string-name>
            <given-names>A.</given-names>
            <surname>Amjad</surname>
          </string-name>
          ,
          <string-name>
            <given-names>N.</given-names>
            <surname>Ashraf</surname>
          </string-name>
          , H.-
          <article-title>T. Chang, Multi-class sentiment analysis of urdu text using multilingual bert</article-title>
          ,
          <source>Scientific Reports</source>
          <volume>12</volume>
          (
          <year>2022</year>
          )
          <fpage>1</fpage>
          -
          <lpage>17</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref10">
        <mixed-citation>
          [10]
          <string-name>
            <given-names>L.</given-names>
            <surname>Khan</surname>
          </string-name>
          ,
          <string-name>
            <given-names>A.</given-names>
            <surname>Amjad</surname>
          </string-name>
          ,
          <string-name>
            <given-names>N.</given-names>
            <surname>Ashraf</surname>
          </string-name>
          , H.-
          <string-name>
            <given-names>T.</given-names>
            <surname>Chang</surname>
          </string-name>
          ,
          <string-name>
            <given-names>A.</given-names>
            <surname>Gelbukh</surname>
          </string-name>
          ,
          <article-title>Urdu sentiment analysis with deep learning methods</article-title>
          ,
          <source>IEEE Access 9</source>
          (
          <year>2021</year>
          )
          <fpage>97803</fpage>
          -
          <lpage>97812</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref11">
        <mixed-citation>
          [11]
          <string-name>
            <given-names>L.</given-names>
            <surname>Khan</surname>
          </string-name>
          ,
          <string-name>
            <given-names>A.</given-names>
            <surname>Amjad</surname>
          </string-name>
          ,
          <string-name>
            <surname>K. M. Afaq</surname>
          </string-name>
          , H.-T. Chang,
          <article-title>Deep sentiment analysis using cnn-lstm architecture of english and roman urdu text shared in social media</article-title>
          ,
          <source>Applied Sciences</source>
          <volume>12</volume>
          (
          <year>2022</year>
          )
          <fpage>2694</fpage>
          .
        </mixed-citation>
      </ref>
    </ref-list>
  </back>
</article>