<!DOCTYPE article PUBLIC "-//NLM//DTD JATS (Z39.96) Journal Archiving and Interchange DTD v1.0 20120330//EN" "JATS-archivearticle1.dtd">
<article xmlns:xlink="http://www.w3.org/1999/xlink">
  <front>
    <journal-meta>
      <journal-title-group>
        <journal-title>Forum for Information Retrieval Evaluation, Decemeber</journal-title>
      </journal-title-group>
    </journal-meta>
    <article-meta>
      <title-group>
        <article-title>Leveraging Sentiment Data for the Detection of Homophobic/Transphobic Content in a Multi-Task, Multi-Lingual Setting Using Transformers</article-title>
      </title-group>
      <contrib-group>
        <contrib contrib-type="author">
          <string-name>Filip Nilsson</string-name>
          <xref ref-type="aff" rid="aff0">0</xref>
          <xref ref-type="aff" rid="aff1">1</xref>
        </contrib>
        <contrib contrib-type="author">
          <string-name>Sana Sabah Al-Azzawi</string-name>
          <email>sana.al-azzawi@ltu.se</email>
          <xref ref-type="aff" rid="aff0">0</xref>
          <xref ref-type="aff" rid="aff1">1</xref>
        </contrib>
        <contrib contrib-type="author">
          <string-name>György Kovács</string-name>
          <xref ref-type="aff" rid="aff0">0</xref>
          <xref ref-type="aff" rid="aff1">1</xref>
        </contrib>
        <contrib contrib-type="editor">
          <string-name>Transphobic Language.</string-name>
        </contrib>
        <aff id="aff0">
          <label>0</label>
          <institution>EISLAB Machine Learning, Luleå University of Technology</institution>
          ,
          <addr-line>977 54 Luleå</addr-line>
          ,
          <country country="SE">Sweden</country>
        </aff>
        <aff id="aff1">
          <label>1</label>
          <institution>Multi-Task</institution>
          ,
          <addr-line>Multi-Language Learning, Hateful Language, Sentiment Analysis, Detecting Homophobic/-</addr-line>
        </aff>
      </contrib-group>
      <pub-date>
        <year>2022</year>
      </pub-date>
      <volume>0</volume>
      <fpage>9</fpage>
      <lpage>13</lpage>
      <abstract>
        <p>Hateful content is published and spread on social media at an increasing rate, harming the user experience. In addition, hateful content targeting particular, marginalized/vulnerable groups (e.g. homophobic/transphobic content) can cause even more harm to members of said groups. Hence, detecting hateful content is crucial, regardless of its origin, or the language used. The large variety of (often underresourced) languages used, however, makes this task daunting, especially as many users use code-mixing in their messages. To help overcome these dificulties, the approach we present here uses a multi-language framework. And to further mitigate the scarcity of labelled data, it also leverages data from the related task of sentiment-analysis to improve the detection of homophobic/transphobic content. We evaluated our system by participating in a sentiment analysis and hate speech detection challenge. Results show that our multi-task model outperforms its single-task counterpart (on average, by 24%) on the detection of homophobic/transphobic content. Moreover, the results achieved in detecting homophobic/transphobic content put our system in 1st or 2nd place for three out of four languages examined.</p>
      </abstract>
      <kwd-group>
        <kwd>Using Transformers</kwd>
      </kwd-group>
    </article-meta>
  </front>
  <body>
    <sec id="sec-1">
      <title>1. Introduction</title>
      <p>
        The increasing use of social media such as Twitter and Youtube has escalated the exploitation of
these platforms to propagate violence [
        <xref ref-type="bibr" rid="ref1">1</xref>
        ]. This violence can take the form of hateful, ofensive,
and abusive language causing harm [
        <xref ref-type="bibr" rid="ref2 ref3">2, 3</xref>
        ]. To help preventing this harm, social media has been
analyzed using various methods designed to detect ofensive language, or more particularly,
detect hateful language [
        <xref ref-type="bibr" rid="ref4">4</xref>
        ], and homophobic/transphobic content [
        <xref ref-type="bibr" rid="ref5">5</xref>
        ]. Homophobic/Transphobic
content is a type of hateful language intending to harm LGBT+ people. Unfortunately, the
shortage of labeled data has limited research in this area, especially in low resource languages [
        <xref ref-type="bibr" rid="ref6">6</xref>
        ].
One approach to overcome this problem is leveraging additional data for improving the detection
of hateful language. Among others, Kovács et al. [
        <xref ref-type="bibr" rid="ref7 ref8">7, 8</xref>
        ] examined this option by (among other
methods) using additional datasets created for the same task. In this paper, we extend this idea
by leveraging datasets from related tasks, as well as datasets in diferent languages to improve
the detection of homophobic/transphobic content in a multi-task, multi-language setting.
https://github.com/flippe3/fire_2022 (F. Nilsson)
      </p>
      <p>© 2022 Copyright for this paper by its authors. Use permitted under Creative Commons License Attribution 4.0 International (CC BY 4.0).</p>
      <p>
        For these experiments, we use the ”Sentiment Analysis and Homophobia detection of YouTube
comments in Code-Mixed Dravidian Languages” (hereinafter DravidianCodeMix) challenge at
FIRE [
        <xref ref-type="bibr" rid="ref9">9</xref>
        ] which presents two shared tasks. One task for sentiment analysis (Task A), and another
one for the detection of homophobic/transphobic content (Task B). The two tasks combined
provided seven datasets with altogether four diferent languages. We participated in both tasks
and used multi-task learning to fine-tune an XLM-RoBERTA (XLM-R) pre-trained language
model to deal with multilingualism [
        <xref ref-type="bibr" rid="ref10">10</xref>
        ].
      </p>
      <p>In this paper, we describe our approach and the results it attained. First, we discuss the related
literature in Section 2. Then, in Section 3 we describe the tasks and datases of the challenge.
This is followed by the description of our methods in Section 4, after which we present our
experiments and results in Section 5. In Section 6 we analyze our experiments, then share our
conclusions and plans for future work in Section 7.</p>
    </sec>
    <sec id="sec-2">
      <title>2. Related Work</title>
      <p>Here, we discuss part of the related literature providing relevant background to the
DravidianCodeMix challenge, including work on detection of hateful language in general, and
homophobic/transphobic language in particular. We also discuss diferent approaches to multi-task
learning, its use in Natural Language Processing (NLP) and for the detection of hateful language.</p>
      <sec id="sec-2-1">
        <title>2.1. Hateful Language</title>
        <p>
          Hateful language is any insult directed against a person or group based on their protected
category that aims to cause damage or stir hatred [
          <xref ref-type="bibr" rid="ref11">11</xref>
          ]. Hateful posts/comments incur the risk
of prompting a shift towards violence. Moreover, this content can potentially cause a harmful
emotional efect to its readers. For these reasons, automatic hate speech identification has
attracted a lot of attention, and many approaches have been investigated for the detection of
hate speech, as well as other relevant areas [
          <xref ref-type="bibr" rid="ref12">12</xref>
          ].
        </p>
        <p>
          Existing methods mainly deal with hateful language detection as a classification task. These
approaches can be categorized into two groups, namely traditional machine learning methods
such as SVM [
          <xref ref-type="bibr" rid="ref11 ref7">7, 11</xref>
          ], logistic regression [
          <xref ref-type="bibr" rid="ref13">13</xref>
          ], random forest [
          <xref ref-type="bibr" rid="ref11">11</xref>
          ], gradient boosting decision
tree models [
          <xref ref-type="bibr" rid="ref14">14</xref>
          ] and naive Bayes [
          <xref ref-type="bibr" rid="ref11">11</xref>
          ]. The other main category is that of deep learning-based
methods, which can be further partitioned into deep learning architecture, word embedding
based methods, and transformer-based methods. Transformer-based methods utilize pre-trained
transformer models (e.g. BERT, ELECTRA, T5), and fine-tune them on datasets annotated for
hateful language detection to detect harmful remarks in social media posts. These methods show
remarkable performance on harm identification across diferent languages [ 15, 16, 17, 18, 19].
        </p>
        <p>
          Homophobic/Transphobic comments are usually categorized as a type of hateful language
directed toward LGBT+ individuals. This phenomenon has been a growing concern. Chakravarthi
et al. [
          <xref ref-type="bibr" rid="ref9">9</xref>
          ] collected and created a dataset in 2021 for Homophobic classification. The dataset
is a collection of 15,141 YouTube multilingual comments on social media. In addition, they
made detection experiments using several classical ML and DL models as baselines, and in 2022
they organized a shared task based on the data collected, at an ACL workshop [
          <xref ref-type="bibr" rid="ref5">5</xref>
          ], to promote
research on identifying homophobic/transphobic content.
        </p>
      </sec>
      <sec id="sec-2-2">
        <title>2.2. Multi-Task Learning</title>
        <p>One method to combat data scarcity in supervised learning is multi-task learning (MTL). In MTL
models are simultaneously trained on multiple related tasks to improve their performance and
generalization ability. MTL has shown promising results in various areas, including computer
vision, reinforcement learning, speech processing [20], and NLP, where MTL has been studied
in various levels of relatedness, goals, and features [21].</p>
        <p>One prominent example of the application of MTL in NLP is the T5 [22] text-to-text
transformer model. This model was primarily trained using a combination of supervised and
unsupervised learning, and an MTL approach, which was shown to have a positive influence.
The authors also experimented with various types of mixing, namely examples-proportional,
temperature-scaled and equal mixing. Results indicated that examples-proportional mixing
leads to the best performance, while equal mixing can degrade the performance due to the model
overfitting on low-resource tasks. Other researchers working on MTL in NLP [ 23] demonstrated
that joint fine-tuning a model on multiple languages could bring substantial improvements to
the performance of a universal language encoder. For this, they used a tree-like structure with
16 heads, one for each language. The authors found that it was only in rare cases when this
joint MTL training hurt the performance of their model.</p>
        <p>MTL has also been used in the context of detecting hate speech [24]. Here, the authors used
a multi-task model to detect hate speech in Spanish using related tasks of polarity and emotion
classification to improve their model. They showed that their MTL system with task-specific
output heads outperformed its single-task counterpart and achieved state-of-the-art results.</p>
      </sec>
    </sec>
    <sec id="sec-3">
      <title>3. Data</title>
      <p>The shared challenge includes two main tasks (Task A and Task B), both coupled with annotated
datasets comprised of Youtube comments in various languages. Task A being a sentiment
analysis problem [25], where the goal is to classify each comment into one of five categories:
Positive, Negative, Mixed feelings, Unknown state, and comment not in the target language.
The detailed statistics and the split of the datasets are shown in Table 1. As can be seen in
Table 1, the data is relatively imbalanced, both in terms of the labels, and languages (the Tamil
dataset having more examples than the other two languages put together). Moreover, we can
also see that the diferent partitions have largely diferent class label distributions.</p>
      <p>
        The main objective of Task B was to identify whether a comment was Transphobic,
Homophobic, or Non-Anti-LGBT (Safe) [
        <xref ref-type="bibr" rid="ref9">9</xref>
        ]. The datasets provided for this task consist of Tamil,
English, Malayalam, and the remaining code-mixed Tamil-English. The detailed statistics and
the split of the datasets are shown in Table 2. As Table 2 shows, the Non-Anti-LGBT (safe)
class outweighs the other two labels. This is most prominent in the English dataset, where
transphobic messages represent less than 0.3% of the full data.
      </p>
      <p>Some examples of the comments are also shown in Table 3. As can be seen from the table,
two kinds of code-mixing is present in the data set. One, where diferent languages are mixed
(e.g. the comment labelled as ”mixed feelings” in row 6). We can see examples of the other
type of code-mixing as well, where native and Roman scripts are mixed (e.g. the Non-LGBT
comment in row 8).</p>
      <p>We hypothesized that the two tasks of the challenge are closely related. That is, negative
sentiment would be more strongly associated with homophobic/transphobic comments than
with safe ones, and the opposite would be true for the positive sentiment. To verify this, we
trained a classifier for the Tamil dataset in Task-A (sentiment analysis), and applied it on the
Tamil dataset in Task-B. Results of these experiments are shown in Figure 1b. As can be seen
here, the negative sentiment on average was higher in homophobic and transphobic content
than in safe comments. While for positive sentiment, on average higher scores were attained
for safe comments than for transphobic ones. Although this was not the case for homophobic
comments, the results partially supporting our hypothesis encouraged us to further examine
the relation between the two tasks.</p>
    </sec>
    <sec id="sec-4">
      <title>4. Methodology</title>
      <p>One of our targets during our experiments was to examine how the task of identifying hateful
comments can benefit from the availability of a sentiment analysis dataset. For this, we decided to
work in a multi-task paradigm to benefit from the connection between the two tasks (supported
by the sentiment scores attained using the data from the task of identifying hateful language in
Figure 1b). For this, as our model, we chose to use a multi-lingual language model pre-trained
on 100 diferent languages that included all languages in our datasets.</p>
      <sec id="sec-4-1">
        <title>4.1. Preprocessing</title>
        <p>
          To handle the irregularities often found in Youtube comments, we applied a preprocessing
step in our pipeline. First, we used the Hugging Face normalizer to remove blank spaces and
emojis. Then, we used the SentencePiece [26] tokenizer. SentecePiece is a language-independent
subword tokenizer we chose to use since we are dealing with diferent types of code-mixing, a
domain where SentencePiece has previously shown promising results [
          <xref ref-type="bibr" rid="ref5">5</xref>
          ].
4.2. Model
We used the XLM-RoBERTA (XLM-R) [
          <xref ref-type="bibr" rid="ref10">10</xref>
          ] multilingual sentence-transformer which is trained
using an XLM-R model as a student model and an SBERT [27] model as the teacher model.
        </p>
      </sec>
      <sec id="sec-4-2">
        <title>4.3. Multi-Task Training</title>
        <p>From our study of the data we hypothesized that hateful content would in general have a higher
degree of negativity (and a lower degree of positivity). Therefore a multi-task model should be
adequate to help improve the result for Task B. To train the model, we split the output heads
into seven diferent RoBERTA [ 28] output heads shown in Figure 2, similar to the architecture
used by Unicoder [23] that produced excellent results. This enabled us to fine-tune the language
model using both tasks and all seven datasets simultaneously.</p>
        <p>An essential part of MTL is the technique used to mix the tasks. In our case, due to the
imbalance among datasets and labels, equal mixing would have incurred the risk of missing
vital data points, negatively afecting the model. Thus, following the findings of [ 22] we used
examples-proportional mixing, which means sampling from tasks proportionally to their dataset
size.</p>
        <p>In our MTL training we experimented using diferent sized output layers. We did not find
any significant improvements when adding more layers, which can also be an indication of the
closely related nature of the two tasks. The final model we used for training uses a RoBERTa
classification head with two dense layers.</p>
        <p>The training of our models was done using 4 epochs with a learning rate of 3e-5. Furthermore,
we used a linear scheduler and an AdamW optimizer. These experiments were trained on a
shared DGX-1 cluster using 2 x 32GB Nvidia V100 GPUs.
5. Experiments and results
We performed four experiments to evaluate our proposed multi-task system, each with a
diferent architecture. In this section will present these architectures, as well as the results of
these experiments. The four diferent architectures we were using in our experiments were as
follows. 1) Single-task architecture: here, a separate model with only one output head was
trained for each individual language and task. 2) Multi-task learning with seven output
heads: here, a joint model was trained for all tasks and languages, equipped with one output
head for each language and task. That is, the model had three output heads for the sentiment
analysis (corresponding to the three languages), and had four more output heads (one for each
language in Task B). 3) Multi-task learning with two output heads: in this approach, similar
to the previous, all tasks and languages were used to train the same model. Here, however,
unlike in the previous case, our goal was to train only one output head for each task. Thus
one output head was trained to predict the labels of Task A and Task B respectively. This is
similar to the arcitecture used by the Spanish hate-speech detection [24]. 4) Language-specific
multi-task learning: similar to the two previous multi-task systems, but here we selected a
specific language and only used datasets containing that language.</p>
        <p>Results of our experiments on Task A are summarized in Table 4. As Table 4 shows, although
the single-language multi-task model attained markedly higher results than that reported
by the winning team, in most cases the use of multi-task architecture did not lead to marked
improvements compared to its single-task counterpart. Our goal with the multi-task architecture,
however, was not to improve the performance of the sentiment analysis model, but rather to
improve the detection of hateful speech. Results of our experiments on Task B are summarized
in Table 5. As Table 5 shows, we attained the best results on the test (and dev) set for all cases
using a multi-language multi-task architecture, or a single-language multi-task architecture.</p>
        <p>Based on the results attained on the dev set (where available at the deadline), we chose the
seven-headed multi-task multi-language model for our final submission. Results attained in
the oficial competition [ 29] by this model are summarized in table Table 6. The table shows
that this model performed relatively well in both tasks. The rankings achieved, however, were
markedly better for Task B, our main target. Thus, for the detection of hateful language, the
multi-task multi-language model beyond attaining an improved performance compared to its
single-task counterpart, also attained a competitive performance in the challenge.</p>
      </sec>
    </sec>
    <sec id="sec-5">
      <title>6. Analysis</title>
      <p>First, we examined our sentiment output heads when probed with hateful data to analyze further
whether our system is viable. In Figure 3a and Figure 3b we show the logits when probed with
hateful data through the matching sentiment head. In Figure 3a we show that the safe quartile is
greater than the homophobic and transphobic quartile. This suggests that the model learned an
expected feature (i.e. safe posts having more positivity, and less negativity than hateful ones).
(a) Positive sentiment
(b) Negative sentiment</p>
      <p>We also examined the confusion matrix, shown in Figure 4. This was an important tool of
analysis, as classifying homophobic comments as transphobic might not be a large problem, but
classifying either as safe is problematic. In Figure 4c we can see that five transphobic comments
got misclassified as safe. However, this is most likely due to the training data since there were
only six labels of transphobic data (see Table 2).</p>
    </sec>
    <sec id="sec-6">
      <title>7. Conclusion and future work</title>
      <p>We introduced a pipeline that fine-tunes a multi-lingual transformer using multi-task learning.
The pipeline uses sentiment data to improve homophobic/transphobic detection. Our
experiments show that the multi-task model outperforms the single-task model and that
languagespecific training can improve the accuracy further. We have also demonstrated that the sentiment
output heads of our model identify hateful content as more negative than safe content.</p>
      <p>Future work could focus on studying solely sentiment labels relevant for hateful speech. One
could also look at other types of task-mixing, such as temporal-scaled mixing. Diferent early
stopping methods should also be considered. Furthermore, data augmentation methods should
also be considered, to counteract the problem of data imbalance. Finally, one can experiment
with various tasks with diferent levels of relatedness and languages with varied similarities.</p>
    </sec>
    <sec id="sec-7">
      <title>8. Acknowledgments</title>
      <p>The work presented here was partially supported by Vinnova, in the project Language models
for Swedish authorities (Språkmodeller för svenska myndigheter) ref. no.: 2019-02996
[15] S. S. Sabry, T. Adewumi, N. Abid, G. Kovács, F. Liwicki, M. Liwicki, Hat5: Hate language
identification using text-to-text transfer transformer, arXiv preprint arXiv:2202.05690
(2022).
[16] T. Adewumi, S. S. Sabry, N. Abid, F. Liwicki, M. Liwicki, T5 for hate speech, augmented
data and ensemble, arXiv preprint arXiv:2210.05480 (2022).
[17] J. S. Malik, G. Pang, A. v. d. Hengel, Deep learning for hate speech detection: A comparative
study, arXiv preprint arXiv:2202.09517 (2022).
[18] P. Alonso, R. Saini, G. Kovács, Hate speech detection using transformer ensembles on the
hasoc dataset, in: International conference on speech and computer, Springer, 2020, pp.
13–21.
[19] E. Lavergne, R. Saini, G. Kovács, K. Murphy, Thenorth@ haspeede 2: Bert-based language
model fine-tuning for italian hate speech detection, 2020.
[20] Y. Zhang, Q. Yang, An overview of multi-task learning, National Science
Review 5 (2017) 30–43. URL: https://doi.org/10.1093/nsr/nwx105. doi:1 0 . 1 0 9 3 / n s r / n w x 1 0 5 .
a r X i v : h t t p s : / / a c a d e m i c . o u p . c o m / n s r / a r t i c l e - p d f / 5 / 1 / 3 0 / 3 1 5 6 7 3 5 8 / n w x 1 0 5 . p d f .
[21] S. Chen, Y. Zhang, Q. Yang, Multi-task learning in natural language processing: An
overview, 2021. URL: https://arxiv.org/abs/2109.09138. doi:1 0 . 4 8 5 5 0 / A R X I V . 2 1 0 9 . 0 9 1 3 8 .
[22] C. Rafel, N. Shazeer, A. Roberts, K. Lee, S. Narang, M. Matena, Y. Zhou, W. Li, P. J. Liu,
Exploring the limits of transfer learning with a unified text-to-text transformer, 2019. URL:
https://arxiv.org/abs/1910.10683. doi:1 0 . 4 8 5 5 0 / A R X I V . 1 9 1 0 . 1 0 6 8 3 .
[23] H. Huang, Y. Liang, N. Duan, M. Gong, L. Shou, D. Jiang, M. Zhou, Unicoder: A universal
language encoder by pre-training with multiple cross-lingual tasks, in: Proceedings
of the 2019 Conference on Empirical Methods in Natural Language Processing and the
9th International Joint Conference on Natural Language Processing (EMNLP-IJCNLP),
Association for Computational Linguistics, Hong Kong, China, 2019, pp. 2485–2494. URL:
https://aclanthology.org/D19-1252. doi:1 0 . 1 8 6 5 3 / v 1 / D 1 9 - 1 2 5 2 .
[24] F. M. Plaza-Del-Arco, M. D. Molina-González, L. A. Ureña-López, M. T. Martín-Valdivia,
A multi-task learning approach to hate speech detection leveraging sentiment analysis,
IEEE Access 9 (2021) 112478–112489. doi:1 0 . 1 1 0 9 / A C C E S S . 2 0 2 1 . 3 1 0 3 6 9 7 .
[25] B. R. Chakravarthi, R. Priyadharshini, V. Muralidaran, N. Jose, S. Suryawanshi, E. Sherly,
J. P. McCrae, Dravidiancodemix: sentiment analysis and ofensive language identification
dataset for dravidian languages in code-mixed text, Language Resources and
Evaluation 56 (2022) 765–806. URL: https://doi.org/10.1007/s10579-022-09583-7. doi:1 0 . 1 0 0 7 /
s 1 0 5 7 9 - 0 2 2 - 0 9 5 8 3 - 7 .
[26] T. Kudo, J. Richardson, SentencePiece: A simple and language independent subword
tokenizer and detokenizer for neural text processing, in: Proceedings of the 2018
Conference on Empirical Methods in Natural Language Processing: System Demonstrations,
Association for Computational Linguistics, Brussels, Belgium, 2018, pp. 66–71. URL:
https://aclanthology.org/D18-2012. doi:1 0 . 1 8 6 5 3 / v 1 / D 1 8 - 2 0 1 2 .
[27] N. Reimers, I. Gurevych, Making monolingual sentence embeddings multilingual using
knowledge distillation, in: Proceedings of the 2020 Conference on Empirical Methods
in Natural Language Processing, Association for Computational Linguistics, 2020. URL:
https://arxiv.org/abs/2004.09813.
[28] Y. Liu, M. Ott, N. Goyal, J. Du, M. Joshi, D. Chen, O. Levy, M. Lewis, L. Zettlemoyer,
V. Stoyanov, Roberta: A robustly optimized bert pretraining approach, 2019. URL: https:
//arxiv.org/abs/1907.11692. doi:1 0 . 4 8 5 5 0 / A R X I V . 1 9 0 7 . 1 1 6 9 2 .
[29] K. Shumugavadivel, M. Subramanian, P. K. Kumaresan, B. R. Chakravarthi, B. B, S.
Chinnaudayar Navaneethakrishnan, L. S.K, T. Mandl, R. Ponnusamy, V. Palanikumar, M. Balaji J,
Overview of the Shared Task on Sentiment Analysis and Homophobia Detection of YouTube
Comments in Code-Mixed Dravidian Languages, in: Working Notes of FIRE 2022 - Forum
for Information Retrieval Evaluation, CEUR, 2022.</p>
    </sec>
  </body>
  <back>
    <ref-list>
      <ref id="ref1">
        <mixed-citation>
          [1]
          <string-name>
            <given-names>A. A.</given-names>
            <surname>Siegel</surname>
          </string-name>
          , Online Hate Speech,
          <source>SSRC Anxieties of Democracy</source>
          , Cambridge University Press,
          <year>2020</year>
          , p.
          <fpage>56</fpage>
          -
          <lpage>88</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref2">
        <mixed-citation>
          [2]
          <string-name>
            <given-names>I.</given-names>
            <surname>Bigoulaeva</surname>
          </string-name>
          ,
          <string-name>
            <given-names>V.</given-names>
            <surname>Hangya</surname>
          </string-name>
          ,
          <string-name>
            <given-names>A.</given-names>
            <surname>Fraser</surname>
          </string-name>
          ,
          <article-title>Cross-lingual transfer learning for hate speech detection</article-title>
          ,
          <source>in: Proceedings of the First Workshop on Language Technology for Equality, Diversity and Inclusion</source>
          ,
          <year>2021</year>
          , pp.
          <fpage>15</fpage>
          -
          <lpage>25</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref3">
        <mixed-citation>
          [3]
          <string-name>
            <given-names>K.</given-names>
            <surname>Gelber</surname>
          </string-name>
          ,
          <string-name>
            <surname>L.</surname>
          </string-name>
          <article-title>McNamara, Evidencing the harms of hate speech</article-title>
          ,
          <source>Social Identities</source>
          <volume>22</volume>
          (
          <year>2016</year>
          )
          <fpage>324</fpage>
          -
          <lpage>341</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref4">
        <mixed-citation>
          [4]
          <string-name>
            <given-names>A.</given-names>
            <surname>Schmidt</surname>
          </string-name>
          ,
          <string-name>
            <given-names>M.</given-names>
            <surname>Wiegand</surname>
          </string-name>
          ,
          <article-title>A survey on hate speech detection using natural language processing</article-title>
          ,
          <source>in: Proceedings of the fith international workshop on natural language processing for social media</source>
          ,
          <year>2017</year>
          , pp.
          <fpage>1</fpage>
          -
          <lpage>10</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref5">
        <mixed-citation>
          [5]
          <string-name>
            <given-names>B. R.</given-names>
            <surname>Chakravarthi</surname>
          </string-name>
          ,
          <string-name>
            <given-names>R.</given-names>
            <surname>Priyadharshini</surname>
          </string-name>
          ,
          <string-name>
            <given-names>T.</given-names>
            <surname>Durairaj</surname>
          </string-name>
          ,
          <string-name>
            <surname>J. McCrae</surname>
          </string-name>
          ,
          <string-name>
            <given-names>P.</given-names>
            <surname>Buitelaar</surname>
          </string-name>
          ,
          <string-name>
            <given-names>P.</given-names>
            <surname>Kumaresan</surname>
          </string-name>
          ,
          <string-name>
            <given-names>R.</given-names>
            <surname>Ponnusamy</surname>
          </string-name>
          ,
          <article-title>Overview of the shared task on homophobia and transphobia detection in social media comments</article-title>
          ,
          <source>in: Proceedings of the Second Workshop on Language Technology for Equality, Diversity and Inclusion</source>
          , Association for Computational Linguistics, Dublin, Ireland,
          <year>2022</year>
          , pp.
          <fpage>369</fpage>
          -
          <lpage>377</lpage>
          . URL: https://aclanthology.org/
          <year>2022</year>
          .ltedi-
          <volume>1</volume>
          .57.
          <article-title>doi:1 0 . 1 8 6 5 3 / v 1 / 2 0 2 2</article-title>
          . l t e
          <source>d i - 1 . 5 7 .</source>
        </mixed-citation>
      </ref>
      <ref id="ref6">
        <mixed-citation>
          [6]
          <string-name>
            <given-names>M.</given-names>
            <surname>Singh</surname>
          </string-name>
          ,
          <string-name>
            <given-names>P.</given-names>
            <surname>Motlicek</surname>
          </string-name>
          ,
          <article-title>Idiap submission@ lt-edi-acl2022: Homophobia/transphobia detection in social media comments</article-title>
          ,
          <source>in: Proceedings of the Second Workshop on Language Technology for Equality, Diversity and Inclusion</source>
          ,
          <year>2022</year>
          , pp.
          <fpage>356</fpage>
          -
          <lpage>361</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref7">
        <mixed-citation>
          [7]
          <string-name>
            <given-names>G.</given-names>
            <surname>Kovács</surname>
          </string-name>
          ,
          <string-name>
            <given-names>P.</given-names>
            <surname>Alonso</surname>
          </string-name>
          ,
          <string-name>
            <given-names>R.</given-names>
            <surname>Saini</surname>
          </string-name>
          ,
          <article-title>Challenges of hate speech detection in social media</article-title>
          ,
          <source>SN Computer Science</source>
          <volume>2</volume>
          (
          <year>2021</year>
          )
          <fpage>1</fpage>
          -
          <lpage>15</lpage>
          .
          <source>doi:1 0 . 1 0 0 7 / s 4 2</source>
          <volume>9 7 9 - 0 2 1 - 0 0 4 5 7 - 3</volume>
          .
        </mixed-citation>
      </ref>
      <ref id="ref8">
        <mixed-citation>
          [8]
          <string-name>
            <given-names>G.</given-names>
            <surname>Kovács</surname>
          </string-name>
          ,
          <string-name>
            <given-names>P.</given-names>
            <surname>Alonso</surname>
          </string-name>
          ,
          <string-name>
            <given-names>R.</given-names>
            <surname>Saini</surname>
          </string-name>
          ,
          <string-name>
            <given-names>M.</given-names>
            <surname>Liwicki</surname>
          </string-name>
          ,
          <article-title>Leveraging external resources for ofensive content detection in social media</article-title>
          ,
          <source>AI Commun</source>
          .
          <volume>35</volume>
          (
          <year>2022</year>
          )
          <fpage>87</fpage>
          -
          <lpage>109</lpage>
          . URL: https://doi.org/10. 3233/AIC-210138.
          <source>doi:1 0 . 3 2 3 3 / A I C - 2</source>
          <volume>1 0 1 3 8 .</volume>
        </mixed-citation>
      </ref>
      <ref id="ref9">
        <mixed-citation>
          [9]
          <string-name>
            <given-names>B. R.</given-names>
            <surname>Chakravarthi</surname>
          </string-name>
          ,
          <string-name>
            <given-names>R.</given-names>
            <surname>Priyadharshini</surname>
          </string-name>
          ,
          <string-name>
            <given-names>R.</given-names>
            <surname>Ponnusamy</surname>
          </string-name>
          ,
          <string-name>
            <given-names>P. K.</given-names>
            <surname>Kumaresan</surname>
          </string-name>
          ,
          <string-name>
            <given-names>K.</given-names>
            <surname>Sampath</surname>
          </string-name>
          ,
          <string-name>
            <given-names>D.</given-names>
            <surname>Thenmozhi</surname>
          </string-name>
          ,
          <string-name>
            <given-names>S.</given-names>
            <surname>Thangasamy</surname>
          </string-name>
          ,
          <string-name>
            <given-names>R.</given-names>
            <surname>Nallathambi</surname>
          </string-name>
          ,
          <string-name>
            <given-names>J. P.</given-names>
            <surname>McCrae</surname>
          </string-name>
          ,
          <article-title>Dataset for identification of homophobia and transophobia in multilingual youtube comments</article-title>
          ,
          <source>arXiv preprint arXiv:2109.00227</source>
          (
          <year>2021</year>
          ).
        </mixed-citation>
      </ref>
      <ref id="ref10">
        <mixed-citation>
          [10]
          <string-name>
            <given-names>A.</given-names>
            <surname>Conneau</surname>
          </string-name>
          ,
          <string-name>
            <given-names>K.</given-names>
            <surname>Khandelwal</surname>
          </string-name>
          ,
          <string-name>
            <given-names>N.</given-names>
            <surname>Goyal</surname>
          </string-name>
          ,
          <string-name>
            <given-names>V.</given-names>
            <surname>Chaudhary</surname>
          </string-name>
          ,
          <string-name>
            <given-names>G.</given-names>
            <surname>Wenzek</surname>
          </string-name>
          ,
          <string-name>
            <given-names>F.</given-names>
            <surname>Guzmán</surname>
          </string-name>
          , E. Grave,
          <string-name>
            <given-names>M.</given-names>
            <surname>Ott</surname>
          </string-name>
          ,
          <string-name>
            <given-names>L.</given-names>
            <surname>Zettlemoyer</surname>
          </string-name>
          ,
          <string-name>
            <given-names>V.</given-names>
            <surname>Stoyanov</surname>
          </string-name>
          ,
          <article-title>Unsupervised cross-lingual representation learning at scale</article-title>
          ,
          <year>2019</year>
          . URL: https://arxiv.org/abs/
          <year>1911</year>
          .02116.
          <source>doi:1 0 . 4 8</source>
          <volume>5 5</volume>
          <fpage>0</fpage>
          <string-name>
            <surname>/ A R X I</surname>
          </string-name>
          <article-title>V . 1 9 1 1 . 0 2 1 1 6</article-title>
          .
        </mixed-citation>
      </ref>
      <ref id="ref11">
        <mixed-citation>
          [11]
          <string-name>
            <given-names>T.</given-names>
            <surname>Davidson</surname>
          </string-name>
          ,
          <string-name>
            <given-names>D.</given-names>
            <surname>Warmsley</surname>
          </string-name>
          ,
          <string-name>
            <given-names>M.</given-names>
            <surname>Macy</surname>
          </string-name>
          ,
          <string-name>
            <surname>I. Weber</surname>
          </string-name>
          ,
          <article-title>Automated hate speech detection and the problem of ofensive language</article-title>
          ,
          <source>in: Proceedings of the international AAAI conference on web and social media</source>
          , volume
          <volume>11</volume>
          ,
          <year>2017</year>
          , pp.
          <fpage>512</fpage>
          -
          <lpage>515</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref12">
        <mixed-citation>
          [12]
          <string-name>
            <given-names>S.</given-names>
            <surname>MacAvaney</surname>
          </string-name>
          , H.
          <string-name>
            <surname>-R. Yao</surname>
            ,
            <given-names>E.</given-names>
          </string-name>
          <string-name>
            <surname>Yang</surname>
            ,
            <given-names>K.</given-names>
          </string-name>
          <string-name>
            <surname>Russell</surname>
            ,
            <given-names>N.</given-names>
          </string-name>
          <string-name>
            <surname>Goharian</surname>
            ,
            <given-names>O.</given-names>
          </string-name>
          <string-name>
            <surname>Frieder</surname>
          </string-name>
          ,
          <article-title>Hate speech detection: Challenges and solutions</article-title>
          ,
          <source>PLOS ONE 14</source>
          (
          <year>2019</year>
          )
          <fpage>1</fpage>
          -
          <lpage>16</lpage>
          . URL: https://doi.org/10.1371
          <source>/journal. pone.0221152. doi:1 0 . 1 3</source>
          <volume>7 1</volume>
          / j o u r n a l .
          <source>p o n e . 0 2</source>
          <volume>2 1 1 5 2 .</volume>
        </mixed-citation>
      </ref>
      <ref id="ref13">
        <mixed-citation>
          [13]
          <string-name>
            <given-names>Z.</given-names>
            <surname>Waseem</surname>
          </string-name>
          ,
          <string-name>
            <given-names>D.</given-names>
            <surname>Hovy</surname>
          </string-name>
          ,
          <article-title>Hateful symbols or hateful people? predictive features for hate speech detection on twitter</article-title>
          ,
          <source>in: Proceedings of the NAACL student research workshop</source>
          ,
          <year>2016</year>
          , pp.
          <fpage>88</fpage>
          -
          <lpage>93</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref14">
        <mixed-citation>
          [14]
          <string-name>
            <given-names>E.</given-names>
            <surname>Katona</surname>
          </string-name>
          ,
          <string-name>
            <given-names>J.</given-names>
            <surname>Buda</surname>
          </string-name>
          ,
          <string-name>
            <given-names>F.</given-names>
            <surname>Bolonyai</surname>
          </string-name>
          ,
          <article-title>Using n-grams and statistical features to identify hate speech spreaders on twitter</article-title>
          .,
          <source>in: CLEF (Working Notes)</source>
          ,
          <year>2021</year>
          , pp.
          <fpage>2025</fpage>
          -
          <lpage>2034</lpage>
          .
        </mixed-citation>
      </ref>
    </ref-list>
  </back>
</article>