<!DOCTYPE article PUBLIC "-//NLM//DTD JATS (Z39.96) Journal Archiving and Interchange DTD v1.0 20120330//EN" "JATS-archivearticle1.dtd">
<article xmlns:xlink="http://www.w3.org/1999/xlink">
  <front>
    <journal-meta />
    <article-meta>
      <title-group>
        <article-title>Toward Multimodal Sentiment Analysis of Historic Plays: A Case Study with Text and Audio for Lessing's Emilia Galotti</article-title>
      </title-group>
      <contrib-group>
        <contrib contrib-type="author">
          <string-name>Thomas Schmidt</string-name>
          <email>thomas.schmidt@ur.de</email>
          <xref ref-type="aff" rid="aff1">1</xref>
        </contrib>
        <contrib contrib-type="author">
          <string-name>Manuel Burghardt</string-name>
          <email>burghardt@informatik.uni-leipzig.de</email>
          <xref ref-type="aff" rid="aff0">0</xref>
        </contrib>
        <contrib contrib-type="author">
          <string-name>Christian Wolff</string-name>
          <email>christian.wolff@ur.de</email>
          <xref ref-type="aff" rid="aff1">1</xref>
        </contrib>
        <aff id="aff0">
          <label>0</label>
          <institution>Computational Humanities Department, University of Leipzig</institution>
          ,
          <addr-line>04109 Leipzig</addr-line>
          ,
          <country country="DE">Germany</country>
        </aff>
        <aff id="aff1">
          <label>1</label>
          <institution>Media Informatics Group, University of Regensburg</institution>
          ,
          <addr-line>93040, Regensburg</addr-line>
          ,
          <country country="DE">Germany</country>
        </aff>
      </contrib-group>
      <fpage>405</fpage>
      <lpage>414</lpage>
      <abstract>
        <p>We present a case study as part of a work-in-progress project about multimodal sentiment analysis on historic German plays, taking Emilia Galotti by G. E. Lessing as our initial use case. We analyze the textual version and an audio version (audiobook). We focus on ready-to-use sentiment analysis methods: For the textual component, we implement a naive lexicon-based approach and another approach that enhances the lexicon by means of several NLP methods. For the audio analysis, we use the free version of the Vokaturi tool. We compare the results of all approaches and evaluate them against the annotations of a human expert, which serves as a gold standard. For our use case, we can show that audio and text sentiment analysis behave very differently: textual sentiment analysis tends to predict sentiment as rather negative and audio sentiment as rather positive. Compared to the gold standard, the textual sentiment analysis achieves accuracies of 56% while the accuracy for audio sentiment analysis is only 32%. We discuss possible reasons for these mediocre results and give an outlook on further steps we want to pursue in the context of multimodal sentiment analysis on historic plays.</p>
      </abstract>
      <kwd-group>
        <kwd>sentiment analysis</kwd>
        <kwd>emotion analysis</kwd>
        <kwd>multimodal</kwd>
        <kwd>multimedia</kwd>
        <kwd>computational literary studies</kwd>
        <kwd>audio</kwd>
        <kwd>audiobooks</kwd>
        <kwd>drama</kwd>
        <kwd>text mining</kwd>
        <kwd>Lessing</kwd>
      </kwd-group>
    </article-meta>
  </front>
  <body>
    <sec id="sec-1">
      <title>-</title>
      <p>
        Sentiment analysis is the area of research that deals with computational methods to
analyze and predict sentiments and emotions in written text
        <xref ref-type="bibr" rid="ref10">(Liu, 2016, p.1)</xref>
        . Although
the majority of work in this area is done with user-generated content like product
reviews and social media
        <xref ref-type="bibr" rid="ref20">(Vinodhini &amp; Chandrasekaran, 2012)</xref>
        , there is growing interest
in exploring the application of sentiment analysis in the digital humanities especially in
computational literary studies. Sentiment analysis is used to analyze fairy tales,
        <xref ref-type="bibr" rid="ref11 ref2">(Alm
et al., 2005; Mohammad, 2011)</xref>
        novels,
        <xref ref-type="bibr" rid="ref6 ref7">(Kakkonen &amp; Kakkonen, 2011; Jockers, 2015)</xref>
        historic plays,
        <xref ref-type="bibr" rid="ref11 ref13">(Mohammad, 2011; Nalisnick &amp; Baird, 2013)</xref>
        or to generate features for
various machine-learning tasks
        <xref ref-type="bibr" rid="ref5 ref9">(Jannidis et al., 2016; Kim, Padó &amp; Klinger, 2017)</xref>
        .
      </p>
      <p>
        Sentiment analysis as a whole and especially the current research in the digital
humanities is predominately focused on the analysis of written text; other media
channels like audio, video or combinations are neglected so far. We propose several reasons
for justifying the broadening of the current focus on written text to other modalities:
First, although systematic performance evaluation in the context of narrative texts is
rare, the existing research shows that when compared to human sentiment annotations,
the accuracy can vary between 20-70% depending on the method and the type of text
        <xref ref-type="bibr" rid="ref16 ref16 ref17 ref17 ref18 ref18 ref19 ref19 ref8">(Kim &amp; Klinger, 2018; Schmidt &amp; Burghardt, 2018a; Schmidt &amp; Burghardt 2018b)</xref>
        .
Therefore, it is far lower than in other areas of sentiment analysis in which accuracies
close to and above 90% are achieved
        <xref ref-type="bibr" rid="ref20">(Vinodhini &amp; Chandrasekaran, 2012)</xref>
        . Multimodal
sentiment analysis has been proven to be more successful than isolated text sentiment
analysis in several areas
        <xref ref-type="bibr" rid="ref1 ref12 ref14">(Morency et al., 2011; Abburi et al., 2016; Poria et al., 2017)</xref>
        and one can observe a general trend from unimodal to multimodal sentiment analysis
approaches
        <xref ref-type="bibr" rid="ref14">(Poria et al., 2017)</xref>
        . We hypothesize that narrative text might be especially
suited for multimodal sentiment analysis. One reason for the current low accuracies
might be that prosodic and voice-based features, which are important parts of the
narrative performance and the expression of emotions, are neglected in text sentiment
analysis. Especially in the context of plays, which are specifically designed for oral
performance in a theatre, the voice and the face of actors are important identifiers for emotion
and sentiment. Furthermore, audio and video/face sentiment analysis are far less
language dependent than the current text-based approaches
        <xref ref-type="bibr" rid="ref4">(Hudlicka, 2003)</xref>
        which is
especially appealing for research in multilingual literary studies.
      </p>
      <p>
        Second, we also see potential for using differing media channels to improve the
annotation of sentiment for narrative texts. Literary texts have been proven to be very
difficult and tedious to annotate due to the historic and complex language
        <xref ref-type="bibr" rid="ref16 ref16 ref17 ref17 ref18 ref18 ref19 ref19 ref2">(Schmidt,
Burghardt &amp; Dennerlein, 2018; Schmidt, Burghardt &amp; Wolff, 2018; Alm et al., 2005)</xref>
        .
The presentation of text material in multimodal form might improve and facilitate the
annotation process concerning sentiment. For example, annotators might not
understand the language and the context of a narrative text unit but the expressed emotion of
an actor in his oral performance. Improvement in the annotation process would enable
us to acquire annotated corpora for evaluation and machine learning purposes on a
larger scale more easily.
      </p>
      <p>In this paper, we contribute to “multimodality in digital humanities” by presenting
first work-in-progress results for multimodal sentiment analysis on narrative texts,
more precisely for the specific use case of the play Emilia Galotti by G. E. Lessing. We
have analyzed and compared existing text sentiment analysis approaches and a
readyto-use audio speech sentiment analysis. In addition, we have evaluated the performance
compared to a gold standard of annotations by a human expert. Finally, we discuss the
results and the limitations of this case study but also formulate an agenda for future
research in this area.</p>
    </sec>
    <sec id="sec-2">
      <title>Research Questions</title>
      <p>
        Sentiment analysis and emotion analysis are often differentiated in research. While
sentiment analysis is regarded as predicting and analyzing the overall affective state
predominantly with the classes positive, negative and neutral (we refer to these classes as
polarity), emotion analysis deals with more complex emotional categories like anger,
surprise or joy
        <xref ref-type="bibr" rid="ref20">(Vinodhini &amp; Chandrasekaran, 2012)</xref>
        . For this case study, we focus
solely on sentiment since its application is in general easier and more successful
        <xref ref-type="bibr" rid="ref10">(Liu,
2016, p.67)</xref>
        especially in the context of literary texts
        <xref ref-type="bibr" rid="ref2">(Alm et al., 2005)</xref>
        .
      </p>
      <p>As modalities for this study, we use the textual and an audio version (audiobook) of
Emilia Galotti. We want to explore the following research questions:</p>
      <p>RQ1: How do text sentiment analysis approaches perform compared to a
ready-to-use audio sentiment analysis approach?
We are focusing on simple and ready-to-use solutions, since not only is this study our
first exploration in this area, but also because it is a common use case in digital
humanities when the focus of research is not the development of new algorithms but the
analysis of cultural artifacts. With RQ1, we want to gather first insights in the possible
similarities and differences of both media channels and analysis types.</p>
      <p>RQ2: How do text sentiment analysis approaches and the audio sentiment
analysis perform against human expert annotations (with text)?
We want to use the sentiment annotations of an expert as gold standard and evaluate
the text and audio approach against each other. In our case study, the expert annotated
the sentiment by being presented with the text. With RQ2, we want to test our
assumption that audio speech conveys more precise emotional information and that audio
sentiment analysis therefore achieves higher accuracies.
3</p>
    </sec>
    <sec id="sec-3">
      <title>Methods and Data</title>
      <p>As use case for our investigations, we chose the play Emilia Galotti by G. E. Lessing
(premiered 1772). The reason for this is that our recent research is focused on Lessing
and Emilia Galotti is one of his most famous plays, which means that audio material
for this play is available, too. All analyses are speech based, thus we are only comparing
singular speeches with each other. A speech is a single utterance of a character
separated by utterances of other characters beforehand and afterwards. Overall, the play
consists of 835 speeches. The longest speech consists of 235 words while the shortest
speech only has one word. On average, a speech in this play consists of 23 words.</p>
      <p>
        As material for the textual sentiment analysis, we gathered an XML-annotated
version of the play from the platform Textgrid1. For the sentiment analysis, we employ two
lexicon-based approaches. A sentiment lexicon is a list of words annotated with
sentiment annotation. Based on simple calculation, text units can be assigned a polarity (e.g.
neutral, positive, negative). In previous work
        <xref ref-type="bibr" rid="ref16 ref17 ref18 ref19">(Schmidt &amp; Burghardt, 2018a, Schmidt
&amp; Burghardt, 2018b)</xref>
        we have evaluated different lexicons and NLP approaches for the
1 https://textgrid.de/ (Note: all URLs mentioned in this article were last checked Feb. 10, 2019)
best performance on a corpus of all of Lessing’s plays. For the use case in this paper
we analyze (1) the existing German sentiment lexicon SentiWS
        <xref ref-type="bibr" rid="ref15">(Remus et al., 2010)</xref>
        as
is and (2) a variant where we enhance SentiWS with additional NLP techniques and
preprocessing methods. The enhanced SentiWS approach achieved the highest
accuracy in a previous study on a Lessing corpus. In this approach lemmatization as well as
the extension of the lexicon with historical variants are employed
        <xref ref-type="bibr" rid="ref16 ref17 ref18 ref19">(for more details see
Schmidt &amp; Burghardt, 2018a)</xref>
        . We refer to the first approach as naive (lexicon-based)
approach and to the second one as optimized (still lexicon-based) approach. Both
approaches produce numerical values with values below zero being assigned as negative
sentiment, over zero as positive and equal to zero as neutral.
      </p>
      <p>For the audio analysis, we at first screened different available commercial and
noncommercial audiobooks. However, we identified several problems: Those audiobooks
use different speakers and a lot of additional sound effects and background music,
which might be problematic for the sentiment analysis without larger preprocessing
steps. Furthermore, audiobooks often differ strongly from the original text, leaving
whole passages out or switching the order of speeches. We found the best material for
this use case via semiprofessional readings of the original text on several platforms on
the web. We used publicly available recordings from YouTube2. The reader is female
and there are no deviations from the original text in the reading. Several piano pieces
are included in the recording, but they are separate from the general reading.</p>
      <p>For the audio sentiment analysis, we use the free version of Vokaturi3. Vokaturi is
an emotion recognition software for spoken language with an easy-to-use API.
GarciaGarcia et al. (2017) recommend Vokaturi as the best free software for spoken language
sentiment analysis. Vokaturi is described as being language independent and it works
via machine learning with two annotated databases. Vokaturi takes audio data of any
length and outputs five values that range from 0 (none) – 1 (a lot) for the categories
neutrality, fear, sadness, anger and happiness.</p>
      <p>The workflow for the audio sentiment analysis is as follows: In a preprocessing
step we trimmed and transformed the audio files from YouTube. We then performed
forced alignment with the free Python library aeneas4. Forced alignment is a method to
align text segments and audio speech and determine precise time stamps of when the
text segments in the audio file start and end. As text segments we used the 835 speeches
of the play. According to the time stamps, we segmented the audio file to get 835
separated audio files, which are finally used with Vokaturi. To map the output of Vokaturi
to the nominal scale used for the textual sentiment analysis we employ a heuristic
mathematical approach. First, we sum up all values for the negative emotions fear, sadness
and anger to get a value for negative sentiment. We regard the value for happiness as
positive sentiment. We then chose the maximum value of negative sentiment, positive
sentiment and the value for neutrality as the final polarity of a speech.
2 The entire playlist of the files are available online:</p>
      <p>https://www.youtube.com/playlist?list=PL06w7wmahre2eMZXDBgGcR2mgAoeI_f8k
3 Available online: https://developers.vokaturi.com/getting-started/overview
4 https://github.com/readbeyond/aeneas
negative
neutral
positive
Text:
Naive</p>
      <p>
        The human annotations we use were gathered with an expert literary scholar who
annotated 200 random speeches of Emilia Galotti. The annotator was presented the
speech to be annotated, the predecessor and successor speech for context and then had
to annotate the polarity as rather positive, neutral or rather negative. The annotator was
instructed to annotate the polarity he feels is most connoted with the speech. Note that
the speeches were solely presented in textual form. The annotator was a male student
of German literary studies who had to write a thesis about Emilia Galotti during the
annotation process and can therefore be regarded as an expert for this specific play. We
restricted the annotation to 200 speeches since sentiment annotation in this context has
been proven to be very tedious and challenging
        <xref ref-type="bibr" rid="ref16 ref17 ref18 ref19 ref2">(Alm et al., 2005; Schmidt, Burghardt
&amp; Dennerlein, 2018)</xref>
        . Therefore, all comparisons with human annotations are done with
those specific 200 speeches. More information about the annotation process can be
found in Schmidt, Burghardt and Dennerlein (2018), where we performed a very
similar annotation study.
4
      </p>
    </sec>
    <sec id="sec-4">
      <title>Results</title>
      <p>We first report all results concerning the comparison of the text and audio sentiment
analysis among all 835 speeches (RQ1). Table 1 shows the overall polarity distribution
of all methods.
The majority of the assignments of the naive approach are neutral while the majority
(51%) of all speeches are assigned as positive by the audio sentiment analysis (see
Table 1). Therefore the proportion of similarly assigned speeches is rather low (32%); 266</p>
      <p>Text:</p>
      <p>Optimized
speeches are assigned the same class. Table 3 shows the same cross table but with the
optimized lexicon-based approach.
The optimized lexicon approach predicts that almost half of the speeches are negative
while the audio sentiment analysis approach produces opposite results with half of the
speeches being assigned as positive (51%; see Table 1). Consequently, a small number
of 253 speeches are assigned the same class. This results in 30%.</p>
      <p>For RQ2 we regard the 200 human annotated speeches and the performance of all
approaches compared to the human annotations. First, Table 4 shows the polarity
distributions of all methods on this subset of speeches.
The polarity distributions for the different computational methods on the subset of the
annotated corpus are similar to the distributions on the entire corpus. Most of the
speeches were annotated as negative by the expert annotator (46%). With the following
cross tables, we show the agreements of the computational methods and the human
annotations. The human annotations are used as gold standard to evaluate the
approaches. The accuracy is the proportion of correctly predicted speeches among all
speeches. First, Table 5 and 6 show the text-based sentiment analysis results.</p>
      <p>Sum
92
60
48
200
Taking the human annotation as benchmark, both textual approaches perform almost
similarly: The naive approach achieves an accuracy of 52% (104 speeches) and the
optimized approach 56 % (112 speeches). Both approaches are over the random
baseline (approx. 36%) and slightly above the majority baseline (46%).
With the human annotation as gold standard, the audio sentiment analysis achieves an
accuracy of 31% with 62 speeches being predicted correctly. This is below the random
baseline. The main difference is that the human annotator chose negative annotations
for most of the speeches while the audio sentiment analysis in contrast predicts positive
sentiment for the majority of times.
5</p>
    </sec>
    <sec id="sec-5">
      <title>Discussion</title>
      <p>With the first research questions, we analyzed and compared basic text and audio
sentiment analysis approaches. We identified that the sentiment analysis on both, text and
audio, produces very different results with the optimized text analysis predicting the
majority of speeches as negative and the audio speech sentiment analysis in contrast
predicting the majority as positive, more specifically as connoted with the emotion
happiness. For our specific use case, we were able to show that both channels seem to work
on very different levels and with very different features, which exemplifies that the
inclusion of other media channels leads to new insights, in this case even contradictory
results.</p>
      <p>
        We assume that there might be some specifics of the audio material that lead to
this tendency of positive sentiment assignments. Firstly, pitch plays an important factor
in the prediction process
        <xref ref-type="bibr" rid="ref3">(Garcia-Garcia et al., 2017)</xref>
        . We analyzed several examples of
speeches falsely assigned as positive and noticed that the reader uses a high pitch voice
for effect reasons in several instances. Furthermore, our speaker is female and therefore
has a general higher pitch. We assume that these factors might falsely direct the
algorithm to the assignment of higher happiness levels. Additionally, note that we used a
heuristic approach to transform the emotion categories of Vokaturi into sentiment.
However, in several research areas emotion and sentiment are regarded as rather
different concepts
        <xref ref-type="bibr" rid="ref10">(Liu, 2016, p.31-39)</xref>
        and a transformation like this might be too simplistic.
      </p>
      <p>
        When evaluating the computational approaches against a gold standard of a human
annotated subset of 200 speeches the general problems and limitations of sentiment
analysis on literary texts are apparent. The textual approaches as well as the
audiobased method perform rather poorly: the textual approach is just slightly above the
majority baseline, the audio approach is below the random baseline. This is in line with
current research on sentiment analysis on literary texts
        <xref ref-type="bibr" rid="ref16 ref17 ref18 ref19 ref8">(Schmidt &amp; Burghardt, 2018a;
Kim &amp; Klinger, 2018)</xref>
        . Historic narrative texts continue to emerge as a very challenging
text sort for sentiment analysis.
      </p>
      <p>Our assumption that the audio speech analysis might improve the results has been
proven wrong on the chosen material. The text sentiment analysis performs far better
on the gold standard. The reason for this is predominately that the human annotator as
well as the text analysis assign the majority of speeches as negative while the audio
sentiment analysis behaves contrarily.</p>
      <p>
        To put these results in perspective one should also note that audio sentiment analysis
is in general known as being more challenging than other media channels like text and
facial expressions via video and performs often lower than those
        <xref ref-type="bibr" rid="ref14 ref4">(Hudlicka, 2003; Poria
et al., 2017)</xref>
        . Also, note that there are far more ready-to-use software and APIs for
textual and facial sentiment analysis than there are for audio
        <xref ref-type="bibr" rid="ref14">(Poria et al., 2017)</xref>
        . Audio
sentiment analysis on standardized corpora achieves accuracies up to 75%
        <xref ref-type="bibr" rid="ref14">(Poria et al.,
2017)</xref>
        . However, in our use case the audio sentiment analysis performs far lower. Note
that the annotator used just the text for annotation and did not annotate or use the audio
speech, which also might be a reason that the text sentiment analysis and the human
annotation are more in line with each other. For example, the interpretation and oral
performance of the reader might have been more positive than the text itself implies.
      </p>
      <p>In future work, we want to investigate how the audio speech influences the human
annotation. Furthermore, bear in mind that the performance comparison is not the main
goal of our efforts. In future research, we rather want to explore possibilities of
combining multiple media channels to improve performance results.</p>
      <p>Overall, there are several limitations in this case study we want to address in future
research. The presented study represents only a singular use case and a
work-in-progress project. Onwards, we want to analyze more examples like different audio material
of one play and in general plays of other writers and eras. We also want to explore
different audio sentiment analysis approaches. The usage of ready-to-use solutions like
Vokaturi in general sentiment analysis research is rather rare; usually an individual
machine-learning algorithm is implemented. For this specific reason, we plan several
annotation studies to acquire annotated audio material on a large scale. In this context, we
also want to explore how the audio or video presentation of speeches can improve the
sentiment annotation process in this area.</p>
      <p>Furthermore, we also want to include the media channel of video especially facial
emotion recognition in our research via video recordings of theatrical performances of
the plays. The inclusion of the oral but also facial expression of an actor provide
necessary interpretation channels for a holistic understanding of the sentiment and emotion
in a play. With these future plans, we want to continue our work towards a multimodal
sentiment analysis and explore possibilities for the technical performance, the
annotation process and the interpretation in literary studies.</p>
    </sec>
  </body>
  <back>
    <ref-list>
      <ref id="ref1">
        <mixed-citation>
          <string-name>
            <surname>Abburi</surname>
            ,
            <given-names>H.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Akkireddy</surname>
            ,
            <given-names>E. S. A.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Gangashetti</surname>
            ,
            <given-names>S.</given-names>
          </string-name>
          , &amp;
          <string-name>
            <surname>Mamidi</surname>
            ,
            <given-names>R.</given-names>
          </string-name>
          (
          <year>2016</year>
          ).
          <article-title>Multimodal Sentiment Analysis of Telugu Songs</article-title>
          .
          <source>In SAAIP@ IJCAI</source>
          (pp.
          <fpage>48</fpage>
          -
          <lpage>52</lpage>
          ).
        </mixed-citation>
      </ref>
      <ref id="ref2">
        <mixed-citation>
          <string-name>
            <surname>Alm</surname>
            ,
            <given-names>C. O.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Roth</surname>
            ,
            <given-names>D.</given-names>
          </string-name>
          , &amp;
          <string-name>
            <surname>Sproat</surname>
            ,
            <given-names>R.</given-names>
          </string-name>
          (
          <year>2005</year>
          ).
          <article-title>Emotions from text: machine learning for text-based emotion prediction</article-title>
          .
          <source>In Proceedings of the conference on human language technology and empirical methods in natural language processing</source>
          (pp.
          <fpage>579</fpage>
          -
          <lpage>586</lpage>
          ).
        </mixed-citation>
      </ref>
      <ref id="ref3">
        <mixed-citation>
          <string-name>
            <surname>Garcia-Garcia</surname>
            ,
            <given-names>J. M.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Penichet</surname>
            ,
            <given-names>V. M.</given-names>
          </string-name>
          , &amp;
          <string-name>
            <surname>Lozano</surname>
            ,
            <given-names>M. D.</given-names>
          </string-name>
          (
          <year>2017</year>
          ).
          <article-title>Emotion detection: a technology review</article-title>
          .
          <source>In Proceedings of the XVIII International Conference on Human Computer Interaction</source>
          (p.
          <fpage>8</fpage>
          ).
        </mixed-citation>
      </ref>
      <ref id="ref4">
        <mixed-citation>
          <string-name>
            <surname>Hudlicka</surname>
            ,
            <given-names>E.</given-names>
          </string-name>
          (
          <year>2003</year>
          ).
          <article-title>To feel or not to feel: The role of affect in human-computer interaction</article-title>
          .
          <source>International journal of human-computer studies</source>
          ,
          <volume>59</volume>
          (
          <issue>1-2</issue>
          ),
          <fpage>1</fpage>
          -
          <lpage>32</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref5">
        <mixed-citation>
          <string-name>
            <surname>Jannidis</surname>
            ,
            <given-names>F.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Reger</surname>
            ,
            <given-names>I.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Zehe</surname>
            ,
            <given-names>A.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Becker</surname>
            ,
            <given-names>M.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Hettinger</surname>
            ,
            <given-names>L.</given-names>
          </string-name>
          &amp;
          <string-name>
            <surname>Hotho</surname>
            ,
            <given-names>A.</given-names>
          </string-name>
          (
          <year>2016</year>
          ).
          <article-title>Analyzing Features for the Detection of Happy Endings in German Novels</article-title>
          .
          <source>arXiv preprint arXiv:1611</source>
          .
          <fpage>09028</fpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref6">
        <mixed-citation>
          <string-name>
            <surname>Jockers</surname>
            ,
            <given-names>M. L.</given-names>
          </string-name>
          (
          <year>2015</year>
          ).
          <article-title>Revealing sentiment and plot arcs with the syuzhet package</article-title>
          . Retrieved from http://www.matthewjockers.net/
          <year>2015</year>
          /02/02/syuzhet/
        </mixed-citation>
      </ref>
      <ref id="ref7">
        <mixed-citation>
          <string-name>
            <surname>Kakkonen</surname>
            ,
            <given-names>T.</given-names>
          </string-name>
          &amp;
          <string-name>
            <surname>Kakkonen</surname>
            ,
            <given-names>G. G.</given-names>
          </string-name>
          (
          <year>2011</year>
          ).
          <article-title>SentiProfiler: creating comparable visual profiles of sentimental content in texts</article-title>
          .
          <source>In Proceedings of Language Technologies for Digital Humanities and Cultural Heritage</source>
          (pp.
          <fpage>62</fpage>
          -
          <lpage>69</lpage>
          ).
        </mixed-citation>
      </ref>
      <ref id="ref8">
        <mixed-citation>
          <string-name>
            <surname>Kim</surname>
            ,
            <given-names>E.</given-names>
          </string-name>
          &amp;
          <string-name>
            <surname>Klinger</surname>
            ,
            <given-names>R.</given-names>
          </string-name>
          (
          <year>2018</year>
          ).
          <article-title>Who feels what and why? annotation of a literature corpus with semantic roles of emotions</article-title>
          .
          <source>In: Proceedings of the 27th International Conference on Computational Linguistics</source>
          (pp.
          <fpage>1345</fpage>
          -
          <lpage>1359</lpage>
          ).
          <article-title>Association for Computational Linguistics</article-title>
          .
        </mixed-citation>
      </ref>
      <ref id="ref9">
        <mixed-citation>
          <string-name>
            <surname>Kim</surname>
            ,
            <given-names>E.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Padó</surname>
            ,
            <given-names>S.</given-names>
          </string-name>
          , &amp;
          <string-name>
            <surname>Klinger</surname>
            ,
            <given-names>R.</given-names>
          </string-name>
          (
          <year>2017</year>
          ).
          <article-title>Prototypical Emotion Developments in Literary Genres</article-title>
          .
          <source>In Pro-ceedings of the Joint SIGHUM Workshop on Computational Linguistics for Cultural Heritage</source>
          ,
          <source>Social Sciences, Humanities and Literature</source>
          (pp.
          <fpage>17</fpage>
          -
          <lpage>26</lpage>
          ).
        </mixed-citation>
      </ref>
      <ref id="ref10">
        <mixed-citation>
          <string-name>
            <surname>Liu</surname>
            ,
            <given-names>B.</given-names>
          </string-name>
          (
          <year>2016</year>
          ).
          <article-title>Sentiment Analysis</article-title>
          .
          <source>Mining Opinions, Sentiments and Emotions</source>
          . New York: Cambridge University Press.
        </mixed-citation>
      </ref>
      <ref id="ref11">
        <mixed-citation>
          <string-name>
            <surname>Mohammad</surname>
            ,
            <given-names>S.</given-names>
          </string-name>
          (
          <year>2011</year>
          ).
          <article-title>From once upon a time to happily ever after: Tracking emotions in novels and fairy tales</article-title>
          .
          <source>In Proceedings of the 5th ACL-HLT Workshop on Language Technology for Cultural Heritage</source>
          ,
          <source>Social Sciences, and Humanities</source>
          (pp.
          <fpage>105</fpage>
          -
          <lpage>114</lpage>
          ).
        </mixed-citation>
      </ref>
      <ref id="ref12">
        <mixed-citation>
          <string-name>
            <surname>Morency</surname>
            ,
            <given-names>L. P.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Mihalcea</surname>
            ,
            <given-names>R.</given-names>
          </string-name>
          , &amp;
          <string-name>
            <surname>Doshi</surname>
            ,
            <given-names>P.</given-names>
          </string-name>
          (
          <year>2011</year>
          , November).
          <article-title>Towards multimodal sentiment analysis: Harvesting opinions from the web</article-title>
          .
          <source>In Proceedings of the 13th international conference on multimodal interfaces</source>
          (pp.
          <fpage>169</fpage>
          -
          <lpage>176</lpage>
          ). ACM.
        </mixed-citation>
      </ref>
      <ref id="ref13">
        <mixed-citation>
          <string-name>
            <surname>Nalisnick</surname>
            ,
            <given-names>E. T.</given-names>
          </string-name>
          &amp;
          <string-name>
            <surname>Baird</surname>
            ,
            <given-names>H. S.</given-names>
          </string-name>
          (
          <year>2013</year>
          ).
          <article-title>Character-to-character sentiment analysis in shakespeare's plays</article-title>
          .
          <source>In Proceedings of the 51st Annual Meeting of the Association for Computational Linguistics</source>
          (pp.
          <fpage>479</fpage>
          -
          <lpage>483</lpage>
          ).
        </mixed-citation>
      </ref>
      <ref id="ref14">
        <mixed-citation>
          <string-name>
            <surname>Poria</surname>
            ,
            <given-names>S.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Cambria</surname>
            ,
            <given-names>E.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Bajpai</surname>
            ,
            <given-names>R.</given-names>
          </string-name>
          , &amp;
          <string-name>
            <surname>Hussain</surname>
            ,
            <given-names>A.</given-names>
          </string-name>
          (
          <year>2017</year>
          ).
          <article-title>A review of affective computing: From unimodal analysis to multimodal fusion</article-title>
          .
          <source>Information Fusion</source>
          ,
          <volume>37</volume>
          ,
          <fpage>98</fpage>
          -
          <lpage>125</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref15">
        <mixed-citation>
          <string-name>
            <surname>Remus</surname>
            ,
            <given-names>R.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Quasthoff</surname>
            ,
            <given-names>U.</given-names>
          </string-name>
          &amp;
          <string-name>
            <surname>Heyer</surname>
            ,
            <given-names>G.</given-names>
          </string-name>
          (
          <year>2010</year>
          ).
          <article-title>SentiWS-A Publicly Available German-language Resource for Sentiment Analysis</article-title>
          .
          <source>In LREC</source>
          (pp.
          <fpage>1168</fpage>
          -
          <lpage>1171</lpage>
          ).
        </mixed-citation>
      </ref>
      <ref id="ref16">
        <mixed-citation>
          <string-name>
            <surname>Schmidt</surname>
            ,
            <given-names>T.</given-names>
          </string-name>
          &amp;
          <string-name>
            <surname>Burghardt</surname>
            ,
            <given-names>M.</given-names>
          </string-name>
          (
          <year>2018a</year>
          ).
          <article-title>An Evaluation of Lexicon-based Sentiment Analysis Techniques for the Plays of Gotthold Ephraim Lessing</article-title>
          .
          <source>In: SIGHUM Workshop on Language Technology for Cultural Heritage</source>
          , Social Sciences, and
          <string-name>
            <surname>Humanities (</surname>
          </string-name>
          LaTeCH-CLfL
          <year>2018</year>
          ) (pp.
          <fpage>139</fpage>
          -
          <lpage>149</lpage>
          ). Retrieved from http://aclweb.org/anthology/W18-4516
        </mixed-citation>
      </ref>
      <ref id="ref17">
        <mixed-citation>
          <string-name>
            <surname>Schmidt</surname>
            ,
            <given-names>T.</given-names>
          </string-name>
          &amp;
          <string-name>
            <surname>Burghardt</surname>
            ,
            <given-names>M.</given-names>
          </string-name>
          (
          <year>2018b</year>
          ).
          <article-title>Toward a Tool for Sentiment Analysis for German Historic Plays</article-title>
          . In: Piotrowski, M. (ed.),
          <source>COMHUM 2018: Book of Abstracts for the Workshop on Computational Methods in the Humanities</source>
          <year>2018</year>
          (pp.
          <fpage>46</fpage>
          -
          <lpage>48</lpage>
          ). Lausanne, Switzerland:
          <article-title>Laboratoire laussannois d'informatique et statistique textuelle</article-title>
          . Retrieved from https://zenodo.org/record/1312779#.XGA0iFxKhPZ
        </mixed-citation>
      </ref>
      <ref id="ref18">
        <mixed-citation>
          <string-name>
            <surname>Schmidt</surname>
            ,
            <given-names>T.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Burghardt</surname>
            ,
            <given-names>M.</given-names>
          </string-name>
          &amp;
          <string-name>
            <surname>Dennerlein</surname>
            ,
            <given-names>K.</given-names>
          </string-name>
          (
          <year>2018</year>
          ).
          <article-title>Sentiment Annotation of Historic German Plays: An Empirical Study on Annotation Behavior</article-title>
          . In: Sandra Kübler, Heike Zinsmeister (eds.),
          <source>Proceedings of the Workshop on Annotation in Digital Humanities (annDH</source>
          <year>2018</year>
          )
          <article-title>(pp</article-title>
          .
          <fpage>47</fpage>
          -
          <lpage>52</lpage>
          ). Sofia, Bulgaria. Retrieved from http://ceur-ws.
          <source>org/</source>
          Vol-
          <volume>2155</volume>
          /schmidt.pdf
        </mixed-citation>
      </ref>
      <ref id="ref19">
        <mixed-citation>
          <string-name>
            <surname>Schmidt</surname>
            ,
            <given-names>T.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Burghardt</surname>
            ,
            <given-names>M.</given-names>
          </string-name>
          &amp;
          <string-name>
            <surname>Wolff</surname>
            ,
            <given-names>C.</given-names>
          </string-name>
          , (
          <year>2018</year>
          ).
          <article-title>Herausforderungen für Sentiment AnalysisVerfahren bei literarischen Texten</article-title>
          . In: Burghardt,
          <string-name>
            <given-names>M.</given-names>
            &amp;
            <surname>Müller-Birn</surname>
          </string-name>
          ,
          <string-name>
            <surname>C.</surname>
          </string-name>
          (Hrsg.),
          <source>INF-DH2018</source>
          .
          <article-title>Bonn: Gesellschaft für Informatik e</article-title>
          .V. Retrieved from https://dl.gi.de/handle/20.500.12116/16996
        </mixed-citation>
      </ref>
      <ref id="ref20">
        <mixed-citation>
          <string-name>
            <surname>Vinodhini</surname>
            ,
            <given-names>G.</given-names>
          </string-name>
          , &amp;
          <string-name>
            <surname>Chandrasekaran</surname>
            ,
            <given-names>R. M.</given-names>
          </string-name>
          (
          <year>2012</year>
          ).
          <article-title>Sentiment analysis and opinion mining: a survey</article-title>
          .
          <source>International Journal of Advanced Research in Computer Science and Software Engineering</source>
          ,
          <volume>2</volume>
          (
          <issue>6</issue>
          ),
          <fpage>282</fpage>
          -
          <lpage>292</lpage>
          .
        </mixed-citation>
      </ref>
    </ref-list>
  </back>
</article>