<!DOCTYPE article PUBLIC "-//NLM//DTD JATS (Z39.96) Journal Archiving and Interchange DTD v1.0 20120330//EN" "JATS-archivearticle1.dtd">
<article xmlns:xlink="http://www.w3.org/1999/xlink">
  <front>
    <journal-meta />
    <article-meta>
      <title-group>
        <article-title>Organization of Information Support for a Bioengineering System of Emotional Response Research</article-title>
      </title-group>
      <contrib-group>
        <contrib contrib-type="author">
          <string-name>© N.N. Filatova</string-name>
          <email>nfilatova99@mail.ru</email>
          <xref ref-type="aff" rid="aff0">0</xref>
          <xref ref-type="aff" rid="aff1">1</xref>
        </contrib>
        <aff id="aff0">
          <label>0</label>
          <institution>Proceedings of the XX International Conference “Data Analytics and Management in Data Intensive Domains” (DAMDID/RCDL'2018)</institution>
          ,
          <addr-line>Moscow</addr-line>
          ,
          <country country="RU">Russia</country>
        </aff>
        <aff id="aff1">
          <label>1</label>
          <institution>N.I. Bodrina © K.V. Sidorov Tver State Technical University</institution>
          ,
          <addr-line>Tver</addr-line>
          ,
          <country country="RU">Russia</country>
        </aff>
      </contrib-group>
      <fpage>90</fpage>
      <lpage>97</lpage>
      <abstract>
        <p>Nowadays studying a mechanism of human emotional responses attracts much attention. Information about a human personality and condition, which is expressed in a manner of speech, is just as important as his statements. However, computer synthesis and speech recognition systems do not currently use this information. It is possible to numerically assess certain physiological characteristics related to emotions (cardiogram, muscle curves, EEG, speech). In order to assess an emotion objectively, it is necessary to use a complex approach including testee's self-evaluation and recording characteristics of certain body functional systems. There are widely distributed databases containing examples of such characteristics. Previous bases contain recordings of scenic speech with imitated emotions. Modern researchers prefer working with natural emotions caused by irritants - incentives. The paper specifies a multi-channel bioengineering system for studying emotions “EEG-Speech+”, which is created in TSTU, and how to work with it. It also describes two series of experiments. The first one includes searching for signs of emotion valence by a speech signal attractor. The second one includes investigating the emotion dynamics by an EEG signal. The authors describe the structure of an extended multimodal emotion base, which stores the results of all experiments. They also consider its open online version.</p>
      </abstract>
    </article-meta>
  </front>
  <body>
    <sec id="sec-1">
      <title>1 Introduction</title>
      <p>Studying human emotional response mechanism refers
to an interdisciplinary field of knowledge, which
attracts more and more attention nowadays. Research
on the works of Anokhin P.K., Simonov P.V.,
Leontyev, Ilyin E.P., Danilova N.N., Izard K., Rusalova
M.N., Ukhtomsky A.A., Fress P., Chomskaya E.D.,
Everly G., Rosenfeld R., Hebb D., etc., allows
identifying several basic conclusions, which are not
disputed by the scientific community at this stage:
 emotions are inherent not only in a human, but in all
intelligent representatives of mammals;
 the emotional response mechanism is innate, some
emotions are shown at the earliest stages of life;
 emotions are most often a reaction in response to an
external or internal irritants – an incentive;
 the system of human emotional reactions is
developing; it is formed in the process of
accumulation of his personal experience and
formation of cognitive functions.</p>
      <p>The mechanism of emotional responses is a further
development of reflex systems of the mammalian
organism. It solves two sets of tasks: improves the
means of adapting an organism to changes in external
conditions and creates an apparatus for implementation
of communicative processes and maintenance of
socially significant contacts. Human emotional
responses are related to brain activity and are revealed
in functioning peculiarities of certain body functional
systems.</p>
      <p>In a colloquial human interaction, extralinguistic
information about speaker’s personality and state,
which is expressed in his manner of speech, is as
important as the text of a statement. However, computer
synthesis and speech recognition systems do not
currently use information about emotions, which is a
very important factor in communication.</p>
      <p>Systems that are capable of generating emotionally
colored speech and recognizing human emotional state
will be in demand in virtual learning, for studying brain
dysfunction, identifying network content and interactive
entertainment. In addition, they will be useful for
people who have different speech deviations. Modern
speech synthesizers do not model emotional speech.
The algorithms for recognizing human emotional state
are only being developed.</p>
      <p>Nowadays there are no objective means of
measuring quantitative characteristics of emotions.
However, there are opportunities for quantitative
assessment of certain physiological characteristics
related to them (cardiogram, galvanic skin response,
muscle curves, electroencephalograms, and speech
patterns).</p>
      <p>Considering testee’s subjective assessments in an
emotion, objective emotion evaluation requires an
integrated complex approach including both testee’s
self-assessment and recording characteristics of certain
functional systems.</p>
      <p>Successful development of emotion recognition
modules by various signals recorded in a person, who is
experiencing an emotion, is possible when there is a big
volume of such signals. Geographical and ethnic studies
show that an emotional expression is formed and
changes with the course of the history of linguistics.
Consequently, the sources of emotional responses
should be carriers of an appropriate language.</p>
      <p>Initially, the bases with the records of emotionally
colored speech have become widespread. They are
gradually expanded. Other biomedical signals
(cardiogram, galvanic skin response, heart rate, muscle
curves, electroencephalograms, etc.) taken at the
moment when a testee demonstrates an emotional
response are added to speech samples.</p>
    </sec>
    <sec id="sec-2">
      <title>2 Modern bases of emotional response examples</title>
      <p>Early studies of emotional responses are based on the
records of scenic speech with imitated emotions [1, 3, 6,
12, 18, and 19]. Usually, exterior listeners recognize
such emotions correctly. The analysis of acoustic
characteristics is based on the records of identic texts.
Nevertheless, it is not known how well an actor is able
to represent all speech characteristics that ordinary
people show when they experience similar emotions.
Imitated emotions are reproduced on assignment and do
not need incentives.</p>
      <p>In studies, the difference between experienced and
expressed emotions is minimal. In everyday social
interactions, it is often appropriate to suppress
emotions. Moreover, it is preferable to express emotions
that people do not really experience at the moment. A
computer synthesizer of an emotional speech, which is
created based on studying only simulated emotions,
might deform user intentions.</p>
      <p>
        Therefore, the majority of modern researchers work
with stimulated emotions (Table 1) instead of using
emotion imitations. Such emotions are natural and are
triggered by specially prepared emotiogenic incentives.
Information support of the bases includes these
incentives or their descriptions. There are some papers
that pay attention to classification, evaluation or
marking of incentives [
        <xref ref-type="bibr" rid="ref14 ref7">7, 14</xref>
        ].
      </p>
      <p>
        The need to confirm the desired emotion in a testee
leads to expanding a list of types of biomedical signals
stored in databases [7, 13, and 17]. In experiments,
testees are usually instructed not to restrain their
emotions, but in real social interactions personal
feelings are not expressed so openly. For this reason,
some researchers use other people as sources of
emotionogenic incentives [
        <xref ref-type="bibr" rid="ref13 ref17">13, 17</xref>
        ]. In an experiment, a
testee together with an assistant must solve some
problem. Interaction, communication with an assistant
is an incentive.
      </p>
    </sec>
    <sec id="sec-3">
      <title>3 Bioengineering system “EEG-Speech+”</title>
      <p>
        A specialized bioengineering system “EEG-Speech+”
has been created and developing at the Department of
Automation of Technological Processes of the Tver
State Technical University [
        <xref ref-type="bibr" rid="ref16">16</xref>
        ]. The system has a
      </p>
    </sec>
    <sec id="sec-4">
      <title>Name, year, language</title>
      <sec id="sec-4-1">
        <title>DEAP data, 2005, Eng. [7]</title>
      </sec>
      <sec id="sec-4-2">
        <title>Film Stim, 2010, Eng., French. [14]</title>
      </sec>
      <sec id="sec-4-3">
        <title>Cognitive</title>
        <p>
          Human
Computer
Interaction
Lab, 2011,
Eng. [
          <xref ref-type="bibr" rid="ref8">8</xref>
          ]
        </p>
      </sec>
      <sec id="sec-4-4">
        <title>MAHNOBHCI, 2012, Eng. [17]</title>
      </sec>
      <sec id="sec-4-5">
        <title>Recola</title>
        <p>
          Databаse,
2013,
French.
[
          <xref ref-type="bibr" rid="ref13">13</xref>
          ]
1-minute
video with
sound (more
than 120)
1–7-minute
video with
sound (more
than 70)
recordings of 
classical
music
video with
sound (more
than 30) and
images
(more than
20)
interaction
















        </p>
        <p>EEG;
physiologic
measuring;
face video;
assessments
assessments
EEG
EEG;
physiologic
measuring;
face and body
video;
speech;
position of
the pupil;
assessments
EEG;
ECG;
speech;
face video;
assessments
multichannel scheme for recording testee's responses to
external emotiogenic incentives. Simultaneous
recording of several types of biomedical signals allows
confirming changes in testee’s emotions according to
the scenario of the experiment.</p>
        <p>Fig. 1 shows the composition and interaction
scheme of the components of the bioengineering system
“EEG-Speech+”. By now, the system has been
expanded to five channels for recording emotional
response (Ch1 - Ch5 in fig. 1): video, sound,
electroencephalogram (EEG), muscle curve (EMG) and
information (testee's report).</p>
        <p>A personal computer B serves to present visual or
acoustic incentives and contains a base of incentives, as
well as all software necessary for their reproduction. A
special device [4, p. 78] delivers olfactory incentives to
a testee. The main workstation A controls the process of
presenting olfactory incentives.</p>
        <p>
          Each experiment session has a specially prepared
scenario (Table 2). The workstation A receives
biomedical signals from all channels used in the current
experiment. The signals are stored in the appropriate
database of testees. The received signals are processed
and cleared of interference and artifacts. The
bioengineering system software includes three groups
of modules (Modules of groups I, II and III in Fig. 1)
[
          <xref ref-type="bibr" rid="ref4">4</xref>
          ]:

registration, processing and saving biomedical
signals;
 formation of attribute models of biomedical
signals;
 monitoring of emotions.
        </p>
        <p>The software modules are implemented in
MATLAB in C# language. The bioengineering system
software is installed on the main workstation A, but can
be used on any personal computer, so that processing of
experimental results can be remote and in a distributed
mode.</p>
      </sec>
    </sec>
    <sec id="sec-5">
      <title>4 Experiments and results</title>
      <p>There are a lot of experiments with the bioengineering
system “EEG-Speech+”. The studies include several
directions:
 search for signs of biomedical signals related to an
emotional response;
 determining the direction of emotion development
(growth, fading).</p>
      <sec id="sec-5-1">
        <title>4.1 Signs of emotion valence in a speech signal</title>
        <p>Most biomedical signals are not stationary and
irregular, i.e. a probability distribution of signal
parameters is random. Therefore, the methods of
nonlinear dynamics become relevant for their processing.</p>
        <p>In particular, in order to identify individual
characteristics of emotions based on the initial
biomedical signal, there is a reconstruction of an
attractor, which becomes an object of research later.
placing electrodes for long-term recording
of biomedical signals;
tuning</p>
        <p>the
experiment;
start recording biomedical signals.</p>
        <p>channels selected
for the</p>
        <sec id="sec-5-1-1">
          <title>Background</title>
          <p>demonstration</p>
        </sec>
        <sec id="sec-5-1-2">
          <title>Incentive demonstration “+”</title>
        </sec>
        <sec id="sec-5-1-3">
          <title>Background demonstration</title>
        </sec>
        <sec id="sec-5-1-4">
          <title>Background</title>
          <p>demonstration</p>
        </sec>
        <sec id="sec-5-1-5">
          <title>Incentive demonstration “-”</title>
        </sec>
        <sec id="sec-5-1-6">
          <title>Background</title>
          <p>demonstration
A short survey of a testee:
confirmation of the expected
emotional response
neutral
positive
fading of positive,
transition to neutral
neutral
negative
fading of negative,
transition to neutral
1 min.</p>
        </sec>
        <sec id="sec-5-1-7">
          <title>A short survey of a testee</title>
          <p>stop recording biomedical signals;
cutting off the channels;
detaching electrodes.</p>
          <p>Detailed survey of a testee: playback
of incentives and their marking</p>
          <p>
            So, a number of authors use an index of a restored
attractor correlation dimension [
            <xref ref-type="bibr" rid="ref11 ref9">9, 11</xref>
            ] to recognize a sign
of emotions.
          </p>
          <p>
            The paper [
            <xref ref-type="bibr" rid="ref11">11</xref>
            ] used this feature when comparing
EEG of the signal recorded for five testee’s states: grief,
joy, time counting, a background with closed eyes and a
background
with
open
eyes.
          </p>
          <p>The
author
notes a
significant increase in the correlation dimension under
conditions of emotional experience comparing with a
neutral state.</p>
          <p>
            One of other signs of emotion recognition is the
Lyapunov exponent. The paper [
            <xref ref-type="bibr" rid="ref9">9</xref>
            ] uses the Lyapunov
exponent to assess testee’s emotional state by certain
phonemes in a speech signal. The author notes a
significant difference between the state of “calmness”
and when there are negative emotions (anger, disgust).
          </p>
          <p>Studying of attractors reconstructed from Russian
speech patterns showed that when a testee experiences
positive emotions, the attractor form expands, in the
case of negative one it gets narrow. Consequently, the
number of points in the center changes. Thus, we can
assume that a correlate of a sign of emotions can be the
point density indicator of attractor trajectories.</p>
          <p>
            The hypothesis was checked through the research
that involved students and postgraduates of the Tver
State Technical University at the age of 18–25. The
testees were offered to watch videos of up to 3 minutes,
which can be conditionally divided into three groups:
1. a positive incentive (k+);
2. a negative incentive (k-);
3. a neutral incentive (N).
challenge phrase.
used the indicator [
            <xref ref-type="bibr" rid="ref5">5</xref>
            ]:
          </p>
          <p>After each video the participants had to say a
As a measure of the attractor density in the center, we
 
= 
 ⁄  , 

= ℎ

+   ⁄2,</p>
          <p>Which is the ratio of the number of attractor points
related to one of the cells of an orthogonal grid covering
an attractor projection, (  ) to the cell area (  ). ℎ is the
number of points inside each j-th cell. The number of
points (  ) on the boundary of the j-th and j+1-th cells is
divided equally between boundary cells.</p>
          <p>The autocorrelation function determines the optimal
value of the time delay τ, which varies depending on a
(1)
testee.</p>
          <p>Attractor properties were analyzed using the first
projection of the attractor, or rather the area of the
greatest cluster of points localized near the origin of
coordinates (Fig. 2).</p>
          <p>Fig. 2. A projection of an attractor, which was
reconstructed from a speech signal, with a selected area
of the greatest cluster of points</p>
          <p>The duration of each received speech record for
analysis was 20,000 readings (≈1 seconds). The records
went through auto-normalization with the removal of
artifacts.</p>
          <p>
            It has been experimentally established that the
presence of a noise component does not affect the
classifying ability of the parameter   [
            <xref ref-type="bibr" rid="ref15">15</xref>
            ].
          </p>
          <p>In total, we analyzed 74 speech signal fragments from
8 testees (3 incentives for each sign of emotions).</p>
          <p>It should be noted that a negative incentive causes an
increase of the   index in relation to a neutral state (from
2 to 55%) almost in all testees. On the contrary, with a
positive video incentive, this parameter tends to decrease
(from 5 to 38%). The obtained result confirms the
hypothesis about the interrelation between an emotional
impact sign and an attractor density.</p>
          <p>
            Similar experiments were performed with samples of
voice recordings from the international database
EmoDB [
            <xref ref-type="bibr" rid="ref1">1</xref>
            ], which contains audio recordings of emotionally
colored speech in German from 10 different speakers.
We analyzed signals with a negative (disgust), positive
(happiness) and neutral incentive. Figure 3b shows the
results of   averaged values for several testees on the
same phrase.
          </p>
          <p>Unlike the samples of Russian speech, German
speech is characterized by an increase (on average by
20%) of the number of points in the attractor center
affected by positive incentives in relation to a neutral
state. Negative incentives also cause an increase in  
density (on average by 10%).</p>
          <p>Conclusion. It is established that the sign of emotions
significantly affects the number of points of the
reconstructed attractor in the center. This is true both for
Russian speech samples and for studying phrases in
German. The density parameter   available from
experiments can be used to construct a classifier.</p>
          <p>In a series of experiments, a testee was consistently
presented with several negative incentives (-E), and then
several positive ones (+E). Before changing an incentive
sign, a testee was presented with neutral frames with a
green background. Each experiment lasted no less than
20 and not more than 25 minutes.</p>
          <p>While watching incentives, testee’s EEG was
continuously recorded. His speech was recorded after
each incentive. The processing of the experimental
results had two stages.</p>
          <p>The first stage included creating fragments of
biomedical signals free from noise (for speech signals)
and artifacts (for EEG signals).</p>
          <p>Perception of incentives of the same sign (-E or +E)
resulted in sequences of EEG fragments. Their
characteristics contain information on changes in testee’s
emotional responses.</p>
          <p>The second stage of processing the experimental
results included identification and quantitative evaluation
of these latent characteristics. The bioengineering system
“EEG-Speech+” provides calculation of signal power
spectral analysis (EEG or speech signals), as well as
attractor reconstruction based on them.</p>
          <p>Figure 4 shows a projection of an attractor
constructed from an EEG fragment (lead C4-A2), which
is correlated with the terminal part of the first negative
incentive.
4d0e0n0sity, ρj
3500
3000
2500
2000
1500
1000
500</p>
          <p>0
9d0e0n0sity, ρj
8500
8000
7500
7000
6500
6000
5500
5000
4500</p>
          <p>Negative</p>
          <p>Neutral
Positive
a
Positive
b
Negative
Neutral</p>
        </sec>
      </sec>
      <sec id="sec-5-2">
        <title>4.2 Research on an emotion dynamics based on the analysis of EEG signals</title>
        <p>A series of experiments included using 2–4-minute video
clips with sound as emotiogenic incentives. The testees
were TSTU students and postgraduates aged 18–25.</p>
        <p>Each video incentive was pre-marked by a testee
according to a sign of an emotional response.</p>
        <p>Fig. 4. A projection of an attractor constructed from an
EEG signal (lead C4-A2)</p>
        <p>The experiments showed that leads F7-A1 and F8-A2
had the strongest changes in power spectra when a testee
was watching positive and negative incentives.</p>
        <p>
          However, the reproducibility of this result was not
high. Therefore, each lead had reconstructed attractors
with their properties depending on the sign of testee’s
emotional response, as shown in previous studies [
          <xref ref-type="bibr" rid="ref10">10</xref>
          ].
        </p>
        <p>
          To characterize attractors, we used the features
proposed in [
          <xref ref-type="bibr" rid="ref5">5</xref>
          ]: an attractor trajectory density near its
center   (1) and a number of empty cells in a grid
covering the attractor projection k0 (Fig. 4). Grid
dimensions are fixed: 196 cells, a step is 50 readings.
        </p>
        <p>Observation of changes in the signs of   and k0
showed that in most experiments there is their correlation
with a sign of an emotional response.</p>
        <p>When a testee experiences positive emotions, k0
decreases. It increases during experiencing negative
emotions.</p>
        <p>Conclusion. Preliminary results show the possibility
of using an attractor density as a sign of EEG signals,
illustrating the development of an emotional state at a
certain time interval. The observation interval does not
have imposed limitations.</p>
      </sec>
    </sec>
    <sec id="sec-6">
      <title>5 A multimodal emotion database and a public emotion database</title>
      <p>
        The experimental results are a basis for a multimodal
emotion database, which contains examples of signals
with a bright and slightly expressed emotional color. At
the first stage, the database has speech patterns and
associated EEG patterns [
        <xref ref-type="bibr" rid="ref4">4</xref>
        ]. The “entity-relation” model
of the extended multimodal emotion database is
supplemented by descriptions of incentives and new
channels (Fig. 5).
      </p>
      <p>The examples of emotional responses in the database
are not labeled with the names of emotions (“anger”,
“fear”, “joy”, etc.). We use only natural emotional
responses, so we determine the valence of an emotion
(positive, negative or neutral) and its level (strong, weak,
etc.).</p>
      <p>The multimodal emotion database includes:
 266 patterns of a challenge phrase lasting 2–6
seconds, pronounced by different speakers who are
not actors in response to a presentation of a video
incentive;
 2660 vowel phonemes lasting 0.025–0.25 seconds
segmented from challenge phrases;
 240 EEG patterns cleared from artifacts lasting for
12 seconds.</p>
      <p>
        Since 2016, there is a public database containing
examples of emotional responses [
        <xref ref-type="bibr" rid="ref2">2</xref>
        ]. The database is
developed in PHP language with MySQL DBMS. There
is a website (http://emotions.tstu.tver.ru) to access the
database using cms joomla. For now, several series of
experiments are available to the public:
1. Recordings of speech signals (in .wav) of 17 testees.
      </p>
      <p>There are up to 10 samples for certain testees.
Emotiogenic incentives are specially prepared videos
with sound, which cause positive, negative and
neutral emotional states.
2. Recordings of speech signals (in .wav) and EEG
signals (in .txt) of 9 testees. There are several
recording sessions for certain testees. Registration of
speech and EEG was parallel. Emotiogenic incentives
were also videos with sound. Parallel recording of
speech signals and EEG allowed objectively fixing
the presence of positive, negative and neutral
responses of testees to incentives.</p>
      <p>Acknowledgments. The reported study has been funded
by RFBR according to the research projects: №
17-0100742, № 18-37-00225.</p>
    </sec>
  </body>
  <back>
    <ref-list>
      <ref id="ref1">
        <mixed-citation>
          [1]
          <string-name>
            <surname>Burkhardt</surname>
            ,
            <given-names>F.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Paeschke</surname>
            ,
            <given-names>A.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Rolfes</surname>
            ,
            <given-names>M.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Sendlmeier</surname>
            ,
            <given-names>W.F.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Weiss</surname>
            ,
            <given-names>B.A.</given-names>
          </string-name>
          :
          <article-title>Database of German Emotional Speech</article-title>
          .
          <source>In: 9th European Conference on Speech Communication and Technology (Interspeech) Proceedings</source>
          , pp.
          <fpage>1517</fpage>
          -
          <lpage>1520</lpage>
          . ISCA. Lisbon, Portugal (
          <year>2005</year>
          ).
        </mixed-citation>
      </ref>
      <ref id="ref2">
        <mixed-citation>
          <article-title>[2] Database of Emotional Response Examples</article-title>
          , http://emotions.tstu.tver.ru,
          <source>last accessed</source>
          <year>2018</year>
          /05/15.
        </mixed-citation>
      </ref>
      <ref id="ref3">
        <mixed-citation>
          [3]
          <string-name>
            <surname>Engberg</surname>
            ,
            <given-names>I.S.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Hansen</surname>
            ,
            <given-names>A.V.</given-names>
          </string-name>
          :
          <article-title>Documentation of the Danish Emotional Speech Database (DES)</article-title>
          . Aalborg University, Denmark (
          <year>1996</year>
          ).
        </mixed-citation>
      </ref>
      <ref id="ref4">
        <mixed-citation>
          [4]
          <string-name>
            <surname>Filatova</surname>
            ,
            <given-names>N.N.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Sidorov</surname>
            ,
            <given-names>K.V.</given-names>
          </string-name>
          :
          <source>Computer Models of Emotions: Construction and Methods of Research. RITs TSTU</source>
          , Tver' (
          <year>2017</year>
          )
          <article-title>(in Russ., Komp'yuternye modeli emotsiy: postroenie i metody issledovaniya: monografiya)</article-title>
        </mixed-citation>
      </ref>
      <ref id="ref5">
        <mixed-citation>
          [5]
          <string-name>
            <surname>Filatova</surname>
            ,
            <given-names>N.N.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Sidorov</surname>
            ,
            <given-names>K.V.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Terekhin</surname>
            ,
            <given-names>S.A.</given-names>
          </string-name>
          :
          <article-title>A Software Package for Interpretation of Nonverbal Information by Analyzing Speech Patterns or Electroencephalogram</article-title>
          .
          <source>Software &amp; Systems</source>
          <volume>111</volume>
          (
          <issue>3</issue>
          ),
          <fpage>22</fpage>
          -
          <lpage>27</lpage>
          (
          <year>2015</year>
          ). doi:
          <volume>10</volume>
          .15827/
          <fpage>0236</fpage>
          -
          <lpage>235X</lpage>
          .
          <fpage>111</fpage>
          .
          <fpage>022</fpage>
          -
          <lpage>027</lpage>
          (in Russ.,
          <source>Programmnye Produkty i Sistemy)</source>
        </mixed-citation>
      </ref>
      <ref id="ref6">
        <mixed-citation>
          [6]
          <string-name>
            <surname>Haq</surname>
            ,
            <given-names>S.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Jackson</surname>
            ,
            <given-names>P.J.B.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Edge</surname>
            ,
            <given-names>J.D.</given-names>
          </string-name>
          :
          <article-title>Audio-Visual Feature Selection and Reduction for Emotion Classification</article-title>
          .
          <source>In: International Conference on Auditory-Visual Speech Processing (AVSP) Proceedings</source>
          , pp.
          <fpage>185</fpage>
          -
          <lpage>190</lpage>
          . ISCA.
          <string-name>
            <surname>Australia</surname>
          </string-name>
          (
          <year>2008</year>
          ).
        </mixed-citation>
      </ref>
      <ref id="ref7">
        <mixed-citation>
          [7]
          <string-name>
            <surname>Koelstra</surname>
            ,
            <given-names>S.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Muehl</surname>
            ,
            <given-names>C.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Soleymani</surname>
            ,
            <given-names>M.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Lee</surname>
            <given-names>J.-S.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Yazdani</surname>
            ,
            <given-names>A.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Ebrahimi</surname>
            ,
            <given-names>T.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Pun</surname>
            ,
            <given-names>T.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Nijholt</surname>
            ,
            <given-names>A.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Patras</surname>
            ,
            <given-names>I.</given-names>
          </string-name>
          :
          <article-title>DEAP: a Database for Emotion Analysis Using Physiological Signals</article-title>
          .
          <source>IEEE Transaction on Affective Computing</source>
          <volume>3</volume>
          (
          <issue>1</issue>
          ),
          <fpage>18</fpage>
          -
          <lpage>31</lpage>
          (
          <year>2012</year>
          ). doi:
          <volume>10</volume>
          .1109/T-AFFC.
          <year>2011</year>
          .15
        </mixed-citation>
      </ref>
      <ref id="ref8">
        <mixed-citation>
          [8]
          <string-name>
            <surname>Liu</surname>
            ,
            <given-names>Y.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Sourina</surname>
            ,
            <given-names>O.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Nguyen</surname>
            ,
            <given-names>M.K.</given-names>
          </string-name>
          :
          <article-title>Real-Time EEG-Based Human Emotion Recognition and Visualization</article-title>
          .
          <source>In: Proceedings of the 2010 International Conference on Cyberworlds</source>
          , pp.
          <fpage>262</fpage>
          -
          <lpage>269</lpage>
          . IEEE Computer Society. Singapore (
          <year>2010</year>
          ). doi:
          <volume>10</volume>
          .1109/CW.
          <year>2010</year>
          .37
        </mixed-citation>
      </ref>
      <ref id="ref9">
        <mixed-citation>
          [9]
          <string-name>
            <surname>Mekler</surname>
            ,
            <given-names>A.A.</given-names>
          </string-name>
          :
          <article-title>The program complex for the analysis of electroencephalograms by methods of the dynamic chaos theory:</article-title>
          <source>Ph.D. Thesis. IHB RAS</source>
          , St. Petersburg.(
          <year>2006</year>
          )
          <article-title>(in Russ., Programmnyy kompleks dlya analiza elektroentsefalogramm metodami teorii dinamicheskogo khaosa)</article-title>
        </mixed-citation>
      </ref>
      <ref id="ref10">
        <mixed-citation>
          [10]
          <string-name>
            <surname>Mekler</surname>
            .
            <given-names>A.A.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Gorbunov</surname>
            ,
            <given-names>I.A.</given-names>
          </string-name>
          :
          <article-title>Relation between the pattern of experienced emotions and characteristics of the EEG complexity</article-title>
          .
          <source>In: The Fifth International Conference on Cognitive Science Proceedings</source>
          , pp.
          <fpage>528</fpage>
          -
          <lpage>529</lpage>
          . Kaliningrad, Russia. (
          <year>2012</year>
          )
          <article-title>(in Russ., Pyataya mezhdunarodnaya konferentsiya po kognitivnoy nauke)</article-title>
        </mixed-citation>
      </ref>
      <ref id="ref11">
        <mixed-citation>
          [11]
          <string-name>
            <surname>Perervenko</surname>
          </string-name>
          , U.C.:
          <article-title>Investigation of invariants of the nonlinear dynamics of speech and principles of building an audio analysis system of the psychophysiological state:</article-title>
          <source>Ph.D. Thesis. TTI UFU</source>
          ,
          <string-name>
            <surname>Taganrog</surname>
          </string-name>
          (
          <year>2009</year>
          )
          <article-title>(in Russ., Issledovaniye invariantov nelineynoy dinamiki rechi i printsipy postroyeniya sistemy audioanaliza psikhofiziologicheskogo sostoyaniya)</article-title>
        </mixed-citation>
      </ref>
      <ref id="ref12">
        <mixed-citation>
          [12] RAVDESS Speech/Song Database, https://smartlaboratory.org/ravdess/,
          <source>last accessed</source>
          <year>2018</year>
          /05/08.
        </mixed-citation>
      </ref>
      <ref id="ref13">
        <mixed-citation>
          [13]
          <string-name>
            <surname>Ringeval</surname>
            ,
            <given-names>F.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Sonderegger</surname>
            ,
            <given-names>A.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Sauer</surname>
            ,
            <given-names>J.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Lalanne</surname>
            ,
            <given-names>D.</given-names>
          </string-name>
          :
          <article-title>Introducing the RECOLA Multimodal Corpus of Remote Collaborative and Affective Interactions</article-title>
          .
          <source>In: Proceedings of 10th IEEE International Conference and Workshops on Automatic Face and Gesture Recognition</source>
          , pp.
          <fpage>1</fpage>
          -
          <lpage>8</lpage>
          . IEEE. Shanghai (
          <year>2013</year>
          ). doi:
          <volume>10</volume>
          .1109/FG.
          <year>2013</year>
          .6553805
        </mixed-citation>
      </ref>
      <ref id="ref14">
        <mixed-citation>
          [14]
          <string-name>
            <surname>Shaefer</surname>
            ,
            <given-names>A.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Nils</surname>
            ,
            <given-names>F.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Sanchez</surname>
            ,
            <given-names>X.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Philippot</surname>
            ,
            <given-names>P.</given-names>
          </string-name>
          :
          <article-title>Assessing the effectiveness of a large database of emotion-eliciting films: A new tool for emotion researches</article-title>
          .
          <source>Cognition and Emotion</source>
          <volume>24</volume>
          (
          <issue>7</issue>
          ),
          <fpage>1153</fpage>
          -
          <lpage>1172</lpage>
          (
          <year>2010</year>
          ).
          <source>doi: 10.1080/02699930903274322</source>
        </mixed-citation>
      </ref>
      <ref id="ref15">
        <mixed-citation>
          [15]
          <string-name>
            <surname>Shemaev</surname>
            ,
            <given-names>P.D.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Filatova</surname>
            ,
            <given-names>N.N.</given-names>
          </string-name>
          :
          <article-title>Investigation of the influence of noise in the voice signal on the recognition of the characteristics of the emotion's valence</article-title>
          .
          <source>In: Proceedings of conference «BIOMEDSYSTEMS-2015»</source>
          , pp.
          <fpage>90</fpage>
          -
          <lpage>93</lpage>
          . RSREU. Ryazan,
          <string-name>
            <surname>Russia</surname>
          </string-name>
          (
          <year>2015</year>
          )
          <article-title>(in Russ</article-title>
          .,
          <source>Vserossiyskaya konferentsiya "BIOMEDSISTEMY-2015")</source>
        </mixed-citation>
      </ref>
      <ref id="ref16">
        <mixed-citation>
          [16]
          <string-name>
            <surname>Sidorov</surname>
            ,
            <given-names>K.V.</given-names>
          </string-name>
          :
          <article-title>Biotechnical System of Human Emotions Monitoring by means of Speech Signals</article-title>
          and Electroencephalogram: Ph.
          <string-name>
            <given-names>D.</given-names>
            <surname>Thesis</surname>
          </string-name>
          . TSTU, Tver' (
          <year>2015</year>
          )
          <article-title>(in Russ., Biotekhnicheskaya sistema monitoring emotsiy cheloveka po rechevym signalam i elektroentsefalogrammam)</article-title>
        </mixed-citation>
      </ref>
      <ref id="ref17">
        <mixed-citation>
          [17]
          <string-name>
            <surname>Soleymani</surname>
            ,
            <given-names>M.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Lichtenauer</surname>
            ,
            <given-names>J.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Pun</surname>
            ,
            <given-names>T.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Pantic</surname>
            ,
            <given-names>M.:</given-names>
          </string-name>
          <article-title>A multimodal database for affect recognition and implicit tagging</article-title>
          .
          <source>IEEE Transactions on Affective Computing</source>
          <volume>3</volume>
          (
          <issue>1</issue>
          ),
          <fpage>42</fpage>
          -
          <lpage>55</lpage>
          (
          <year>2012</year>
          ). doi:
          <volume>10</volume>
          .1109/T-AFFG.
          <year>2011</year>
          .25
        </mixed-citation>
      </ref>
      <ref id="ref18">
        <mixed-citation>
          [18]
          <string-name>
            <surname>Tillmann</surname>
            ,
            <given-names>H.G.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Draxler</surname>
          </string-name>
          , Chr.,
          <string-name>
            <surname>Kotten</surname>
            ,
            <given-names>K.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Schiel</surname>
            ,
            <given-names>F.</given-names>
          </string-name>
          :
          <article-title>The Phonetic Goals of the new Bavarian Archive for Speech Signals</article-title>
          . In: Elenius,
          <string-name>
            <given-names>K.</given-names>
            ,
            <surname>Peter Branderud</surname>
          </string-name>
          , P. (eds.) 13th
          <source>International Congress of Phonetic Sciences Proceedings</source>
          , vol.
          <volume>4</volume>
          , pp.
          <fpage>550</fpage>
          -
          <lpage>553</lpage>
          . Congress organizers at KTH and Stockholm University. Stockholm, Sweden (
          <year>1995</year>
          ).
        </mixed-citation>
      </ref>
      <ref id="ref19">
        <mixed-citation>
          [19]
          <string-name>
            <surname>Wang</surname>
            ,
            <given-names>Y.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Guan</surname>
            ,
            <given-names>L.</given-names>
          </string-name>
          :
          <article-title>Recognizing human emotional state from audiovisual signals</article-title>
          .
          <source>IEEE Transactions on Multimedia</source>
          <volume>10</volume>
          (
          <issue>5</issue>
          ),
          <fpage>936</fpage>
          -
          <lpage>946</lpage>
          (
          <year>2008</year>
          ). doi:
          <volume>10</volume>
          .1109/TMM.
          <year>2008</year>
          .927665
        </mixed-citation>
      </ref>
    </ref-list>
  </back>
</article>