<!DOCTYPE article PUBLIC "-//NLM//DTD JATS (Z39.96) Journal Archiving and Interchange DTD v1.0 20120330//EN" "JATS-archivearticle1.dtd">
<article xmlns:xlink="http://www.w3.org/1999/xlink">
  <front>
    <journal-meta />
    <article-meta>
      <title-group>
        <article-title>Pragmatic Markers and Parts of Speech: on the Problems of Annotation of the Speech Corpus</article-title>
      </title-group>
      <contrib-group>
        <contrib contrib-type="author">
          <string-name>Kristin</string-name>
          <xref ref-type="aff" rid="aff0">0</xref>
        </contrib>
        <aff id="aff0">
          <label>0</label>
          <institution>St. Petersburg State University</institution>
          ,
          <addr-line>7/9 Universitetskaya nab., 199034, St. Petersburg</addr-line>
          ,
          <country country="RU">Russia</country>
        </aff>
      </contrib-group>
      <pub-date>
        <year>2020</year>
      </pub-date>
      <fpage>129</fpage>
      <lpage>139</lpage>
      <abstract>
        <p>The article considers the range of possibilities of pragmatic markers (PM) annotation: from the speaker's code to the speaker's commentaries for all difficult cases. The research is based on the material of two corpora of everyday Russian speech - “One Day of Speech” (ORD; dialogues / polylogues) and “Balanced Annotated Text Collection” (SAT; monologues). Two main annotation levels have become the objects of research in this paper: the part of speech of the original lexical unit, from which the basic version of the PM had derived (POS), and the model of formation of the PM which consist of more than one word (Model). The research shows the low feasibility of trying to fit PM into the system of traditional parts of speech, and, conversely, the importance and role of defining a model of formation of PM for their systematic description. In any case, the automatic annotation of corpus material turns out to be considerably difficult.</p>
      </abstract>
      <kwd-group>
        <kwd>Spoken Speech</kwd>
        <kwd>Speech Corpus</kwd>
        <kwd>Pragmatic Marker</kwd>
        <kwd>Pragmaticalization</kwd>
        <kwd>Part of Speech</kwd>
        <kwd>Model of Formation</kwd>
      </kwd-group>
    </article-meta>
  </front>
  <body>
    <sec id="sec-1">
      <title>Introduction</title>
      <p>
        A speech corpus, by definition, should include not only a set of texts, but also their
annotation [Zakharov 2005: 4; Plungian 2005: 6]. Two corpora of everyday Russian
speech, which became the sources of observations for this research, are annotated:
“One Day of Speech” (ORD; dialogues / polylogues)
        <xref ref-type="bibr" rid="ref8 ref9">(see about that: [Russkij yazyk
… 2016; Bogdanova-Beglarian et al. 2016 a, b)</xref>
        and “Balanced Annotated Text
Collection” (SAT; monologues) (see: [Zvukovoj korpus … 2013]). The corpus “One Day
of Speech” was formed using the method of long hours monitoring. This method is
traditionally used in Japanese field linguistics studies
        <xref ref-type="bibr" rid="ref13 ref31">(see: [Shibata 1983; Campbell
2004])</xref>
        , and, furthermore, was implemented during data collection for the spoken
language part of the British National Corpus [Burnard 2020]. The advantage of this
method lies in the receiving for the analysis such spoken material which is the closest
to the natural everyday speech. During the process of the corpus development, this
method has been first-time used as applied to the Russian language. The specificity of
the corpus “Balanced Annotated Text Collection” is that it includes monologues
received from the socially balanced groups of native speakers. The monologues follow
4 most common communicative scenarios: reading of a text, retelling of a text,
de
      </p>
      <p>Copyright ©2020 for this paper by its authors.</p>
      <p>Use permitted under Creative Commons License Attribution 4.0 International (CC BY 4.0).
scription of a picture and a story. This balancing allows comparing monologues of
one speaker, produced in different communicative situations, and monologues of
different speakers, produced in similar communicative situations.</p>
      <p>Apart from other types of annotation, in both corpora the pragmatic markers (PM)
were annotated. PM constitute a significant part of structural units of any oral text:
2,77 % in the whole material, 2,83 and 2,57 % in dialogue and monologue,
respectively.
1</p>
    </sec>
    <sec id="sec-2">
      <title>Annotation of PM: Literature Background</title>
      <p>In this paper, the term “pragmatic markers” describes particular discourse units
(words, expressions, and phrases) with a weakened (sometimes even vanished)
referential meaning, which have a variety of functions in the discourse: marking the start
and the end of the speech, pause-filling, speech-reflection, etc. The term “pragmatic
marker” was developed by B. Fraser who defined them as a class of words that signal
some important for the speaker messages towards the speech [Fraser 1996]. The
majority of the researchers more often use the term “discourse markers” (DM), referring
to a group of discourse units which structure the text or label different kinds of
relations between its parts. [Baranov et al. 1993; Shiffrin 1996; Lenk 1998; Kiseleva,
Paillard 1998, 2003; Schourup 1999].</p>
      <p>However, there are some important differences between discourse and pragmatic
markers as the objects of the investigation [Bogdanova-Beglarian 2018]: DM are put
in the text consciously by the speaker, and PM are inserted unconsciously as speech
automatisms; DM in some cases have lexical and grammatical meanings, and PM
usually lose both lexical meaning and grammatical properties completely; DM, in
general, convey speaker’s attitude toward the speech, and PM express speaker’s
relation to the speech process itself and verbalize difficulties in speech production;
moreover, most of DM can be found in the language dictionaries, while PM are left out of
the lexicographic description.</p>
      <p>Thus, the meanings of the terms PM and DM do not fully coincide. However, the
specificity of their annotation in the corpus material is similar in many respects. This
paper presents several results of the first attempt of pragmatic markers annotation
understood as specific elements of merely oral speech. The annotation of discourse
markers, a wider range of units, as shown, was carried out in the different corpora.</p>
      <p>D. Verdonik, M. Rojc, and M. Stabej annotated DMs in the corpus of Slovenian
telephone conversations TURDIS and analyzed as DM, among the DM in the
conventional understanding, hesitations and backchannel expressions. For most cases, as the
researchers noticed, “it is not possible to say that a discourse marker performs only
one of &lt;…&gt; pragmatic functions” [Verdonik, Rojc, Stabej 2007: 162]. The authors
suggest manual annotation of markers since even if the development of an algorithm,
trained on the manually annotated data, is possible, subsequent manual correction is
needed because of the existing ambiguation of markers and content words.</p>
      <p>L. Crible and S. Zufferey implemented the annotation of DM in French and
English spoken speech and written texts using the structure of four domains — ideational,
rhetorical, sequential, and interpersonal [Crible, Zufferey 2015]. The researchers
stated that the inter-annotator agreement of manual annotation was from 34% (for
English texts) to 52% (for French speech). Several issues of such annotation arose: e.g.,
the distinction of similar DM functions, the discovery of new functions and their
clustering, and the ambiguation of markers and other words.</p>
      <p>L. Crible and M.-J. Cuenca [Crible, Cuenca 2017] reported that most annotation
models of DMs were developed for annotation of written discourse: the Rhetorical
Structure Theory (RST) [Mann, Thompson 1988], the Penn Discourse Treebank
[Prasad et al. 2007], and the Cognitive approach to Cognitive Relations (CCR) [Sanders
et al. 1992]. The researchers annotated discourse markers in the French-English
corpus DisFrEn without applying the prescribed DM list [Ibid.]. The following problems
appeared during the annotation: the presence of truncated structures in spoken speech,
the ambiguity of some DM, and the multifunctionality of markers. Therefore, the
authors concluded that the automatic annotation of DM is not possible.</p>
      <p>Regarding the semi-automatic annotation of pragmatic units, the EXMARaLDA
annotation tools, for instance, should be mentioned. It allows marking two or more
functions for each discourse marker in different contexts manually or
semiautomatically, using prescribed list of discourse functions, but does not allow the
process of annotation being completely automatic [Crible 2018]. The automatic tool
for the annotation of discourse markers is provided in the MDMA (Model for
Discourse Marker Annotation) project, which uses the methodology named
“back-andforth from theory to data” [Université catholique… 2020]. Within this project, manual
selection of DM in the spoken speech and their further semantic, syntactic and
pragmatic annotation is made for the NLP-tasks. The results of the research showed that
only the initial position of the marker in the sentence let the algorithm based on
statistical modeling identify the marker and its particular function [Bolly et al. 2017].</p>
      <p>PRAGMATEXT model of annotation includes the list of tagged pragmatic
functions, e.g., labelling emotional language, discourse relations, discourse modality,
speech act, etc. The researchers used this model for the first attempt of discourse
markers annotation in the multilingual parallel corpus (Arabic-Spanish-English)
[Samy, Gonzalez-Ledesma 2008]. At the first stage, the Spanish part of the corpus
was annotated at the discourse markers level. At the second stage, the comparison
with the DM in texts in another two languages was made using a bilingual dictionary.
Non-ambiguous DM in Arabic and English texts were automatically tagged the same
way as in the Spanish texts. The ambiguous markers were disambiguated manually
considering their prosodic features and position within the sentence. The authors
intend to develop the automatic disambiguation tool for the purposes of the DM
annotation; however, it is prevented by such factors as DM categorical, syntactic, and
discursive ambiguity, as well as the absence of the clear distinction between DMs and
idiomatic expressions since the DM tend to form the lasts.</p>
      <p>Thus, as it was shown above, pragmatic markers annotation of corpus spoken
speech data can be performed manually and semi-automatically with necessary
checking.</p>
    </sec>
    <sec id="sec-3">
      <title>Annotation of PM and Types of Pragmatic Markers</title>
      <p>The annotation was implemented at the several levels: the particular usage of PM
(PM); its functions in this particular usage (Function PM); the commentaries for
introducing the optional information and marking the difficult cases which show the
troubles in the detection of PM and their functions (Comment PM); the basic version
of PM (excluding its structural versions and/or inflectional paradigm) (Standard); the
parts-of-speech tagging of the source lexical unit, from which the basic version of the
PM derived (POS); the model of formation of the PM which consist of more than one
word (Model); the speaker’s code (Speaker PM); and phrase commentary (Phrase)
[Bogdanova-Beglarian et al. 2018, 2019b].</p>
      <p>The main functional types of PM in oral discourse turned to be the following: A –
marker-approximator (vrode ‘like’, kak by ‘kinda’), G – boundary marker (starting,
final, and navigational), D – deictic marker (vot (…) vot ‘like … this’), Z – all types of
replacement markers (for someone else’s speech, whole set or its parts: bla-bla-bla, i
vse dela ‘and all that’, i vs’o takoe prochee ‘and all that stuff’), K – “xeno” marker
(tipa (togo chto) ‘sort of’, takoj ‘like’), M – meta-communicative marker (da ‘yeah’,
(ja) ne znaju ‘(I) don’t know’, znaesh’ ‘you know’, smotri ‘look’), F – reflexive
marker (skazhem tak ‘let’s say’, ili kak tam? ‘or whatever’), H – hesitation marker (eto
‘what’, tam ‘em’, eto samoe ‘whatchamacallit’) [Ibid.].</p>
      <p>The automatic annotation of such material seems almost impossible: the very
specificity of oral spontaneous speech, which is difficult to any systematization, causes too
many problems. For instance, the syntagmatic division of spontaneous speech itself is
problematic, which is relevant for the distinction of various boundary PM (G):
starting, navigational, and final. It is also difficult to establish a distinction between the
formally similar PM and meaningful units of discourse, that are pragmaticalized in the
speech and sometimes are at the different stages of the pragmaticalization from the
lexemes to the pragmatemes (on prishol, a tam nikogo net ‘he came but no one was
there’ (adverb of place) – on tam prishol tam, a nikogo net ‘he em came em but no
one was there’ (two PM used in hesitative and rhythm-forming functions)).</p>
      <p>The most PM of Russian speech are polyfunctional, which leads to the necessity of
identifying the main and additional functions of each marker in every particular case
(on tam prishol tam ‘he em came em’ – HR). At last, spontaneous speech reveals such
feature of PM as their “magnetism”, attraction of one PM to another if they have one
common (synonymous) function. Consequently, the need to distinguish different PM
which consist of more than one word, on the one side, and a chain of markers, on the
other side, appears: eto kak jego ‘what whatchamacallit’ (one marker) or eto + kak
jego ‘er + whatchamacallit’ (a chain of markers) [Bogdanova-Beglarian et al. 2019a].
3</p>
    </sec>
    <sec id="sec-4">
      <title>Pragmatic Markers and Parts of Speech</title>
      <p>The set of the PM revealed in the material shows that PM have different “origins” in
the field of parts of speech: particles (vot ‘here’, von ‘there’), verbs (znat’ ‘to know’,
govorit’ ‘to speak’, smotret’ ‘to look’, dumat’ ‘to think’), including gerund (govor’a
‘speaking’), adverbs (tam ‘there’, tak ‘that way’, kuda ‘where’), pronouns (etot ‘this’,
samyj ‘the most’, on ‘he’, ona ‘she’), conjunctions/prepositions (tipa ‘kind of’, vrode
‘like’).</p>
      <p>The parts-of-speech tagging of corpus data was initially made automatically with
the software “MyStem” (Yandex Technologies) and then checked manually. Only
particular usages of PM were annotated. In the table 1, the results of this annotation in
the ORD-corpus for the top of the frequency list of PM (first 20 ranks) are
demonstrated. It can be already seen that certain difficulties during the automatic annotation
with a help of “MyStem” software application arise, as well as insignificant
divergence of two annotation types.</p>
      <p>Thus, the program does not identify the integrity of the unit kak by ‘kinda’,
marking it as “adverbial pronoun + particle” (ADVPRO&amp;PART), while, during the manual
annotation, the experts choose the option which is closer to the nature of this unit –
“particle / conjunction”.</p>
      <p>The software attributes to the marker znachit ‘well’ the tag “adverb and
parenthesis” (ADV, parenth), whereas the manual annotation gives a variant “verb /
parenthesis”. The element eto ‘what’ in all PM (eto ‘what’ and eto samoe ‘whatchamacallit’)
is marked by the software as the “subject pronoun” (SPRO), although the traditional
grammar, which became a base of manual annotation, treats this unit as the “adjective
pronoun”, that nominalized in the particular cases. The adjective nature of this unit is
supported, for instance, by the ability of gender inflection (eta ‘what (fem.)’, eto ‘what
(neutr.)’, etot ‘what (masc.)’). The marker tipa ‘kind of’ in the automatic annotation is
merely the “particle” (PART), although in the manual annotation it is “noun /
preposition”, which is required by the dictionaries in the first place.</p>
      <p>However, even considering revealed inaccuracy of the automatic POS annotation
of material, it is clear that the information about the POS of the original units, which
have pragmaticalized and became pragmatic markers in oral speech, is rather a
historical background which does not really describe new discourse units. For instance, the
markers tam and tak as a PM lose all the adverbial properties [Turchanenko 2018], the
word da as a meta-communicative marker falls into category of neither particle, nor
conjunction [Shershneva 2015].</p>
      <p>The verbal meta-communicative markers similarly lose the majority of their verbal
characteristics in their new usage: verbs in indicative mood like znaesh’/znaete ‘you
know’, vidish’/vidite ‘you see’, ponimaesh’/ponimaete ‘you know’ and verbs in
imperative mood as slushaj/slushajte ‘listen’, predstav’/predstav’te ‘imagine’ leave merely
formal number inflection [Bogdanova-Beglarian, Maslova 2019], the markers (ja) ne
znaju ‘(I) don’t know’ и znachit ‘well’ completely lose any grammatical inflection and
are used only in one fixed form [Bogdanova-Beglarian 2019], and the pragmatic
“xeno” marker govorit ‘says’ is presented in the spoken speech solely in the present
tense forms, more often phonetically reduced (grit, gyt, gr’u, grim, etc.) [Stojka
2019].
It could hardly be correct to refer all these pragmaticalized forms to the certain
traditional POS categories.
4</p>
    </sec>
    <sec id="sec-5">
      <title>Basic Versions of PM and Parts of Speech</title>
      <p>The top of the frequency list of basic (standard) versions of PM seems slightly
different than the one of particular usages of PM in the table (the data from the two corpora
altogether):
(...) vot ‘(…) er’ (IPM here and hereinafter – 7119),
(...) tam ‘(…) em’ (2970),
(...) eto, eta, eti… (…) ‘(…) what… (…)’ (1827),
(...) da/da da da ‘(…) yeah/ yeah yeah yeah’ (1572),
(...) tak/tak tak tak ‘(…) well/well well well (1357),
(...) kak by ‘(…) kinda’ (1353),
govorit/govor’u/govorim... ‘says/say…’ (1337),
znachit (...) ‘well (…)’ (1062),
takoj/takaja, takie ‘like’ (1033),
eto samoe/eti samye, etot samyj… ‘whatchamacallit…’ (879),
(...) znaesh’ (...)/(...) znaete (...) ‘(…) you know (…)’ (839),
vot (...) vot ‘like (…) this’ (778),
(...) (po)slushaj /(...) (po)slushajte ‘(…) listen’ (750),
(...) ne znaju ‘(…) don’t know’ (498),
(...) koroche govor’a ‘(…) long story short’ (462),
(...) tipa/tipa togo/tipa togo chto ‘(…) sort of’ (458),
(...) ponimaesh’ / (...) ponimaete ‘(…) you know (…)’ (405),
(...) vs’o ‘that’s all’ (357),
(...) vidish’ (...) / vidite ‘(…) you see (…)’ (255),
voobshche ‘generally’ (231),
(...) dumaju (...) ‘(…) think (…)’ (223),
(...) skazhem (...) ‘(…) let’s say (…)’ (211),
(...) v principe ‘(…) basically’ (207),
vrode (...) ‘like (…)’ (150),
(...) v obshchem ‘(…) anyway’ (130),
smotri/smotrite ‘look’ (122),
na samom dele ‘actually’ (122),
(ty) predstavlyaesh’ ‘you know’ (113),
shchas/shchas shchas shchas ‘one moment’ (93),
(…) tak dalee ‘(…) so on’ (89).</p>
      <p>One could see that the majority of markers has only generalized structure of basic
version, with potential extension or restricted grammatical flexibility, cf.: VOT – i vot
‘and er’, da vot ‘well er’; ZNAESH – ty znaesh’ ‘you know’, vot znaes’h ‘er you
know’, nu znajete ‘well you know’, etc.; GOVORIT – govor’u ‘I say’, govorish’ ‘you
say’, govorim ‘we say’, etc.; VRODE – nu vrode ‘well like’, vrode kak ‘like as’, vrode
by ‘like well’. Rare PM from the whole list of PM, annotated in the material, do not
show such structural variability: von ‘err’, prikin’ ‘guess’, i tak dalee ‘and so on’, po
idee ‘normally’ and a few others. However, the deictic marker VOT (…) VOT ‘like
… this’ exists merely as a structural model, which is filled by a new unit each time:
vot tak vot ‘like this’, vot takoj vot ‘like this’, vot ots’uda vot ‘like this’, etc. In fact,
this marker does not have some single basic (standard) form. The automatic
parts-ofspeech tagging of such material not only seems difficult, but also has rather inaccurate
results since it cannot consider the specificity of possible extensions.
5</p>
    </sec>
    <sec id="sec-6">
      <title>Models of Formation of PM</title>
      <p>The annotation of corpus material at the level of models of formation of the PM,
which consist of more than one word (Model), is supposed to be the most informative
and scientifically valuable. At least 12 such models have been identified:
1. PM, which initially consist of more than one word, that are basic versions (but
not the source “lexicographic” version): eto samoe ‘whatchamacallit’, kak jego
(jejo, ikh) ‘whatchamacallit’, kak eto? ‘whatchamacallit?’ kak skazat’? ‘how can
I say?’ kak eto nazyvaets’a? ‘what am I call it?’ chto jeshcho? ‘what else?’ kak
(by) skazat’? ‘how can I say it?’
2. combination of two or more PM, which consist of one word: nu vot ‘well er’, nu
tam ‘well em’, nu tak ‘well um’, vot tak ‘er um’, nu znaesh’ ‘well you know’, tak
skazhem ‘let’s say’, skazhem tak ‘let’s say’, skazhem tam ‘let’s say em’, vot
skazhem ‘er let’s say’, znaesh’ tam ‘you know em’, vot kak by ‘er kinda’, vot
skazhem tak ‘er let’s say’, nu ne znaju ‘well don’t know’, tam tipa ‘em sort of’, nu
koroche ‘well in short’, znachit vot ‘well er’, v principe vs’o ‘basically that’s all’;
3. combination of PM, which consist of one and more than one word: nu kak
skazat’? ‘well how can I say?’ kak jego tam? ‘em whatchamacallit?’ nu vot eti vot
‘well these ones’, nu (ja) ne znaju tam ‘well (I) don’t know em’;
4. addition of the personal pronoun with a weakened lexical and grammatical
meaning: (ja) ne znaju ‘(I) don’t know’; (ja) (ne) dumaju (chto) ‘(I) don’t think (that)’;
(ty) znaesh’, ponimaesh’, vidish’… ‘(you) know’; (ty) predstav’, prikin’… ‘(you)
imagine’; chto (tebe) jeshchyo skazat’? ‘what else can I say (to you)?’
5. addition of emphatic particles/conjunctions: i vse dela ‘and all that’, i vs’o takoe
‘and all that’, a vot ‘and er’, nu i vs’o ‘well that’s all’, i to i s’o ‘this and that’, ja
uzh ne znaju tam ‘I even don’t know em’, ty zh ponimaesh’ ‘you really know’;
6. addition of non-personal pronoun: vs’o takoe prochee ‘all that stuff’, tipa
togo/etogo ‘sort of’, vrode togo ‘like’, takoj kakoj-to ‘like that one’, kak (by) eto
skazat’? ‘how can I say that?’
7. addition of the conjunction CHTO (CHEGO) ‘that’: dumaju chto ‘think that’,
bojus’ chto ‘am afraid that’, tipa togo chto ‘sort of that’, vrode togo chto ‘like
that’, znaesh’ chto/chego ‘you know that’;
8. addition of parentheses: vs’o navernoe ‘that’s all probably’, vs’o pozhaluj ‘that’s
all perhaps’;
9. addition of interjection: oj slushaj ‘ooh listen’;
10. loss of the gerund GOVORYA ‘speaking’: koroche ‘in short’, sobstvenno
‘strictly’, voobshche ‘generally’;
11. reduplication: da-da-da ‘yeah-yeah-yeah’, na-na-na, shchas-shchas-shchas ‘one
moment-one moment-one moment’, te-te-te, op-op-op, bla-bla-bla, tak-tak-tak
‘em-em-em’, eto-eto-eto ‘what-what-what’;
12. ILI ‘or’ + (more often) the rhetorical question: ili kak jego? ‘or whatchamacallit?’
ili kak tam? ‘or whatchamacallit?’ ili chto? ‘or what?’ ili kak skazat’? ‘or how to
say?’ ili etot? ‘or what?’ nu ili ne znaju ‘well or I don’t know’.</p>
    </sec>
    <sec id="sec-7">
      <title>Conclusion</title>
      <p>Previous works have shown that in the speech recognition process POS-tagging of
some markers (excluding multi-word markers and phrases) can be useful for the task
of prediction of the following words [Heemant et al. 1998].</p>
      <p>However, the automatically derived classification algorithm of DM POS-tagging
showed an error rate of 37,3%, in comparison to, for instance, the error rate of 45,3%
for the algorithm of J. Hirschberg and D. Litman [Hirschberg, Litman 1993]. In other
words, using this automatic algorithm (decision tree), only for 4 from 10 particular
pragmatic markers the correct tag could be assigned. Presumably, this POS-tagging
heuristic may be improved by the expansion of data, and only after implemented for
the objectives of this investigation.</p>
      <p>
        Our paper provides the theoretical basis of the relevant PM POS-tagging and the
classification of PM-models for further linguistic elaboration. Anyway, the result of
speech corpora annotation at the level of pragmatic markers can become a systematic
description of PM as the inherent structural components of oral discourse. The
description should be done considering PM functions, polyfunctionality, and possible
“synonymic” relations, their formal grammar
        <xref ref-type="bibr" rid="ref10 ref29 ref3 ref5 ref6 ref7">(and not only at the parts-of-speech
level, but also, for example, at the level of predicative units [Bogdanova-Beglarian,
Zaides 2019])</xref>
        , the specificity of their usage, and the possible correlation with
speaker’s characteristics, type of speech (monologue/dialogue) or communicative situation.
      </p>
    </sec>
  </body>
  <back>
    <ref-list>
      <ref id="ref1">
        <mixed-citation>
          1.
          <string-name>
            <surname>Baranov</surname>
            ,
            <given-names>A. N.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Plungian</surname>
            ,
            <given-names>V. A.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Rakhilina</surname>
            ,
            <given-names>E. V.</given-names>
          </string-name>
          : Guide to Discursive
          <source>Words of Russian. Pomovskij i Partnery</source>
          , Moscow (
          <year>1993</year>
          ). [In Rissian].
        </mixed-citation>
      </ref>
      <ref id="ref2">
        <mixed-citation>
          2.
          <string-name>
            <surname>Bogdanova-Beglarian</surname>
            ,
            <given-names>N. V.</given-names>
          </string-name>
          :
          <article-title>On the Possible Communicative Interference in CrossCultural Oral Communication</article-title>
          .
          <source>Mir russkogo slova</source>
          ,
          <volume>3</volume>
          ,
          <fpage>93</fpage>
          -
          <lpage>99</lpage>
          (
          <year>2018</year>
          ). [In Rissian].
        </mixed-citation>
      </ref>
      <ref id="ref3">
        <mixed-citation>
          3.
          <string-name>
            <surname>Bogdanova-Beglarian</surname>
            ,
            <given-names>N. V.</given-names>
          </string-name>
          :
          <article-title>Grammatical “Atavisms” of Pragmatic Markers of Russian Oral Speech</article-title>
          . In: Glazunova,
          <string-name>
            <given-names>O. I.</given-names>
            ,
            <surname>Rogova</surname>
          </string-name>
          ,
          <string-name>
            <surname>K. A</surname>
          </string-name>
          . (eds.).
          <source>Russian Grammar: Structural Language Organization and Processes of Language Functioning</source>
          , pp.
          <fpage>436</fpage>
          -
          <lpage>446</lpage>
          . Moscow (
          <year>2019</year>
          ). [In Rissian].
        </mixed-citation>
      </ref>
      <ref id="ref4">
        <mixed-citation>
          4.
          <string-name>
            <surname>Bogdanova-Beglarian</surname>
            ,
            <given-names>N.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Blinova</surname>
            ,
            <given-names>O.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Martynenko</surname>
            ,
            <given-names>G.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Sherstinova</surname>
          </string-name>
          . T.,
          <string-name>
            <surname>Zaides</surname>
            ,
            <given-names>K.</given-names>
          </string-name>
          :
          <article-title>Pragmatic Markers in Russian Spoken Speech: an Experience of Systematization and Annotation for the Improvement of NLP Tasks</article-title>
          . In: Balandin,
          <string-name>
            <given-names>S.</given-names>
            ,
            <surname>Cinotti</surname>
          </string-name>
          ,
          <string-name>
            <given-names>T.</given-names>
            ,
            <surname>Viola</surname>
          </string-name>
          ,
          <string-name>
            <given-names>F.</given-names>
            ,
            <surname>Tyutina</surname>
          </string-name>
          , T. (eds.).
          <source>Proceedings of the FRUCT'23</source>
          . Bologna, Italy,
          <fpage>13</fpage>
          -16
          <source>November</source>
          <year>2018</year>
          , pp.
          <fpage>69</fpage>
          -
          <lpage>77</lpage>
          . FRUCT Oy,
          <string-name>
            <surname>Finland</surname>
          </string-name>
          (
          <year>2018</year>
          ).
        </mixed-citation>
      </ref>
      <ref id="ref5">
        <mixed-citation>
          5.
          <string-name>
            <surname>Bogdanova-Beglarian</surname>
            ,
            <given-names>N. V.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Blinova</surname>
            ,
            <given-names>O. V.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Martynenko</surname>
            ,
            <given-names>G. Ja.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Sherstinova</surname>
            ,
            <given-names>T. Ju.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Zaides</surname>
            ,
            <given-names>K. D.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Popova</surname>
            ,
            <given-names>T. I.</given-names>
          </string-name>
          :
          <article-title>Pragmatic Markers Annotation in Russian Speech Corpus: Research Problem, Approaches and Results</article-title>
          . In: Selegej,
          <string-name>
            <surname>V. P.</surname>
          </string-name>
          (ed.).
          <source>“Dialogue” Conference Proceedings “Computational Linguistics and Intellectual Technologies”. Moscow</source>
          , 29 May - 1
          <source>June</source>
          <year>2019</year>
          ,
          <volume>18</volume>
          (
          <issue>25</issue>
          ), pp.
          <fpage>72</fpage>
          -
          <lpage>85</lpage>
          (
          <year>2019a</year>
          ). [In Rissian].
        </mixed-citation>
      </ref>
      <ref id="ref6">
        <mixed-citation>
          6.
          <string-name>
            <surname>Bogdanova-Beglarian</surname>
            ,
            <given-names>N. V.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Blinova</surname>
            ,
            <given-names>O. V.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Sherstinova</surname>
            ,
            <given-names>T. Ju.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Troshchenkova</surname>
            ,
            <given-names>E. V.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Gorbunova</surname>
            ,
            <given-names>D. A.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Zaides</surname>
          </string-name>
          , K. D.:
          <article-title>Pragmatic Markers of Russian Everyday Speech: The Revised Typology and Corpus-Based Study</article-title>
          . In: Balandin,
          <string-name>
            <given-names>S.</given-names>
            ,
            <surname>Niemi</surname>
          </string-name>
          ,
          <string-name>
            <given-names>V.</given-names>
            ,
            <surname>Tuytina</surname>
          </string-name>
          , T. (eds.).
          <source>Proceedings of the 25th Conference of Open Innovations Association FRUCT</source>
          , pp.
          <fpage>57</fpage>
          -
          <lpage>63</lpage>
          . Helsinki, Finland (
          <year>2019</year>
          ).
        </mixed-citation>
      </ref>
      <ref id="ref7">
        <mixed-citation>
          7.
          <string-name>
            <surname>Bogdanova-Beglarian</surname>
            ,
            <given-names>N. V.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Maslova</surname>
          </string-name>
          , Je. R.:
          <article-title>Russian Contact Verbs in Oral Spontaneous Speech: Dictionary Volume and Functional and Semantical Diversity</article-title>
          .
          <source>In: Acta Linguistica Petropolitana. Proceedings of the Institute of Linguistics RAS</source>
          ,
          <volume>3</volume>
          (
          <issue>15</issue>
          ), pp.
          <fpage>115</fpage>
          -
          <lpage>135</lpage>
          . St.
          <string-name>
            <surname>Petersburg</surname>
          </string-name>
          (
          <year>2019</year>
          ). [In Rissian].
        </mixed-citation>
      </ref>
      <ref id="ref8">
        <mixed-citation>
          8.
          <string-name>
            <surname>Bogdanova-Beglarian</surname>
            ,
            <given-names>N.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Sherstinova</surname>
            ,
            <given-names>T.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Blinova</surname>
            ,
            <given-names>O.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Baeva</surname>
            ,
            <given-names>E.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Martynenko</surname>
            ,
            <given-names>G.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Ryko</surname>
            ,
            <given-names>A.</given-names>
          </string-name>
          :
          <article-title>Sociolinguistic Extension of the ORD Corpus of Russian Everyday Speech</article-title>
          .
          <source>In: SPECOM 2016, Lecture Notes in Artificial Intelligence, LNAI</source>
          , vol.
          <volume>9811</volume>
          , pp.
          <fpage>659</fpage>
          -
          <lpage>666</lpage>
          . Springer, Switzerland (
          <year>2016</year>
          ).
        </mixed-citation>
      </ref>
      <ref id="ref9">
        <mixed-citation>
          9.
          <string-name>
            <surname>Bogdanova-Beglarian</surname>
            ,
            <given-names>N.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Sherstinova</surname>
            ,
            <given-names>T.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Blinova</surname>
            ,
            <given-names>O.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Martynenko</surname>
            ,
            <given-names>G.</given-names>
          </string-name>
          :
          <article-title>An Exploratory Study on Sociolinguistic Variation of Spoken Russian</article-title>
          .
          <source>In: SPECOM 2016. Lecture Notes in Artificial Intelligence, LNAI</source>
          , vol.
          <volume>9811</volume>
          , pp.
          <fpage>100</fpage>
          -
          <lpage>107</lpage>
          . Springer, Switzerland (
          <year>2016</year>
          ).
        </mixed-citation>
      </ref>
      <ref id="ref10">
        <mixed-citation>
          10.
          <string-name>
            <surname>Bogdanova-Beglarian</surname>
            ,
            <given-names>N. V.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Zaides</surname>
          </string-name>
          , K. D.:
          <article-title>Corpus of Monological Speech: Real and Formal Predictivity of Russian Oral Discourse Units</article-title>
          . In: Kocharov,
          <string-name>
            <given-names>D. A.</given-names>
            ,
            <surname>Skrelin</surname>
          </string-name>
          , P. A. (eds.).
          <source>Analysis of Spoken Russian Speech (AR3-2019): Proceedings of the 8th Interdisciplinary Seminar</source>
          , pp.
          <fpage>11</fpage>
          -
          <lpage>16</lpage>
          . St.
          <string-name>
            <surname>Petersburg</surname>
          </string-name>
          (
          <year>2019</year>
          ). [In Rissian].
        </mixed-citation>
      </ref>
      <ref id="ref11">
        <mixed-citation>
          11.
          <string-name>
            <surname>Bolly</surname>
            ,
            <given-names>C. T.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Crible</surname>
            ,
            <given-names>L.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Degand</surname>
            ,
            <given-names>L.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Uygur-Distexhe</surname>
            ,
            <given-names>D.</given-names>
          </string-name>
          :
          <article-title>Towards a Model for Discourse Marker Annotation: From Potential to Feature-based Discourse Markers</article-title>
          . In: Fedriani,
          <string-name>
            <given-names>Ch.</given-names>
            ,
            <surname>Sansó</surname>
          </string-name>
          ,
          <string-name>
            <surname>A</surname>
          </string-name>
          . (eds.).
          <source>Pragmatic Markers</source>
          , Discourse Markers and Modal Particles:
          <article-title>New perspectives</article-title>
          . Pp.
          <volume>71</volume>
          -
          <fpage>98</fpage>
          . John Benjamins, Amsterdam (
          <year>2017</year>
          ).
        </mixed-citation>
      </ref>
      <ref id="ref12">
        <mixed-citation>
          12.
          <string-name>
            <surname>Burnard</surname>
            ,
            <given-names>L</given-names>
          </string-name>
          . (ed.):
          <article-title>Reference Guide for the British National Corpus (XML edition)</article-title>
          .
          <source>Published for the British National Corpus Consortium</source>
          by Oxford University Computing Services, [Electronic resource], http://www.natcorp.ox.ac.uk/docs/URG/, last accessed 01/05/
          <year>2020</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref13">
        <mixed-citation>
          13.
          <string-name>
            <surname>Campbell</surname>
          </string-name>
          , N.:
          <string-name>
            <surname>Speech</surname>
          </string-name>
          &amp;
          <article-title>Expression; the Value of a Longitudinal Corpus</article-title>
          .
          <source>In: Proceedings of the Fourth International Conference on Language Resources and Evaluation LREC</source>
          <year>2004</year>
          , pp.
          <fpage>183</fpage>
          -
          <lpage>186</lpage>
          . ELRA, Lisbon, Portugal (
          <year>2004</year>
          ).
        </mixed-citation>
      </ref>
      <ref id="ref14">
        <mixed-citation>
          14.
          <string-name>
            <surname>Crible</surname>
            ,
            <given-names>L.</given-names>
          </string-name>
          :
          <article-title>Discourse Markers and (Dis)fluency. Forms and Functions across Languages and Registers</article-title>
          . John Benjamins, Amsterdam (
          <year>2018</year>
          ).
        </mixed-citation>
      </ref>
      <ref id="ref15">
        <mixed-citation>
          15.
          <string-name>
            <surname>Crible</surname>
            ,
            <given-names>L.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Cuenca</surname>
          </string-name>
          , M.-J.:
          <article-title>Discourse Markers in Speech: Characteristics and Challenges for Corpus Annotation</article-title>
          .
          <source>Dialogue and Discourse</source>
          ,
          <volume>8</volume>
          (
          <issue>2</issue>
          ),
          <fpage>149</fpage>
          -
          <lpage>166</lpage>
          (
          <year>2017</year>
          ).
        </mixed-citation>
      </ref>
      <ref id="ref16">
        <mixed-citation>
          16.
          <string-name>
            <surname>Crible</surname>
            ,
            <given-names>L.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Zufferey</surname>
            ,
            <given-names>S.</given-names>
          </string-name>
          :
          <article-title>Using a Unified Taxonomy to Annotate Discourse Markers in Speech and Writing</article-title>
          .
          <source>In: Proceedings of the 11th Joint ISO-ACL/SIGSEM Workshop on Interoperable Semantic Annotation</source>
          , pp.
          <fpage>14</fpage>
          -
          <lpage>22</lpage>
          . London, UK (
          <year>2015</year>
          ).
        </mixed-citation>
      </ref>
      <ref id="ref17">
        <mixed-citation>
          17.
          <string-name>
            <surname>Kiseleva</surname>
            ,
            <given-names>K.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Paillard</surname>
            ,
            <given-names>D</given-names>
          </string-name>
          . (eds).
          <source>Discursive Words of Russian: Experience of Contextual and Semantic Description</source>
          . Moscow, Metatext (
          <year>1998</year>
          ) [In Rissian].
        </mixed-citation>
      </ref>
      <ref id="ref18">
        <mixed-citation>
          18.
          <string-name>
            <surname>Kiseleva</surname>
            ,
            <given-names>K.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Paillard</surname>
            ,
            <given-names>D</given-names>
          </string-name>
          . (eds).
          <source>Discourse Words of Russian: Contextual Variation and Semantic Units</source>
          . Moscow, Azbukovnik (
          <year>2003</year>
          ) [In Rissian].
        </mixed-citation>
      </ref>
      <ref id="ref19">
        <mixed-citation>
          19.
          <string-name>
            <surname>Fraser</surname>
            ,
            <given-names>B.: Pragmatic</given-names>
          </string-name>
          <string-name>
            <surname>Markers</surname>
          </string-name>
          . Pragmatics,
          <volume>6</volume>
          (
          <issue>2</issue>
          ),
          <fpage>167</fpage>
          -
          <lpage>190</lpage>
          (
          <year>1996</year>
          ).
        </mixed-citation>
      </ref>
      <ref id="ref20">
        <mixed-citation>
          20.
          <string-name>
            <surname>Heemant</surname>
            ,
            <given-names>P. A.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Byron</surname>
            ,
            <given-names>D.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Allen</surname>
            ,
            <given-names>J. F.</given-names>
          </string-name>
          :
          <article-title>Identifying Discourse Markers in Spoken Dialog</article-title>
          .
          <source>In: AAAI 1998 Spring Symposium on Applying Machine Learning to Discourse Processing</source>
          , pp.
          <fpage>44</fpage>
          -
          <lpage>51</lpage>
          . The AAAI Press, California, Menlo Park (
          <year>1998</year>
          ).
        </mixed-citation>
      </ref>
      <ref id="ref21">
        <mixed-citation>
          21.
          <string-name>
            <surname>Hirschberg</surname>
            ,
            <given-names>J.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Litman</surname>
            ,
            <given-names>D.</given-names>
          </string-name>
          :
          <article-title>Empirical Studies on the Disambiguation of Cue Phrases</article-title>
          .
          <source>Computational Linguistics</source>
          ,
          <volume>19</volume>
          (
          <issue>3</issue>
          ),
          <fpage>501</fpage>
          -
          <lpage>530</lpage>
          (
          <year>1993</year>
          ).
        </mixed-citation>
      </ref>
      <ref id="ref22">
        <mixed-citation>
          22.
          <string-name>
            <surname>Lenk</surname>
            ,
            <given-names>U.</given-names>
          </string-name>
          :
          <article-title>Marking Discourse Coherence: Functions of Discourse Markers in Spoken English</article-title>
          . Narr,
          <string-name>
            <surname>Tuebingen</surname>
          </string-name>
          (
          <year>1998</year>
          ).
        </mixed-citation>
      </ref>
      <ref id="ref23">
        <mixed-citation>
          23.
          <string-name>
            <surname>Mann</surname>
            ,
            <given-names>W. C.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Thompson</surname>
            ,
            <given-names>S. A.</given-names>
          </string-name>
          :
          <article-title>Rhetorical Structure Theory: Toward a Functional Theory of Text Organization</article-title>
          . Text,
          <volume>8</volume>
          (
          <issue>3</issue>
          ),
          <fpage>243</fpage>
          -
          <lpage>281</lpage>
          (
          <year>1988</year>
          ).
        </mixed-citation>
      </ref>
      <ref id="ref24">
        <mixed-citation>
          24.
          <string-name>
            <surname>Plungian</surname>
            ,
            <given-names>V. A.</given-names>
          </string-name>
          :
          <article-title>Why We Need the Russian National Corpus: Informal Introduction</article-title>
          .
          <source>Russian National Corpus</source>
          ,
          <fpage>2003</fpage>
          -
          <lpage>2005</lpage>
          , pp.
          <fpage>6</fpage>
          -
          <lpage>20</lpage>
          . Moscow (
          <year>2005</year>
          ). [In Rissian].
        </mixed-citation>
      </ref>
      <ref id="ref25">
        <mixed-citation>
          25.
          <string-name>
            <surname>Prasad</surname>
            ,
            <given-names>R.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Miltsakaki</surname>
            ,
            <given-names>E.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Dinesh</surname>
            ,
            <given-names>N.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Lee</surname>
            ,
            <given-names>A.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Joshi</surname>
            ,
            <given-names>A.</given-names>
          </string-name>
          :
          <article-title>The Penn Discourse Treebank 2.0 annotation manual</article-title>
          .
          <source>Technical report, Institute for Research in Cognitive Science</source>
          ,
          <year>2007</year>
          , [Electronic resource], https://repository.upenn.edu/cgi/viewcontent.cgi?article= 1203&amp;context=ircs_reports, last accessed 01/05/
          <year>2020</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref26">
        <mixed-citation>
          26.
          <article-title>The Everyday Russian Language: Functioning Features in Different Social Groups</article-title>
          .
          <string-name>
            <surname>Bogdanova-Beglarian</surname>
            ,
            <given-names>N. V. (ed.). Collective</given-names>
          </string-name>
          <string-name>
            <surname>Monograph</surname>
          </string-name>
          . St.
          <string-name>
            <surname>Petersburg</surname>
          </string-name>
          (
          <year>2016</year>
          ).
          <article-title>(in Russian)</article-title>
        </mixed-citation>
      </ref>
      <ref id="ref27">
        <mixed-citation>
          27.
          <string-name>
            <surname>Samy</surname>
            ,
            <given-names>D.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Gonzalez-Ledesma</surname>
            ,
            <given-names>A.</given-names>
          </string-name>
          :
          <article-title>Pragmatic Annotation of Discourse Markers in a Multilingual Parallel Corpus (Arabic-Spanish-English)</article-title>
          .
          <source>In: Proceedings of the International Conference on Language Resources and Evaluation</source>
          ,
          <string-name>
            <surname>LREC</surname>
          </string-name>
          <year>2008</year>
          , Marrakech, Morocco, [Electronic resource], http://www.analedesma.es/wp-content/uploads/2010/04/ doaingles.pdf, last accessed 01/05/
          <year>2020</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref28">
        <mixed-citation>
          28.
          <string-name>
            <surname>Sanders</surname>
            ,
            <given-names>T. J. M.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Spooren</surname>
            ,
            <given-names>W. P. M. S.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Noordman</surname>
            ,
            <given-names>L. G. M.</given-names>
          </string-name>
          :
          <article-title>Toward a Taxonomy of Coherence Relations</article-title>
          .
          <source>Discourse Processes</source>
          ,
          <volume>15</volume>
          (
          <issue>1</issue>
          ),
          <fpage>1</fpage>
          -
          <lpage>35</lpage>
          (
          <year>1992</year>
          ).
        </mixed-citation>
      </ref>
      <ref id="ref29">
        <mixed-citation>
          29.
          <string-name>
            <surname>Stojka</surname>
            ,
            <given-names>D. A.</given-names>
          </string-name>
          :
          <article-title>The Dictionary of Reduced Forms of Russian Speech</article-title>
          . BogdanovaBeglarian, N. V. (ed.).
          <source>St. Petersburg</source>
          (
          <year>2019</year>
          ). [In Rissian].
        </mixed-citation>
      </ref>
      <ref id="ref30">
        <mixed-citation>
          30.
          <string-name>
            <surname>Shershneva</surname>
            ,
            <given-names>D. M.</given-names>
          </string-name>
          :
          <article-title>Da as a Lexical and Functional Unit of Russian Speech</article-title>
          . Studia
          <string-name>
            <surname>Slavica</surname>
            <given-names>XVIII</given-names>
          </string-name>
          , Tallinn,
          <fpage>270</fpage>
          -
          <lpage>278</lpage>
          (
          <year>2015</year>
          ). [In Rissian].
        </mixed-citation>
      </ref>
      <ref id="ref31">
        <mixed-citation>
          31.
          <string-name>
            <surname>Shibata</surname>
            ,
            <given-names>T.</given-names>
          </string-name>
          :
          <article-title>Investigation of Language Living within 24 Hours</article-title>
          . In: Linguistics in Japan, pp.
          <fpage>134</fpage>
          -
          <lpage>141</lpage>
          . Moscow (
          <year>1983</year>
          ). [In Rissian].
        </mixed-citation>
      </ref>
      <ref id="ref32">
        <mixed-citation>
          32.
          <string-name>
            <surname>Shiffrin</surname>
            ,
            <given-names>D.</given-names>
          </string-name>
          : Discourse Markers. Cambridge University Press, Cambridge (
          <year>1996</year>
          ).
        </mixed-citation>
      </ref>
      <ref id="ref33">
        <mixed-citation>
          33.
          <string-name>
            <surname>Schourup</surname>
            ,
            <given-names>L.</given-names>
          </string-name>
          :
          <article-title>Discourse Markers</article-title>
          . Lingua,
          <volume>107</volume>
          ,
          <fpage>227</fpage>
          -
          <lpage>265</lpage>
          . Elsevier, The UK (
          <year>1999</year>
          ).
        </mixed-citation>
      </ref>
      <ref id="ref34">
        <mixed-citation>
          34.
          <string-name>
            <surname>Turchanenko</surname>
          </string-name>
          , V. V.: “Tam”
          <article-title>-analysis: Functioning of a Unit in Russian Oral Spontaneous Speech</article-title>
          .
          <source>In: Proceedings of the XX Open Conference for Students-Philologists. St. Petersburg</source>
          ,
          <volume>17</volume>
          -21
          <source>April</source>
          <year>2017</year>
          , pp.
          <fpage>60</fpage>
          -
          <lpage>65</lpage>
          . St.
          <string-name>
            <surname>Petersburg</surname>
          </string-name>
          (
          <year>2018</year>
          ). [In Rissian].
        </mixed-citation>
      </ref>
      <ref id="ref35">
        <mixed-citation>
          35. Université catholique de Louvain official website, MDMA - Model for Discourse Marker Annotation, [Electronic resource], https://uclouvain.be/fr/instituts-recherche/ilc/valibel/ mdma
          <article-title>-model-for-discourse-marker-annotation</article-title>
          .html, last accessed 01/05/
          <year>2020</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref36">
        <mixed-citation>
          36.
          <string-name>
            <surname>Verdonik</surname>
            ,
            <given-names>D.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Rojc</surname>
            ,
            <given-names>M.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Stabej</surname>
            ,
            <given-names>M.</given-names>
          </string-name>
          :
          <article-title>Annotating Discourse Markers in Spontaneous Speech Corpora on an Example for the Slovenian Language</article-title>
          .
          <source>Language Resources and Evaluation</source>
          ,
          <volume>41</volume>
          (
          <issue>2</issue>
          ),
          <fpage>147</fpage>
          -
          <lpage>180</lpage>
          . The Netherlands (
          <year>2007</year>
          ).
        </mixed-citation>
      </ref>
      <ref id="ref37">
        <mixed-citation>
          37.
          <string-name>
            <surname>Zakharov</surname>
            ,
            <given-names>V. P.</given-names>
          </string-name>
          : Corpus Linguistics: Teaching Aid. St.
          <string-name>
            <surname>Petersburg</surname>
          </string-name>
          (
          <year>2005</year>
          ). [In Rissian].
        </mixed-citation>
      </ref>
      <ref id="ref38">
        <mixed-citation>
          38.
          <string-name>
            <surname>Sound</surname>
          </string-name>
          <article-title>Corpus as a Base for Analysis of Russian Speech: Collective Monograph</article-title>
          . Part I. Reading. Retelling. Description.
          <string-name>
            <surname>Bogdanova-Beglarian</surname>
            ,
            <given-names>N. V. (ed.).</given-names>
          </string-name>
          <string-name>
            <surname>St. Petersburg</surname>
          </string-name>
          (
          <year>2013</year>
          ). [In Rissian].
        </mixed-citation>
      </ref>
    </ref-list>
  </back>
</article>