<!DOCTYPE article PUBLIC "-//NLM//DTD JATS (Z39.96) Journal Archiving and Interchange DTD v1.0 20120330//EN" "JATS-archivearticle1.dtd">
<article xmlns:xlink="http://www.w3.org/1999/xlink">
  <front>
    <journal-meta>
      <journal-title-group>
        <journal-title>TPDL</journal-title>
      </journal-title-group>
    </journal-meta>
    <article-meta>
      <title-group>
        <article-title>of the Bilingual Word Indices to the Ninth-Century Uchitel'noe evangelie</article-title>
      </title-group>
      <contrib-group>
        <contrib contrib-type="author">
          <string-name>Martin Ruskov</string-name>
          <email>martin.ruskov@unimi.it</email>
          <xref ref-type="aff" rid="aff1">1</xref>
        </contrib>
        <contrib contrib-type="author">
          <string-name>Lora Taseva</string-name>
          <email>lorataseva@balkanstudies.bg</email>
          <xref ref-type="aff" rid="aff0">0</xref>
        </contrib>
        <contrib contrib-type="editor">
          <string-name>Corpus Linguistics, Computer Assisted Philology, Palaeoslavistics</string-name>
        </contrib>
        <aff id="aff0">
          <label>0</label>
          <institution>Institute of Balkan Studies and Centrе of Thracology, Bulgarian Academy of Sciences</institution>
          ,
          <addr-line>1000 Sofia</addr-line>
          ,
          <country country="BG">Bulgaria</country>
        </aff>
        <aff id="aff1">
          <label>1</label>
          <institution>University of Milan</institution>
          ,
          <addr-line>20100 Milan</addr-line>
          ,
          <country country="IT">Italy</country>
        </aff>
      </contrib-group>
      <pub-date>
        <year>2022</year>
      </pub-date>
      <volume>26</volume>
      <fpage>20</fpage>
      <lpage>23</lpage>
      <abstract>
        <p>The development of bilingual dictionaries to medieval translations presents diverse dificulties. These result from two types of philological circumstances: a) the asymmetry between the source language and the target language; and b) the varying available sources of both the original and translated texts. In particular, the full critical edition of Tihova of Constantine of Preslav's Uchitel'noe evangelie ('Didactic Gospel') gives a relatively good idea of the Old Church Slavonic translation but not of its Greek source text. This is due to the fact that Cramer's edition of the catenae - used as the parallel text in it - is based on several codices whose text does not fully coincide with the Slavonic. This leads to the addition of the newly-discovered parallels from Byzantine manuscripts and John Chrysostom's homilies. Our approach to these issues is a step-wise process with two main goals: a) to facilitate the philological annotation of input data and b) to consider the manifestations of the mentioned challenges, first, separately in order to simplify their resolution, and, then, in their combination. We demonstrate how we model various types of asymmetric translation correlates and the variability resulting from the pluralism of sources. We also demonstrate how all these constructions are being modelled and processed into the final indices. Our approach is designed with generalisation in mind and is intended to be applicable also for other translations from Greek into Old Church Slavonic.</p>
      </abstract>
    </article-meta>
  </front>
  <body>
    <sec id="sec-1">
      <title>1. Introduction</title>
      <p>The creation of the Slavonic alphabet in the second half of the 9th century marked the beginning
of a completely new written tradition in Europe – first only in Glagolitic script, then in Cyrillic.
The main part of this literary production consisted of translated texts from Greek, which were
linked to the needs of the church practice. The studies on the relationship between source
and target texts have occupied a significant place in the philological Slavic Medieval Studies
(Palaeoslavistics) and an important tool for their analysis have been the bilingual dictionaries to
specific scholarly edited writings. These diachronic dictionaries provide the scholars not only
with a generalised view to the vocabulary of the respective work but also with an opportunity
to study various specific issues concerning the Greek-Slavonic translation correlates.</p>
      <p>The very process of developing such dictionaries is complicated not just from the philological
perspective but also from the perspective of their computer support. A major technical challenge
in this respect is the asymmetry between the two languages, i.e., the cases in which there is
no one-to-one correlation of the words of the original and their translations. We present the
approach we have developed for computer-aided creation of such dictionaries, together with the
main problems we have encountered and the decisions we have made to overcome them. This
paper is structured by the following sections: philological basis, technical approach, overview
of specific problems and their solutions, and conclusion.</p>
    </sec>
    <sec id="sec-2">
      <title>2. Philological basis</title>
      <p>The philological background of every bilingual dictionary to a translated text is made up of the
two respective sources – original and target ones. In most cases the medieval translations are
not preserved as authentic documents, so scholars build their understanding of them on the
basis of the preserved (or available) today copies and/or editions. Therefore, the text-critical
tradition of the original text and its translation, as well as the degree of the research on them, are
factors which directly afect both the content and appearance of the end lexicographic product.
Here follows a brief overview of the sources in relation to the bilingual word indices to the Old
Church Slavonic Uchitel’noe evangelie (UE).</p>
      <sec id="sec-2-1">
        <title>2.1. Slavonic tradition</title>
        <p>UE was created by Constantine of Preslav, a prominent Old Bulgarian man of letters, probably
between the years 886 and 893 [1: 3–4, 10; 2: 86]. It consists of 51 sermons commenting on the
respective Sunday Gospel readings for the whole year. The main part of this commentaries is
translated from Greek, but the prologue to the collection and the most of the introductory and
concluding words to each sermon are authored by Constantine of Preslav.</p>
        <p>
          The manuscript tradition of UE is not complicated – 4 full copies came down to us [
          <xref ref-type="bibr" rid="ref3">3</xref>
          ]: one
Russian – MS ГИМ, Син. 262 (henceforth S) dated to the late 11th [4: LXIX] or mid-12th
century [
          <xref ref-type="bibr" rid="ref5">5</xref>
          ] and three Serbian ones – MS РНБ, Гильф. 32 (henceforth G) of the year 1286 [6:
214], MS ÖNB, Cod. Slav. 12 (henceforth W) of the 14th century [7: 126–127] and MS Hil.
385 (henceforth H) of the year 1344 [8: 151–152]. None of them preserves the original text of
Constantine in its authentic form, yet they correct and add to each other.
        </p>
        <p>
          This means that in order to select the most probable original translation correspondences in
the word indices, the material from all preserved manuscripts needs to be critically used. The
diplomatic edition of the oldest witness by Tihova which includes the variant readings after the
other three copies [
          <xref ref-type="bibr" rid="ref9">9</xref>
          ] together with the corrections published in the review by Krys’ko [
          <xref ref-type="bibr" rid="ref10">10</xref>
          ]
give a solid basis for such a lexicographic work.
        </p>
      </sec>
      <sec id="sec-2-2">
        <title>2.2. Greek tradition</title>
        <p>
          No Byzantine collection, identical to UE, is known, but even the early researchers Gorsky and
Nevostruyev assume that Constantine of Preslav used some kind of an abridgement of John
Chrysostom’s Gospel homilies [11: 412, 423–424], and Antonij the archimandrite [12: 37, 40–50]
points out the presence of parallels in Greek Gospel catenae published by Cramer [13]. Using
the same multivolume publication, Tihova adds to her edition the Greek counterparts for the
majority of the translated parts in UE. Her omissions (mainly due to the fact that Cramer’s
Supplementum was unavailable to her) are added by Krys’ko [
          <xref ref-type="bibr" rid="ref10">10</xref>
          ]. Unfortunately, Cramer’s
edition is based on a limited source basis for the Greek catenae with main evidence the Parisian
manuscript Cod. Coislin. Gr. 23 (concerning the manuscripts used by Cramer see the work
of Lamb [14: 281–290]). The project team, therefore, has broadened the range of sources
by adding to it, on the one hand, certain Byzantine catenae manuscripts which are available
in both the database of the Münster Institute for New Testament Textual Research and the
websites of major libraries such as the ones in Paris, Munich and Vatican, and, on the other,
the full text of John Chrysostom’s homilies [15]. Kotova [16] and Petrov [17] found in such
sources more accurate correspondences for individual words and expressions, parallels for
hitherto unidentified fragments, including even some parts of sermons 19, 20 and 42 which were
previously considered original. These discoveries reveal that the Greek texts to be processed
for the sake of the index should include not just the main text in Cramer’s edition and in its
Supplementum but also the additions and the better readings from the manuscript traditions of
the catenae and the full Crhysostomian homilies. What is more, the final index is supposed to
include the information about the source from which a particular reading was taken.
        </p>
      </sec>
      <sec id="sec-2-3">
        <title>2.3. Language and text asymmetry</title>
        <p>The complexity of these dictionaries, imposed by the two textual traditions and the plurality
of sources for each, is combined with the usual dificulties brought about by the asymmetry
between the two languages. It is the reason that the lexical units from the original and its
translation do not always correspond unambiguously, and also that often a given grammatical
meaning is conveyed by lexical means or vice versa. UE, just like other earlier Old Church
Slavonic translations, is freer than the later ones because the literary language is still in process
of creation, which also means that the stable correlates are rare and the possibilities for variations
(both qualitative and quantitative) are much greater. This suggests a significant number of cases
beyond the standard models which need specific solving.</p>
      </sec>
    </sec>
    <sec id="sec-3">
      <title>3. Approach</title>
      <p>
        Automatic alignment between original and translated text is a common task in corpus
linguistics [18, 19] . It is performed on languages that are rich with linguistic resources and most
often it is limited to alignment at the sentence level. Word-level alignment is a dificult task
with limited success rates [19]. When it comes to languages with less resources available,
especially old and regional ones, automation is not possible and the use of technology is limited
to supporting the work of philologists [20, 21, 22] [
        <xref ref-type="bibr" rid="ref10 ref6">6, 10, 11</xref>
        ]. This is especially true for texts
such as UE, where there is a rich variability both in the Greek original (because of the absence
of an exact source text), and its Slavonic translation (because the initial translation is judged by
the extant later copies).
      </p>
      <p>Our approach is based on the scheme for the production of dictionaries to the Church Slavonic
translations of the anti-Latin polemic of Gregory Palamas and Barlaam of Calabria, developed by
von Waldenfels and Taseva [21]. However, due to the greater complexity of the indices to the UE,
the approach has had to be revised and expanded to meet the philological goal of approximating,
as close as possible, between the Greek sources and their translation. The solution presented
here uses two particular open technologies: the Python programming language and the Open
Ofice XML file standard. The result of the process is the generation of two word-indices:
Slavonic-Greek and Greek-Slavonic. Each of these maps each lexeme from one of the texts to
all the word usages corresponding to it in the other language.</p>
      <p>This publication examines in detail the last crucial step of a wider process whose initial stages
are as follows. It begins with a transcription of the source text (in this case Tihova’s edition of
UE) carried out by Rabus with the program Transkribus (for the program see [23]). The text
document is then corrected and enriched by a philologist who manualy introduces positional
variants from the Slavonic transcriptions. Then it is automatically transformed into a spreadsheet
table where each line contains a word from the text and its possible variants. The project team
then enriches this table, adding Greek correspondences (including variants), lemmatising the
word usages (bringing them into standard dictionary form), and adding necessary grammatical
and lexicographical annotations. As a result, the completed table (see Fig. 1) contains the
positional correspondences between the source and its target, annotated with the necessary
philological information. This table is the input solution described here.</p>
      <p>The structure of this table should allow for the representation of variant readings, yet be
easily maintainable for philologists working in teams. The proposed solution for this trade-of
is summarised in Table 1 and detailed here. Since the focal point of this work is the Slavonic
translation, the used Slavonic word form (column F) in the main text, its lemma (columns H)
and their sublemmas (in case where it is necessary to explain the correspondence in the other
language, sequentially in columns I-K) are central. Next to these are their addresses in the text,
consisting of word, page, column, and line (column E), and their contexts, containing the entire
line from which the corresponding word was extracted (column G). The information in column
G is not used by the software, but is useful for the annotation work of philologists. To the
right of the columns for the main Slavic text is the corresponding information for the Greek
main text – word usages (column L) and their adjacent lemmas (columns M-P). Where variant
information is available, it is added in the free columns on the appropriate side. In the columns
to the left of those for the main Slavic text is the information on Slavic variants – word usages
(column A) and lemmas (columns B-D) respectively, and to the right of the main Greek text is
the information on Greek variants – word usages (column Q) and lemmas (columns R-T). When
more than one variant reading is present, the information is combined in the same variant
columns, as illustrated in the last example of Section 4.2.</p>
      <p>In the last step of the wider process, the two types of dictionaries are created: lists and indices
(see Fig. 2). The lists contain all the positional correspondences in the two texts united under a
common lemma. Entries show specific word usages and a precise address. These lists are used
by philologists to verify alignment and lemmatisation, to correct possible inconsistencies and
errors in the collaborative enrichment. The indices present the final word indices in a form ready
for publication. They contain only the summary information about the lexical correspondences,
a list of the exact address of the lemma occurrences and information about their frequency in
the text. For each of the two types of dictionaries, both Slavonic-Greek and Greek-Slavonic
versions are created.</p>
      <p>Two dedicated software tools are used to create the lists and indices using shared logic in
the software implementation, the main elements of which are described in this section. The
lists are generated by a program called integrator. The indices are created by a program called
generator. Each of the two tools performs three steps, executed sequentially for each of the two
dictionaries – Slavonic-Greek and Greek-Slavonic. The first two of these steps are common, the
third is diferent for each of the tools:
1. Adaptation
2. Aggregation
3. Export
Adaptation of the content extraction table. This step takes into account which direction
the dictionary is currently being built for and creates in-memory tables similar to the input
table, but with transformations that allow all the information needed for a single word use to be
contained in a single row. This is relevant in cases of quantitative asymmetry (see next section),
where this condition is not satisfied for the input table. If there are variants in the information
that enters the new row, the step considers separately the columns for the main text and the
ones for the variants. This way the necessary adaptive transformation can be applied for each
of the two cases separately. This transformation is an application of the divide-and-conquer
principle. In this case it is important, firstly, for performance reasons. Secondly, it reduces the
task of constructing the list or index to the simpler subtask of generating entries from a single
row in the adapted table. The simplified subtask is solved in the following step.
Aggregation of alphabetically ordered indices. This step is implemented using nested sorted
mapping1 standard data structures, which allow the construction of specialised alphabetically
sorted reference tables functionally similar to the desired end result. Because this structure is
shared by the integrator and the generator, it must collect the necessary information for both
tools to allow the data needed for each to be retrieved and displayed. To make this possible, a
grouping is created in the sorted mapping structure according to the following hierarchy:
 → ( → (2 → (3 →)))
where the parenthesis indicate optional sublemmas, included only if present in the input table.</p>
      <p>We call Alignment the structure shown in Fig. 3. It consists of the combination of Usage (i.e.
the word usage) in the source and target languages, an Address, and information about whether
the text is a Biblical quotation. The Address allows to locate the word usage up to the line
(containing also page and column) or up to a line span (represented by the self-reference in
address indicating the end of the span in Fig. 3). Biblical quotations are noted in the input tables
in bold and italics, and accordingly appear in the same way in the output dictionaries, an example
being present in Fig. 7. Usage in turn has several important properties. First, these include
1In Python called SortedDict, available as part of https://pypi.org/project/sortedcontainers/
information about the respective word, its lemmas and, if there are repetitions at this address,
a counter specifying the particular repetition of the word in the line, if any. The language of
the Usage is also stored because of the direction-agnostic approach (see last paragraph of this
section), which applies the same logic to the source-to-target and target-to-source dictionaries.
Last but not least, the structure stores information about the source (Source) of the usage, and
possible alternative spellings (Alternative), divided into main and a mapping (var ) of variants to
their alternative spellings. This structure contains the necessary information for the needs of
the dictionaries created in the next step.</p>
      <p>Export of the dictionaries to a word processing document formatted according to their
respective needs – for the integrator and the generator. The result is grouped according to the
hierarchy of the previous step. The lemmas are at its highest level. They are sorted alphabetically.
They list the usages and possible references to other corresponding variants. In the case of
sublemmas at diferent levels, they divide the usages in the corresponding lemma into separate
lines formatted with indentation. Since the integrator ’s goal is to allow for the verification of
the manual preparation of the input table, in addition to the lemmas, the specific word usages
are displayed in the generated lists. However, for the indices these word usages are superfluous
information. Rather than that, the generator adds to the indices a frequency count of occurrences
for each of the languages – both in the main text and in the variants (the latter are indicated with
var in superscript). In Fig. 2 the result from the sample input from Fig. 1 is shown. Note that in
the example, the word in the Greek main text is missing (denoted by “om.”) and, accordingly, in
the Greek dictionaries (the respective list and index), only one lemma appears (παρά). And the
two lemmas in the Slavonic dictionaries (въ and оу) are paralleled with only one Greek word
from an unspecified variant (παρά).</p>
      <p>Our approach applies one more simplifying technique – the Slavonic-Greek indices and the
Greek-Slavonic indices are treated as symmetric. In other words, the same programming logic is
used to generate the indices in both directions. The only diference between the two is that the
table columns of the source and target languages are swapped when provided to the program
for the Slavonic-Greek index and for the Greek-Slavonic index, respectively. For shortness, in
the next section we will only show the two-way result in one of the dictionaries – the one that
better illustrates the diferences.</p>
    </sec>
    <sec id="sec-4">
      <title>4. Solutions</title>
      <p>In order to illustrate in detail how our approach works, here we look at specific problems and
our corresponding solutions. We address two categories of problems: 1) quantitative asymmetry,
which we illustrate with the generated lists, and 2) variability in transcripts, better illustrated
by the resulting indices.</p>
      <sec id="sec-4-1">
        <title>4.1. Quantitative asymmetry</title>
        <p>As mentioned in the previous section, with our approach and the program implementing it, the
quantitative asymmetry is solved at the adaptation step. Several examples that demonstrate
diferent cases of asymmetry follow.
One-to-many (1:n) is the base example that occurs when a Greek word is translated with a
Slavonic phrase. To indicate such occurrences, we use background colouring of the
corresponding rows in the word-use column (F), as shown in Fig. 4. In cases where two groups immediately
follow each other, they are delineated by using an empty line between the two.</p>
        <p>What the adaptation step does is to collect translation information from the whole phrase
and associate it with each of the lexemes by adding it in its corresponding row in the table.
In the example below, according to the currently generated dictionary, the target words (for
Slavonic in column F) remain as they are, but the lemmas (H) are given as a phrase – together –
and repeated for each of resulting entries.</p>
      </sec>
      <sec id="sec-4-2">
        <title>One-to-many, including grammatical value (1:n*) is a more complicated case where one</title>
        <p>of the words from the phrase has only grammatical meaning. The example in Fig. 5 uses an
analytic verb form for a past participle.</p>
        <p>To indicate such occurances we use “gramm.” in the second lemma (column I) with a coloured
background. The particularity in this case is that the grammatical value is valid only for the
lexeme on the same line, but not for the rest of the phrase. In the dictionaries, therefore, this
information should be shown only for the corresponding lemma, but not for the others in the
group.</p>
        <p>In this case, the adaptation step does not add the sublemmas of the lexeme with grammatical
value to the information about the other lexemes. Thus, the corresponding feature is present
only in the translation from Slavonic to Greek and leads to an asymmetry between the two
dictionaries in Fig. 6 (see sublemma “gramm.”). This is the reason why diferent adapted tables
are needed for the aggregation step from Slavonic to Greek and from Greek to Slavonic.
Many-to-many (n:m) asymmetry is an occurrence when a Greek phrase is translated with
a Slavonic phrase but there is no direct corresponding between the separate words of these
phrases in the two languages as shown in Fig. 6. Our approach solves this case using the same
logic. In the adaptation step, the program collects in a single row all the information needed for
an entry in the dictionaries. As a consequence, the aggregation step, does not need to perform
any processing other than just reading the data row by row and adding it to the sorted mapping
data structure.</p>
      </sec>
      <sec id="sec-4-3">
        <title>4.2. Variant readings in manuscript copies</title>
        <p>As a rule, the indices and lists need to include all word uses from the main text and those
variants from the other copies that represent either better readings than the main text or equally
possible readings in the given context. Missing, wrong or less precise variants from the Slavonic
transcripts W, G and H are ignored, and not lemmatised, so the program does not include them
in the dictionaries.</p>
        <p>Diferent translation correspondences (seen in Fig. 1 and Fig. 2) is a scenario where a
word other than the main one is used in the additional sources. In such circumstances, the
dictionaries show the usage in both lexemes used (in the main version and in the variant),
adding references to the variant lexemes to the addresses of the usages. This functionality is
the reason why – in the aggregation step – the structure describing the usages also contains
information about the alternatives.</p>
        <p>Diferent copies suggest diferent correspondences is a fundamentally similar but more
complicated occurrence. In the example in Fig. 7, the rows indicate four diferent combinations
delineating three diferent translations ( но ѧдъ, д нородъ and д но ѧдъ ) of the Greek word
μονογενής, as well as the possibility of a missing translation. Similar to the previous example,
this diversity should be traceable in the generated lists and indices through the corresponding
references. Variant readings in copies are relatively rare, and it is even rarer to have two diferent
variant readings to the same lexeme. The solution should therefore be able to handle such cases
without complicating the manual enrichment and annotation performed by philologists. The
solution adopted here combines the diferent variant readings in column A and their respective
lemmata in column B for the Slavonic variant. Such cases are interpreted in the aggregation
step, where the combinations of variants are recorded as separate Alignment data objects.</p>
      </sec>
    </sec>
    <sec id="sec-5">
      <title>5. Conclusion</title>
      <p>This paper presents our work on creating dictionaries for UE. In it, we show the combined
solution to two particular methodological challenges: quantitative asymmetry and variation in
the sources. We solve these problems applying a stepwise approach of adaptation, aggregation
and export. We demonstrate a divide-and-conquer approach (implemented in the adaptation step)
that allows solving even complex cases of many-to-many quantitative asymmetry with diferent
grammatical values. We also propose alignment modelling for the purpose of generating
dictionaries with cross-references for text variants. As a result, in the dictionaries it is possible
to indicate not only the alignment between original and translation, but also between these two
and a large number of variants.</p>
      <p>The approach has been applied to the 52 constituent texts of the UE (the introduction and
the 51 sermons). More complex combinations of asymmetry and variation in the sources were
also successfully modelled for this purpose. Such cases represent combinations of the examples
provided in this paper. The program was written with applicability to other Greek-Slavonic
translations in mind. In order to allow for this applicability, and for future extensions, we
commit to publishing the software code under a free license.</p>
    </sec>
    <sec id="sec-6">
      <title>Acknowledgments</title>
      <p>The article is written within the project The Vocabulary of Constantine of Preslav’s Uchitel’noe
evangelie (’Didactic Gospel’): Old Bulgarian-Greek and Greek-Old Bulgarian Word Indices, financed
by the Bulgarian National Science Fund (contract КП-06-Н50/2 of 30.11.2020).
didactic gospel of constantine of bulgaria: editio princeps.], Bulletin of the RAS: Studies
in Literature and Language 76 (2017) 52–62.
[11] A. V. Gorsky, K. I. Nevostruyev, Описание славянских рукописей Московской
синодальной библиотеки. Отдел 2. Писания святых отцев. 2. Писания догматическия
и духовно- нравственныя [Description of the Slavic manuscripts of the Moscow
Synodal Library. Section 2. Scriptures of the Holy Fathers. 2. Dogmatic and Spiritual-Moral
Scriptures], Synodal Printing House, Moscow, 1859.
[12] archim. Antonij (Vadkovskij), Из истории древнеболгарской церковной проповеди.
Константин, епископ болгарский и его Учительное евангелие [From the history of
Old Bulgarian church preaching. Konstantin, Bishop of Bulgaria and his Didactic Gospel],
Typogr. Imp. Univ., Kazan, 1885.
[13] J. A. Cramer, Catenae graecorum patrum in novum testamentum, number v. 4 in Catenae
graecorum patrum in novum testamentum, E Typographeo Academico, 1844.
[14] W. Lamb, Conservation and conversation: New testament catenae in byzantium, in:
D. Krueger, R. S. Nelson (Eds.), The New Testament in Byzantium, Dumbarton Oaks
Research Library and Collection, Washington, D.C., 2016, pp. 277–299.
[15] J.-P. Migne (Ed.), Patrologiae cursus completus: Series Graeca, T. 58. Joannes Chrysostomus,</p>
      <p>Paris, 1862.
[16] D. Kotova, “Слово 19 от Учителното евангелие и неговите гръцки източници”
[“Sermon 19 in Constantine of Preslav’s Didactic Gospel and Its Greek Sources”],
Palaeobulgarica 46 (2022) 3–28.
[17] I. P. Petrov, “The Greek Sources of Učitel’noe Evangelie Revisited: Sermon 20”,
Palaeobulgarica 46(2) (2022) 3–28.
[18] Y.-C. Chiao, O. Kraif, D. Laurent, T. M. H. Nguyen, N. Semmar, F. Stuck, J. Véronis, W.
Zaghouani, Evaluation of multilingual text alignment systems: the ARCADE II project, in:
Proc. of the 5th Language Resources and Evaluation Conference, ELRA, Genoa, 2006.
[19] F. J. Och, H. Ney, “A Systematic Comparison of Various Statistical Alignment Models”,</p>
      <p>Computational Linguistics 29 (2003) 19–51.
[20] T. Pataridze, B. Kindt, “Text Alignment in Ancient Greek and Georgian: A Case-Study on
the First Homily of Gregory of Nazianzus”, Journal of Data Mining &amp; Digital Humanities
Special Issue on Computer-Aided Processing of Intertextuality in Ancient Languages
(2018).
[21] R. von Waldenfels, L. Taseva, An der Schnittstelle von Korpuslinguistik und Paläoslavistik:
Wörterverzeichnisse zu einer mittelalterlichen Handschrift als Keimzelle eines anotierten
digitalen Korpus, in: B. Hansen (Ed.), Diachrone Aspekte slavischer Sprachen: für Ernst
Hansack zum 65. Geburtstag, Otto Sagner, 2012, pp. 243–258.
[22] A. Yli-Jyrä, J. Purhonen, M. Liljeqvist, A. Antturi, P. Nieminen, K. M. Räntilä, V.
Luoto, HELFI: a Hebrew-Greek-Finnish parallel Bible corpus with cross-lingual morpheme
alignment, in: Proc. of the 12th Language Resources and Evaluation Conference, ELRA,
Marseille, 2020, pp. 4229–4236.
[23] A. Rabus, “Recognizing Handwritten Text in Slavic Manuscripts: a Neural-Network
Approach Using Transkribus”, Scripta &amp; e-Scripta 19 (2019) 9–32.</p>
    </sec>
  </body>
  <back>
    <ref-list>
      <ref id="ref1">
        <mixed-citation>
          [1]
          <string-name>
            <given-names>E.</given-names>
            <surname>Gallucci</surname>
          </string-name>
          , “
          <article-title>Учительное евангелие Константина Преславского (ІХ-Х в.) и последовательность воскресных евангельских чтений церковного года” [“The Lectionary of Constantine of Preslav (9th-10th с.) and the Order of the Sunday Gospel Readings of the Church Year”]</article-title>
          ,
          <source>Palaeobulgarica</source>
          <volume>25</volume>
          (
          <year>2001</year>
          )
          <fpage>3</fpage>
          -
          <lpage>20</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref2">
        <mixed-citation>
          [2]
          <string-name>
            <given-names>M.</given-names>
            <surname>Spasova</surname>
          </string-name>
          , “
          <article-title>На коя дата и през кой месец се е провел Преславският събор от 893 година” [“On what date and in what month was the Council of Preslav of 893 held”], Preslavska knizhovna shkola 8 (</article-title>
          <year>2005</year>
          )
          <fpage>84</fpage>
          -
          <lpage>101</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref3">
        <mixed-citation>
          [3]
          <string-name>
            <given-names>E.</given-names>
            <surname>Ukhanova</surname>
          </string-name>
          ,
          <article-title>Палеографические и кодикологические особенности древнейшего списка Учительного Евангелия Константина Преславского (ГИМ, Син. 262) и история его создания [Palaeographic and codicological features of the oldest copy of the Didactic Gospel of Constantine of Preslav (GIM, Sin. 262) and the history of its creation]</article-title>
          , in: M.
          <string-name>
            <surname>Tihova</surname>
          </string-name>
          (Ed.),
          <article-title>Старобългарското Учително евангелие на Константин Преславски [The Didactic Gospel of Constantin of Preslav], Weiher, Freiburg i</article-title>
          . Br.,
          <year>2012</year>
          , p.
          <article-title>LVI-LXXVII.</article-title>
        </mixed-citation>
      </ref>
      <ref id="ref4">
        <mixed-citation>
          [4]
          <string-name>
            <given-names>M.</given-names>
            <surname>Tihova</surname>
          </string-name>
          , “
          <article-title>Учителното евангелие на Константин Преславски и неговите преписи” [“The Didactic gospel of Constantin of Preslav and its copies”], Preslavska knizhovna shkola 7 (</article-title>
          <year>2003</year>
          )
          <fpage>127</fpage>
          -
          <lpage>135</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref5">
        <mixed-citation>
          [5]
          <string-name>
            <given-names>V. B.</given-names>
            <surname>Krysko</surname>
          </string-name>
          ,
          <string-name>
            <given-names>G. A.</given-names>
            <surname>Molkov</surname>
          </string-name>
          , “
          <article-title>Языковые особенности Учительного евангелия Константина Преславского и его древнейшего списка” [“Linguistic Features of the Teaching Gospel of Constantin of Preslav and Its oldest copy”]</article-title>
          ,
          <source>Zeitschrift für Slavische Philologie</source>
          <volume>73</volume>
          (
          <year>2017</year>
          )
          <fpage>331</fpage>
          -
          <lpage>395</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref6">
        <mixed-citation>
          [6]
          <string-name>
            <given-names>S. O.</given-names>
            <surname>Shmidt</surname>
          </string-name>
          , et al. (Eds.),
          <article-title>Сводный каталог славяно-русских рукописьных книг, хранящихся в СССР. XI-XIII вв</article-title>
          . [
          <article-title>General catalog of Slavonic-Russian handwritten books stored in the USSR</article-title>
          . 11th-13th c.], Nauka, Moscow,
          <year>1984</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref7">
        <mixed-citation>
          [7]
          <string-name>
            <given-names>G.</given-names>
            <surname>Birkfellner</surname>
          </string-name>
          ,
          <article-title>Glagolitische und kyrillische Handschriften in Österreich, Österreichische Akademie der Wissenschaften. Philosophisch-Historische Klasse</article-title>
          . Schriften der Balkankommission. Linguistische
          <string-name>
            <surname>Abteilung</surname>
            <given-names>XXIII</given-names>
          </string-name>
          , Verl. d. ÖAW, Wien,
          <year>1975</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref8">
        <mixed-citation>
          [8]
          <string-name>
            <given-names>D.</given-names>
            <surname>Bogdanović</surname>
          </string-name>
          ,
          <article-title>Katalog ćirilskih rukopisa manastira Hilandara, SANU</article-title>
          , Beograd,
          <year>1978</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref9">
        <mixed-citation>
          [9]
          <string-name>
            <given-names>M.</given-names>
            <surname>Tihova</surname>
          </string-name>
          (Ed.),
          <article-title>Старобългарското Учително евангелие на Константин Преславски [The Didactic Gospel of Constantin of Preslav], number 58 in Monumenta linguae slavicae dialecti veteris: fontes et dissertationes, Weiher, Freiburg i</article-title>
          . Br.,
          <year>2012</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref10">
        <mixed-citation>
          [10]
          <string-name>
            <given-names>V. B.</given-names>
            <surname>Krysko</surname>
          </string-name>
          ,
          <article-title>Учительное евангелие Константина Болгарского: editio princeps</article-title>
          . [the
        </mixed-citation>
      </ref>
    </ref-list>
  </back>
</article>