<!DOCTYPE article PUBLIC "-//NLM//DTD JATS (Z39.96) Journal Archiving and Interchange DTD v1.0 20120330//EN" "JATS-archivearticle1.dtd">
<article xmlns:xlink="http://www.w3.org/1999/xlink">
  <front>
    <journal-meta />
    <article-meta>
      <title-group>
        <article-title>Linking the Lewis &amp; Short Dictionary to the LiLa Knowledge Base of Interoperable Linguistic Resources for Latin</article-title>
      </title-group>
      <contrib-group>
        <aff id="aff0">
          <label>0</label>
          <institution>Francesco Mambrini, Eleonora Litta, Marco Passarotti, Paolo Ruffolo CIRCSE Research Centre Universita` Cattolica del Sacro Cuore Largo Gemelli</institution>
          ,
          <addr-line>1 - 20123 Milan</addr-line>
          ,
          <country country="IT">Italy</country>
        </aff>
      </contrib-group>
      <abstract>
        <p>This paper describes the steps taken to include data from the Lewis &amp; Short bilingual Latin-English dictionary into the Knowledge Base of linguistic resources for Latin LiLa. First, data were extracted from the original XML and matched with entries in LiLa, overcoming ambiguities and structural inconsistencies in the source. Subsequently, senses were modelled using the Ontolex Lemon Lexicographic module (lexicog), so that they could be included in the LiLa Knowledge Base and thus made interoperable with the (meta)data of the linguistic resources for Latin therein interlinked.</p>
      </abstract>
    </article-meta>
  </front>
  <body>
    <sec id="sec-1">
      <title>1 Introduction</title>
      <p>
        Since the pioneering times of 1949, when the
Jesuit Roberto Busa persuaded Thomas Watson Sr.,
CEO of IBM, to fund his project aimed at
processing the Latin texts of Thomas Aquinas with
computers
        <xref ref-type="bibr" rid="ref16">(Jones, 2016)</xref>
        , scholars in the areas
of Computational Linguistics, Literary
Computing and Digital Humanities have built a plethora
of linguistic resources for both modern and
historical languages.
      </p>
      <p>Particularly over the last two decades, many and
diverse linguistic resources have been made
available for Latin. These consist in corpora of texts
spanning different eras and genres1, dependency</p>
      <p>Copyright © 2021 for this paper by its authors. Use
permitted under Creative Commons License Attribution 4.0
International (CC BY 4.0).</p>
      <p>
        1See, for example, Musisque deoque for Classical Latin
poetry
        <xref ref-type="bibr" rid="ref24">(Manca et al., 2011)</xref>
        , CLaSSES, containing epigraphic
material
        <xref ref-type="bibr" rid="ref8">(De Felice et al., 2015)</xref>
        , the large corpus of Classical
Latin prose and poetic texts by LASLA
        <xref ref-type="bibr" rid="ref9">(Denooz, 2007)</xref>
        and
CroALa, which brings together writings by Croatian authors
produced between the 10th and 20th centuries
        <xref ref-type="bibr" rid="ref17">(Jovanovic´,
2012)</xref>
        .
treebanks2 and lexica3. These digital resources
join the large set of textual and lexical resources
that were created over the centuries for Latin:
textual collections, thesauri, lexica, glossaries and
mono/bilingual dictionaries. Among the latter,
we could mention, for instance, the Oxford Latin
Dictionary
        <xref ref-type="bibr" rid="ref14">(Glare, 1968)</xref>
        , the Dictionary of
medieval Latin from British sources
        <xref ref-type="bibr" rid="ref1">(Ashdowne et al.,
1975)</xref>
        , the Forcellini lexicon
        <xref ref-type="bibr" rid="ref12">(Forcellini and
Facciolati, 1871)</xref>
        and the still under construction
Thesaurus Linguae Latinae
        <xref ref-type="bibr" rid="ref11">(Ehlers, 1968)</xref>
        , many of
which are today accessible also in digital format.
      </p>
      <p>
        However, the impact of these digital resources
on the everyday work of classicists is still limited.
On the one side, this is due to the still existing
divisive dichotomy between “traditional”
Humanities and computational approaches. On the other,
it is a matter of fact that classicists are not yet
put in the best condition to fully exploit all
available resources for ancient languages, as these are
currently scattered across the web in
uncommunicative blocks, using different query languages,
data formats, annotation criteria and tagsets. The
last decade has seen a number of exploratory
solutions to tackle the sparseness of linguistic
resources. Among them, the European
infrastructure CLARIN4 represents a common hub where
data and metadata of resources collected in
single repositories (at national level) can be searched
(through the so-called Virtual Language
Observatory) and processed with different tools (through
the CLARIN Language Resource Switchboard).
As for Classical languages, Logeion5 is a
meta2Index Thomisticus Treebank
        <xref ref-type="bibr" rid="ref31">(Passarotti, 2019)</xref>
        , Late
Latin Charter Treebank
        <xref ref-type="bibr" rid="ref5 ref6">(Cecchini et al., 2020a)</xref>
        , UDante
        <xref ref-type="bibr" rid="ref5 ref6">(Cecchini et al., 2020b)</xref>
        , PROIEL
        <xref ref-type="bibr" rid="ref10">(Eckhoff et al., 2018)</xref>
        and
Latin Dependency Treebank
        <xref ref-type="bibr" rid="ref2">(Bamman and Crane, 2011)</xref>
        .
      </p>
      <p>
        3Such as, for instance, valency and subcategorisation
lexica
        <xref ref-type="bibr" rid="ref26 ref28 ref8">(Passarotti et al., 2016; McGillivray and Vatri, 2015)</xref>
        ,
the Latin WordNet
        <xref ref-type="bibr" rid="ref27">(Minozzi, 2017)</xref>
        and word lists
        <xref ref-type="bibr" rid="ref33 ref36">(Tombeur,
1998; Ramminger, 2008)</xref>
        .
      </p>
      <p>4https://www.clarin.eu.
5https://logeion.uchicago.edu/lexidium.
dictionary that allows to query together the lexical
entries of several dictionaries for Ancient Greek
and Latin, while Corpus Corporum6 is a
metacollection that allows searches across more than
twenty different corpora for Latin. However, what
such initiatives still lack is to provide a real
interoperability between distributed resources, which
would result in interaction at both syntactic
(structural) and semantic (conceptual) level.</p>
      <p>
        Syntactic interoperability is defined as ‘the
ability of different systems to process (read)
exchanged data either directly or via trivial
conversion’, using a common data model consisting of
shared protocols and data formats. Semantic
interoperability, on the other hand, is ‘the ability
to automatically interpret exchanged information
meaningfully and accurately in order to produce
useful results’, by using a set of common linguistic
data categories defined in ad-hoc ontologies
        <xref ref-type="bibr" rid="ref15">(Ide
and Pustejovsky, 2010)</xref>
        .
      </p>
      <p>
        Attaining syntactic and semantic
interoperability between distributed linguistic resources is the
objective of the Linguistic Linked Open Data
(LLOD) community, which applies the
principles of the Linked Data paradigm
        <xref ref-type="bibr" rid="ref3">(Bizer et al.,
2008)</xref>
        to the (meta)data contained in linguistic
resources. As for Classical languages, the LiLa
Knowledge Base (KB)7
        <xref ref-type="bibr" rid="ref21 ref22 ref30 ref5">(Passarotti et al., 2020)</xref>
        makes textual and lexical resources for Latin
interact through a commonly used data model, called
the Resource Description Framework (RDF)
        <xref ref-type="bibr" rid="ref18">(Lassila et al., 1998)</xref>
        , and ontologies developed and
shared by the LLOD community. In this way, the
linked resources become interoperable with each
other as well as with those for other languages
described following the same structural and
conceptual principles.
      </p>
      <p>Based on a large collection of “canonical
forms” (lemmas) - the so-called “Lemma Bank”,
LiLa achieves interoperability between resources
by linking all those entries in lexical resources and
tokens in corpora that point to the same lemma in
the LiLa collection.</p>
      <p>
        The lexical resources for Latin linked so far
to LiLa include a word formation lexicon
        <xref ref-type="bibr" rid="ref32">(Pellegrini et al., 2021)</xref>
        , a polarity lexicon
        <xref ref-type="bibr" rid="ref35 ref6">(Sprugnoli et
al., 2020)</xref>
        , an etymological dictionary
        <xref ref-type="bibr" rid="ref21 ref22 ref30 ref35 ref5">(Mambrini
and Passarotti, 2020)</xref>
        and a joint resource
providing a manually checked subset of the Latin
Word6http://www.mlat.uzh.ch/MLS/.
7https://lila-erc.eu.
      </p>
      <p>
        Net and a valency lexicon
        <xref ref-type="bibr" rid="ref23">(Mambrini et al., 2021)</xref>
        .
The most recent among the LiLa connections is
the bilingual Latin-English dictionary by Charlton
Lewis and Charles Short (1879). The inclusion of
this type of lexicon in LiLa was much needed, as
no resource providing semantic information
consisting of translations and definitions was
available in the network of connected resources before.
Since Lewis &amp; Short is the first lexical resource of
its kind included in LiLa, the process of its
linking to the KB opened a number of LLOD-related
challenges.
      </p>
      <p>This paper describes how such challenges have
been tackled and is organised as follows: Section 2
describes the Lewis &amp; Short dictionary in its main
characteristics. Section 3 discusses the ontologies
involved in the modelling phase, the challenges
that need to be overcome in the representation of
the linguistic data as LLOD (3.1), and the
strategies adopted to represent the dictionary entries
using the chosen vocabularies (3.2). Finally, Section
4 discusses conclusions and highlights directions
for future work.
2
2.1</p>
    </sec>
    <sec id="sec-2">
      <title>The “Lewis &amp; Short” Dictionary</title>
      <sec id="sec-2-1">
        <title>The Printed and Digital Dictionary</title>
        <p>
          The Latin Dictionary, curated by Ch. T. Lewis
and Ch. Short and commonly referred to as the
“Lewis &amp; Short” (L&amp;S), was published by Harper
and Oxford University Press in 1879
          <xref ref-type="bibr" rid="ref19">(Lewis and
Short, 1879)</xref>
          . Though based on previous work by
German scholars, it remained a standard in Latin
lexicography in the English-speaking world until
it was superseded by the Oxford Latin Dictionary
          <xref ref-type="bibr" rid="ref14">(Glare, 1968)</xref>
          .
        </p>
        <p>
          In the digital age, its importance rests on two
grounds. On the one hand, its relevance for the
history of Classical Scholarship is undeniable. On
the other hand, also on account of its copyright
status, as the dictionary belongs now to the
public domain, the L&amp;S has quickly become one of
the most used and best curated digital Latin
dictionaries on the web. Following the same
worklfow used for the Greek-English Lexicon
          <xref ref-type="bibr" rid="ref20">(Liddell
et al., 1940)</xref>
          , the Perseus Project has developed a
widely used digital edition of the dictionary based
on the standards of the Text Encoding Initiative
(TEI)
          <xref ref-type="bibr" rid="ref34">(Rydberg-Cox, 2002)</xref>
          . The digital L&amp;S has
been incorporated in the word-search tools
available on the Perseus website and in a series of other
desktop and web applications.8
        </p>
        <p>Perseus’ TEI edition is the point of departure of
our work.9 Though its publication was a
remarkable achievement, this electronic text is not exempt
from occasional flaws and inconsistencies, which
had to be taken into account.</p>
        <p>In the digital edition, entries from the L&amp;S are
based on an XML encoding of the whole
dictionary. The XML structure, albeit not always
consistent, offers the following information about
each word:
1. Entry: the headword. Entries are encoded
within the TEI element &lt;entryFree&gt; and
are 51,596 in total.10
2. Information about inflection, encoded as
attributes in the XML and visualised in the
output reproducing the customary descriptions
for Latin dictionaries, e.g. a masculine noun
of the second declension (e.g. gallus ‘cock’)
is followed by the genitive singular ending of
the word (‘i’), and the abbreviation for
gender ‘m.’ (e.g. gallus, i, m.).
3. Etymological or derivational information,
encoded within the same element &lt;etym&gt;.
4. Sense(s): these act as containers where the
meaning of the word is matched with a
number of representative citations from Classical
Latin sources. Each citation is accompanied
by its canonical reference (e.g. “Cic. Sen. 8,
26” for a reference to Cicero, De Senectute,
chapter 8, paragraph 26).</p>
        <p>Entries can contain what we call “sub-entries”,
words that are not given a record of their own, but
are discussed within another entry. Usually, these
sub-entries consist of lexicalised present and past
participles like, for example, adolescens ‘young
man’ – sub-entry of adolesco ‘to grow up’;
another instance is the substantivised forms of
adjectives, such as verum ‘the truth’ – sub-entry of
verus ‘true’. Sub-entries are encoded within the
&lt;sense&gt; element and followed by the same type
of inflectional information structured as the main
entries.</p>
        <p>8One example is the app Diogenes for querying corpora
of Greek and Latin texts: https://d.iogen.es/.</p>
        <p>9The digital edition is available from the repository of the
Perseus DL and is distributed under a CC BY SA 4.0 license:
https://github.com/PerseusDL/lexica.</p>
        <p>10See
https://tei-c.org/release/doc/teip5-doc/en/html/ref-entryFree.html.
2.2</p>
      </sec>
      <sec id="sec-2-2">
        <title>Linking the L&amp;S to LiLa</title>
        <p>The LiLa KB includes about 200,000 canonical
forms, each of which is described by a series of
properties that record the part of speech (PoS),
the full morphological description and the
inflectional category. Also, the data property “written
representation”, defined in the ontology Ontolex
(see Section 3.1), registers all the attested spellings
of any lemma. Publishing a lexical resource as
LLOD within LiLa means to both represent its
information using the appropriate standards and
vocabularies (Section 3.1) and to link the dictionary
entries to the right form in LiLa by matching the
lemmas used to index the records to the
appropriate form in the KB.</p>
        <p>
          In order to achieve the latter goal, firstly we
had to normalise the spelling of the L&amp;S
dictionary lemmas by removing upper case initials and
substituting j with i and v with u in order to
mirror LiLa’s conventions. Then, after mapping
partof-speech and inflectional information between
resources, we extracted 31,142 1:1 matches, 2,998
1:N matches and 4,553 1:0 matches, on the basis
of the tuple written representation - PoS. The
latter group was subsequently matched only on the
basis of graphical representation, at which point
we obtained 946 1:1 matches and 50 1:N matches.
Of the remaining 3,557 unmatched entries, 1,289
were successfully analysed by the morphological
analyser Lemlat
          <xref ref-type="bibr" rid="ref29">(Passarotti et al., 2017)</xref>
          , leaving
2,239 definitely unmatched entries. After
resolving multi-word spellings and graphical variants,
the unmatched entries were all added to the LiLa
Lemma Bank, while 1:N matches were manually
disambiguated and matched to the relevant
lemmas.
3
3.1
        </p>
      </sec>
    </sec>
    <sec id="sec-3">
      <title>Modelling Lexical Entries</title>
      <sec id="sec-3-1">
        <title>LiLa, Ontolex and lexicog</title>
        <p>
          As said, the LiLa KB for Latin resources is built
around a collection of canonical forms that can be
used both as head words of dictionaries or as
“targets” for the lemmatisation of corpora
          <xref ref-type="bibr" rid="ref21 ref22 ref30 ref5">(Passarotti
et al., 2020)</xref>
          . These lemmas are modelled using
the Ontolex ontology, a now de facto standard of
the LLOD community
          <xref ref-type="bibr" rid="ref25 ref7">(Cimiano et al., 2020;
McCrae et al., 2017)</xref>
          . In particular, lemmas in the
LiLa KB are defined as forms of words that are
linked (or are ready to be linked) to lexical entries
via the property “canonical form” of the Ontolex
ontology.11
        </p>
        <p>Ontolex provides several classes and properties
to describe the relationships that lexical entries
have with, on the one hand, the grammatical forms
attested in language and, on the other, the senses
and the meanings of words. The core Ontolex
module, however, imposes a series of restrictions
that make its classes and properties ill-suited to
represent the information in most standard
dictionaries. The class Lexical Entry from the core
Ontolex module, for instance, is inadequate to
represent entries that license multiple syntactic
interpretations, such as words that are registered in a
dictionary as both adverb and conjunction.
Subentries like the noun verum from the adjective verus,
formed by a process of substantivisation from the
word in the main entry, would also produce a
mismatch between the dictionary and the lexical entry.
Finally, the L&amp;S, as most dictionaries, defines the
senses of all but the most simple words by
grouping them in sense clusters; those clusters are
generally organized into hierarchies with multiple
levels of nesting, from the most general to the most
specific sense, a structure for which Ontolex has
no suitable representation.</p>
        <p>
          In order to overcome these issues, the Ontolex
community has developed a specific extension of
the ontology called the “OntoLex lexicography
module” or lexicog
          <xref ref-type="bibr" rid="ref4">(Bosque-Gil and Gracia,
2019)</xref>
          .12 The module is explicitly designed to
capture the structural information expressed in a
lexicographic resource and is primarily intended to
support the conversion of lexicographic data that
are not native to Ontolex. Retro-digitised
dictionaries like the L&amp;S are thus a perfect use case.
        </p>
        <p>
          As said, lexicog focuses on the structural
properties of dictionaries and does not attempt to
convey any lexical, or indeed linguistic
information, which are left to the classes and properties
of Ontolex. The most important of these structural
elements introduced in the vocabulary is that of the
Lexicographic Entry. In lexicog, an entry is a
container that represents a lexicographic article or
record as it is arranged in the source
          <xref ref-type="bibr" rid="ref4">(Bosque-Gil
and Gracia, 2019)</xref>
          . Thus, while a lexical entry (as
defined in Ontolex) is an item in the lexicon of a
given language, a lexicographic entry is a record in
a linguistic resource that documents or discusses
some properties of a given lexical item.
        </p>
        <p>11http://www.w3.org/ns/lemon/ontolex#c
anonicalForm.</p>
        <p>12https://www.w3.org/ns/lemon/lexicog#.</p>
        <p>Lexicographic entries are a special subset of
a larger class called Lexicographic Component.
Apart from whole dictionary articles (the
entries), components can be used to represent senses,
sense groups or subentries (like the substantivised
verum) within lexicographic entries.</p>
        <p>It is important to stress once again that
components represent only structural units; all
linguistic information that is conveyed within these units
must be expressed using Ontolex. The property
lexicog:describes provides a link between
the two dimensions, so that a lexicographic entry
can be said to describe a lexical entry (as defined
in Ontolex). In the same way, the lexicographic
components that discuss a sense of a word or
introduce a subentry, describe that specific lexical
sense (as defined in Ontolex) or another lexical
entry.
3.2</p>
      </sec>
      <sec id="sec-3-2">
        <title>Lexicographic and Lexical Entries in the L&amp;S</title>
        <p>The LLOD version of the L&amp;S linked to LiLa is
now available online in the LiLa KB.13 The entries
can also be searched using LiLa’s query interface
and SPARQL endpoint.14</p>
        <p>Figure 1 shows a visualisation of how the
information from a sample entry, the adjective hosticus
in the L&amp;S dictionary, is represented in LiLa. In
particular, the interplay between the linguistic and
structural information is reflected in the complex
relation between the lexical and lexicographic
entries.</p>
        <p>The L&amp;S distinguishes two senses for the word:
“belonging to an enemy, hostile” and “belonging
to a stranger, foreign”. Following the Ontolex
approach, these meanings are represented by the
two ‘triangles’ between the lexical entry (the light
green node on the left), the concepts evoked by the
word (gray-blue nodes), and the senses, labeled 0
ad 1, that mediate between them (greenish-yellow
nodes).</p>
        <p>The lexical entry is described by a lexicographic
entry, identified by the id n21014 (inherited from
the TEI XML file of the Perseus DL), while a
specific lexicographic component describes each of
the two senses (n21014 0 and n21014 1,
respectively). What is particularly relevant is that the
component n21014 0, which corresponds to the
13http://lila-erc.eu/data/lexicalResour
ces/LewisShort/Lexicon.</p>
        <p>14https://lila-erc.eu/query/, and https:
//lila-erc.eu/sparql/.
sense “hostile”, is linked to a sub-component that
describes the lexical entry of the noun hosticum,
a substantivised usage of the neuter adjective that
means “the enemy’s territory”. That section of
the entry that discusses the subentry “hosticum”,
which is itself a section of the paragraph dedicated
to the first sense, is thus linked (via the “describes”
property) to a different lexical entry.
4</p>
      </sec>
    </sec>
    <sec id="sec-4">
      <title>Conclusions and Future Work</title>
      <p>Perhaps even more than for any other modern
language, a great number of lexical resources, either
bi- or monolingual, is available for Latin, many
of which have already been digitised and
disseminated on the web. In this paper, we described
a model of how this huge wealth of information
can be published using the modern standards of
the Semantic Web. The greatest advantage of this
approach is that all the lexical resources published
according to the same data model can be integrated
in a wider network of linguistic information, along
with the other digital resources that are connected
to it. In the case of the L&amp;S in LiLa, the Latin
lexical entries of the bilingual dictionary can be
queried together with the information about the
same words provided by the other linguistic
resources linked to the lemmas in the KB.</p>
      <p>
        One example of the fruitful interactions
between resources is the possibility to investigate
the polysemy of words in relation to their
derivation, as recorded in the Word Formation Latin
resource, which is also linked to LiLa
        <xref ref-type="bibr" rid="ref21">(Litta et al.,
2020)</xref>
        . The adjective hosticus of Figure 1, for
instance, clearly inherits its two main senses
(‘hostile’ and ‘foreign’) from the same polysemy of the
noun hostis ’stranger’ or ’enemy’, from which it is
derived. At the same time, while other resources
in LiLa describe the senses of words, such as the
Latin WordNet
        <xref ref-type="bibr" rid="ref13 ref23">(Franzini et al., 2019; Mambrini
et al., 2021)</xref>
        , the complex relations between those
senses (whether, for instance, one sense is
interpreted as a specialised derivation from another) is
generally available only in traditional lexical
resources like the L&amp;S.
      </p>
      <p>The solutions we found to address the
challenges raised by the representation of the L&amp;S in
LLOD will be reused when we will link further
bilingual, as well as monolingual, dictionaries of
Latin to the KB. Including such lexical resources
in LiLa is an important achievement, as it makes
it possible for the KB to interact with linguistic
(meta)data for languages other than Latin.
Undoubtedly, such an inter-linguistic (re)use of
distributed resources is one of the objectives of the
LLOD community, to which LiLa contributes by
steadily providing it also with new (kinds of)
linguistic resources represented in LLOD.</p>
    </sec>
    <sec id="sec-5">
      <title>Acknowledgments</title>
      <p>This project has received funding from the
European Research Council (ERC) under the
European Union’s Horizon 2020 research and
innovation programme – Grant Agreement No. 769994.</p>
    </sec>
  </body>
  <back>
    <ref-list>
      <ref id="ref1">
        <mixed-citation>
          <string-name>
            <given-names>Richard</given-names>
            <surname>Ashdowne</surname>
          </string-name>
          , David R Howlett, and Ronald Edward Latham.
          <year>1975</year>
          .
          <article-title>Dictionary of medieval Latin from British sources</article-title>
          . Oxford University Press.
        </mixed-citation>
      </ref>
      <ref id="ref2">
        <mixed-citation>
          <string-name>
            <given-names>David</given-names>
            <surname>Bamman</surname>
          </string-name>
          and
          <string-name>
            <given-names>Gregory</given-names>
            <surname>Crane</surname>
          </string-name>
          .
          <year>2011</year>
          .
          <article-title>The ancient greek and latin dependency treebanks</article-title>
          .
          <source>In Language technology for cultural heritage</source>
          , pages
          <fpage>79</fpage>
          -
          <lpage>98</lpage>
          . Springer.
        </mixed-citation>
      </ref>
      <ref id="ref3">
        <mixed-citation>
          <string-name>
            <given-names>Christian</given-names>
            <surname>Bizer</surname>
          </string-name>
          , Tom Heath,
          <string-name>
            <given-names>Kingsley</given-names>
            <surname>Idehen</surname>
          </string-name>
          , and
          <string-name>
            <surname>Tim</surname>
          </string-name>
          Berners-Lee.
          <year>2008</year>
          .
          <article-title>Linked data on the web (ldow2008)</article-title>
          .
          <source>In Proceedings of the 17th international conference on World Wide Web</source>
          , pages
          <fpage>1265</fpage>
          -
          <lpage>1266</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref4">
        <mixed-citation>
          <string-name>
            <given-names>Julia</given-names>
            <surname>Bosque-Gil</surname>
          </string-name>
          and
          <string-name>
            <given-names>Jorge</given-names>
            <surname>Gracia</surname>
          </string-name>
          .
          <year>2019</year>
          .
          <article-title>The OntoLex lemon lexicography module</article-title>
          . https://on tolex.github.io/lexicog/.
        </mixed-citation>
      </ref>
      <ref id="ref5">
        <mixed-citation>
          <string-name>
            <given-names>Flavio</given-names>
            <surname>Massimiliano</surname>
          </string-name>
          <string-name>
            <surname>Cecchini</surname>
          </string-name>
          , Timo Korkiakangas, and
          <string-name>
            <given-names>Marco</given-names>
            <surname>Passarotti</surname>
          </string-name>
          .
          <year>2020a</year>
          .
          <article-title>A new latin treebank for universal dependencies: Charters between ancient latin and romance languages</article-title>
          .
          <source>In Proceedings of The 12th Language Resources and Evaluation Conference</source>
          , pages
          <fpage>933</fpage>
          -
          <lpage>942</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref6">
        <mixed-citation>
          <string-name>
            <given-names>Flavio</given-names>
            <surname>Massimiliano</surname>
          </string-name>
          <string-name>
            <surname>Cecchini</surname>
          </string-name>
          , Rachele Sprugnoli, Giovanni Moretti, and
          <string-name>
            <given-names>Marco</given-names>
            <surname>Passarotti</surname>
          </string-name>
          . 2020b.
          <article-title>Udante: First steps towards the universal dependencies treebank of dante's latin works</article-title>
          . In CLiC-it.
        </mixed-citation>
      </ref>
      <ref id="ref7">
        <mixed-citation>
          <string-name>
            <given-names>Philipp</given-names>
            <surname>Cimiano</surname>
          </string-name>
          ,
          <string-name>
            <given-names>Christian</given-names>
            <surname>Chiarcos</surname>
          </string-name>
          ,
          <string-name>
            <surname>John P. McCrae</surname>
            ,
            <given-names>and Jorge</given-names>
          </string-name>
          <string-name>
            <surname>Gracia</surname>
          </string-name>
          .
          <year>2020</year>
          .
          <article-title>Linguistic Linked Data: Representation, Generation</article-title>
          and Applications. Springer, Cham.
        </mixed-citation>
      </ref>
      <ref id="ref8">
        <mixed-citation>
          <string-name>
            <surname>Irene De Felice</surname>
            , Giovanna Marotta, and
            <given-names>Margherita</given-names>
          </string-name>
          <string-name>
            <surname>Donati</surname>
          </string-name>
          .
          <year>2015</year>
          .
          <article-title>Classes: A new digital resource for latin epigraphy</article-title>
          .
          <source>IJCoL</source>
          .
          <source>Italian Journal of Computational Linguistics</source>
          ,
          <volume>1</volume>
          (
          <issue>1</issue>
          -1):
          <fpage>125</fpage>
          -
          <lpage>136</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref9">
        <mixed-citation>
          <string-name>
            <given-names>Joseph</given-names>
            <surname>Denooz</surname>
          </string-name>
          .
          <year>2007</year>
          .
          <article-title>Opera latina: le nouveau site internet du lasla</article-title>
          .
          <source>Journal of Latin Linguistics</source>
          ,
          <volume>9</volume>
          (
          <issue>3</issue>
          ):
          <fpage>21</fpage>
          -
          <lpage>34</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref10">
        <mixed-citation>
          <string-name>
            <given-names>Hanne</given-names>
            <surname>Eckhoff</surname>
          </string-name>
          , Kristin Bech, Gerlof Bouma, Kristine Eide, Dag Haug, Odd Einar Haugen, and
          <string-name>
            <given-names>Marius</given-names>
            <surname>Jøhndal</surname>
          </string-name>
          .
          <year>2018</year>
          .
          <article-title>The proiel treebank family: a standard for early attestations of indo-european languages</article-title>
          .
          <source>Language Resources and Evaluation</source>
          ,
          <volume>52</volume>
          (
          <issue>1</issue>
          ):
          <fpage>29</fpage>
          -
          <lpage>65</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref11">
        <mixed-citation>
          <string-name>
            <given-names>Wilhelm</given-names>
            <surname>Ehlers</surname>
          </string-name>
          .
          <year>1968</year>
          .
          <article-title>Der thesaurus linguae latinae. prinzipien und erfahrungen</article-title>
          .
          <source>Antike und Abendland</source>
          ,
          <volume>14</volume>
          (
          <issue>1</issue>
          ):
          <fpage>172</fpage>
          -
          <lpage>184</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref12">
        <mixed-citation>
          <string-name>
            <given-names>Egidio</given-names>
            <surname>Forcellini</surname>
          </string-name>
          and
          <string-name>
            <given-names>Jacobo</given-names>
            <surname>Facciolati</surname>
          </string-name>
          .
          <year>1871</year>
          .
          <article-title>Lexicon totius latinitatis</article-title>
          , volume
          <volume>3</volume>
          . Typis seminarii.
        </mixed-citation>
      </ref>
      <ref id="ref13">
        <mixed-citation>
          <string-name>
            <given-names>Greta</given-names>
            <surname>Franzini</surname>
          </string-name>
          , Andrea Peverelli, Paolo Ruffolo, Marco Passarotti, Helena Sanna, Edoardo Signoroni, Viviana Ventura, and
          <string-name>
            <given-names>Federica</given-names>
            <surname>Zampedri</surname>
          </string-name>
          .
          <year>2019</year>
          .
          <article-title>Nunc Est Aestimandum</article-title>
          .
          <article-title>Towards an evaluation of the Latin WordNet</article-title>
          . In Raffaella Bernardi, Roberto Navigli, and Giovanni Semeraro, editors,
          <source>Sixth Italian Conference on Computational Linguistics</source>
          (CLiC-it
          <year>2019</year>
          ), pages
          <fpage>1</fpage>
          -
          <lpage>8</lpage>
          , Bari, Italy.
          <source>CEUR-WS.org.</source>
        </mixed-citation>
      </ref>
      <ref id="ref14">
        <mixed-citation>
          <string-name>
            <surname>Peter GW Glare</surname>
          </string-name>
          .
          <year>1968</year>
          . Clarendon Press, Oxford.
        </mixed-citation>
      </ref>
      <ref id="ref15">
        <mixed-citation>
          <string-name>
            <given-names>Nancy</given-names>
            <surname>Ide</surname>
          </string-name>
          and
          <string-name>
            <given-names>James</given-names>
            <surname>Pustejovsky</surname>
          </string-name>
          .
          <year>2010</year>
          .
          <article-title>What does interoperability mean, anyway? toward an operational definition of interoperability for language technology</article-title>
          .
          <source>In Proceedings of the Second International Conference on Global Interoperability for Language Resources. Hong Kong</source>
          , China.
        </mixed-citation>
      </ref>
      <ref id="ref16">
        <mixed-citation>
          <string-name>
            <given-names>Steven E</given-names>
            <surname>Jones</surname>
          </string-name>
          .
          <year>2016</year>
          .
          <article-title>Roberto Busa, SJ, and the emergence of humanities computing: the priest and the punched cards</article-title>
          .
          <source>Routledge.</source>
        </mixed-citation>
      </ref>
      <ref id="ref17">
        <mixed-citation>
          <string-name>
            <surname>Neven</surname>
            <given-names>Jovanovic´.</given-names>
          </string-name>
          <year>2012</year>
          .
          <article-title>Croala. enhancing a teiencoded text collection</article-title>
          .
          <source>Journal of the Text Encoding Initiative</source>
          , (
          <volume>2</volume>
          ).
        </mixed-citation>
      </ref>
      <ref id="ref18">
        <mixed-citation>
          <string-name>
            <given-names>Ora</given-names>
            <surname>Lassila</surname>
          </string-name>
          ,
          <string-name>
            <surname>Ralph R. Swick</surname>
            , World Wide, and
            <given-names>Web</given-names>
          </string-name>
          <string-name>
            <surname>Consortium</surname>
          </string-name>
          .
          <year>1998</year>
          .
          <article-title>Resource description framework (rdf) model and syntax specification</article-title>
          .
        </mixed-citation>
      </ref>
      <ref id="ref19">
        <mixed-citation>
          <string-name>
            <surname>Charlton</surname>
            <given-names>T.</given-names>
          </string-name>
          <string-name>
            <surname>Lewis</surname>
            and
            <given-names>Charles</given-names>
          </string-name>
          <string-name>
            <surname>Short</surname>
          </string-name>
          .
          <year>1879</year>
          .
          <article-title>A Latin Dictionary</article-title>
          .
          <article-title>Founded on Andrews' edition of Freund's Latin dictionary</article-title>
          . Clarendon Press, Oxford.
        </mixed-citation>
      </ref>
      <ref id="ref20">
        <mixed-citation>
          <string-name>
            <given-names>Henry</given-names>
            <surname>Liddell</surname>
          </string-name>
          ,
          <string-name>
            <given-names>Robert</given-names>
            <surname>Scott</surname>
          </string-name>
          , and Henry Stuart Jones.
          <year>1940</year>
          .
          <string-name>
            <given-names>A</given-names>
            <surname>Greek-English Lexicon</surname>
          </string-name>
          . Clarendon Press, Oxford,
          <volume>9</volume>
          <fpage>edition</fpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref21">
        <mixed-citation>
          <string-name>
            <given-names>Eleonora</given-names>
            <surname>Litta</surname>
          </string-name>
          , Marco Passarotti, and
          <string-name>
            <given-names>Francesco</given-names>
            <surname>Mambrini</surname>
          </string-name>
          .
          <year>2020</year>
          .
          <article-title>Derivations and Connections: Word Formation in the LiLa Knowledge Base of Linguistic Resources for Latin</article-title>
          .
          <source>The Prague Bulletin Of Mathematical Linguistics</source>
          ,
          <volume>115</volume>
          :
          <fpage>163</fpage>
          -
          <lpage>186</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref22">
        <mixed-citation>
          <string-name>
            <given-names>Francesco</given-names>
            <surname>Mambrini</surname>
          </string-name>
          and
          <string-name>
            <given-names>Marco</given-names>
            <surname>Passarotti</surname>
          </string-name>
          .
          <year>2020</year>
          .
          <article-title>Representing etymology in the lila knowledge base of linguistic resources for latin</article-title>
          .
          <source>In Proceedings of the 2020 Globalex Workshop on Linked Lexicography</source>
          , pages
          <fpage>20</fpage>
          -
          <lpage>28</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref23">
        <mixed-citation>
          <string-name>
            <given-names>Francesco</given-names>
            <surname>Mambrini</surname>
          </string-name>
          , Marco Passarotti, Eleonora Litta, and
          <string-name>
            <given-names>Giovanni</given-names>
            <surname>Moretti</surname>
          </string-name>
          .
          <year>2021</year>
          .
          <article-title>Interlinking valency frames and wordnet synsets in the lila knowledge base of linguistic resources for latin</article-title>
          .
          <source>In Further with Knowledge Graphs</source>
          , pages
          <fpage>16</fpage>
          -
          <lpage>28</lpage>
          . IOS Press.
        </mixed-citation>
      </ref>
      <ref id="ref24">
        <mixed-citation>
          <string-name>
            <given-names>Massimo</given-names>
            <surname>Manca</surname>
          </string-name>
          , Linda Spinazze`,
          <string-name>
            <surname>Paolo</surname>
            <given-names>Mastandrea</given-names>
          </string-name>
          , Luigi Tessarolo, and
          <string-name>
            <given-names>Federico</given-names>
            <surname>Boschetti</surname>
          </string-name>
          .
          <year>2011</year>
          .
          <article-title>Musisque deoque: Text retrieval on critical editionse</article-title>
          .
          <source>J. Lang. Technol. Comput. Linguistics</source>
          ,
          <volume>26</volume>
          (
          <issue>2</issue>
          ):
          <fpage>127</fpage>
          -
          <lpage>138</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref25">
        <mixed-citation>
          <string-name>
            <surname>John P. McCrae</surname>
          </string-name>
          ,
          <string-name>
            <surname>Julia</surname>
            Bosque-Gil, Jorge Gracia, Paul Buitelaar, and
            <given-names>Philipp</given-names>
          </string-name>
          <string-name>
            <surname>Cimiano</surname>
          </string-name>
          .
          <year>2017</year>
          .
          <article-title>The OntoLex-Lemon Model: development and applications</article-title>
          .
          <source>In Proceedings of eLex 2017</source>
          , pages
          <fpage>587</fpage>
          -
          <lpage>597</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref26">
        <mixed-citation>
          <string-name>
            <given-names>Barbara</given-names>
            <surname>McGillivray</surname>
          </string-name>
          and
          <string-name>
            <given-names>Alessandro</given-names>
            <surname>Vatri</surname>
          </string-name>
          .
          <year>2015</year>
          .
          <article-title>Computational valency lexica for latin and greek in use: a case study of syntactic ambiguity</article-title>
          .
          <source>Journal of Latin Linguistics</source>
          ,
          <volume>14</volume>
          (
          <issue>1</issue>
          ):
          <fpage>101</fpage>
          -
          <lpage>126</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref27">
        <mixed-citation>
          <string-name>
            <given-names>Stefano</given-names>
            <surname>Minozzi</surname>
          </string-name>
          .
          <year>2017</year>
          .
          <article-title>Latin wordnet, una rete di conoscenza semantica per il latino e alcune ipotesi di utilizzo nel campo dell'information retrieval. Strumenti digitali e collaborativi per le Scienze dell'Antichita`</article-title>
          , (
          <volume>14</volume>
          ):
          <fpage>123</fpage>
          -
          <lpage>134</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref28">
        <mixed-citation>
          <string-name>
            <given-names>Marco</given-names>
            <surname>Passarotti</surname>
          </string-name>
          , Berta Gonza´lez Saavedra, and
          <string-name>
            <given-names>Christophe</given-names>
            <surname>Onambele</surname>
          </string-name>
          .
          <year>2016</year>
          .
          <article-title>Latin vallex. a treebank-based semantic valency lexicon for latin</article-title>
          .
          <source>In Proceedings of the Tenth International Conference on Language Resources and Evaluation (LREC'16)</source>
          , pages
          <fpage>2599</fpage>
          -
          <lpage>2606</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref29">
        <mixed-citation>
          <string-name>
            <given-names>Marco</given-names>
            <surname>Passarotti</surname>
          </string-name>
          , Marco Budassi, Eleonora Litta, and
          <string-name>
            <given-names>Paolo</given-names>
            <surname>Ruffolo</surname>
          </string-name>
          .
          <year>2017</year>
          .
          <article-title>The Lemlat 3.0 Package for Morphological Analysis of Latin</article-title>
          . In Gerlof Bouma and Yvonne Adesam, editors,
          <source>Proceedings of the NoDaLiDa 2017 Workshop on Processing Historical Language</source>
          , volume
          <volume>133</volume>
          , pages
          <fpage>24</fpage>
          -
          <lpage>31</lpage>
          , Gothenburg. Linko¨ping University Electronic Press.
        </mixed-citation>
      </ref>
      <ref id="ref30">
        <mixed-citation>
          <string-name>
            <given-names>Marco</given-names>
            <surname>Passarotti</surname>
          </string-name>
          , Francesco Mambrini, Greta Franzini, Flavio Massimiliano Cecchini, Eleonora Litta, Giovanni Moretti, Paolo Ruffolo, and
          <string-name>
            <given-names>Rachele</given-names>
            <surname>Sprugnoli</surname>
          </string-name>
          .
          <year>2020</year>
          .
          <article-title>Interlinking through lemmas. the lexical collection of the lila knowledge base of linguistic resources for latin</article-title>
          .
          <source>Studi e Saggi Linguistici</source>
          ,
          <volume>58</volume>
          (
          <issue>1</issue>
          ):
          <fpage>177</fpage>
          -
          <lpage>212</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref31">
        <mixed-citation>
          <string-name>
            <given-names>Marco</given-names>
            <surname>Passarotti</surname>
          </string-name>
          .
          <year>2019</year>
          .
          <article-title>The project of the index thomisticus treebank</article-title>
          .
          <source>In Digital Classical Philology</source>
          , pages
          <fpage>299</fpage>
          -
          <lpage>320</lpage>
          . De Gruyter Saur.
        </mixed-citation>
      </ref>
      <ref id="ref32">
        <mixed-citation>
          <string-name>
            <given-names>Matteo</given-names>
            <surname>Pellegrini</surname>
          </string-name>
          , Eleonora Litta, Marco Passarotti, Francesco Mambrini, and
          <string-name>
            <given-names>Giovanni</given-names>
            <surname>Moretti</surname>
          </string-name>
          .
          <year>2021</year>
          .
          <article-title>The two approaches to word formation in the lila knowledge base of latin resources</article-title>
          .
          <source>In Proceedings of the Third International Workshop on Resources and Tools for Derivational Morphology (DeriMo</source>
          <year>2021</year>
          ), pages
          <fpage>101</fpage>
          -
          <lpage>109</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref33">
        <mixed-citation>
          <string-name>
            <given-names>Johann</given-names>
            <surname>Ramminger</surname>
          </string-name>
          .
          <year>2008</year>
          .
          <article-title>Neulateinische Wortliste. Ein Wo¨rterbuch der Lateinischen von Petrarca bis 1700</article-title>
          . Thesaurus Linguae Latinae.
        </mixed-citation>
      </ref>
      <ref id="ref34">
        <mixed-citation>
          <string-name>
            <surname>Jeffrey A Rydberg-Cox</surname>
          </string-name>
          .
          <year>2002</year>
          .
          <article-title>an Electronic Greek Lexicon</article-title>
          .
          <volume>98</volume>
          (
          <issue>2</issue>
          ):
          <fpage>183</fpage>
          -
          <lpage>188</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref35">
        <mixed-citation>
          <string-name>
            <given-names>Rachele</given-names>
            <surname>Sprugnoli</surname>
          </string-name>
          , Francesco Mambrini, Giovanni Moretti, and
          <string-name>
            <given-names>Marco</given-names>
            <surname>Passarotti</surname>
          </string-name>
          .
          <year>2020</year>
          .
          <article-title>Towards the modeling of polarity in a latin knowledge base</article-title>
          .
          <source>In WHiSe@ ESWC</source>
          , pages
          <fpage>59</fpage>
          -
          <lpage>70</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref36">
        <mixed-citation>
          <string-name>
            <given-names>Paul</given-names>
            <surname>Tombeur</surname>
          </string-name>
          .
          <year>1998</year>
          .
          <article-title>Thesaurus formarum totius Latinitatis: a Plauto usque ad saeculum XXum; TF.[2]. CETEDOC Index of Latin forms: database for the study of the vocabulary of the entire Latin world; base de donne´es pour l'e´tude du vocabulaire de toute la latinite´</article-title>
          . Brepols.
        </mixed-citation>
      </ref>
    </ref-list>
  </back>
</article>