<!DOCTYPE article PUBLIC "-//NLM//DTD JATS (Z39.96) Journal Archiving and Interchange DTD v1.0 20120330//EN" "JATS-archivearticle1.dtd">
<article xmlns:xlink="http://www.w3.org/1999/xlink">
  <front>
    <journal-meta />
    <article-meta>
      <title-group>
        <article-title>Italian Sign Language (LIS) and Natural Language Processing: an Overview</article-title>
      </title-group>
      <contrib-group>
        <contrib contrib-type="author">
          <string-name>University of Catania</string-name>
        </contrib>
        <contrib contrib-type="author">
          <string-name>Department of Humanities</string-name>
        </contrib>
        <contrib contrib-type="author">
          <string-name>Italy sabina.fontana@unict.it</string-name>
        </contrib>
        <aff id="aff0">
          <label>0</label>
          <institution>University of Catania, Department of Humanities</institution>
          ,
          <country country="IT">Italy</country>
        </aff>
      </contrib-group>
      <pub-date>
        <year>1819</year>
      </pub-date>
      <fpage>0000</fpage>
      <lpage>0003</lpage>
      <abstract>
        <p>The past decade has seen an increase in studies conducted on the interaction between NLP and sign languages. In this paper, we mainly focus on LIS while discussing the current state of the art, possible future developments and the ethical implications of this growing research context. In the history of NLP, human/computer interaction has been mainly based on the transcription of spoken languages. We investigate how existing resources for spoken language processing can be applied to SLs and combined with language-specific tools, providing examples of recent resources. We discuss novel strategies for sign transcription that consider both the need for standardized writing forms to enable NLP, as well as the language-specific features of SLs that are conveyed through the visual-manual channel. Deaf contributors are fundamental within this research. When NLP and SLs interact, we find a shift from a user-centric approach towards a user-based one to be essential: Deaf end-users of the resulting resources thus become part of the designing process1.</p>
      </abstract>
      <kwd-group>
        <kwd>Italian Sign Language (LIS)</kwd>
        <kwd>Natural Language Processing</kwd>
        <kwd>Translation</kwd>
      </kwd-group>
    </article-meta>
  </front>
  <body>
    <sec id="sec-1">
      <title>-</title>
      <p>
        the WHO (World Health Organization) predicts that by 2050 nearly 2.5 billion people in
the world will have some degree of hearing loss [
        <xref ref-type="bibr" rid="ref3">3</xref>
        ].
      </p>
      <p>
        Whether they are born deaf or lose their hearing in their early infancy, due to illness
or an accident, individuals with different levels of deafness (moderate, severe, profound)
will be facing difficulties in communicating with hearing speakers. Such difficulties are
of different kinds, depending on whether deafness is experienced in early infancy during
language development or if it is acquired at an adult age, or if deaf children are exposed
to a sign language in early infancy or not. About 10% of the deaf population is born deaf
to deaf parents. The remaining 90% have hearing parents who do not know a sign
language. This raises an issue concerning the transmission of sign languages, which tends
to occur in a horizontal (among peers) rather than vertical way (from generation to
generation). Recently, mainstream and technology have further influenced the educational
path of deaf children and have led to delay or deny access to a sign language. Families
tend to prefer normalization, choosing a spoken language and resisting bilingualism.
Consequently, sign language is accessed and learnt very often at an adult age.
Consequently, Deaf signers may acquire different levels of proficiency in spoken language and
sign language [
        <xref ref-type="bibr" rid="ref4">4</xref>
        ].
      </p>
      <p>
        Sign languages and spoken languages vastly differ since they employ different
modalities: auditory-oral and visual-manual. For this reason, the bilingualism of Deaf
signers is often referred to as bimodal [
        <xref ref-type="bibr" rid="ref5">5</xref>
        ]. Deaf people's bilingualism has some specific
features. Firstly, it is bimodal [
        <xref ref-type="bibr" rid="ref6">6</xref>
        ] because they learn or acquire two languages exploiting
different modalities through different paths: spoken language through speech therapy
starting from early infancy and sign language through exposure, rarely during infancy,
more frequently at a young age or even later in their lives. The two languages are used
alternately but are in contact, so they influence each other while occurring
simultaneously, since there are no serial order constraints [
        <xref ref-type="bibr" rid="ref7">7</xref>
        ]. Secondly, it is generally an
unbalanced bilingualism: even if learnt later, the sign language appears to be their natural
language and skills in spoken language are rarely comparable to natives2. Lastly, the two
languages have different sociolinguistic statuses. On the one hand, the spoken language
– that is the majority language – is widely shared, institutional and used in education, in
media communication and many other formal contexts. On the other hand, the sign
language is a minority language that has been long stigmatized and only recently is being
used in formal contexts, but still not enough in education. Although many schools and
universities provide support in sign language through professional interpreters or
communication assistants, there are only a few bilingual sign/spoken language schools in
Italy (in Biella, Cossato and Rome) where LIS is studied like any other subject and taught
to all students [
        <xref ref-type="bibr" rid="ref8">8</xref>
        ].
      </p>
      <p>
        It is important to highlight that this condition of bilingualism influences the usage
and the perception of sign languages from Deaf users [
        <xref ref-type="bibr" rid="ref9">9</xref>
        ]. It also influences the research
work that must be carried out. It is fundamental to keep in mind that the target group is
far from being homogeneous as it consists of people who have vastly different
2 We are not using the categories ‘first language’, ‘second language’, and ‘mother tongue’ as
they do not exactly mirror the specificities of deaf bilingualism.
      </p>
      <p>Italian Sign Language (LIS) and Natural Language Processing: an Overview 3
experiences of Deafness and language acquisition. The experience of Deaf signers who
present bimodal bilingualism will be different from that of individuals who lost their
hearing later in life and are mainly faced with communication problems related to access
rather than comprehension of a spoken language.
1.1</p>
    </sec>
    <sec id="sec-2">
      <title>The Italian Sign Language (LIS)</title>
      <p>As for spoken languages, when analyzing a sign language, adopting an interlinguistic
perspective and a broader, more international frame of reference, will surely provide a
rounder understanding of linguistic and social phenomena. Given our research interests,
in this paper we will mainly focus on LIS. Therefore, all images and mentioned resources
will be looked at as part of the LIS framework.</p>
      <p>
        As for the signing population, the database Ethnologue [
        <xref ref-type="bibr" rid="ref10">10</xref>
        ] lists 121 sign languages
used by 70 million Deaf people with 60.000 users of LIS. But what are the characteristics
that determine the basic traits of LIS? Studying sign language from a phonocentric
perspective can lead to the research of the same categories that define spoken languages.
While this has been the preferred strategy in the past, it is fundamental to keep in mind
that the different modalities (visual-manual/auditory-vocal) will call for different
descriptions of a language that may not mirror each other. For this reason, we like to start
describing LIS by observing how gestuality is systematically organized to create
meaning. LIS employs the visual-manual modality, involving both manual and non-manual
elements in the construction of signs. Manual elements are usually defined using the
following major parameters: the shape one’s hand (or hands) acquire while performing
a sign, its place of articulation (also referred to as ‘location’), the way hands move in
space and hand orientation, i.e., the position of the palm of the dominant hand [
        <xref ref-type="bibr" rid="ref11">11</xref>
        ].
Nonmanual elements play an equivalently crucial role in the construction of meaning and
include head and body movements, facial expressions and mouth gestures [
        <xref ref-type="bibr" rid="ref11">11</xref>
        ]. These
elements can simultaneously convey semantic information or have pronominal value
during role-shift, a complex process that allows signers to ‘become’ the person, animal
or object they are representing.
The systematic creation of meaning in LIS can generate signs with different degrees of
arbitrariness or iconicity. Gestuality is an essential aspect of human communication: if
hands become the core elements of a language, there will be a continuity between
gestuality and signs, which originates iconic phenomena. The presence of iconicity does not
exclude arbitrariness. The spontaneous evolution of an arbitrary sign is arbitrary and
unpredictable and takes place within a community that is impacted by specific cultural,
social and geographical influences.
      </p>
      <p>
        More than half a century has passed since the publication of the first
methodological studies on sign languages. During this time, a shift in linguistics has slowly taken
place, leading to a transformation in the perception of sign languages that are now
universally accepted as natural ones and have been legally recognized as official languages
in numerous countries. Changes in attitude and increased visibility forced signers to think
upon their language and define a notion of correctness. At the same time, when sign
languages became visible and accessible to a larger Deaf and hearing public (also
through professional interpreting services in TV and other official contexts), Deaf people
realized that it lacked many lexical items and functions to meet the different
communicative needs [
        <xref ref-type="bibr" rid="ref13">13</xref>
        ].
      </p>
      <p>
        The recognition of sign languages has been gradual but steady and has seen a fast
development in the last 20 years. Within the Italian political and social context, the 2020
pandemic made LIS (Lingua dei Segni Italiana–Italian Sign Language) increasingly
visible and public, as interpreters translated presidential speeches and conferences [
        <xref ref-type="bibr" rid="ref14">14</xref>
        ].
This newly acquired visibility– together with the widespread use of social media by
Italian Deaf people – only sped up the ongoing process of standardization. At the same time,
as LIS becomes more prominent and is used in more contexts, linguistic growth becomes
necessary, leading to a rapid expansion of its vocabulary. Furthermore, the pandemic
highlighted the hardship of Deaf people in contexts where face-to-face communication
is limited, culminating in the official recognition of LIS and LIST (Tactile Italian Sign
Language) by the Italian Government in May of 2021 [
        <xref ref-type="bibr" rid="ref15">15</xref>
        ].
      </p>
      <p>
        As for the analysis of LIS structures, the main issue is that, up until the end of 2020,
there were no grammars that extensively described its structural phenomena. This is due
to different factors. LIS was first described in 1987 by Virginia Volterra [
        <xref ref-type="bibr" rid="ref16">16</xref>
        ]. At the
time, the objective was the ‘dignification’ of the language, leading to a pursuit of the
same parameters that define spoken languages within an assimilationist perspective.
These first stages of LIS studies were characterized by a search for categories such as
phonemes or minimal couples. After the first decade of research on LIS, signers
themselves started using LIS in more formal environments [
        <xref ref-type="bibr" rid="ref17">17</xref>
        ]: the removal of LIS from the
domestic and informal contexts led to an expansion and standardization of the language,
which is still ongoing. With that came the newfound awareness of Italian signers who
now see their language as a vehicle of Deaf pride and object in need of protection and
preservation.
      </p>
      <p>We take it for granted that sign languages are natural, meaning that they evolve
spontaneously within a community. They are different from spoken languages in that they are
not a manual translation of spoken languages and do not have standard written forms.
Since the members of these communities are Deaf, they will make use of their bodies to</p>
      <p>
        Italian Sign Language (LIS) and Natural Language Processing: an Overview 5
convey meaning. As we have said before, at present, sign languages are considered ‘oral’
in that there is no formal system for their transcription. As bimodal bilinguals, signers
usually rely on the written form of spoken languages. This has largely influenced
research on sign languages, leading to the use of glosses for sign transcription: the
translation of the sign into spoken language written in all capitals. Of course, an external writing
system that relies on translation does not convey the expressiveness of signs, hence the
centrality of the need to move beyond glosses in sign language studies. Within this
context, several attempts have been made at designing an independent signing system, the
most successful being SignWriting [
        <xref ref-type="bibr" rid="ref18">18</xref>
        ].
2
      </p>
      <sec id="sec-2-1">
        <title>Natural Language Processing and Italian Sign Language:</title>
      </sec>
      <sec id="sec-2-2">
        <title>Issues and Resources</title>
        <p>We focus our attention on the possible contacts between NLP and sign languages to
facilitate interactions between Deaf and hearing people and reflect on the practical and
ethical challenges researchers are faced with when these two worlds come into contact.
The sign language we focus on is Italian Sign Language (LIS). However, the
observations we make in this paper refer to topics that lay at the core of all sign languages. We
also make some observations on the implications of translating automatically sign
languages into vocal languages and vice-versa.</p>
        <p>In this section, we first focus on linguistic issues to be considered when working with
LIS. We then move on to the current state of the art, describing projects for text to video
synthesis in the Italian context. After that, we discuss available options for glossing LIS.
We also provide Italian and international examples of different strategies employed to
build datasets by converting text into sign language and sign language videos into text,
through processes of manual and automatic recognition.
2.1</p>
      </sec>
    </sec>
    <sec id="sec-3">
      <title>Linguistic and Ethical Issues</title>
      <p>
        In shaping a dataset, different types of issues must be taken into consideration. First,
we consider those issues that are related to the specific characteristics of LIS. We
mentioned that sign languages should be looked at and analyzed using independent categories
that do not necessarily mirror those of spoken languages. However, given that a
computational analysis must go through a written message, it is necessary to reflect upon sign
language transcription, its linguistic, social and ethical implications. Like other sign
languages, LIS is an oral language and does not have a standardized writing system, given
the visual-manual modality it employs and the fact that it relies on written Italian as an
external writing system. An international, language-specific transcription system is
SignWriting [
        <xref ref-type="bibr" rid="ref18">18</xref>
        ] which consists of a database of iconic symbols representing orientation
and handshape of one or both hands, facial expression, movements, contact and location
of the sign. These symbols are combined to represent signs. Due to its iconic nature,
SignWriting allows for a detailed and simultaneous representation of the multilinearity
of signed discourse but, despite the positive response from researchers, it has failed to
become the annotation system of choosing for signers. As discussed in subsection 1.1.,
only recently sign languages have been analyzed, described and recognized as natural
languages. For these reasons, some signs could have not been developed yet. From a
social standpoint, dataset collection should involve Deaf people not only as informants
but as participants in that their living experience and linguistic knowledge is crucial for
the development of the system that should account for their needs. The contribution of
Deaf researchers is also important to define the setting for data elicitation to create a
more natural dataset. Very often, when data are elicited in a very artificial setting, Deaf
informants may shift to spoken language or mixed spoken and signed language varieties,
i.e., contact signing [
        <xref ref-type="bibr" rid="ref19">19</xref>
        ]. We come to the second issue that should be considered in
collecting datasets: the necessity to account for the variability deriving from different ages
of acquisition and skills in spoken and sign languages. Ideally, to correctly identify
linguistic structures, datasets should be collected in a naturalistic setting and involve Deaf
people as participants.
2.2
      </p>
    </sec>
    <sec id="sec-4">
      <title>Automatic Processing</title>
      <p>
        Different issues arise once automatic processing is implemented on the language.
Once data is obtained, researchers will be faced with the need for data annotation. A
basic requirement is, of course, a readable transcription. This requirement is problematic
for sign languages since current methodologies rely primarily on word labels:
translations taken from spoken languages [
        <xref ref-type="bibr" rid="ref20">20</xref>
        ]. For this reason, in this section we also discuss
tasks of word segmentation, reflecting on the issues one comes across when combining
it with LIS. Word segmentation is a basic step, fundamental to sign annotation. When it
comes to spoken languages, the content of each segment will naturally be one word,
usually divided by the remaining string of written language by a space delimiter [
        <xref ref-type="bibr" rid="ref21">21</xref>
        ].
However, this step is not as straightforward when it comes to sign languages since it calls
for the recognition of the beginning and end of a sign as well as the design of a
transcription strategy. For this reason, gloss annotation is one of the main and most
time-consuming issues [
        <xref ref-type="bibr" rid="ref22">22</xref>
        ].
2.3
      </p>
    </sec>
    <sec id="sec-5">
      <title>State of the Art</title>
      <p>
        Several models have been used to automatically transfer information from sign
language to spoken language and vice versa. The general research problem is sign language
recognition, as well as the processes leading to it [
        <xref ref-type="bibr" rid="ref23">23</xref>
        ]. Sign language recognition is
concerned with the identification, segmentation and definition of signs. Given the
multimodality and multilinearity of sign languages, data must be collected through video or
images. To achieve said recognition, multi-channel approaches have been used, combining
hardware-based or software-based resources [
        <xref ref-type="bibr" rid="ref24">24</xref>
        ]. In the past year, among the recent
technologies employed for data collection for sign recognition, we find camera images
[
        <xref ref-type="bibr" rid="ref25">25</xref>
        ] – in some instances acquired with RGB and RGB-D sensors [
        <xref ref-type="bibr" rid="ref26">26</xref>
        ] – fused with radar
sensor data [
        <xref ref-type="bibr" rid="ref27">27</xref>
        ].
      </p>
      <p>Italian Sign Language (LIS) and Natural Language Processing: an Overview 7
2.4</p>
    </sec>
    <sec id="sec-6">
      <title>Text to Video Synthesis Projects for LIS</title>
      <p>
        Within the Italian context, different projects have been carried out on the application
of NLP and MT on LIS. We will now describe two related projects developed in the past
years. The LIS4ALL project (2012–2014) [
        <xref ref-type="bibr" rid="ref28">28</xref>
        ]: funded in 2012 by Piedmont to create a
prototype system for ITA-LIS automatic translation, providing a service for displaying
information in LIS on mobile devices in train station terminals. The output was LIS
utterances signed by an animated interpreter [
        <xref ref-type="bibr" rid="ref20">20</xref>
        ]. The LIS4ALL designers found that a
small number of templates covered most announcements [
        <xref ref-type="bibr" rid="ref29">29</xref>
        ] and thus built three regular
expressions that matched three templates. Following a process of sentence
simplification, non-mandatory components (for example deictic signs) were not translated into
LIS. The project resulted in the translation of 63 announcements. The main sources of
error identified during translation were lexical gaps and the inability to handle the
doubling of subject [
        <xref ref-type="bibr" rid="ref29">29</xref>
        ]. Given the novelty of the project, central aspects of LIS, such as
non-manual elements and the morphological aspects of sign movement and location,
were not considered in the project.
      </p>
      <p>
        The LIS4ALL system was based on an existing one developed in the context of the
ATLAS project for the translation of weather forecasts [
        <xref ref-type="bibr" rid="ref30 ref31">30–31</xref>
        ] through the creation of
a lexicon of 2350 signs [
        <xref ref-type="bibr" rid="ref30">30</xref>
        ]. Moreover, the translation strategies used for LIS4ALL were
interlingua rule-based and statistical translation [
        <xref ref-type="bibr" rid="ref28">28</xref>
        ], an evolution of the strategies that
had been employed in the ATLAS project (Automatic Translation into sign LAnguageS,
2009–2012) for the description and identification of the relations between signs. For
ATLAS, glossed LIS utterances had been analyzed using a set of lexical items and
several combinatorial rules [
        <xref ref-type="bibr" rid="ref28">28</xref>
        ] that were elaborated by a LIS generator that ‘[…] builds a
tree representing the generic LIS lexical items and some generic syntactic relations
among them […]’ [
        <xref ref-type="bibr" rid="ref28">28</xref>
        ]. Additionally, a set of values was associated with each sign gloss
and provided information on the database name of the sign, its ID, the number of hands
used to perform the sign and the part of speech of the sign [
        <xref ref-type="bibr" rid="ref31">31</xref>
        ] with the final aim of
obtaining a signed target text performed by an avatar.
2.5
      </p>
    </sec>
    <sec id="sec-7">
      <title>Going Beyond Glosses: Available Options for LIS</title>
      <p>Traditionally, signs have been transcribed using glosses, i.e., a translation of the sign
into the spoken target language. Therefore, in the context of sign transcription, if a LIS
signer is describing something that happened involving something their dog did at home,
we will surely have to transcribe ‘CANE’ (DOG) and ‘CASA’ (HOUSE). As can be
inferred, despite their widespread use, glosses are flattening to the complexity of LIS. In
fact, to a non-signer, the mentioned glosses provide no information on the production of
the sign itself, only specifying its meaning. Additionally, signs may vary for many
reasons, such as the geographical origin of a signer. Therefore, glosses make it virtually
impossible for non-signers to obtain information on the sign once removed from the
signed context.</p>
      <p>
        Within the Italian framework, different strategies have been adopted to work around
this issue. Generally, the most successful solution is the combination of video files and
univocal glosses. In the recently online-published resource A Grammar of Italian Sign
Language (LIS) [
        <xref ref-type="bibr" rid="ref32">32</xref>
        ] the authors opted for a variation of gloss annotation by included
videos of signers performing an isolated sign or an utterance in LIS, thus showing how
signs combine and interact, as well as different variations of the same sign. The
Grammar was published in the context of the SIGN-HUB project which aims at preserving
and researching the linguistic, historical and cultural heritage of European Deaf signing
communities with an integral resource [
        <xref ref-type="bibr" rid="ref33">33</xref>
        ].
      </p>
      <p>
        By associating glosses and video files, the Grammar certainly shows an aptitude towards
a rework of the widespread glossing methodology. However, it does not collect its signs
and utterances within a dictionary and, even if it did, the amount of data would be
extremely limited. As for LIS dictionaries, at present, the only digital resources available
to LIS researchers are the Dizionario Bilingue Elementare della Lingua dei Segni
Italiana (LIS)3 [
        <xref ref-type="bibr" rid="ref34">34</xref>
        ] in its digital version, and the website SpreadTheSign [
        <xref ref-type="bibr" rid="ref35 ref36">35–36</xref>
        ].
      </p>
      <p>
        The Dictionary is one of the most renowned and retrievable resources for LIS and it
includes more than 2500 videos of signs performed by native signers. Each video is
marked by a specific code that includes the translation of the sign into Italian and a
sequence of numbers and/or letters. As mentioned, the richness of sign languages cannot
be reproduced through capital letters [
        <xref ref-type="bibr" rid="ref37">37</xref>
        ] and might even lead to the creation of
ambiguity, not providing information on the variation of the sign. For this reason, the
Dictionary can be an excellent resource for LIS transcription since it provides an
unambiguous code for each sign, thus creating an unequivocal association between the sign and
its translation.
      </p>
      <p>Despite its usefulness and the vastness of the resource, the Dictionary is limited in
that it was concluded in 1992. For this reason, the multi-language dictionaries available
on the website SpreadTheSign provide an invaluable contribution. The website is the
result of an EU-funded project created for Deaf education and provides multilingual
dictionaries for several sign languages from around the world.</p>
      <p>In the previous section, we mentioned existing online dictionaries for LIS. We will
now discuss different annotation tools and strategies used for other European and
Northern American sign languages.</p>
      <p>
        With regards to available resources for the segmentation and analysis of sign
language videos, at present, ELAN is the most used resource in this field. It was initially
released in 2000 by The Max Planck Institute for Psycholinguistics in Nijmegen,
Netherlands, as a tool to annotate audiovisual files on different levels. ELAN is a useful tool
for multimodality research on sign languages since it allows users to create time-aligned
annotation levels where simultaneous information on manual and non-manual elements
can be included [
        <xref ref-type="bibr" rid="ref38">38</xref>
        ].
      </p>
      <p>
        Sign language translation requires glossing, which is a time-consuming, yet
necessary, annotation process. Tokenization on the gloss level is widely used since a
recognition of each sign that makes up an utterance facilitates the translation process. An
example of a tokenized sign language corpus is the Swedish Sign language Corpus (SSLC), a
resource developed at the Department of Linguistics of the University of Stockholm. The
SSLC was compiled between 2009 and 2001 and included video-recorded conversations
of 42 Swedish Sign Language Signers from the age of 20 to 82. The tokens collected for
the SSLC are 33,600 [
        <xref ref-type="bibr" rid="ref39">39</xref>
        ]. The tokenization process was led on the ELAN platform.
Annotators developed different levels (or tiers) to provide information on sign glosses
taken from the Swedish Sign Language Dictionary [
        <xref ref-type="bibr" rid="ref40">40</xref>
        ], adding specifically developed
tags and symbols to signal phenomena such as overlapping, merging, fingerspelling, or
gesture-like sign. Another corpus partly annotated and tagged using ELAN is the British
Sign Language Corpus Project (BSLCP) [
        <xref ref-type="bibr" rid="ref41">41</xref>
        ]. For the creation of the corpus, 249 BSL
Deaf signers were recorded in conversational contexts. The ELAN software was used to
provide information on what was being signed by the participants [
        <xref ref-type="bibr" rid="ref42">42</xref>
        ]. Within the
context of the SIGN-HUB Project, Pfau [
        <xref ref-type="bibr" rid="ref43">43</xref>
        ] and other researchers from Germany, Italy, The
Netherlands, Spain and Turkey, collaborated on a shared project for sign language
annotation on ELAN. The goal was the inclusion of an annotation that went beyond
translation or glossing, including aspects on non-manual productions and even comments from
annotators. As a result, more than 5 hours of footage were annotated by Deaf and hearing
researchers.
      </p>
      <p>
        All sign language data of the projects discussed in this section were annotated
manually. As mentioned, annotation is a time-consuming process. To minimize the amount
of time spent on it as well as facilitate the growth of large, annotated corpora for sign
languages research, Drew and Ney [
        <xref ref-type="bibr" rid="ref44">44</xref>
        ] created a new interface for ELAN able to
automatically recognize signs and annotate simultaneous tiers. The information included in
the tiers is: glosses corresponding to the translations into spoken language as well as
information on the features of each annotated sign. The richness of gloss annotation was
designed to be modelled by users depending on their interests.
      </p>
      <p>
        Given that most of the research on sign language recognition has been mainly
focused on the tasks of gesture recognition, the linguistic qualities of sign languages have
been overlooked. However, recent papers suggest a new approach to sign language
translation problems. Sign language recognition uses contact [
        <xref ref-type="bibr" rid="ref45">45</xref>
        ] or vision-based systems
[
        <xref ref-type="bibr" rid="ref46">46</xref>
        ]. The former is cumbersome, while the latter is mainly focused on identifying
individual signs. Instead, real-world sign language recognition should be continuous to
process the signing flow accurately. In 2018, Camgoz et al. [
        <xref ref-type="bibr" rid="ref47">47</xref>
        ] proposed the generation of
spoken language translation from German sign language videos, mirroring steps of
standard Neural Machine Translation which resulted in the PHOENIX4T dataset. In this
dataset, gloss information makes up the data. Further development came with the
introduction of sign language transformers [
        <xref ref-type="bibr" rid="ref48">48</xref>
        ] which can translate from spoken language
sentences to a 3D skeleton [
        <xref ref-type="bibr" rid="ref49">49</xref>
        ].
      </p>
      <p>
        To our knowledge, several attempts have been made at collecting sign language
data by combining computer vision and information captured through gloves. Using the
CopyCat system, Zafrulla et al. [
        <xref ref-type="bibr" rid="ref50">50</xref>
        ] collected 320 utterances from ASL signers. The
tools for data collection used were a camcorder and colored gloves containing
accelerometers providing information on acceleration, direction and rotation of hands.
Another tool is the AcceleGlove, an electronic glove placed on a signer’s hand and arm.
The glove collected information on hand movement, orientation and location in relation
to the body and recognized 176 signs in isolation [
        <xref ref-type="bibr" rid="ref51">51</xref>
        ] The AcceleGlove was combined
with a gesture recognition toolkit [
        <xref ref-type="bibr" rid="ref52">52</xref>
        ] by McGuire et al. [
        <xref ref-type="bibr" rid="ref45">45</xref>
        ] to record 665 utterances in
ASL to establish a pattern recognition framework to be expanded.
      </p>
      <p>
        The idea of associating glosses taken from a set of values, together with additional
information on the manual qualities of each sign, was developed in a 2020 master’s thesis
[
        <xref ref-type="bibr" rid="ref53">53</xref>
        ] and article [
        <xref ref-type="bibr" rid="ref54">54</xref>
        ]. The goal was the creation of a Universal Dependencies-compliant
resource for the syntactic annotation of LIS, i.e., the first LIS treebank. To create this
treebank, tasks of segmentation, annotation, POS tagging and parsing had to be carried
out. Particular attention was paid to the inclusion of unambiguous information in the
segmentation and annotation process. LIS videos were segmented on ELAN, after that,
each sign was analyzed on different levels. The first tier provided an unambiguous gloss
taken from the Dictionary or SpreadTheSign. The following tiers included information
on sign location and Universal Dependencies POS tag. Utterances were then transferred
into CoNLL-U format for the construction of dependency trees. During that step, gloss,
sign location and POS tag were transferred in the columns, together with a translation
Italian Sign Language (LIS) and Natural Language Processing: an Overview
into Italian. The treebank can be found on GitHub [
        <xref ref-type="bibr" rid="ref55">55</xref>
        ]. The main issue encountered in
this project is, once again, the lack of information on the annotation of non-manual
elements. Annotating on ELAN, which allows for simultaneous viewing of video and
annotation, is convenient. However, once we move to the level of syntactic annotation in
CoNLL-U format, the video material is no longer visible. Therefore, all information that
could be not codified in the tiers through word labels is virtually inaccessible.
      </p>
      <p>
        As regards available datasets, a list of sign languages recognition datasets with
information on data type and annotation systems, as well as related papers can be found in the
bibliography [
        <xref ref-type="bibr" rid="ref56">56</xref>
        ].
3
3.1
      </p>
      <sec id="sec-7-1">
        <title>Additional Challenges for LIS–IT and IT–LIS Translation</title>
      </sec>
    </sec>
    <sec id="sec-8">
      <title>Subdivision of Manual and Non-Manual Elements</title>
      <p>Sign transcription for data collection is not the only relevant problem to be faced for
sign language processing. On the one hand, the manual elements that make up a sign in
LIS have been divided into different parameters: handshape, place of articulation,
orientation and movement. On the other hand, non-manual elements are equally important in
utterance construction and sign disambiguation and include facial expression
(movements of eyebrows, eye, mouth, nose), body posture and movements.</p>
      <p>
        Once a methodology for sign transcription is identified, the second aspect to be
considered is the identification of the relevant elements in sign construction. Can sign
annotation be limited to only manual ones? Or are non-manual elements fundamental and
cannot be taken out of the discourse? This issue holds a central spot in Italian and
international contexts. Two main perspectives provide different answers to this issue.
‘Assimilationists’ show a tendency to focus on those aspects that draw sign languages closer
to spoken ones, thus giving priority to ‘standard signs’, which are easier to define and
mainly constituted by manual elements. ‘Non-assimilationists’ want to highlight the
language-specific properties of sign languages, such as the relevance of non-manual
elements [
        <xref ref-type="bibr" rid="ref57">57</xref>
        ].
      </p>
      <p>Whichever approach one may choose, the search for relevance remains central.
Ideally, manual and non-manual elements should be taken into consideration, thus providing
information on every aspect of the sign. However – given the current technologies
employed in this field and discussed in section 2.3., such as RGB cameras or radar sensors
– manual elements hold a central position in the investigation, temporarily winning over
non-manual ones.
3.2</p>
    </sec>
    <sec id="sec-9">
      <title>Challenges of LIS Utterance Structure Description</title>
      <p>Despite our non-assimilationist inclinations, it is undeniable that if the aim of a
process of data collection and annotation is MT, a dismissal of the pre-existing POS tagging
systems is counterproductive. For this reason, in this subsection, we examine recent
assimilationist observations on LIS utterances, by taking into consideration the manual and
non-manual levels. Utterances are based on signs, the entire body and non-manual
features such as facial expression, mouth actions, movements of the torso and eye gaze,
which is used to point at, describe or depict the referent.</p>
      <p>
        Sign languages have specific constructions: the structure of an utterance in LIS will
not mirror that of spoken Italian. It has been observed that the unmarked order of signs
in LIS is Subject-Object-Verb, describing LIS as a head-final language, where the most
meaningful element is found in final position [
        <xref ref-type="bibr" rid="ref58">58</xref>
        ]. Utterances are in most cases much
more complex as they function following pragmatic constraints based on iconicity and
the multifaceted structures of LIS have not been comprehensively defined up to this
point.
      </p>
      <p>Non-manual elements represent the hardest challenge in the representation and
recognition of sign languages for their non-segmentable nature. Facial expressions that also
includes eye gaze and mouth actions convey relevant information in co-occurrence with
signing that is hard to process as relevant but is crucial for the understanding of the
meaning of the single sign and the utterance.
4</p>
      <sec id="sec-9-1">
        <title>Conclusion</title>
        <p>This paper explores some of the main challenges in the field of sign language
processing. We have discussed the socio-political situation of the Deaf community and
considered how, behind the label of Deafness, there can be different users with different
needs. This highlights that the involvement of users is essential in every phase of research
and development. When creating sign language datasets, Deaf people should be involved
in collecting reliable data that represent sign language usage, making possible the
creation of appropriate computational models, interface design and, finally, of the overall
systems. Furthermore, the creation of datasets based on the processing of natural
conversation and long utterances will be necessary to go beyond the state of the art. Another
crucial step is annotation. The lack of a standard written form and the necessity of fluent
signers who annotate to produce the machine-readable inputs for training algorithms,
represent the main difficulties in applying NLP methods to sign languages. Only through
a multidimensional approach that combines NLP with computer vision and radar-based
technologies, as well as with the involvement of Deaf participants, could it be possible
to design effective technologies that could have an impact both at the social and the
linguistic level. On a social level, an effective translation could support the interactions
between hearing and Deaf people when the interpreting service is not available, for
example. On a linguistic level, sign language processing would play an important role in
sign language acquisition and learning.</p>
        <p>The proper development of this new technology will innovate along two main
directions: technological and social. Under the technological aspect, the collaboration of
different research approaches (radar technologies, artificial intelligence, and computer
vision research) can open new insights in understanding the shaping of machines to
respond to human beings’ needs. On the other hand, it will have high benefits in terms of
the inclusion of Deaf people.</p>
        <p>Italian Sign Language (LIS) and Natural Language Processing: an Overview 13</p>
        <p>Italian Sign Language (LIS) and Natural Language Processing: an Overview 15</p>
      </sec>
    </sec>
  </body>
  <back>
    <ref-list>
      <ref id="ref1">
        <mixed-citation>
          1. Istituto Nazionale di Statistica: Conoscere il Mondo della Disabilità: Persone, Relazioni e Istituzioni.
          <source>Istituto Nazionale di Statistica</source>
          , Roma (
          <year>2019</year>
          ).
        </mixed-citation>
      </ref>
      <ref id="ref2">
        <mixed-citation>
          2. Deafness and Hearing Loss https://www.who.int/news-room/fact-sheets/detail/deafness-and
          <string-name>
            <surname>-</surname>
          </string-name>
          hearing-loss,
          <source>last accessed</source>
          <year>2021</year>
          /10/02.
        </mixed-citation>
      </ref>
      <ref id="ref3">
        <mixed-citation>
          3.
          <string-name>
            <surname>Malerba</surname>
            ,
            <given-names>D.</given-names>
          </string-name>
          :
          <article-title>Sordità Percezione e Realtà nell'approccio pedagogico</article-title>
          .
          <source>Sapienza Università Editrice</source>
          , Roma (
          <year>2020</year>
          ).
        </mixed-citation>
      </ref>
      <ref id="ref4">
        <mixed-citation>
          4.
          <string-name>
            <surname>Fontana</surname>
            ,
            <given-names>S.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Mignosi</surname>
          </string-name>
          , E.: Segnare, Parlare, Intendersi: Modalità e forme.
          <source>Mimesis, Milano and Udine</source>
          (
          <year>2012</year>
          ).
        </mixed-citation>
      </ref>
      <ref id="ref5">
        <mixed-citation>
          5.
          <string-name>
            <surname>Onofrio</surname>
            ,
            <given-names>D.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Rinaldi</surname>
            ,
            <given-names>P.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Caselli</surname>
            ,
            <given-names>M.C.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Volterra</surname>
          </string-name>
          , V.:
          <article-title>Il Bilinguismo Bimodale dei Bambini Sordi: Aspetti Teorici ed Esperienze di Ricerca. Rivista di psicolinguistica applicate XIV (1</article-title>
          ),
          <fpage>25</fpage>
          -
          <lpage>42</lpage>
          . (
          <year>2014</year>
          ).
        </mixed-citation>
      </ref>
      <ref id="ref6">
        <mixed-citation>
          6.
          <string-name>
            <surname>Pinto</surname>
            <given-names>M. A.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Volterra</surname>
            <given-names>V.</given-names>
          </string-name>
          :
          <article-title>Bilinguismo lingue dei segni / lingue vocali: Aspetti educativi e psicolinguistici - Sign languages, spoken languages and bilingualism: educational and psycholinguistics issues</article-title>
          .
          <source>Italian Journal of Applied</source>
          Psycholinguistics VIII (
          <year>2008</year>
          ).
        </mixed-citation>
      </ref>
      <ref id="ref7">
        <mixed-citation>
          7.
          <string-name>
            <surname>Capek</surname>
            ,
            <given-names>C. M.</given-names>
          </string-name>
          et al.:
          <article-title>Hand and Mouth: Cortical Correlates of Lexical Processing in British Sign Language and Speechreading English</article-title>
          .
          <source>Journal of Cognitive Neuroscience</source>
          <volume>20</volume>
          (
          <issue>7</issue>
          ), pp.
          <fpage>1220</fpage>
          -
          <lpage>1233</lpage>
          (
          <year>2008</year>
          ).
        </mixed-citation>
      </ref>
      <ref id="ref8">
        <mixed-citation>
          8.
          <string-name>
            <surname>Maragna</surname>
            <given-names>S.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Roccaforte</surname>
            <given-names>M.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Tomasuolo</surname>
            <given-names>E.</given-names>
          </string-name>
          :
          <article-title>Una Didattica innovativa per l'apprendente sordo</article-title>
          .
          <source>FrancoAngeli</source>
          ,
          <string-name>
            <surname>Milano</surname>
          </string-name>
          (
          <year>2013</year>
          ).
        </mixed-citation>
      </ref>
      <ref id="ref9">
        <mixed-citation>
          9.
          <string-name>
            <surname>Fontana</surname>
            <given-names>S</given-names>
          </string-name>
          :
          <article-title>Metalinguistic Awareness in sign language: epistemological considerations, in Pinto M.A</article-title>
          .,
          <string-name>
            <surname>Rinaldi</surname>
            <given-names>P</given-names>
          </string-name>
          . (eds)
          <article-title>Metalinguistic Awareness and Bimodal Bilingualism: Studies on Deaf and Hearing Subjects</article-title>
          ,
          <source>Italian Journal of Applied Pysholinguistics XVI(2)</source>
          (
          <year>2016</year>
          ).
        </mixed-citation>
      </ref>
      <ref id="ref10">
        <mixed-citation>
          10. Ethnologue Homepage, https://www.ethnologue.com/,
          <source>last accessed</source>
          <year>2021</year>
          /10/08.
        </mixed-citation>
      </ref>
      <ref id="ref11">
        <mixed-citation>
          11.
          <string-name>
            <surname>Scelzi</surname>
          </string-name>
          , R.:
          <article-title>Le component non manuali della LIS</article-title>
          .
          <source>Studi di Glottodidattica (1)</source>
          ,
          <fpage>261</fpage>
          -
          <lpage>291</lpage>
          (
          <year>2010</year>
          ).
        </mixed-citation>
      </ref>
      <ref id="ref12">
        <mixed-citation>
          12.
          <string-name>
            <surname>Romeo</surname>
            ,
            <given-names>O.</given-names>
          </string-name>
          :
          <article-title>Dizionario dei Segni</article-title>
          . Zanichelli Editore
          <string-name>
            <given-names>S. p. A.</given-names>
            ,
            <surname>Bologna</surname>
          </string-name>
          (
          <year>1991</year>
          ).
        </mixed-citation>
      </ref>
      <ref id="ref13">
        <mixed-citation>
          13.
          <string-name>
            <surname>Fontana</surname>
            <given-names>S.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Corazza</surname>
            <given-names>S.</given-names>
          </string-name>
          ,
          <string-name>
            <given-names>Boyes Braem P.</given-names>
            ,
            <surname>Volterra</surname>
          </string-name>
          <string-name>
            <surname>V.</surname>
          </string-name>
          :
          <article-title>Language research and language community change: Italian Sign Language 1981-2013</article-title>
          , in
          <source>International Journal of the Sociology of Language</source>
          , pp.
          <fpage>1</fpage>
          -
          <lpage>30</lpage>
          . De Gruyter Mouton, Berlin (
          <year>2015</year>
          ).
        </mixed-citation>
      </ref>
      <ref id="ref14">
        <mixed-citation>
          14.
          <string-name>
            <surname>Tomasuolo</surname>
            <given-names>E.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Gulli</surname>
            <given-names>T.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Volterra</surname>
            <given-names>V.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Fontana</surname>
            <given-names>S.:</given-names>
          </string-name>
          <article-title>The Italian Deaf Community at the time of Coronavirus. Frontiers in Sociology (</article-title>
          <year>2021</year>
          ). DOI: https://doi.org/10.3389/fsoc.
          <year>2020</year>
          .612559
        </mixed-citation>
      </ref>
      <ref id="ref15">
        <mixed-citation>
          15. Italian Official Gazette,
          <source>Decree Law of March 22nd</source>
          <year>2021</year>
          , n.41, https://www.gazzettaufficiale.it/atto/serie_generale/caricaDettaglioAtto/originario?atto.
          <source>dataPubblicazioneGazzetta=2021-05-21&amp;atto.codiceRedazionale=21A03181&amp;elenco30giorni=false, last accessed</source>
          <year>2021</year>
          /10/01.
        </mixed-citation>
      </ref>
      <ref id="ref16">
        <mixed-citation>
          16.
          <string-name>
            <surname>Volterra</surname>
          </string-name>
          , V.:
          <article-title>La Lingua Italiana dei Segni</article-title>
          .
          <article-title>La comunicazione visivo-gestuale dei sordi</article-title>
          .
          <source>Il Mulino</source>
          , Bologna (
          <year>1987</year>
          ).
        </mixed-citation>
      </ref>
      <ref id="ref17">
        <mixed-citation>
          17.
          <string-name>
            <surname>Fontana</surname>
            ,
            <given-names>S.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Roccaforte</surname>
            ,
            <given-names>M.</given-names>
          </string-name>
          :
          <article-title>Oltre l'approccio assimilazionista nella descrizione LIS</article-title>
          .
          <article-title>Quando la prassi comunicativa diventa norma</article-title>
          .
          <source>In: LII Congresso Internazionale di Studi della Società di Linguistica Italiana</source>
          , pp.
          <fpage>273</fpage>
          -
          <lpage>286</lpage>
          . Bern (
          <year>2019</year>
          ).
        </mixed-citation>
      </ref>
      <ref id="ref18">
        <mixed-citation>
          18.
          <string-name>
            <given-names>Di</given-names>
            <surname>Renzo</surname>
          </string-name>
          ,
          <string-name>
            <given-names>A.</given-names>
            ,
            <surname>Lamano</surname>
          </string-name>
          ,
          <string-name>
            <given-names>L.</given-names>
            ,
            <surname>Lucioli</surname>
          </string-name>
          ,
          <string-name>
            <given-names>T.</given-names>
            ,
            <surname>Pennacchi</surname>
          </string-name>
          ,
          <string-name>
            <given-names>B.</given-names>
            ,
            <surname>Ponzo</surname>
          </string-name>
          ,
          <string-name>
            <surname>L.</surname>
          </string-name>
          :
          <article-title>Italian Sign Language (LIS): can we write it and transcribe it with SignWriting</article-title>
          ? In: Editor, F.,
          <string-name>
            <surname>Editor</surname>
          </string-name>
          ,
          <source>S. (eds.) 2nd Workshop on the Representation and processing of Sign Languages: Lexicografic matters and didactic scenarios</source>
          ,
          <source>International Conference on Language Resources and Evaluation -LREC</source>
          <year>2006</year>
          ,
          <string-name>
            <surname>Genoa</surname>
          </string-name>
          (
          <year>2006</year>
          ).
        </mixed-citation>
      </ref>
      <ref id="ref19">
        <mixed-citation>
          19.
          <string-name>
            <surname>Volterra</surname>
            <given-names>V.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Roccaforte</surname>
            <given-names>M.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Di</surname>
            <given-names>Renzo A.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Fontana</surname>
            <given-names>S.</given-names>
          </string-name>
          : Descrivere Lingue dei Segni.
          <article-title>Una prospettiva sociosemiotica</article-title>
          .
          <source>Il Mulino</source>
          , Bologna (
          <year>2019</year>
          ).
        </mixed-citation>
      </ref>
      <ref id="ref20">
        <mixed-citation>
          20.
          <string-name>
            <given-names>Antinoro</given-names>
            <surname>Pizzuto</surname>
          </string-name>
          ,
          <string-name>
            <given-names>E.</given-names>
            ,
            <surname>Chiari</surname>
          </string-name>
          ,
          <string-name>
            <given-names>I.</given-names>
            ,
            <surname>Rossini</surname>
          </string-name>
          ,
          <string-name>
            <surname>P:</surname>
          </string-name>
          <article-title>The Representation Issue and its Multifaceted Aspects in Constructing Sign Language Corpora: Questions, Answers, Further Problems</article-title>
          .
          <source>In: Proceedings of the LREC2008 3rd Workshop on the Representation and Processing of Sign Languages: Construction and Exploitation of Sign Language Corpora</source>
          , pp.
          <fpage>150</fpage>
          -
          <lpage>150</lpage>
          . ELRA,
          <string-name>
            <surname>Marrakech</surname>
          </string-name>
          (
          <year>2009</year>
          ).
        </mixed-citation>
      </ref>
      <ref id="ref21">
        <mixed-citation>
          21.
          <string-name>
            <surname>Mitkov</surname>
            ,
            <given-names>R.</given-names>
          </string-name>
          :
          <source>The Oxford Handbook of Computational Linguistics</source>
          . Oxford University Press, New York (
          <year>2003</year>
          ).
        </mixed-citation>
      </ref>
      <ref id="ref22">
        <mixed-citation>
          22.
          <string-name>
            <surname>Camgöz</surname>
            ,
            <given-names>N.C.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Koller</surname>
            ,
            <given-names>O.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Hadfield</surname>
            ,
            <given-names>S.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Bowden</surname>
            ,
            <given-names>R.</given-names>
          </string-name>
          <article-title>: Multi-channel Transformers for Multi-articulatory Sign Language Translation</article-title>
          . (
          <year>2020</year>
          ).
        </mixed-citation>
      </ref>
      <ref id="ref23">
        <mixed-citation>
          23.
          <string-name>
            <surname>Manaris</surname>
            ,
            <given-names>B.</given-names>
          </string-name>
          :
          <string-name>
            <surname>Natural Language Processing: A Human-Computer Interaction</surname>
          </string-name>
          Perspective.
          <source>Advances in Computes 47</source>
          ,
          <fpage>1</fpage>
          -
          <lpage>66</lpage>
          (
          <year>1998</year>
          ).
        </mixed-citation>
      </ref>
      <ref id="ref24">
        <mixed-citation>
          24.
          <string-name>
            <surname>Farooq</surname>
            ,
            <given-names>U.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Rahim</surname>
            ,
            <given-names>M.S.M.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Sabir</surname>
            ,
            <given-names>N.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Hussain</surname>
            ,
            <given-names>A.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Abid</surname>
            ,
            <given-names>A.</given-names>
          </string-name>
          :
          <article-title>Advances in machine translation for sign language: approaches, limitations, and challenges</article-title>
          .
          <source>Neural Computing and Applications</source>
          (
          <year>2021</year>
          ).
        </mixed-citation>
      </ref>
      <ref id="ref25">
        <mixed-citation>
          25.
          <string-name>
            <surname>Nobis</surname>
            ,
            <given-names>F.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Geisslinger</surname>
            ,
            <given-names>M.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Weber</surname>
            ,
            <given-names>M.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Johannes</surname>
            ,
            <given-names>B.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Lienkamp</surname>
            ,
            <given-names>M.:</given-names>
          </string-name>
          <article-title>A Deep Learning-based Radar and Camera Sensor Fusion Architecture for Object Detection</article-title>
          . (
          <year>2020</year>
          ).
        </mixed-citation>
      </ref>
      <ref id="ref26">
        <mixed-citation>
          26.
          <string-name>
            <surname>Huang</surname>
            ,
            <given-names>J.</given-names>
          </string-name>
          , Wen-gang,
          <string-name>
            <given-names>Z.</given-names>
            ,
            <surname>Qilin</surname>
          </string-name>
          ,
          <string-name>
            <given-names>Z.</given-names>
            ,
            <surname>Houqiang</surname>
          </string-name>
          ,
          <string-name>
            <given-names>L.</given-names>
            ,
            <surname>Weiping</surname>
          </string-name>
          ,
          <string-name>
            <surname>L.</surname>
          </string-name>
          :
          <article-title>Video-based Sign Language Recognition without Temporal Segmentation</article-title>
          .
          <source>AAAI</source>
          (
          <year>2018</year>
          ).
        </mixed-citation>
      </ref>
      <ref id="ref27">
        <mixed-citation>
          27.
          <string-name>
            <surname>Santhalingam</surname>
            ,
            <given-names>P.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Du</surname>
            ,
            <given-names>R.</given-names>
          </string-name>
          <string-name>
            <surname>Wilkerson</surname>
            ,
            <given-names>R.</given-names>
          </string-name>
          , Al Amin,
          <string-name>
            <given-names>H.</given-names>
            ,
            <surname>Ding</surname>
          </string-name>
          ,
          <string-name>
            <given-names>Z.</given-names>
            ,
            <surname>Parth</surname>
          </string-name>
          ,
          <string-name>
            <given-names>P.</given-names>
            ,
            <surname>Rangwala</surname>
          </string-name>
          ,
          <string-name>
            <given-names>H.</given-names>
            ,
            <surname>Kushalnagar</surname>
          </string-name>
          , R.:
          <article-title>Expressive ASL Recognition using Millimeter-wave Wireless Signals</article-title>
          .
          <year>2020</year>
          17th Annual IEEE International Conference on Sensing, Communication, and
          <string-name>
            <surname>Networking</surname>
          </string-name>
          (SECON), pp.
          <fpage>1</fpage>
          -
          <lpage>9</lpage>
          . (
          <year>2020</year>
          ).
        </mixed-citation>
      </ref>
      <ref id="ref28">
        <mixed-citation>
          28.
          <string-name>
            <surname>Lombardo</surname>
            ,
            <given-names>V.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Battaglino</surname>
            ,
            <given-names>C.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Damiano</surname>
            ,
            <given-names>R.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Nunnari</surname>
            ,
            <given-names>F.</given-names>
          </string-name>
          :
          <article-title>An Avatar-based Interface for the Italian Sign Language</article-title>
          . International Conference on Complex,
          <source>Intelligent and Software Intensive Systems</source>
          ,
          <string-name>
            <surname>CISIS</surname>
          </string-name>
          <year>2011</year>
          . Korean Bible University, Seoul (
          <year>2011</year>
          ).
        </mixed-citation>
      </ref>
      <ref id="ref29">
        <mixed-citation>
          29.
          <string-name>
            <surname>Battaglino</surname>
            ,
            <given-names>C.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Geraci</surname>
            ,
            <given-names>C.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Lombardo</surname>
            ,
            <given-names>V.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Mazzei</surname>
            ,
            <given-names>A.</given-names>
          </string-name>
          :
          <article-title>Prototyping and Preliminary Evaluation of Sign Language Translation System in the Railway Domain</article-title>
          .
          <source>International Conference on Universal Access in Human-Computer Interaction</source>
          , pp.
          <fpage>229</fpage>
          -
          <lpage>350</lpage>
          . (
          <year>2015</year>
          ).
        </mixed-citation>
      </ref>
      <ref id="ref30">
        <mixed-citation>
          30.
          <string-name>
            <surname>Mazzei</surname>
            ,
            <given-names>A.</given-names>
          </string-name>
          :
          <article-title>Translating Italian to LIS in the Rail Stations</article-title>
          .
          <source>Proceedings of the 15th European Workshop on Natural Language Generation</source>
          , ENLG. Association for Computational Linguistics,
          <string-name>
            <surname>Brighton</surname>
          </string-name>
          (
          <year>2015</year>
          ).
        </mixed-citation>
      </ref>
      <ref id="ref31">
        <mixed-citation>
          31.
          <string-name>
            <surname>Mazzei</surname>
            ,
            <given-names>A.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Lesmo</surname>
            ,
            <given-names>L.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Battaglino</surname>
            ,
            <given-names>C.</given-names>
          </string-name>
          <string-name>
            <surname>Vendrame</surname>
            ,
            <given-names>M.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Bucciarelli</surname>
            ,
            <given-names>M.</given-names>
          </string-name>
          :
          <article-title>Deep Natural Language Processing for Italian Sign Language Translation</article-title>
          . In:
          <string-name>
            <surname>Baldoni</surname>
            <given-names>M.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Baroglio</surname>
            <given-names>C.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Boella</surname>
            <given-names>G.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Micalizio</surname>
            <given-names>R</given-names>
          </string-name>
          . (eds) Advances in Artificial Intelligence,
          <source>AI*IA</source>
          <year>2013</year>
          , vol
          <volume>8249</volume>
          . Springer, Cham.
        </mixed-citation>
      </ref>
      <ref id="ref32">
        <mixed-citation>
          32.
          <string-name>
            <surname>Branchini</surname>
            ,
            <given-names>C.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Mantovan</surname>
            ,
            <given-names>L.</given-names>
          </string-name>
          ,
          <article-title>A Grammar of Italian Sign Language (LIS). Published with open access at https://www.sign-hub</article-title>
          .eu/grammardetail/UUIDGRMM-e0adecd1
          <string-name>
            <surname>-</surname>
          </string-name>
          c01e
          <string-name>
            <surname>-</surname>
          </string-name>
          47ef
          <string-name>
            <surname>-</surname>
          </string-name>
          b2c0-
          <fpage>c2d6a4ce45dc</fpage>
          (
          <year>2020</year>
          ).
        </mixed-citation>
      </ref>
      <ref id="ref33">
        <mixed-citation>
          33.
          <string-name>
            <surname>The</surname>
            <given-names>SIGN</given-names>
          </string-name>
          -HUB Project Homepage, https://www.sign-hub.
          <source>eu/project, last accessed</source>
          <year>2021</year>
          /10/04.
        </mixed-citation>
      </ref>
      <ref id="ref34">
        <mixed-citation>
          34.
          <string-name>
            <surname>Radutzky</surname>
          </string-name>
          , E.:
          <article-title>Dizionario bilingue elementare della lingua italiana dei segni</article-title>
          .
          <source>Oltre 2500 significati. Con DVD-ROM. Kappa</source>
          (
          <year>1992</year>
          ).
        </mixed-citation>
      </ref>
      <ref id="ref35">
        <mixed-citation>
          35. SpreadTheSign Homepage, https://www.spreadthesign.com/it.it/search/, last accessed
          <year>2021</year>
          /10/04.
        </mixed-citation>
      </ref>
      <ref id="ref36">
        <mixed-citation>
          36.
          <string-name>
            <surname>Hilzensauer</surname>
            ,
            <given-names>M.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Krammer</surname>
            ,
            <given-names>K.</given-names>
          </string-name>
          :
          <article-title>A multilingual dictionary for sign languages: “SpreadTheSign”</article-title>
          .
          <source>The 8th annual International Conference of Education, Research and Innovation, ICERI2015</source>
          ,
          <string-name>
            <surname>Seville</surname>
          </string-name>
          (
          <year>2015</year>
          ).
        </mixed-citation>
      </ref>
      <ref id="ref37">
        <mixed-citation>
          37.
          <string-name>
            <surname>Chesi</surname>
            ,
            <given-names>C.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Geraci</surname>
            ,
            <given-names>C.</given-names>
          </string-name>
          : Segni al computer.
          <article-title>Manuale di documentazione della Lingua Italiana dei Segni e alcune applicazioni computazionali</article-title>
          .
          <source>Siena</source>
          (
          <year>2009</year>
          ).
        </mixed-citation>
      </ref>
      <ref id="ref38">
        <mixed-citation>
          38.
          <string-name>
            <surname>Wittenburg</surname>
            ,
            <given-names>P.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Brugman</surname>
            ,
            <given-names>H.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Russel</surname>
            ,
            <given-names>A.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Klassmann</surname>
            ,
            <given-names>A.</given-names>
          </string-name>
          , and
          <string-name>
            <surname>Sloetjes</surname>
          </string-name>
          , H.:
          <article-title>ELAN: a professional framework for multimodality research</article-title>
          .
          <source>Proceedings of the Fifth International Conference on Language Resources and Evaluation</source>
          , LREC'
          <fpage>06</fpage>
          .
          <string-name>
            <surname>European Language Resources Association</surname>
          </string-name>
          (ELRA),
          <source>Genoa</source>
          (
          <year>2006</year>
          ).
        </mixed-citation>
      </ref>
      <ref id="ref39">
        <mixed-citation>
          39.
          <string-name>
            <surname>Mesch</surname>
            ,
            <given-names>J.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Wallin</surname>
            ,
            <given-names>L.</given-names>
          </string-name>
          :
          <article-title>Gloss annotations in the Swedish Sign Language Corpus</article-title>
          .
          <source>International Journal of Corpus Linguistics (20)</source>
          , pp.
          <fpage>102</fpage>
          -
          <lpage>120</lpage>
          (
          <year>2015</year>
          ).
        </mixed-citation>
      </ref>
      <ref id="ref40">
        <mixed-citation>
          40.
          <article-title>Svenskt teckenspråkslexikon (Swedish Sign Language Dictionary)</article-title>
          , http:// teckensprakslexikon.su.se/, last accessed
          <year>2021</year>
          /10/08.
        </mixed-citation>
      </ref>
      <ref id="ref41">
        <mixed-citation>
          41. BSLCP Homepage, https://bslcorpusproject.org/,
          <source>last accessed</source>
          <year>2021</year>
          /10/08.
        </mixed-citation>
      </ref>
      <ref id="ref42">
        <mixed-citation>
          42.
          <string-name>
            <surname>Schembri</surname>
            ,
            <given-names>A.</given-names>
          </string-name>
          :
          <article-title>British Sign Language Corpus Project: Open Access Archives and the Observer's</article-title>
          <string-name>
            <surname>Paradox.</surname>
          </string-name>
          (
          <year>2015</year>
          ).
        </mixed-citation>
      </ref>
      <ref id="ref43">
        <mixed-citation>
          43.
          <string-name>
            <surname>Pfau</surname>
          </string-name>
          , R.:
          <source>Annotation in ELAN, Version</source>
          <volume>1</volume>
          .1.
          <string-name>
            <surname>SIGN-HUB</surname>
            <given-names>Project</given-names>
          </string-name>
          , European
          <string-name>
            <surname>Commission</surname>
          </string-name>
          (
          <year>2018</year>
          ).
        </mixed-citation>
      </ref>
      <ref id="ref44">
        <mixed-citation>
          44.
          <string-name>
            <surname>Drew</surname>
            ,
            <given-names>P.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Ney</surname>
          </string-name>
          , H.:
          <article-title>Towards Automatic Sign Language Annotation for the ELAN Tool</article-title>
          .
          <article-title>(</article-title>
          <year>2008</year>
          ).
        </mixed-citation>
      </ref>
      <ref id="ref45">
        <mixed-citation>
          45.
          <string-name>
            <surname>McGuire</surname>
          </string-name>
          et al.:
          <article-title>Towards a One-way American sign language translator</article-title>
          .
          <source>In Proceedings of the Sixth IEEE International Conference on Automatic Face and Gesture Recognition</source>
          , pp.
          <fpage>620</fpage>
          -
          <lpage>625</lpage>
          . Seoul (
          <year>2004</year>
          ).
        </mixed-citation>
      </ref>
      <ref id="ref46">
        <mixed-citation>
          46.
          <string-name>
            <surname>Santhalingam</surname>
            ,
            <given-names>P.S.</given-names>
          </string-name>
          , et al.:
          <article-title>Expressive ASL Recognition using Millimeterwave Wireless Signals</article-title>
          .
          <source>In 17th Annual IEEE International Conference on Sensing, Communication and Networking</source>
          , pp.
          <fpage>1</fpage>
          -
          <lpage>9</lpage>
          . SECON,
          <string-name>
            <surname>Como</surname>
          </string-name>
          (
          <year>2020</year>
          ).
        </mixed-citation>
      </ref>
      <ref id="ref47">
        <mixed-citation>
          47.
          <string-name>
            <surname>Camgoz</surname>
            ,
            <given-names>N.C.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Hadfield</surname>
            ,
            <given-names>S.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Koller</surname>
            ,
            <given-names>O.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Ney</surname>
            ,
            <given-names>H.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Bowden</surname>
            ,
            <given-names>R.</given-names>
          </string-name>
          :
          <article-title>Neural Sign Language Translation</article-title>
          .
          <source>In: Conference on Computer Vision and Pattern Recognition</source>
          , pp.
          <fpage>7784</fpage>
          -
          <lpage>7793</lpage>
          . (
          <year>2018</year>
          ).
        </mixed-citation>
      </ref>
      <ref id="ref48">
        <mixed-citation>
          48.
          <string-name>
            <surname>Camgoz</surname>
            ,
            <given-names>N.C.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Hadfield</surname>
            ,
            <given-names>S.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Koller</surname>
            ,
            <given-names>O.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Bowden</surname>
          </string-name>
          , R.:
          <article-title>Sign Language Transformers: Joint End-to-End Sign Language Recognition and Translation</article-title>
          .
          <source>In: Conference on Computer Vision and Pattern Recognition</source>
          . (
          <year>2020</year>
          ).
        </mixed-citation>
      </ref>
      <ref id="ref49">
        <mixed-citation>
          49.
          <string-name>
            <surname>Saunders</surname>
            ,
            <given-names>B.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Camgoz</surname>
            ,
            <given-names>N.C.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Bowden</surname>
          </string-name>
          , R.:
          <article-title>Progressive Transformers for Endto-End Sign Language Production</article-title>
          .
          <source>In: European Conference on Computer Vision</source>
          (ECCV) (
          <year>2020</year>
          ).
        </mixed-citation>
      </ref>
      <ref id="ref50">
        <mixed-citation>
          50.
          <string-name>
            <surname>Zafrulla</surname>
          </string-name>
          et al.:
          <article-title>American Sign Language Recognition with the Kinect</article-title>
          .
          <source>In: Proceedings of the 13th International Conference on Multimodal Interfaces</source>
          ,
          <string-name>
            <surname>ICMI</surname>
          </string-name>
          <year>2011</year>
          .
          <string-name>
            <surname>Alicante</surname>
          </string-name>
          (
          <year>2011</year>
          ).
        </mixed-citation>
      </ref>
      <ref id="ref51">
        <mixed-citation>
          51.
          <string-name>
            <surname>Hernandez-Rebollar</surname>
            ,
            <given-names>J.L.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Kyriakopoulos</surname>
            ,
            <given-names>N.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Lindeman</surname>
            ,
            <given-names>R.W.:</given-names>
          </string-name>
          <article-title>A New Instrumented Approach for Translating American Sign Language into Sound And Tex</article-title>
          .
          <source>In: Proceedings of the Sixth IEEE International Conference on Automatic Face and Gesture Recognition</source>
          . Seoul (
          <year>2004</year>
          ).
        </mixed-citation>
      </ref>
      <ref id="ref52">
        <mixed-citation>
          52.
          <string-name>
            <surname>Westeyn</surname>
            ,
            <given-names>T.L.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Brashear</surname>
            ,
            <given-names>H.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Atrash</surname>
            ,
            <given-names>A.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Starner</surname>
            ,
            <given-names>T.</given-names>
          </string-name>
          :
          <article-title>Georgia tech gesture toolkit: supporting experiments in gesture recognition</article-title>
          .
          <source>In Proceedings of the 5th International Conference on Multimodal Interfaces</source>
          ,
          <string-name>
            <surname>ICMI</surname>
          </string-name>
          <year>2003</year>
          . Vancouver (
          <year>2003</year>
          ).
        </mixed-citation>
      </ref>
      <ref id="ref53">
        <mixed-citation>
          53.
          <string-name>
            <surname>Caligiore</surname>
          </string-name>
          , G.:
          <article-title>Universal Dependencies for Italian Sign Language: a treebank from the storytelling domain</article-title>
          . University of Turin, Turin (
          <year>2020</year>
          ).
        </mixed-citation>
      </ref>
      <ref id="ref54">
        <mixed-citation>
          54.
          <string-name>
            <surname>Caligiore</surname>
            ,
            <given-names>G.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Bosco</surname>
            ,
            <given-names>C.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Mazzei</surname>
            ,
            <given-names>A.</given-names>
          </string-name>
          :
          <article-title>Building a treebank in Universal Dependencies for Italian Sign Language</article-title>
          .
          <source>Proceedings of Seventh Italian Conference on Computational Linguistics, CLIC-2020</source>
          , vol.
          <volume>2769</volume>
          , (
          <year>2021</year>
          ).
        </mixed-citation>
      </ref>
      <ref id="ref55">
        <mixed-citation>
          55.
          <string-name>
            <surname>Caligiore</surname>
            ,
            <given-names>G.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Bosco</surname>
            ,
            <given-names>C.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Mazzei</surname>
            ,
            <given-names>A.</given-names>
          </string-name>
          :
          <article-title>LIS-UD</article-title>
          . https://github.com/alexmazzei/LIS-UD.git, (
          <year>2020</year>
          ),
          <source>last accessed</source>
          <year>2020</year>
          /10/06.
        </mixed-citation>
      </ref>
      <ref id="ref56">
        <mixed-citation>
          56. Sign language datasets http://facundoq.github.io/guides/sign_language_datasets/slr, last accessed
          <year>2021</year>
          /10/03.
        </mixed-citation>
      </ref>
      <ref id="ref57">
        <mixed-citation>
          57.
          <string-name>
            <surname>Scelzi</surname>
          </string-name>
          , R.:
          <article-title>Le Componenti non Manuali (CNM) della LIS</article-title>
          .
          <source>In: Studi di Glottodidattica</source>
          , pp.
          <fpage>261</fpage>
          -
          <lpage>291</lpage>
          . (
          <year>2010</year>
          ).
        </mixed-citation>
      </ref>
      <ref id="ref58">
        <mixed-citation>
          58.
          <string-name>
            <surname>Geraci</surname>
            ,
            <given-names>C.:</given-names>
          </string-name>
          <article-title>L'ordine dei segni nella LIS (Lingua dei Segni Italiana)</article-title>
          . University of Milan, Milan (
          <year>2002</year>
          ).
        </mixed-citation>
      </ref>
    </ref-list>
  </back>
</article>