<!DOCTYPE article PUBLIC "-//NLM//DTD JATS (Z39.96) Journal Archiving and Interchange DTD v1.0 20120330//EN" "JATS-archivearticle1.dtd">
<article xmlns:xlink="http://www.w3.org/1999/xlink">
  <front>
    <journal-meta />
    <article-meta>
      <title-group>
        <article-title>Raising Interest and Collecting Suggestions on the EVALITA Evaluation Campaign</article-title>
      </title-group>
      <contrib-group>
        <contrib contrib-type="author">
          <string-name>Rachele Sprugnoli</string-name>
          <email>sprugnoli@fbk.eu</email>
          <xref ref-type="aff" rid="aff0">0</xref>
        </contrib>
        <contrib contrib-type="author">
          <string-name>Viviana Patti</string-name>
          <email>patti@di.unito.it</email>
          <xref ref-type="aff" rid="aff2">2</xref>
        </contrib>
        <contrib contrib-type="author">
          <string-name>Franco Cutugno</string-name>
          <email>cutugno@unina.it</email>
          <xref ref-type="aff" rid="aff1">1</xref>
        </contrib>
        <aff id="aff0">
          <label>0</label>
          <institution>FBK and University of Trento</institution>
          ,
          <addr-line>Via Sommarive, 38123 Trento</addr-line>
          ,
          <country country="IT">Italy</country>
        </aff>
        <aff id="aff1">
          <label>1</label>
          <institution>University of Naples Federico II</institution>
          ,
          <addr-line>Via Claudio 21, 80126 Naples</addr-line>
          ,
          <country country="IT">Italy</country>
        </aff>
        <aff id="aff2">
          <label>2</label>
          <institution>University of Turin</institution>
          ,
          <addr-line>c.so Svizzera 185, I-10149 Torino</addr-line>
          ,
          <country country="IT">Italy</country>
        </aff>
      </contrib-group>
      <abstract>
        <p>This paper describes the design and reports the results of two questionnaires. The first of these questionnaires was created to collect information about the interest of industrial companies in the field of Italian text/speech analytics towards the evaluation campaign EVALITA; the second to gather comments and suggestions for the future of the evaluation and of its final workshop from the participants and the organizers of the campaign on the last two editions (2011 and 2014). Novelties introduced in the organization of EVALITA 2016 on the basis of the questionnaires results are also reported.</p>
      </abstract>
    </article-meta>
  </front>
  <body>
    <sec id="sec-1">
      <title>Introduction</title>
      <p>
        (a)
(b)
The EUROMAP analysis, dated back to 2003, detected structural limits in the Italian situation regarding
the Human Language Technology (HLT) market
        <xref ref-type="bibr" rid="ref20">(Joscelyne and Lockwood, 2003)</xref>
        . Within that study, 26
Italian HLT suppliers are listed: in 2015, at the time of questionnaire development, only 13 of them were
still active. The dynamism of the HLT market in Italy is confirmed by a more recent survey where the
activities of 35 Italian enterprises are reported
        <xref ref-type="bibr" rid="ref15 ref26 ref5">(Di Carlo and Paoloni, 2009)</xref>
        : only 18 were still operative
in 2015.
      </p>
      <p>
        Starting from the active private companies present in the aforementioned surveys, we created a
repository of enterprises working on Italian text and speech technologies in Italy and abroad. In order to find
new enterprises not listed in previous surveys, we took advantage of online repositories (e.g. AngelList2
and CrunchBase3) and of extensive searches on the Web. Our final list included 115 companies among
which 57 are not based in Italy. This high number of enterprises dealing with Italian also outside the
national borders, reinforces one of the findings of the 2014 Alta Plana survey
        <xref ref-type="bibr" rid="ref17">(Grimes, 2014)</xref>
        4 that
provides a detailed analysis of text analytics market thanks to the answers given to a questionnaire dedicated
to technology and solution providers. No Italian company took part in that investigation but Italian
resulted as the fourth most analyzed language other than English (after Spanish, German and French) and
registered an estimated growth of +11% in two years.
      </p>
      <p>All the companies in our repository were directly contacted via email and asked to fill in the online
questionnaire. After an introductory description and the privacy statement, the questionnaire was
divided into three sections and included 18 questions. In the first section, we collected information about
the company such as its size and nationality; the second had the aim of assessing the interest towards
evaluation campaigns in general and towards a possible future participation in EVALITA. Finally, in the
third section we collected suggestions for the next edition of EVALITA.</p>
      <p>We collected responses from 39 private companies (response rate of 33.9%)5: 25 based in Italy
(especially in north and central regions) and the rest in other 9 countries6. 27 companies work on text
technologies, 2 on speech technologies and the remaining declares to do business in both sectors. The
great majority of companies (84.6%) has less than 50 employees and, more specifically, 43.6% of them
are start-up.</p>
      <p>
        Around 80% of respondents thinks that initiatives for the evaluation of NLP and speech tools are useful
for companies and expresses the interest in participating in EVALITA in the future. Motivations behind
the negative responses to this last point are related to the fact that the participation to a campaign is
considered very time-consuming and also a reputation risk in case of bad results. In addition, EVALITA
is perceived as too academically oriented, too focused on general (i.e. non application-oriented) tasks
2https://angel.co/
3https://www.crunchbase.com/
4http://altaplana.com/TA2014
5This response rate is in line with the rates reported in the literature on surveys distributed through emails, see
        <xref ref-type="bibr" rid="ref21 ref7">(Kaplowitz
et al., 2004; Baruch and Holtom, 2008)</xref>
        among others, and with the ones reported in the papers cited in Section 1.
      </p>
      <p>6Belgium, Finland, France, Netherlands, Russia, Spain, USA, Sweden, and Switzerland.
and with a limited impact on media. This last problem seems to be confirmed by the percentage of
respondents who were not aware of the existence of EVALITA before starting the questionnaire, i.e.
38.5% with 24.1% among Italian companies.</p>
      <p>For each of the questions regarding the suggestions for the next campaign (third section), we provided
a list of pre-defined options, so to speed up the questionnaire completion, together with a open field
for optional additional feedback. Participants could select more than one option. First of all we asked
what would encourage and what would prevent the company from participating in the next EVALITA
campaign. The possibility of using training and test data also for commercial purposes and the presence
of tasks related to the domains of interest for the company have been the most voted options followed
by the possibility of advertising for the company during the final workshop (for example by means of
exhibition stands or round tables) and the anonymisation of the results so avoiding negative effects on
the company image. On the contrary, the lack of time and/or funds is seen as the major obstacle.</p>
      <p>Favorite domains and tasks for companies participating in the questionnaire are shown in Figure 1.
Social media and news resulted to be the most popular among the domains of interest, followed by
humanities, politics and public administration. Domains included in the “Other” category are survey
analysis, financial but also public transport and information technology. For what concerns the tasks
of interest, sentiment analysis and named entity recognition were the top voted tasks, but a significant
interest has been expressed also about content analytics and Word Sense Disambiguation (WSD). In the
“Other” category, respondents suggested new tasks such as dialogue analysis, social-network analysis,
speaker verification and text classification.
3</p>
    </sec>
    <sec id="sec-2">
      <title>Questionnaire for EVALITA Participants and Organizers</title>
      <p>The questionnaire for participants and organizers of past EVALITA campaigns was divided into 3 parts.
In the first part respondents were required to provide general information such as job position and type
of affiliation. In the second part we collected comments about the tasks of past editions asking to rate
the level of satisfaction related to four dimensions: (i) the clarity of the guidelines,; (ii) the amount of
training data; (iii) the data format; and (iv) the adopted evaluation methods. An open field was also
available to add supplementary feedback. Finally, the third section aimed at gathering suggestions for
the future of EVALITA posing questions on different aspects, e.g. application domains, type of tasks,
structure of the final workshop, evaluation methodologies, dissemination of the results.</p>
      <p>The link to the questionnaire was sent to 90 persons who participated in or organized a task in at
least one of the last two EVALITA editions. After two weeks we received 39 answers (43.3% response
rate) from researchers, Phd candidates and technologists belonging to universities (61.54%) but also to
public (25.64%) and private (12.82%) research institutes. No answer from former participants affiliated
to private companies was received.</p>
      <p>Fifteen out of seventeen tasks of the past have been commented. All the four dimensions taken into
consideration obtained positive rates of satisfaction: in particular, 81% of respondents declared to be very
o somewhat satisfied by the guidelines and 76% by the format of distributed data. A small percentage of
unsatisfied responses (about 13%) were registered on the quantity of training data and on the evaluation.
In the open field, the most recurring concern was about the low number of participants in some tasks,
sometimes just one or two, especially in the speech ones.</p>
      <p>
        Respondents expressed the will to see some of the old tasks proposed again in the next EVALITA
campaign: sentiment polarity classification
        <xref ref-type="bibr" rid="ref2 ref8">(Basile et al., 2014)</xref>
        , parsing
        <xref ref-type="bibr" rid="ref12 ref2">(Bosco et al., 2014)</xref>
        , frame labeling
        <xref ref-type="bibr" rid="ref11">(Basili et al., 2013)</xref>
        , emotion recognition in speech
        <xref ref-type="bibr" rid="ref14 ref19 ref24">(Origlia and Galata`, 2014)</xref>
        , temporal information
processing
        <xref ref-type="bibr" rid="ref14">(Caselli et al., 2014)</xref>
        , and speaker identity verification
        <xref ref-type="bibr" rid="ref5">(Aversano et al., 2009)</xref>
        . As for the
domains of interest, the choices made by participants and organizers are in line with the ones made by
industrial companies showing a clear preference for social media (27), news (15), and humanities (13).
      </p>
      <p>
        The diverging stacked bar chart
        <xref ref-type="bibr" rid="ref14 ref19 ref24">(Heiberger and Robbins, 2014)</xref>
        in Figure 2, shows how the respondents
ranked their level of agreement with a set of statements related to the organization of the final workshop,
the performed evaluation and the campaign in general. The majority of respondents agree with almost
all statements: in particular, there is a strong consensus about having a demo session during the
workshop and also about taking into consideration, during the evaluation, not only systems’ effectiveness but
also their replicability. Providing the community with a web-based platform to share publicly available
resources and systems seems to be another important need as well as enhancing the international
visibility of EVALITA. A more neutral, or even negative, feedback was given regarding the possibility of
organizing more round tables during the workshop.
4
      </p>
    </sec>
    <sec id="sec-3">
      <title>Lessons Learnt and Impact on EVALITA 2016</title>
      <p>Both questionnaires provided us with useful information for the future of EVALITA: they allowed us
to acquire input on different aspects of the campaign and also to raise interest towards the initiative
engaging two different sectors, the research community and the enterprise community.</p>
      <p>Thanks to the questionnaire for industrial companies, we had the possibility to reach and get in touch
with a segment of potential participants who weren’t aware about the existence of EVALITA or had
little knowledge about it. Some of the suggestions coming from enterprises are actually feasible, for
example by proposing more application-oriented tasks and by covering domains that are important for
them. As for this last point, it is worth noting that the preferred domains are the same for both enterprises
and former participants and organizers: this facilitate the design of future tasks based on the collected
suggestions. Another issue emerged from both questionnaires is the need of improving the dissemination
of EVALITA results in Italy and abroad, in particular outside the boarders of the research community.</p>
      <p>The questionnaire for former participants and organizers gave us insights also on practical aspects
related to the organization of the final workshop and ideas on how to change the systems evaluation
approach taking into consideration different aspects such as replicability and usability.</p>
      <p>The results of the questionnaires were presented and discussed during the panel “Raising Interest and
Collecting Suggestions on the EVALITA Evaluation Campaign”7 organized in the context of the second
Italian Computational Linguistics Conference8 (CLiC-it 2015). The panel has sparked an interesting
debate on the participation of industrial companies to the campaign, which led to the decision of exploring
new avenues for involving industrial stakeholders in EVALITA, as the possibility to call for tasks of
industrial interest, that are proposed, and financially supported by the proponent companies. At the same
time, the need for a greater internationalization of the campaign, looking for tasks linked to the ones
proposed outside Italy, was highlighted. Panelists also wished for an effort in future campaigns towards
7http://www.evalita.it/towards2016
8https://clic2015.fbk.eu/
the development of shared datasets. Being the manual annotation of data a cost-consuming activity, the
monetary contribution of the Italian Association for Computational Linguistics9 (AILC) was solicited.</p>
      <p>
        The chairs of EVALITA 201610 introduced in the organization of the new campaign novel elements,
aimed at addressing most of the issues raised by both the questionnaires and the panel
        <xref ref-type="bibr" rid="ref10 ref6 ref9">(Basile et al.,
2016b)</xref>
        .
      </p>
      <p>
        EVALITA 2016 has an application-oriented task (i.e., QA4FAQ) in which representatives of three
companies11 are involved as organizers
        <xref ref-type="bibr" rid="ref13 ref9">(Caputo et al., 2016)</xref>
        . Another industrial company12 is part of the
organization of another task, i.e., PoSTWITA
        <xref ref-type="bibr" rid="ref25">(Tamburini et al., 2016)</xref>
        . Moreover, IBM Italy runs, for the
first time in the history of the campaign, a challenge for the development of an app providing monetary
awards for the best submissions: the evaluation follows various criteria, not only systems’ effectiveness
but also other aspects such as intuitiveness and creativity13. Given the widespread interest in social
media, a particular effort has been put in providing tasks dealing with texts in that domain. Three tasks
focus on the processing of tweets (i.e., NEEL-it, PoSTWITA, and SENTIPOLC) and part of the test
set is shared among 4 different tasks (i.e., FacTA, NEEL-it, PoSTWITA, and SENTIPOLC)
        <xref ref-type="bibr" rid="ref10 ref22 ref6 ref6 ref9">(Minard et
al., 2016; Basile et al., 2016a; Barbieri et al., 2016)</xref>
        . Part of the SENTIPOLC data was annotated via
Crowdflower14 thanks to funds allocated by AILC.
      </p>
      <p>
        For what concerns the internationalization issue, in the 2016 edition we had two EVALITA tasks
having an explicit link to other shared tasks proposed for English in the context of other evaluation
campaigns: the re-run of SENTIPOLC, with an explicit link to the Sentiment analysis in Twitter task
at SEMEVAL15, and the new NEEL-it, which is linked to the Named Entity rEcognition and Linking
(NEEL) Challenge proposed for English tweets at the 6th Making Sense of Microposts Workshop
        <xref ref-type="bibr" rid="ref4">(#Microposts2016, co-located with WWW 2016)</xref>
        16. Both tasks have been proposed with the aim to establish
a reference evaluation framework in the context of Italian tweets.
      </p>
      <p>We also used social media such as Twitter and Facebook, in order to improve dissemination of
information on EVALITA, with the twofold aim to reach a wider audience and to ensure timely communication
about various stages of the evaluation campaign.</p>
      <p>As for the organization of the final workshop, a demo session is scheduled for the systems participating
to the IBM challenge, as a first try to address the request from the community to have new participatory
modalities of interacting with systems and teams during the workshop.</p>
    </sec>
    <sec id="sec-4">
      <title>Acknowledgments</title>
      <p>We are thankful to the panelists and to the audience of the panel ‘Raising Interest and Collecting
Suggestions on the EVALITA Evaluation campaign’ at CLiC-it 2015, for the inspiring and passionate debate.
We are also very grateful to Malvina Nissim and Pierpaolo Basile, who accepted with us the challenge
to rethink EVALITA and to co-organize the edition 2016 of the evaluation campaign.</p>
    </sec>
  </body>
  <back>
    <ref-list>
      <ref id="ref1">
        <mixed-citation>
          <string-name>
            <given-names>Jeffrey</given-names>
            <surname>Allen</surname>
          </string-name>
          and
          <string-name>
            <given-names>Khalid</given-names>
            <surname>Choukri</surname>
          </string-name>
          .
          <year>2000</year>
          .
          <article-title>Survey of language engineering needs: a language resources perspective</article-title>
          .
          <source>In Proceedings of the Second International Conference on Language Resources and Evaluation (LREC</source>
          <year>2000</year>
          ).
        </mixed-citation>
      </ref>
      <ref id="ref2">
        <mixed-citation>
          <string-name>
            <given-names>Giuseppe</given-names>
            <surname>Attardi</surname>
          </string-name>
          , Valerio Basile, Cristina Bosco, Tommaso Caselli, Felice Dell'Orletta, Simonetta Montemagni, Viviana Patti, Maria Simi, and
          <string-name>
            <given-names>Rachele</given-names>
            <surname>Sprugnoli</surname>
          </string-name>
          .
          <year>2015</year>
          .
          <article-title>State of the Art Language Technologies for Italian: The EVALITA 2014 Perspective</article-title>
          . Intelligenza Artificiale,
          <volume>9</volume>
          (
          <issue>1</issue>
          ):
          <fpage>43</fpage>
          -
          <lpage>61</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref3">
        <mixed-citation>
          9http://www.ai
          <article-title>-lc.it/ 10They include Malvina Nissim and Pierpaolo Basile, in addition to the authors of this paper</article-title>
          . 11QuestionCube:http://www.questioncube.com; AQP:www.aqp.it;
        </mixed-citation>
      </ref>
      <ref id="ref4">
        <mixed-citation>SudSistemi: http://www.sudsistemi.eu 12CELI: https://www.celi.it/ 13http://www.evalita.it/2016/tasks/ibm-challenge 14https://www.crowdflower.com/ 15http://alt.qcri.org/semeval2016/task4/ 16http://microposts2016.seas.upenn.edu/challenge.html</mixed-citation>
      </ref>
      <ref id="ref5">
        <mixed-citation>
          <string-name>
            <given-names>Guido</given-names>
            <surname>Aversano</surname>
          </string-name>
          , Niko Bru¨mmer, and
          <string-name>
            <given-names>Mauro</given-names>
            <surname>Falcone</surname>
          </string-name>
          .
          <year>2009</year>
          .
          <article-title>EVALITA 2009 Speaker Identity Verification Application Track -</article-title>
          <source>Organizer's Report. Proceedings of EVALITA</source>
          ,
          <string-name>
            <surname>Reggio</surname>
            <given-names>Emilia</given-names>
          </string-name>
          , Italy.
        </mixed-citation>
      </ref>
      <ref id="ref6">
        <mixed-citation>
          <string-name>
            <given-names>Francesco</given-names>
            <surname>Barbieri</surname>
          </string-name>
          , Valerio Basile, Danilo Croce, Malvina Nissim, Nicole Novielli, and
          <string-name>
            <given-names>Viviana</given-names>
            <surname>Patti</surname>
          </string-name>
          .
          <year>2016</year>
          .
          <article-title>Overview of the EVALITA 2016 SENTiment POLarity Classification Task</article-title>
          . In Pierpaolo Basile, Anna Corazza, Franco Cutugno, Simonetta Montemagni, Malvina Nissim, Viviana Patti, Giovanni Semeraro, and Rachele Sprugnoli, editors,
          <source>Proceedings of Third Italian Conference on Computational Linguistics</source>
          (CLiC-it
          <year>2016</year>
          ) &amp;
          <article-title>Fifth Evaluation Campaign of Natural Language Processing and Speech Tools for Italian</article-title>
          .
          <source>Final Workshop (EVALITA</source>
          <year>2016</year>
          ).
          <article-title>Associazione Italiana di Linguistica Computazionale (AILC).</article-title>
        </mixed-citation>
      </ref>
      <ref id="ref7">
        <mixed-citation>
          <string-name>
            <given-names>Yehuda</given-names>
            <surname>Baruch and Brooks C Holtom</surname>
          </string-name>
          .
          <year>2008</year>
          .
          <article-title>Survey response rate levels and trends in organizational research</article-title>
          .
          <source>Human Relations</source>
          ,
          <volume>61</volume>
          (
          <issue>8</issue>
          ):
          <fpage>1139</fpage>
          -
          <lpage>1160</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref8">
        <mixed-citation>
          <string-name>
            <given-names>Valerio</given-names>
            <surname>Basile</surname>
          </string-name>
          , Andrea Bolioli, Malvina Nissim, Viviana Patti, and
          <string-name>
            <given-names>Paolo</given-names>
            <surname>Rosso</surname>
          </string-name>
          .
          <year>2014</year>
          .
          <article-title>Overview of the Evalita 2014 sentiment polarity classification task</article-title>
          .
          <source>In Proceedings of the 4th evaluation campaign of Natural Language Processing and Speech tools for Italian (EVALITA'14)</source>
          . Pisa, Italy, pages
          <fpage>50</fpage>
          -
          <lpage>57</lpage>
          . Pisa University Press.
        </mixed-citation>
      </ref>
      <ref id="ref9">
        <mixed-citation>
          <string-name>
            <given-names>Pierpaolo</given-names>
            <surname>Basile</surname>
          </string-name>
          , Annalina Caputo, Anna Lisa Gentile, and
          <string-name>
            <given-names>Giuseppe</given-names>
            <surname>Rizzo</surname>
          </string-name>
          . 2016a.
          <article-title>Overview of the EVALITA 2016 Named Entity rEcognition and Linking in Italian Tweets (NEEL-IT) Task</article-title>
          . In Pierpaolo Basile, Anna Corazza, Franco Cutugno, Simonetta Montemagni, Malvina Nissim, Viviana Patti, Giovanni Semeraro, and Rachele Sprugnoli, editors,
          <source>Proceedings of Third Italian Conference on Computational Linguistics</source>
          (CLiC-it
          <year>2016</year>
          ) &amp;
          <article-title>Fifth Evaluation Campaign of Natural Language Processing and Speech Tools for Italian</article-title>
          .
          <source>Final Workshop (EVALITA</source>
          <year>2016</year>
          ).
          <article-title>Associazione Italiana di Linguistica Computazionale (AILC).</article-title>
        </mixed-citation>
      </ref>
      <ref id="ref10">
        <mixed-citation>
          <string-name>
            <given-names>Pierpaolo</given-names>
            <surname>Basile</surname>
          </string-name>
          , Franco Cutugno, Malvina Nissim,
          <source>Viviana Patti, and Rachele Sprugnoli. 2016b. EVALITA</source>
          <year>2016</year>
          :
          <article-title>Overview of the 5th Evaluation Campaign of Natural Language Processing and Speech Tools for Italian. Associazione Italiana di Linguistica Computazionale (AILC).</article-title>
        </mixed-citation>
      </ref>
      <ref id="ref11">
        <mixed-citation>
          <string-name>
            <given-names>Roberto</given-names>
            <surname>Basili</surname>
          </string-name>
          , Diego De Cao, Alessandro Lenci, Alessandro Moschitti, and
          <string-name>
            <given-names>Giulia</given-names>
            <surname>Venturi</surname>
          </string-name>
          .
          <year>2013</year>
          .
          <article-title>Evalita 2011: the frame labeling over Italian texts task</article-title>
          .
          <source>In Evaluation of Natural Language and Speech Tools for Italian</source>
          , pages
          <fpage>195</fpage>
          -
          <lpage>204</lpage>
          . Springer.
        </mixed-citation>
      </ref>
      <ref id="ref12">
        <mixed-citation>
          <string-name>
            <given-names>Cristina</given-names>
            <surname>Bosco</surname>
          </string-name>
          , Felice Dell'Orletta, Simonetta Montemagni, Manuela Sanguinetti, and
          <string-name>
            <given-names>Maria</given-names>
            <surname>Simi</surname>
          </string-name>
          .
          <year>2014</year>
          .
          <article-title>The EVALITA 2014 dependency parsing task</article-title>
          .
          <source>Proceedings of EVALITA.</source>
        </mixed-citation>
      </ref>
      <ref id="ref13">
        <mixed-citation>
          <string-name>
            <given-names>Annalina</given-names>
            <surname>Caputo</surname>
          </string-name>
          , Marco de Gemmis, Pasquale Lops, Franco Lovecchio, and
          <string-name>
            <given-names>Vito</given-names>
            <surname>Manzari</surname>
          </string-name>
          .
          <year>2016</year>
          .
          <article-title>Overview of the EVALITA 2016 Question Answering for Frequently Asked Questions (QA4FAQ) Task</article-title>
          . In Pierpaolo Basile, Anna Corazza, Franco Cutugno, Simonetta Montemagni, Malvina Nissim, Viviana Patti, Giovanni Semeraro, and Rachele Sprugnoli, editors,
          <source>Proceedings of Third Italian Conference on Computational Linguistics (CLiCit</source>
          <year>2016</year>
          ) &amp;
          <article-title>Fifth Evaluation Campaign of Natural Language Processing and Speech Tools for Italian</article-title>
          .
          <source>Final Workshop (EVALITA</source>
          <year>2016</year>
          ).
          <article-title>Associazione Italiana di Linguistica Computazionale (AILC).</article-title>
        </mixed-citation>
      </ref>
      <ref id="ref14">
        <mixed-citation>
          <string-name>
            <given-names>Tommaso</given-names>
            <surname>Caselli</surname>
          </string-name>
          , Rachele Sprugnoli, Manuela Speranza, and
          <string-name>
            <given-names>Monica</given-names>
            <surname>Monachini</surname>
          </string-name>
          .
          <year>2014</year>
          .
          <article-title>EVENTI: EValuation of Events and Temporal INformation at Evalita 2014</article-title>
          .
          <source>In Proceedings of the First Italian Conference on Computational Linguistics CLiC-it 2014 &amp; and of the Fourth International Workshop EVALITA</source>
          <year>2014</year>
          , pages
          <fpage>27</fpage>
          -
          <lpage>34</lpage>
          . Pisa University Press.
        </mixed-citation>
      </ref>
      <ref id="ref15">
        <mixed-citation>
          <string-name>
            <given-names>Andrea</given-names>
            <surname>Di</surname>
          </string-name>
          Carlo and
          <string-name>
            <given-names>Andrea</given-names>
            <surname>Paoloni</surname>
          </string-name>
          .
          <year>2009</year>
          .
          <article-title>Libro Bianco sul Trattamento Automatico della Lingua</article-title>
          ., volume
          <volume>1</volume>
          .
          <string-name>
            <given-names>Fondazione</given-names>
            <surname>Ugo</surname>
          </string-name>
          <string-name>
            <surname>Bordoni</surname>
          </string-name>
          , Roma.
        </mixed-citation>
      </ref>
      <ref id="ref16">
        <mixed-citation>
          <string-name>
            <given-names>Julio</given-names>
            <surname>Gonzalo</surname>
          </string-name>
          , Felisa Verdejo, Anselmo Pen˜as, Carol Peters, Khalid Choukri, and
          <string-name>
            <given-names>Michael</given-names>
            <surname>Kluck</surname>
          </string-name>
          .
          <year>2002</year>
          .
          <article-title>Cross Language Evaluation Forum - User Needs: Deliverable 1.1.1</article-title>
          .
          <source>Technical report.</source>
        </mixed-citation>
      </ref>
      <ref id="ref17">
        <mixed-citation>
          <string-name>
            <given-names>Seth</given-names>
            <surname>Grimes</surname>
          </string-name>
          .
          <year>2014</year>
          .
          <article-title>Text analytics 2014: User perspectives on solutions and providers</article-title>
          . Alta Plana.
        </mixed-citation>
      </ref>
      <ref id="ref18">
        <mixed-citation>
          FLaReNet Working Group.
          <year>2010</year>
          .
          <article-title>Results of the questionnaire on the priorities in the field of language resources</article-title>
          .
          <source>Technical report</source>
          , Department of Computer Science, Michigan State University, September.
        </mixed-citation>
      </ref>
      <ref id="ref19">
        <mixed-citation>
          <string-name>
            <surname>Richard M Heiberger and Naomi B Robbins</surname>
          </string-name>
          .
          <year>2014</year>
          .
          <article-title>Design of diverging stacked bar charts for likert scales and other applications</article-title>
          .
          <source>Journal of Statistical Software</source>
          ,
          <volume>57</volume>
          (
          <issue>5</issue>
          ):
          <fpage>1</fpage>
          -
          <lpage>32</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref20">
        <mixed-citation>
          <string-name>
            <given-names>Andrew</given-names>
            <surname>Joscelyne</surname>
          </string-name>
          and
          <string-name>
            <given-names>Rose</given-names>
            <surname>Lockwood</surname>
          </string-name>
          .
          <year>2003</year>
          .
          <article-title>Benchmarking HLT progress in Europe</article-title>
          ., volume
          <volume>1</volume>
          .
          <source>The EUROMAP Study</source>
          , Copenhagen.
        </mixed-citation>
      </ref>
      <ref id="ref21">
        <mixed-citation>
          <string-name>
            <surname>Michael D Kaplowitz</surname>
          </string-name>
          ,
          <string-name>
            <surname>Timothy D Hadlock</surname>
            , and
            <given-names>Ralph</given-names>
          </string-name>
          <string-name>
            <surname>Levine</surname>
          </string-name>
          .
          <year>2004</year>
          .
          <article-title>A comparison of web and mail survey response rates</article-title>
          .
          <source>Public opinion quarterly</source>
          ,
          <volume>68</volume>
          (
          <issue>1</issue>
          ):
          <fpage>94</fpage>
          -
          <lpage>101</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref22">
        <mixed-citation>
          <string-name>
            <surname>Anne-Lyse</surname>
            <given-names>Minard</given-names>
          </string-name>
          , Manuela Speranza, and
          <string-name>
            <given-names>Tommaso</given-names>
            <surname>Caselli</surname>
          </string-name>
          .
          <year>2016</year>
          .
          <article-title>The EVALITA 2016 Event Factuality Annotation Task (FactA)</article-title>
          . In Pierpaolo Basile, Anna Corazza, Franco Cutugno, Simonetta Montemagni, Malvina Nissim, Viviana Patti, Giovanni Semeraro, and Rachele Sprugnoli, editors,
          <source>Proceedings of Third Italian Conference on Computational Linguistics</source>
          (CLiC-it
          <year>2016</year>
          ) &amp;
          <article-title>Fifth Evaluation Campaign of Natural Language Processing and Speech Tools for Italian</article-title>
          .
          <source>Final Workshop (EVALITA</source>
          <year>2016</year>
          ).
          <article-title>Associazione Italiana di Linguistica Computazionale (AILC).</article-title>
        </mixed-citation>
      </ref>
      <ref id="ref23">
        <mixed-citation>
          <string-name>
            <given-names>Nelleke</given-names>
            <surname>Oostdijk</surname>
          </string-name>
          and
          <string-name>
            <given-names>Lou</given-names>
            <surname>Boves</surname>
          </string-name>
          .
          <year>2006</year>
          .
          <article-title>User requirements analysis for the design of a reference corpus of written dutch</article-title>
          .
          <source>In Proceedings of the Fifth International Conference on Language Resources and Evaluation(LREC</source>
          <year>2006</year>
          ).
        </mixed-citation>
      </ref>
      <ref id="ref24">
        <mixed-citation>
          <string-name>
            <given-names>Antonio</given-names>
            <surname>Origlia</surname>
          </string-name>
          and Vincenzo Galata`.
          <year>2014</year>
          .
          <article-title>EVALITA 2014: Emotion Recognition Task (ERT)</article-title>
          .
          <source>In Proceedings of the First Italian Conference on Computational Linguistics CLiC-it 2014 &amp; and of the Fourth International Workshop EVALITA</source>
          <year>2014</year>
          , pages
          <fpage>112</fpage>
          -
          <lpage>115</lpage>
          . Pisa University Press.
        </mixed-citation>
      </ref>
      <ref id="ref25">
        <mixed-citation>
          <string-name>
            <given-names>Fabio</given-names>
            <surname>Tamburini</surname>
          </string-name>
          , Cristina Bosco, Alessandro Mazzei, and
          <string-name>
            <given-names>Andrea</given-names>
            <surname>Bolioli</surname>
          </string-name>
          .
          <year>2016</year>
          .
          <article-title>Overview of the EVALITA 2016 Part Of Speech on TWitter for ITAlian Task</article-title>
          . In Pierpaolo Basile, Anna Corazza, Franco Cutugno, Simonetta Montemagni, Malvina Nissim, Viviana Patti, Giovanni Semeraro, and Rachele Sprugnoli, editors,
          <source>Proceedings of Third Italian Conference on Computational Linguistics</source>
          (CLiC-it
          <year>2016</year>
          ) &amp;
          <article-title>Fifth Evaluation Campaign of Natural Language Processing and Speech Tools for Italian</article-title>
          .
          <source>Final Workshop (EVALITA</source>
          <year>2016</year>
          ).
          <article-title>Associazione Italiana di Linguistica Computazionale (AILC).</article-title>
        </mixed-citation>
      </ref>
      <ref id="ref26">
        <mixed-citation>
          <string-name>
            <given-names>Katrin</given-names>
            <surname>Tomanek</surname>
          </string-name>
          and
          <string-name>
            <given-names>Fredrik</given-names>
            <surname>Olsson</surname>
          </string-name>
          .
          <year>2009</year>
          .
          <article-title>A web survey on the use of active learning to support annotation of text data</article-title>
          .
          <source>In Proceedings of the NAACL HLT 2009 Workshop on Active Learning for Natural Language Processing</source>
          , pages
          <fpage>45</fpage>
          -
          <lpage>48</lpage>
          . Association for Computational Linguistics.
        </mixed-citation>
      </ref>
    </ref-list>
  </back>
</article>