<!DOCTYPE article PUBLIC "-//NLM//DTD JATS (Z39.96) Journal Archiving and Interchange DTD v1.0 20120330//EN" "JATS-archivearticle1.dtd">
<article xmlns:xlink="http://www.w3.org/1999/xlink">
  <front>
    <journal-meta />
    <article-meta>
      <title-group>
        <article-title>RDF triples extraction from company web pages: comparison of state-of-the-art Deep Models1</article-title>
      </title-group>
      <contrib-group>
        <contrib contrib-type="author">
          <string-name>Wouter BAES</string-name>
          <xref ref-type="aff" rid="aff0">0</xref>
          <xref ref-type="aff" rid="aff1">1</xref>
        </contrib>
        <contrib contrib-type="author">
          <string-name>Franc¸ois PORTET</string-name>
          <xref ref-type="aff" rid="aff1">1</xref>
        </contrib>
        <contrib contrib-type="author">
          <string-name>Hamid MIRISAEE</string-name>
          <xref ref-type="aff" rid="aff0">0</xref>
        </contrib>
        <contrib contrib-type="author">
          <string-name>Cyril LABB E´</string-name>
          <xref ref-type="aff" rid="aff1">1</xref>
        </contrib>
        <aff id="aff0">
          <label>0</label>
          <institution>Skopai</institution>
          ,
          <addr-line>38400 Saint-Martin-d'He`res</addr-line>
          ,
          <country country="FR">France</country>
        </aff>
        <aff id="aff1">
          <label>1</label>
          <institution>Univ. Grenoble Alpes</institution>
          ,
          <addr-line>CNRS, Grenoble INP, LIG, F-38000 Grenoble</addr-line>
          ,
          <country country="FR">France</country>
        </aff>
      </contrib-group>
      <abstract>
        <p>Relation extraction (RE) is a promising way to extend the semantic web from web pages. However, it is unclear how RE can deal with the several challenges of web pages such as noise, data sparsity and conflicting information. In this paper, we benchmark state-of-the-art RE approaches on the particular case of company web pages, since company web pages are important source of information for Fintech and BusinnessTech. To this end, we present a method to build a corpus mimicking web pages characteristics. This corpus was used to evaluate several deep learning RE models and compared to another benchmark corpus.</p>
      </abstract>
      <kwd-group>
        <kwd />
        <kwd>relation extraction</kwd>
        <kwd>NLP</kwd>
        <kwd>linked data</kwd>
        <kwd>Deep Learning</kwd>
      </kwd-group>
    </article-meta>
  </front>
  <body>
    <sec id="sec-1">
      <title>1. Introduction</title>
      <p>
        Relation Extraction (RE) refers to the process of identifying semantic links between
entities in a sentence [
        <xref ref-type="bibr" rid="ref1">1</xref>
        ]. As an example, Bill Gates founded Microsoft, has Bill Gates and
Microsoft as entities and the founder relation as a semantic link between those two. RE
has been successfully applied to a wide range of domains such as knowledge base
enrichment [
        <xref ref-type="bibr" rid="ref2">2</xref>
        ] and Question-Answering [
        <xref ref-type="bibr" rid="ref3">3</xref>
        ]. With the extremely fast growth of the
internet, web pages are now considered as a very rich source for populating knowledge bases.
Those pages, however, contain information in plain text, or in a poorly structured form.
Extracting this information is not easy as they suffer from noise and data sparsity.
Accordingly, the extracted information can be incomplete, or in conflict with other
information. Although RE is a mature technology which has been evaluated on some
benchmarks, it is still difficult to predict how it will behave on new datasets different from
those of the benchmarks. For instance, in Skopai2, a company that uses deep learning
techniques to analyze and classify startups, one of the objectives is to extract
information from company web pages in a form that is exploitable for reasoning. Such company
needs an efficient semantics extraction from web-pages to reduce the amount of
corrections to be performed by human experts. Furthermore, storing the extracted relations as
RDF-triples in an ontology allows for reasoning, deduction of implicit information and
automatic updating of the information, all of which help with the company’s objective.
      </p>
      <p>In this paper, we present the results of a study aiming at comparing current
state-ofthe-art deep learning RE models to the specific domain of company web pages. To do
this, we built a corpus which mimics the characteristics of the desired data.</p>
      <p>The contributions of this paper are (1) the construction of a dataset for the task of
relation extraction (with a focus on RE from company webpages) and (2) the comparison
of several state-of-the-art relation extraction models on different benchmarks.</p>
    </sec>
    <sec id="sec-2">
      <title>2. State of the art</title>
      <p>
        Relation Extraction (RE) task is to detect and classify semantic relationship mentions
from plain free-text. Several techniques for RE from patterns matching to statistical
models have been proposed. However, recent advance in deep learning has made it the
current state-of-the-art [
        <xref ref-type="bibr" rid="ref1">1</xref>
        ]. In the literature, the task of RE has been studied both in the
supervised and unsupervised paradigm. For instance, [
        <xref ref-type="bibr" rid="ref4">4</xref>
        ] proposes a CNN-based
technique to extract lexical features for relation classification. The biggest drawback of this
approach is the need for large amount of high-quality, manually labeled training and test
data, which is costly and time-consuming to make [
        <xref ref-type="bibr" rid="ref5">5</xref>
        ]. Unsupervised techniques do not
require human labor for labeling, but they usually lead to inferior results [
        <xref ref-type="bibr" rid="ref6">6</xref>
        ].
      </p>
      <p>
        To assess the progress of the domain, several benchmarks and challenges have
emerged this last decade. In the domain of supervised RE, a widely used benchmark
is the SemEval 2010 Task 8 [
        <xref ref-type="bibr" rid="ref7">7</xref>
        ] challenge. The dataset used during the challenge, was
composed 10k sentences each of which annotated with one of 19 possible relations (9
bi-directional relations, and one Other).
      </p>
      <p>
        In supervised RE, popular Deep Neural Network (DNN) architectures are
convolutional neural networks (CNN) and recurrent neural networks (RNN). For instance, [
        <xref ref-type="bibr" rid="ref4">4</xref>
        ]
proposed a ‘CNN + Softmax’ model which reached 78.9 % of F-measure on the SemEval
2010 Task 8 challenge. Since then, BERT-based models have shown a definite
improvement. For instance, [
        <xref ref-type="bibr" rid="ref8">8</xref>
        ] proposed a model called ‘BERT-Entity + MTB’ which used the
representation of entity markers (more specifically, the end markers of those entities) as
output of the final hidden transformer layer. MTB signifies that the transformer model
used is not regular BERT, but one that has been pre-trained to ‘Match the Blanks’ (MTB),
meaning it got fed sentences with words blanked out, where the goal was to predict what
these blanked out words were. This model reached 89.5 % of F-measure on the SemEval
2010 Task 8 challenge far above the CNN model.
      </p>
      <p>
        These DNN models generally perform well but heavily depend on the availability
of a large amount of high-quality, manually labeled training and test data. This is costly
and time-consuming in human labor [
        <xref ref-type="bibr" rid="ref5">5</xref>
        ]. We partially address this problem by creating
semi-automatically a corpus dedicated to company web pages.
      </p>
    </sec>
    <sec id="sec-3">
      <title>3. Method</title>
      <p>A general overview of the approach to acquire a new dataset and train RE models from
it is given, with the different steps laid out in a schema. The ontology definition and the
alignment process are then detailed.</p>
      <sec id="sec-3-1">
        <title>3.1. Overview of the approach</title>
        <p>Figure 1 shows the different steps undertaken in this study. To extract semantic
information, the list of the concepts and their relations is first defined within an Ontology.
This process is explained in Section 3.2. At the beginning of the process, we consider
a free-text corpus and a set of semantic relations none of which being aligned. For
instance a fact such as founder(Bill Gates,Microsoft) is given but it is not known
which sentence in the corpus describes this fact. These facts, together with the free text
sentences, compose the Unaligned dataset.</p>
        <p>Using the ontology terminology, the dataset facts are processed to populate the
ontology. This step can be seen as the transformation of an arbitrary semantic information
into RDF-triples.</p>
        <p>Once the set of RDF-triples are processed, the alignment step seeks the sentences
that are the most probably associated to each triple. The aim is to produce an aligned
dataset where sentences are annotated with the triples. This process is detailed in
Section 3.2. The aligned output can then be used to train and evaluate some of the deep
learning models described in Section 2.</p>
      </sec>
      <sec id="sec-3-2">
        <title>3.2. Ontology definition</title>
        <p>To build the reference ontology for description of companies we used the DBpedia 3
OWL structure, in particular the relations linked to the Organisation and Company
entities since it already contains most of the needed relations. It also plays the role of a
top-ontology where the links between the classes are established and could be used to
infer further information. Moreover, using DBpedia makes it much easier to ensure the
interoperability.</p>
        <p>The ontology was then confronted to the professional Skopai database, by looking
at the possible attributes that can be present in the collections. Not all of these attributes
can be modeled using the relations already extracted from DBpedia. Hence, some extra
predicates such as those related to patents, funding and awards were added to the final
ontology.</p>
      </sec>
      <sec id="sec-3-3">
        <title>3.3. Data to Triple alignment</title>
        <p>The alignment consists of matching two sources of information, each with a list of
sentences and a list of RDF-triples. Each of those sentences may or may not actually
describe one or more of the triples. Hence, the objective is to align those sentences with
triple(s) they describe, discarding those which describe no relation. For example,
consider two sentences: Bill gates founded Microsoft and Microsoft was founded in 1975 and
two RDF triples: founder(Microsoft, Bill Gates) and location(Microsoft,
Redmond, USA). From all of this information we know that Microsoft was founded in
1975 by Bill Gates, and has its headquarters in Redmond, USA. However, only the fact
that Bill Gates is the founder is present in both the sentences and the triples. So the only
alignment that can be made is Bill gates founded Microsoft with founder(Microsoft,
Bill Gates).</p>
        <p>
          To perform this task we used the alignment tool4 built in the context T-REx [
          <xref ref-type="bibr" rid="ref9">9</xref>
          ].
T-REx is a large aligned dataset of 3.09 million Wikipedia abstracts (6.2 million
sentences) with 11 million Wikidata5 triples. The tool aligns the sentences with triples using
distantly supervised learning triple aligners, more specifically those specified in [
          <xref ref-type="bibr" rid="ref5">5</xref>
          ].
        </p>
      </sec>
    </sec>
    <sec id="sec-4">
      <title>4. Experiments</title>
      <sec id="sec-4-1">
        <title>4.1. The Corpora</title>
        <p>
          In this section, we present the corpora that have been used, the result of the alignment
process of one corpus and the performance of the RE models presented in Section 2.
To assess the performances of RE on company texts, we used the Wikipedia Company
Corpus (WCC)6 [
          <xref ref-type="bibr" rid="ref10">10,11</xref>
          ]. This was constructed for the automatic generation of company
descriptions from a set of relations. The WCC consists of a 43 980 companies extracted
from Wikipedia. Each company example comes with an abstract and a list of at least two
attribute-value pairs. The total amount of sentences in the corpus is 159 710, an average
of just under four sentences per abstract. However, the attribute-value pairs were not
aligned with the text. Even worse, since it is a real noisy corpus, some attribute-value
pairs are not present in the Wikipedia text and vice-versa. Hence, this corpus needs to be
aligned.
        </p>
        <p>To evaluate the models on clean conditions, we also used the dataset released with
the WebNLG 2020 challenge7, which already comes aligned. The dataset contains a
total of 16 categories including the Company category. An example of text and its
corresponding triples is shown Figure 2. The amount of triples per sentence ranges from 1 to
7. As of the time of writing, the test set has not yet been released, only a training and
a development set. The training set consists of 13 229 entries (in the form of a set of
triples) with 35 415 texts (3 to 5 texts per entry). The development set consists of 1669
set of triples with 4 468 texts. The amount of instances per category ranges from 299 for</p>
      </sec>
      <sec id="sec-4-2">
        <title>Monument to 1591 for Food.</title>
      </sec>
      <sec id="sec-4-3">
        <title>4.2. Results of the alignment</title>
        <p>The output produced by the T-Rex pipeline on WCC consists of 193 203 triples over 108
227 sentences, concerning 34 299 companies. The distantly supervised approach was
4https://github.com/hadyelsahar/RE-NLG-Dataset
5https://www.wikidata.org/wiki/Wikidata:Main_Page
6https://gricad-gitlab.univ-grenoble-alpes.fr/getalp/wikipediacompanycorpus
7https://webnlg-challenge.loria.fr/challenge_2020/
able to recognize and align the majority of the sentences and the semantic facts. As the
WCC corpus is noisy, perfect alignment was not expected. Looking at the distribution
of the triples in Table 1, it is unsurprising that the biggest part, describes the location
relation, as it is available for almost every company. At the other end of the spectrum, the
numberOfEmployees relation is often present as fact but rarely in the abstract, so there is
very little alignment possible.</p>
        <p>Hereafter, when referring to the Wikipediacompanycorpus (or WCC), we are
referring to the aligned version described in this section.</p>
      </sec>
      <sec id="sec-4-4">
        <title>4.3. Evaluation</title>
        <p>The models that were evaluated are the ones mentioned in Section 2, with three different
variants of the BERT-based model. BERT does not use the entity markings that
BERTEntity uses, and only BERT-Entity + MTB uses the MTB pre-trained model.</p>
        <p>The results are reported in Table 2. The WCC corpus was randomly split into
64/16/20 % for training/development and testing. For the WebNLG 2020 challenge, since
the test set has not yet been released, the reported results are those of the development
set. This means that the result presented should be higher than when using a true test set.
For WebNLG, it can be seen that BERT-Entity performs the best, followed by
BERTEntity+MTB then BERT and then CNN. The results obtained from experimenting on
WebNLG points to BERT-Entity as the best model to use. For WCC, CNN + Softmax
performs the worst overall as well, while BERT-Entity + MTB performs the best over all
metrics.</p>
        <p>Trane, which was founded on January 1st 1913 in La Crosse, Wisconsin, is based in Ireland. It has 29,000
employees.
&lt;entry category="Company" eid="Id21" shape="(X (X) (X) (X) (X))" shape_type="sibling" size="4"&gt;
&lt;modifiedtripleset&gt;
&lt;mtriple&gt;Trane | foundingDate | 1913-01-01&lt;/mtriple&gt;
&lt;mtriple&gt;Trane | location | Ireland&lt;/mtriple&gt;
&lt;mtriple&gt;Trane | foundationPlace | La_Crosse,_Wisconsin&lt;/mtriple&gt;
&lt;mtriple&gt;Trane | numberOfEmployees | 29000&lt;/mtriple&gt;
&lt;/modifiedtripleset&gt;
&lt;/entry&gt;
5. Discussion and further work
In this study, we have explored the performance of several RE models on corpus about
company information. For this purpose, we have created an aligned corpus using the
Wikipediacompanycorpus and augmenting it through the T-REx alignment pipeline. This
gave a set of aligned sentences and RDF-triples, that are specific to companies. A big
shortcoming of the corpus in its current form is the absence of negative training samples,
sentences that do not contain any relations, or irrelevant ones. This is needed because
there is an overwhelming amount of noise as wells as irrelevant information in company
web pages.</p>
        <p>After evaluating on WebNLG and WCC, BERT-Entity comes forward as the most
accurate and most consistent model. In some cases, using the MTB pre-trained model
improves results, but not enough to warrant its use (taking into account the need for
memory and time intensive pre-training).</p>
        <p>Future work for this project includes the improvement of the constructed corpus
(with e.g. negative training samples and data from Skopai database), the implementation
of the best model to be used by Skopai as well as the possibility to change language focus
with different variants of BERT and a way to automatically handle ontology population
and ontology evolution.
[11] Qader R, Portet F, Labbe C. Semi-Supervised Neural Text Generation by Joint Learning of Natural
Language Generation and Natural Language Understanding Models. In: 12th International Conference
on Natural Language Generation (INLG 2019). Tokyo, Japan; 2019. Available from: https://hal.
archives-ouvertes.fr/hal-02371384.
[12] Gardent C, Shimorina A, Narayan S, Perez-Beltrachini L. Creating Training Corpora for NLG
MicroPlanners. In: Proceedings of the 55th Annual Meeting of the Association for Computational Linguistics
(Volume 1: Long Papers). Association for Computational Linguistics; 2017. p. 179–188. Available from:
http://www.aclweb.org/anthology/P17-1017.</p>
      </sec>
    </sec>
  </body>
  <back>
    <ref-list>
      <ref id="ref1">
        <mixed-citation>
          [1]
          <string-name>
            <surname>Kumar</surname>
            <given-names>S.</given-names>
          </string-name>
          <article-title>A survey of deep learning methods for relation extraction; 2017</article-title>
          . ArXiv preprint arXiv:
          <volume>1705</volume>
          .
          <fpage>03645</fpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref2">
        <mixed-citation>
          [2]
          <string-name>
            <surname>Trisedya</surname>
            <given-names>BD</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Weikum</surname>
            <given-names>G</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Qi</surname>
            <given-names>J</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Zhang</surname>
            <given-names>R</given-names>
          </string-name>
          .
          <article-title>Neural Relation Extraction for Knowledge Base Enrichment</article-title>
          .
          <source>In: Proceedings of the 57th Annual Meeting of the Association for Computational Linguistics; 2019</source>
          . p.
          <fpage>229</fpage>
          -
          <lpage>240</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref3">
        <mixed-citation>
          [3]
          <string-name>
            <surname>Li</surname>
            <given-names>X</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Yin</surname>
            <given-names>F</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Sun</surname>
            <given-names>Z</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Li</surname>
            <given-names>X</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Yuan</surname>
            <given-names>A</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Chai</surname>
            <given-names>D</given-names>
          </string-name>
          , et al.
          <article-title>Entity-Relation Extraction as Multi-Turn Question Answering</article-title>
          . In:
          <article-title>Proceedings of the 57th Annual Meeting of the Association for Computational Linguistics</article-title>
          . Florence, Italy: Association for Computational Linguistics;
          <year>2019</year>
          . p.
          <fpage>1340</fpage>
          -
          <lpage>1350</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref4">
        <mixed-citation>
          [4]
          <string-name>
            <surname>Zeng</surname>
            <given-names>D</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Liu</surname>
            <given-names>K</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Lai</surname>
            <given-names>S</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Zhou</surname>
            <given-names>G</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Zhao</surname>
            <given-names>J</given-names>
          </string-name>
          .
          <string-name>
            <surname>Relation</surname>
          </string-name>
          <article-title>Classification via Convolutional Deep Neural Network</article-title>
          .
          <source>In: Proceedings of COLING</source>
          <year>2014</year>
          ,
          <source>the 25th International Conference on Computational Linguistics: Technical Papers</source>
          . Dublin, Ireland: Dublin City University and Association for Computational Linguistics;
          <year>2014</year>
          . p.
          <fpage>2335</fpage>
          -
          <lpage>2344</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref5">
        <mixed-citation>
          [5]
          <string-name>
            <surname>Augenstein</surname>
            <given-names>I</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Maynard</surname>
            <given-names>D</given-names>
          </string-name>
          , Ciravegna F.
          <article-title>Distantly supervised web relation extraction for knowledge base population</article-title>
          .
          <source>Semantic Web</source>
          .
          <year>2016</year>
          ;
          <volume>7</volume>
          (
          <issue>4</issue>
          ):
          <fpage>335</fpage>
          -
          <lpage>349</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref6">
        <mixed-citation>
          [6]
          <string-name>
            <surname>Elsahar</surname>
            <given-names>H</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Demidova</surname>
            <given-names>E</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Gottschalk</surname>
            <given-names>S</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Gravier</surname>
            <given-names>C</given-names>
          </string-name>
          , Laforest F.
          <article-title>Unsupervised open relation extraction</article-title>
          . In: European Semantic Web Conference;
          <year>2017</year>
          . p.
          <fpage>12</fpage>
          -
          <lpage>16</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref7">
        <mixed-citation>
          [7]
          <string-name>
            <surname>Hendrickx</surname>
            <given-names>I</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Kim</surname>
            <given-names>SN</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Kozareva</surname>
            <given-names>Z</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Nakov</surname>
            <given-names>P</given-names>
          </string-name>
          , O´ Se´aghdha
          <string-name>
            <surname>D</surname>
          </string-name>
          , Pado´ S, et al. SemEval
          <article-title>-2010 Task 8: Multi-Way Classification of Semantic Relations between Pairs of Nominals</article-title>
          .
          <source>In: Proceedings of the 5th International Workshop on Semantic Evaluation. Uppsala</source>
          , Sweden: Association for Computational Linguistics;
          <year>2010</year>
          . p.
          <fpage>33</fpage>
          -
          <lpage>38</lpage>
          . Available from: https://www.aclweb.org/anthology/S10-1006.
        </mixed-citation>
      </ref>
      <ref id="ref8">
        <mixed-citation>
          [8]
          <string-name>
            <surname>Baldini Soares</surname>
            <given-names>L</given-names>
          </string-name>
          ,
          <string-name>
            <surname>FitzGerald</surname>
            <given-names>N</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Ling</surname>
            <given-names>J</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Kwiatkowski</surname>
            <given-names>T.</given-names>
          </string-name>
          <article-title>Matching the Blanks: Distributional Similarity for Relation Learning</article-title>
          . In:
          <article-title>Proceedings of the 57th Annual Meeting of the Association for Computational Linguistics</article-title>
          . Florence, Italy: Association for Computational Linguistics;
          <year>2019</year>
          . p.
          <fpage>2895</fpage>
          -
          <lpage>2905</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref9">
        <mixed-citation>
          [9]
          <string-name>
            <surname>Elsahar</surname>
            <given-names>H</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Vougiouklis</surname>
            <given-names>P</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Remaci</surname>
            <given-names>A</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Gravier</surname>
            <given-names>C</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Hare</surname>
            <given-names>J</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Laforest</surname>
            <given-names>F</given-names>
          </string-name>
          , et al.
          <article-title>T-REx: A Large Scale Alignment of Natural Language with Knowledge Base Triples</article-title>
          .
          <source>In: Proceedings of the Eleventh International Conference on Language Resources and Evaluation (LREC</source>
          <year>2018</year>
          ). Miyazaki, Japan: European Language Resources Association (ELRA);
          <year>2018</year>
          . p.
          <fpage>3448</fpage>
          -
          <lpage>3452</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref10">
        <mixed-citation>
          [10]
          <string-name>
            <surname>Qader</surname>
            <given-names>R</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Jneid</surname>
            <given-names>K</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Portet</surname>
            <given-names>F</given-names>
          </string-name>
          , Labbe´ C.
          <article-title>Generation of Company descriptions using concept-to-text and textto-text deep models: dataset collection and systems evaluation</article-title>
          .
          <source>In: Proceedings of the 11th International Conference on Natural Language Generation; 2018</source>
          . p.
          <fpage>254</fpage>
          -
          <lpage>263</lpage>
          .
        </mixed-citation>
      </ref>
    </ref-list>
  </back>
</article>