<!DOCTYPE article PUBLIC "-//NLM//DTD JATS (Z39.96) Journal Archiving and Interchange DTD v1.0 20120330//EN" "JATS-archivearticle1.dtd">
<article xmlns:xlink="http://www.w3.org/1999/xlink">
  <front>
    <journal-meta />
    <article-meta>
      <title-group>
        <article-title>Preface of MEPDaW 2021: Managing the Evolution and Preservation of the Data Web</article-title>
      </title-group>
      <contrib-group>
        <contrib contrib-type="author">
          <string-name>Fabrizio Orlandi</string-name>
          <email>orlandif@tcd.ie</email>
          <xref ref-type="aff" rid="aff0">0</xref>
        </contrib>
        <contrib contrib-type="author">
          <string-name>Damien Graux</string-name>
          <email>damien.graux@inria.fr</email>
          <xref ref-type="aff" rid="aff1">1</xref>
        </contrib>
        <contrib contrib-type="author">
          <string-name>Julio Cesar dos Reis</string-name>
          <email>jreis@ic.unicamp.br</email>
          <xref ref-type="aff" rid="aff2">2</xref>
        </contrib>
        <contrib contrib-type="author">
          <string-name>Maria-Esther Vidal</string-name>
          <email>mvidal@umiacs.umd.edu</email>
          <xref ref-type="aff" rid="aff3">3</xref>
        </contrib>
        <aff id="aff0">
          <label>0</label>
          <institution>ADAPT SFI Centre, Trinity College Dublin</institution>
          ,
          <country country="IE">Ireland</country>
        </aff>
        <aff id="aff1">
          <label>1</label>
          <institution>Inria, Université Côte d'Azur</institution>
          ,
          <addr-line>CNRS, I3S</addr-line>
          ,
          <country country="FR">France</country>
        </aff>
        <aff id="aff2">
          <label>2</label>
          <institution>Inst. of Computing, Univ. of Campinas (UNICAMP)</institution>
          ,
          <country country="BR">Brazil</country>
        </aff>
        <aff id="aff3">
          <label>3</label>
          <institution>Technische Informationsbibliothek (TIB)</institution>
          ,
          <country country="DE">Germany</country>
        </aff>
      </contrib-group>
      <pub-date>
        <year>2021</year>
      </pub-date>
      <abstract>
        <p>The MEPDaW workshop series targets one of the emerging and fundamental problems of the Web, specifically the management and preservation of evolving knowledge graphs. During the past seven years, the workshop series has been gathering a community of researchers and practitioners around these challenges. To date, the series has successfully published more than 35 articles allowing more than 50 individual authors to present and share their ideas. This 7th edition, virtually co-located with the International Semantic Web Conference (ISWC 2021), gathered the community around six research publications and one invited keynote presentation. The event took place online on the 25th of October, 2021.</p>
      </abstract>
      <kwd-group>
        <kwd>Web Data evolution</kwd>
        <kwd>Data preservation</kwd>
        <kwd>provenance and lineage</kwd>
        <kwd>Temporal &amp; Evolving Knowledge Graphs</kwd>
        <kwd>RDF archiving and versioning</kwd>
      </kwd-group>
    </article-meta>
  </front>
  <body>
    <sec id="sec-1">
      <title>-</title>
      <p>
        There is a vast and rapidly increasing quantity of scientific, corporate,
government, and crowd-sourced data openly published on the Web. Open Data
plays a catalyst role in the way structured information is exploited on a large
scale. A traditional view of digitally preserving these datasets by “pickling and
locking them away” for future use, like groceries, conflicts with their evolution.
There are several approaches and frameworks (e.g. Linked Data Stack [
        <xref ref-type="bibr" rid="ref1">7</xref>
        ],
PoolParty Suite1, Metaphactory2, etc.) targeted at managing the life-cycle of the
Data Web. More specifically, these solutions are expected to tackle major issues
such as the synchronisation problem (monitoring changes) [
        <xref ref-type="bibr" rid="ref3 ref8">9,14</xref>
        ], the curation
problem (repairing data imperfections) [
        <xref ref-type="bibr" rid="ref5">11</xref>
        ], the appraisal problem (assessing
the quality of a dataset) [
        <xref ref-type="bibr" rid="ref2">8</xref>
        ], the citation problem (how to cite a particular
version of a dataset) [
        <xref ref-type="bibr" rid="ref6">12</xref>
        ], the archiving problem (retrieving a specific version of a
      </p>
      <p>
        dataset) [
        <xref ref-type="bibr" rid="ref4 ref7">10,13</xref>
        ], and the sustainability problem (preserving at scale, ensuring
long-term access) [
        <xref ref-type="bibr" rid="ref6">12</xref>
        ].
      </p>
      <p>The seventh edition of this workshop was organised for the second time at
the International Semantic Web Conference (ISWC) and followed the structure
of the previous editions. We invited a number of experts in the field of Linked
Data and Data Evolution &amp; Preservation in order to suggest and advise on the
diferent topics that our workshop covered this year. This year, at ISWC 2022,
we successfully gathered more than 50 participants for our half-day event. In
line with most academic events, this year MEPDaW was held as a virtual event
and we had to re-think the interactions between participants.</p>
    </sec>
    <sec id="sec-2">
      <title>MEPDaW Scientific programme</title>
      <p>The workshop started with the keynote entitled “How can we fix the Web of
Data?” given by Prof. Katja Hose3 from the Department of Computer Science
of the Aalborg University (Denmark). She initiated her presentation from the
observation that Semantic Web practitioners typically consider the Web of Data
as a static corpus of information always available and unmutable; however, “in
real life settings”, a broad range of problems hits the practitioners such as
unavailability of entire knowledge graphs or dead-links for the associated SPARQL
endpoints. And more generally, the current Semantic Web tools and paradigms
(almost-) completely miss the concept of versioning and provenance of metadata.
During her keynote, Professor Hose highlighted some of the solutions her group
developed to mitigate these problems. She first showed how to keep knowledge
available for continuous and scalable querying. Then, she presented the
attendees an approach that enables community-driven updates so that mistakes can
be corrected or missing information can be added. And finally, she described
how learning from RDF archiving can be done using solutions to better support
evolving knowledge graphs. Overall, this keynote [2] gave the audience in-depth
details on practical (and industrial) use cases backed by cutting-edge research
techniques.</p>
      <p>The first article presented dealt with an approach which helps SPARQL
practitioners to know which SPARQL endpoints has been updated when they
run complex pipelines relying on several RDF sources [1]. It was followed by [5]
which proposes the use of a visual interface to explore and fix multi-dimensional
metadata bases, in particular she showed how she will apply these ideas in the
context of popular music data during her PhD. Finally, the first paper-session
ended-up with the presentation of TrieDF [3]: a solution to index
metadataaugmented RDF datasets inspired by the trie data structure.</p>
      <p>The second session started with an industrial talk from J. Fernández who
described how clinical data standards at Roche benefit from RDF version
management.The next efort [6] focused on provided the audience with several
application use-cases where our the eforts of our community could contribute to.
3 http://people.cs.aau.dk/~khose/About_me.html
Finally, the last article of the workshop described UpLOD [4], a tool to repair
broken links in the linked-open data.</p>
      <sec id="sec-2-1">
        <title>Organizing Committee</title>
        <p>– Fabrizio Orlandi, ADAPT SFI Centre, Trinity College Dublin, Ireland
– Damien Graux, Inria, Université Côte d’Azur, CNRS, I3S, France
– Julio Cesar dos Reis, Inst. of Computing, Univ. of Campinas, Brazil
– Maria-Esther Vidal, TIB, Hannover, Germany</p>
      </sec>
      <sec id="sec-2-2">
        <title>Advisory Board</title>
        <p>– Philippe Cudré-Mauroux, eXascale Infolab, Univ. of Fribourg, Switzerland
– Jeremy Debattista, TopQuadrant Inc
– Javier D. Fernández, Information Architect at Roche, Switzerland
– Fabien Gandon, Inria, Université Côte d’Azur, CNRS, I3S, France
– Axel Polleres, Vienna University of Economics and Business, Austria
Programme Committee</p>
      </sec>
    </sec>
    <sec id="sec-3">
      <title>Acknowledgements</title>
      <p>We would like to thank all the authors, reviewers, committee members and the
speakers for their contributions, support and commitment.</p>
      <p>These research activities were conducted with the financial support of the
European Union’s Horizon 2020 research and innovation programme under the Marie
Skłodowska-Curie Grant Agreement No. 713567 at the ADAPT SFI Research
Centre at Trinity College Dublin. The ADAPT SFI Centre for Digital Media
Technology is funded by Science Foundation Ireland through the SFI Research
Centres Programme and is co-funded under the European Regional Development
Fund (ERDF) through Grant #13/RC/2106_P2.</p>
    </sec>
  </body>
  <back>
    <ref-list>
      <ref id="ref1">
        <mixed-citation>
          7.
          <string-name>
            <surname>Auer</surname>
            ,
            <given-names>S.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Bühmann</surname>
            ,
            <given-names>L.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Dirschl</surname>
            ,
            <given-names>C.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Erling</surname>
            ,
            <given-names>O.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Hausenblas</surname>
            ,
            <given-names>M.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Isele</surname>
            ,
            <given-names>R.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Lehmann</surname>
            ,
            <given-names>J.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Martin</surname>
            ,
            <given-names>M.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Mendes</surname>
            ,
            <given-names>P.N.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Van Nufelen</surname>
            ,
            <given-names>B.</given-names>
          </string-name>
          , et al.:
          <article-title>Managing the life-cycle of linked data with the LOD2 stack</article-title>
          . In: International semantic Web conference. pp.
          <fpage>1</fpage>
          -
          <lpage>16</lpage>
          . Springer (
          <year>2012</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref2">
        <mixed-citation>
          8.
          <string-name>
            <surname>Debattista</surname>
            ,
            <given-names>J.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Auer</surname>
            ,
            <given-names>S.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Lange</surname>
            ,
            <given-names>C.</given-names>
          </string-name>
          :
          <article-title>Luzzu-a methodology and framework for linked data quality assessment</article-title>
          .
          <source>J. Data and Information Quality</source>
          <volume>8</volume>
          (
          <issue>1</issue>
          ) (Oct
          <year>2016</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref3">
        <mixed-citation>
          9.
          <string-name>
            <surname>Endris</surname>
            ,
            <given-names>K.M.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Faisal</surname>
            ,
            <given-names>S.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Orlandi</surname>
            ,
            <given-names>F.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Auer</surname>
            ,
            <given-names>S.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Scerri</surname>
            ,
            <given-names>S.</given-names>
          </string-name>
          :
          <article-title>Interest-based RDF update propagation</article-title>
          .
          <source>In: Proceedings of the 14th International Conference on The Semantic Web - ISWC 2015 - Volume</source>
          <volume>9366</volume>
          . p.
          <fpage>513</fpage>
          -
          <lpage>529</lpage>
          . Springer-Verlag, Berlin, Heidelberg (
          <year>2015</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref4">
        <mixed-citation>
          10.
          <string-name>
            <surname>Fernández</surname>
            ,
            <given-names>J.D.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Polleres</surname>
            ,
            <given-names>A.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Umbrich</surname>
          </string-name>
          , J.:
          <article-title>Towards eficient archiving of dynamic linked open data</article-title>
          . In: MEPDaW workshop at ESWC'
          <volume>15</volume>
          (
          <year>2015</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref5">
        <mixed-citation>
          11.
          <string-name>
            <surname>Freitas</surname>
            ,
            <given-names>A.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Curry</surname>
          </string-name>
          , E.:
          <article-title>Big data curation</article-title>
          . In:
          <article-title>New Horizons for a Data-Driven Economy (</article-title>
          <year>2016</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref6">
        <mixed-citation>
          12.
          <string-name>
            <surname>Gleim</surname>
            ,
            <given-names>L.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Decker</surname>
            ,
            <given-names>S.</given-names>
          </string-name>
          :
          <article-title>Timestamped URLs as persistent identifiers</article-title>
          .
          <source>In: Proceedings of the 6th Workshop on Managing the Evolution and Preservation of the Data Web (MEPDaW)</source>
          (
          <year>2020</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref7">
        <mixed-citation>
          13.
          <string-name>
            <surname>Pelgrin</surname>
            ,
            <given-names>O.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Galárraga</surname>
            ,
            <given-names>L.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Hose</surname>
            ,
            <given-names>K.</given-names>
          </string-name>
          :
          <article-title>Towards fully-fledged archiving for RDF datasets</article-title>
          .
          <source>Semantic Web (Preprint)</source>
          ,
          <fpage>1</fpage>
          -
          <lpage>24</lpage>
          (
          <year>2020</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref8">
        <mixed-citation>
          14.
          <string-name>
            <surname>Tasnim</surname>
            ,
            <given-names>M.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Collarana</surname>
            ,
            <given-names>D.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Graux</surname>
            ,
            <given-names>D.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Orlandi</surname>
            ,
            <given-names>F.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Vidal</surname>
            ,
            <given-names>M.E.</given-names>
          </string-name>
          :
          <article-title>Summarizing entity temporal evolution in knowledge graphs</article-title>
          .
          <source>In: Companion Proceedings of The 2019 World Wide Web Conference</source>
          . p.
          <fpage>961</fpage>
          -
          <lpage>965</lpage>
          . WWW '
          <volume>19</volume>
          ,
          <string-name>
            <surname>Association</surname>
          </string-name>
          for Computing Machinery, New York, NY, USA (
          <year>2019</year>
          )
        </mixed-citation>
      </ref>
    </ref-list>
  </back>
</article>