<!DOCTYPE article PUBLIC "-//NLM//DTD JATS (Z39.96) Journal Archiving and Interchange DTD v1.0 20120330//EN" "JATS-archivearticle1.dtd">
<article xmlns:xlink="http://www.w3.org/1999/xlink">
  <front>
    <journal-meta />
    <article-meta>
      <title-group>
        <article-title>Mining Social Networks to Learn about Rumors, Hate Speech, Bias and Polarization - Abstract</article-title>
      </title-group>
      <contrib-group>
        <contrib contrib-type="author">
          <string-name>Bárbara Poblete</string-name>
          <xref ref-type="aff" rid="aff0">0</xref>
          <xref ref-type="aff" rid="aff1">1</xref>
        </contrib>
        <aff id="aff0">
          <label>0</label>
          <institution>OHARS'20: Workshop on Online Misinformationand Harm-Aware Recommender Systems</institution>
        </aff>
        <aff id="aff1">
          <label>1</label>
          <institution>University of Chile</institution>
          ,
          <country country="CL">Chile</country>
        </aff>
      </contrib-group>
      <abstract>
        <p>Online social networks are a rich resource of unedited user-generated multimedia content. Buried within their day-to-day chatter, we can find breaking news, opinions and valuable insight into human behaviour, including the articulation of emerging social movements. Nevertheless, in recent years social platforms have become fertile ground for diverse information disorders and hate speech expressions. This situation poses an important challenge to the extraction of useful and trustworthy information from social media. In this talk I provide an overview of existing work in the area of social media information credibility, starting with our research in 2011 on rumor propagation during the massive earthquake in Chile in 2010 [1]. I discuss, as well, the complex problem of automatic hate speech detection in online social networks. In particular, how our review of the existing literature in the area shows important experimental errors and dataset biases that produce an overestimation of current state-of-the-art techniques [2]. Especifically, these issues become evident at the moment of attempting to apply these models to more diverse scenarios or to transfer this knowledge to languages other than English. As a particular way of dealing with the need to extract reliable information from online social media, I talk about two applications, Twically [3] and Galean [4]. These applications harvest collective signals created from social media text to provide a broad view of natural disasters and real-world news, respectively.</p>
      </abstract>
      <kwd-group>
        <kwd>eol&gt;Online social networks</kwd>
        <kwd>information credibility</kwd>
        <kwd>hate speech</kwd>
      </kwd-group>
    </article-meta>
  </front>
  <body />
  <back>
    <ref-list>
      <ref id="ref1">
        <mixed-citation>
          [1]
          <string-name>
            <given-names>C.</given-names>
            <surname>Castillo</surname>
          </string-name>
          ,
          <string-name>
            <given-names>M.</given-names>
            <surname>Mendoza</surname>
          </string-name>
          ,
          <string-name>
            <given-names>B.</given-names>
            <surname>Poblete</surname>
          </string-name>
          , Information credibility on twitter,
          <source>in: Proceedings of the 20th International Conference on World Wide Web, WWW '11</source>
          ,
          <string-name>
            <surname>Association</surname>
          </string-name>
          for Computing Machinery, New York, NY, USA,
          <year>2011</year>
          , p.
          <fpage>675</fpage>
          -
          <lpage>684</lpage>
          .
          <source>doi:1 0 . 1 1</source>
          <volume>4 5 / 1 9 6 3 4 0 5 . 1 9 6 3 5 0 0 .</volume>
        </mixed-citation>
      </ref>
      <ref id="ref2">
        <mixed-citation>
          [2]
          <string-name>
            <given-names>A.</given-names>
            <surname>Arango</surname>
          </string-name>
          ,
          <string-name>
            <given-names>J.</given-names>
            <surname>Pérez</surname>
          </string-name>
          ,
          <string-name>
            <given-names>B.</given-names>
            <surname>Poblete</surname>
          </string-name>
          ,
          <article-title>Hate speech detection is not as easy as you may think: A closer look at model validation (extended version)</article-title>
          ,
          <source>Information Systems (2020) 101584. doi:1 0 . 1 0 1 6 / j . i s . 2 0</source>
          <volume>2 0 . 1 0 1 5 8 4 .</volume>
        </mixed-citation>
      </ref>
      <ref id="ref3">
        <mixed-citation>
          [3]
          <string-name>
            <given-names>B.</given-names>
            <surname>Poblete</surname>
          </string-name>
          ,
          <string-name>
            <given-names>J.</given-names>
            <surname>Guzmán</surname>
          </string-name>
          ,
          <string-name>
            <given-names>J.</given-names>
            <surname>Maldonado</surname>
          </string-name>
          ,
          <string-name>
            <given-names>F.</given-names>
            <surname>Tobar</surname>
          </string-name>
          ,
          <article-title>Robust detection of extreme events using twitter: Worldwide earthquake monitoring</article-title>
          ,
          <source>IEEE Transactions on Multimedia</source>
          <volume>20</volume>
          (
          <year>2018</year>
          )
          <fpage>2551</fpage>
          -
          <lpage>2561</lpage>
          .
          <source>doi:1 0 . 1 1 0 9 / T M M . 2 0</source>
          <volume>1 8 . 2 8 5 5 1 0 7 .</volume>
        </mixed-citation>
      </ref>
      <ref id="ref4">
        <mixed-citation>
          [4]
          <string-name>
            <given-names>V.</given-names>
            <surname>Peña-Araya</surname>
          </string-name>
          ,
          <string-name>
            <given-names>M.</given-names>
            <surname>Quezada</surname>
          </string-name>
          ,
          <string-name>
            <given-names>B.</given-names>
            <surname>Poblete</surname>
          </string-name>
          ,
          <string-name>
            <given-names>D.</given-names>
            <surname>Parra</surname>
          </string-name>
          ,
          <article-title>Gaining historical and international relations insights from social media: spatio-temporal real-world news analysis using twitter</article-title>
          ,
          <source>EPJ Data Science</source>
          <volume>6</volume>
          (
          <year>2017</year>
          )
          <fpage>25</fpage>
          .
        </mixed-citation>
      </ref>
    </ref-list>
  </back>
</article>