<!DOCTYPE article PUBLIC "-//NLM//DTD JATS (Z39.96) Journal Archiving and Interchange DTD v1.0 20120330//EN" "JATS-archivearticle1.dtd">
<article xmlns:xlink="http://www.w3.org/1999/xlink">
  <front>
    <journal-meta />
    <article-meta>
      <title-group>
        <article-title>Filtering and Polarity Detection for Reputation Management on Tweets</article-title>
      </title-group>
      <contrib-group>
        <contrib contrib-type="author">
          <string-name>Viktor Hangya</string-name>
          <email>hangyav@gmail.com</email>
          <xref ref-type="aff" rid="aff0">0</xref>
        </contrib>
        <contrib contrib-type="author">
          <string-name>Richárd Farkas</string-name>
          <email>rfarkas@inf.u-szeged.hu</email>
          <xref ref-type="aff" rid="aff0">0</xref>
        </contrib>
        <aff id="aff0">
          <label>0</label>
          <institution>University of Szeged, Department of Informatics</institution>
        </aff>
      </contrib-group>
      <abstract>
        <p>In this paper we introduce our contribution to the RepLab 2013 - An evaluation campaign for Online Reputation Management Systems challenge. We participated in the filtering and polarity detection subtasks. The task of filtering is to determine whether a tweet is related to an entity. Then we classify tweets into positive, negative or neutral classes from the entity point of view. To solve these problems we employed supervised machine learning techniques. We applied several Twitter specific text preprocessing and features engineering methods. Besides supervised methods, we experimented with incorporating clustering information as well. Our system was ranked 2nd in the filtering task and 1st in the polarity detection task.</p>
      </abstract>
      <kwd-group>
        <kwd>Reputation management</kwd>
        <kwd>Sentiment analysis</kwd>
        <kwd>Twitter</kwd>
      </kwd-group>
    </article-meta>
  </front>
  <body>
    <sec id="sec-1">
      <title>Introduction</title>
      <p>
        In the past few years, the popularity of social media has increased. People post
messages on a variety of topics for example products, political issues, etc. Thus a
big amount of user generated data is created day-by-day. The manual processing
of this data is impossible, therefore automatic procedures are needed. Many
studies have been made in the area [
        <xref ref-type="bibr" rid="ref13 ref6">6, 13</xref>
        ], for example predicting the results of
elections [
        <xref ref-type="bibr" rid="ref14">14</xref>
        ], monitoring brands [
        <xref ref-type="bibr" rid="ref9">9</xref>
        ] or predicting stock price changes [
        <xref ref-type="bibr" rid="ref15">15</xref>
        ].
      </p>
      <p>
        In this paper we introduce our contribution for the RepLab 2013 An
evaluation campaign for Online Reputation Management Systems challenge [
        <xref ref-type="bibr" rid="ref3">3</xref>
        ]. Here,
the end-user application is monitoring the reputation of several entities, like
companies, organizations, celebrities, etc. from Twitter messages. The
organizers defined four tasks, namely filtering, polarity classification, topic detection
and assigning priority, from which we take part in the first two ones. In the case
of the filtering task the goal was to determine which tweets are related to a given
entity and which are not, for instance, distinguishing between tweets that
contain the word “Stanford” referring to the University of Stanford or to Stanford
as a place. This step is necessary as we can ignore the unrelated tweets, so this
step could be considered as a preprocessing step. The task of polarity detection
is to decide whether a given tweet carries positive, negative or neutral message
towards the entity in question.
The data provided by the organizers of the challenge consists of English and
Spanish messages. The data was crawled from 61 entities each from automotive,
banking, universities or music/artists domains. The data was collected using the
entities canonical names as queries. Besides the train and test databases a set of
background tweets were also provided which could be used while preparing our
system.
      </p>
      <p>
        We created a similar system for the filtering and polarity detection tasks. It
is an n-gram based supervised model as it has been shown that it works well
on short messages like tweets [
        <xref ref-type="bibr" rid="ref1 ref10 ref11 ref5">1, 5, 10, 11</xref>
        ]. We reduced the size of the dictionary
by normalizing the messages. Furthermore we examined novel methods which
increased the precision of our classifiers for example specially weighting our
features and using the results of topic modeling technologies. In case of the filtering
task the official evaluation metric was an F measure calculated from the
reliability and sensitivity values [
        <xref ref-type="bibr" rid="ref2">2</xref>
        ]. Besides the organizers provided the accuracy of
the participated systems as well. Our system achieved 0.438 F measure in this
task which is the second best result from all off the participated systems, and
achieved 0.928 accuracy which is the best. In the polarity detection task, the
official measure was the accuracy. We achieved an accuracy of 0.685, the highest
value among the participants.
2
      </p>
    </sec>
    <sec id="sec-2">
      <title>Approach</title>
      <p>
        We employed a unigram model along with tweet-specific normalization
techniques. We investigated novel features as well, which increased the accuracy
of our classifier. For implementation we used the MALLET toolkit, which is a
Java-based package for natural language processing [
        <xref ref-type="bibr" rid="ref12">12</xref>
        ].
2.1
      </p>
      <sec id="sec-2-1">
        <title>Normalization</title>
        <p>One reason for the unusually big dictionary size of the standard unigram model
is that it contains one word in many forms, for example in upper and lower
case, in a misspelled form, with character repetition, etc. On the other hand,
it contains various special annotations which are typical for blogging, such as
Twitter-specific annotations, URL’s, smiles, etc. Keeping these in mind we made
the following normalization steps:</p>
        <p>First, in order to get rid of the multiple forms of a single word we converted
them into lower case form then we stemmed them. For this purpose we used
the Porter Stemming Algorithm.</p>
        <p>We replaced the @ Twitter-specific tag and every URL with the [USER] and
[URL] notations, respectively. Besides, in case of a hash tag we deleted the
hash mark from it, for example we converted #funny to funny. This way we
do not distinguish Twitter specific tag from other words.</p>
        <p>Smileys in messages play an important role in polarity classification. For
this reason we grouped them into positive and negative smiley classes. We
considered :), :-),: ), :D, =), ;), ; ), (: and :(, :-(, : (, ):, ) : smileys as
positive and negative, respectively.</p>
        <p>Since numbers do not contain information regarding a message polarity, we
converted them as well to the [NUMBER] form. In addition, we replaced the
question and exclamation marks with the [QUESTION_MARK] and
[EXCLAMATION_MARK] notations. After this we removed the unnecessary
characters ’"#$%&amp;()*+,./:;&lt;=&gt;\^{}~, with the exception that we removed
the ’ character only if a word started or ended with it.</p>
        <p>In the case of words which contained character repetitions – more precisely
those which contained the same character at least three times in a row –, we
reduced the length of this sequence to three. For instance, in the case of the
word yeeeeahhhhhhh we got the form yeeeahhh. This way we unified these
character repetitions, but we did not loose this extra information.</p>
        <p>Before the normalization step, the dictionary contained approximately 113; 000
words. After the above introduced steps we managed to reduce the size of the
dictionary to 38; 000 words. It is important to mention that we handle English
and Spanish tweets the same way.
2.2</p>
      </sec>
      <sec id="sec-2-2">
        <title>Features</title>
        <p>After normalizing Twitter messages, we investigated feature space for a
supervised classifier. In many cases, phrases are important because they can catch
aspects of messages that simple unigrams can’t. For example “don’t like” if we
handle the two words separately we lose the knowledge that the negation word
refers to the word “like”. From this reason we examined the effects of bigrams
and trigrams and we realized that bigrams can improve the accuracy of our
classifier. However trigrams did not improve our result significantly so we used only
bigrams besides unigrams. Furthermore, we added a new feature as well which
is the number of negation words in a message.</p>
        <p>
          We searched for special features which characterize the polarity of the tweets.
One such feature is the polarity of each word in a message. To determine the
polarity of a word, we used the SentiWordNet sentiment lexicon [
          <xref ref-type="bibr" rid="ref4">4</xref>
          ]. In this
lexicon, a positive, negative and an objective real value belong to each word,
which describes the polarity of the given word. We created three new features
for each tweet which are the sum of the positive, negative and objective values
divided by the number of words in a message.
        </p>
        <p>For handling acronyms, we used an acronym lexicon which can be found on
the www.internetslang.com website. For each acronym we separately summed
up the positive and negative values of each word in the description of the acronym
and we normalized them by the number of words in the description. Then for each
tweet we added two new features which are the sums of the positive and negative
values of the acronyms in the message divided by the number of acronyms.</p>
        <p>Our intuition was that people like to use character repetitions in their
words for expressing their happiness or sadness. Besides normalizing these tokens
(see Section 2.1), we created a new feature as well, which represents the number
of this kind of words in a tweet.</p>
        <p>Beyond character repetitions people like to write words or a part of the text
in upper case in order to call the reader’s attention. Because of this we created
another feature which is the number of upper case words in the given text.</p>
        <p>Since the task is to filter tweets related to given entity and to classify them
by their sentiments, we found it important to sign whether the message contains
the mention of the entity or not (e.g. the username can also contains the name
of the entity). For this purpose we created a binary feature which indicates this
aspect.</p>
        <p>Furthermore it could be helpful to take into consideration the distance
between the token in question and the mention of the target entity. The
closer a token is to an entity the more the possibility that the given token is
related to the entity. For example consider following message where the first
sentence does not refer to BMW at all:</p>
        <p>I do agree that money can’t buy happiness. But somehow, it’s more
comfortable to sit and cry in a BMW than on a bicycle!
For this reason we weighted each word in the message by its distance from the
mention of given entities:</p>
        <p>1
e n1 ji jj
(1)
where n is the length of the message, i and j are the position of the actual word
and the mention of the entity in the message. Besides this, because different
properties could be positive or negative to different entities we created a new
feature which is the name of the entity, this way the classifier could handle
entities differently.</p>
        <p>
          Beyond the above supervised steps we experimented with leveraging
unlabeled data as well. We used Latent Dirichlet Allocation (LDA) [
          <xref ref-type="bibr" rid="ref7">7</xref>
          ] for detecting
topics on the train, test and background tweets provided by the RepLab 2013
organizers. The goal of topic modeling is to discover abstract topics that occur
in a collection of documents. As a result of LDA, we get the topic distribution
for each tweet and the topic for each word. We used the topic distributions over
each tweet as features. Each word belongs to a topic, so for a given message we
calculated the number of each topic by its content and used it as a feature as
well.
3
        </p>
      </sec>
    </sec>
    <sec id="sec-3">
      <title>Results</title>
      <p>
        In this section we report the results of the several systems which we
experimented with and their parametrization and the official results of the RepLab
2013 challenge. The training data which was provided by the organizers consists
of 36; 940 English and 8; 731 Spanish tweets. From this 15; 123 tweets are from
automotive, 7; 774 from banking, 6; 964 from universities and 15; 814 from
music/artists domains. The test data consists of 79; 981 English and 25; 118 Spanish
tweets which are distributed over the four domains similarly like the train data.
In our system we used Maximum Entropy classifier because we earlier showed
that it works well in document classification [
        <xref ref-type="bibr" rid="ref8">8</xref>
        ].
      </p>
      <p>Below we will show the effects of the methods which was introduced in section
2. On figures 1 and 2 the accuracy of our system for both filtering and polarity
detection tasks can be seen while adding more and more features to it. Our
baseline system is named naive which uses simple unigram features without any
normalization steps or extra features. In the norm version of our system we used
all of the normalization steps which we introduced earlier. It can be seen that this
step increased the accuracy for both tasks with a relatively large value. The next
step was to use the bigrams as features besides unigrams and normalization in the
N-gram system. This step increased the accuracy as well. The next system which
we named features uses several abstract features too, which are the polarity
of words and acronyms, the presence of character repetitions and upper case
words and the number of negation words. These features increased our results
marginally. In the so called weighting system we applied the distance based
weighting for each word in the message and in the entities system we used the
presence or absence of the given entity’s name in the messages. These steps
further increased our accuracy. In the last two systems we used the results of
LDA topic modeling with 50 topics. In the topic50 we used the topic distributions
over the messages as features and in the topic50-num we used the number of
topics feature as well. These methods slightly improved our classifier only for
the polarity task. In total, the introduced methods and features improved our
results by 3-4 percents compared to our baseline.</p>
      <p>In table 1 we show the effects of LDA with different number of topics. Here
we used a system that uses all of the above introduced methods and features. It
can be seen that LDA increased our results in both tasks, mainly the F measures.
The best topic number was 50 and 100 for the filtering and polarity detection
tasks, respectively.</p>
      <p>The RepLab 2013 participants were allowed to send up to ten different runs
per subtask. In case of the filtering task we achieved our best result with the
following system. We used all of the above mentioned methods and features, we
detected 50 topics with LDA on the train and test data. Furthermore we run
our classifier separately on the four domains (automotive, banking, university,
music/artists). This way we reached 0.438 F measure comparing to the best
system which reached 0.488 and the baseline provided by the organizers was 0.325.
With this system we reached 0.928 accuracy which turns out to be the highest
value. In case of the polarity detection task our best system was parametrized
as follows. Like before, we used all the technologies which we developed, we
detected only 20 topics on the train and test data. Unlike the filtering task, here
we run our classifier on the four domains at the same time. We achieved 0.685
accuracy which is the highest among all of the participated systems, the given
baseline was 0.584. We achieved our highest F measure with a different system
which similar to the previous but it did not use the polarity of the words and
the acronyms. This way we reached 0.381 F measure, whilst the given baseline
was 0.297.</p>
      <p>From this analysis we can conclude that the normalization of the messages
yielded a considerable increase in the accuracy of our classifier. We discussed
above that this step also significantly reduced the size of the dictionary. The
features and other methods increased the precision as well. During our experiments
we realized that in some cases LDA did not detect topics well which can cause
the low improvements in accuracy by LDA features. For example consider the
following topic which contains these words “the to for a in of and year a”. In
future it would be worth trying to improve the performance of topic modeling.
4</p>
    </sec>
    <sec id="sec-4">
      <title>Conclusions and Future Work</title>
      <p>Recently, sentiment analysis on Twitter messages has gained a lot of attention
due to the huge amount of Twitter users and their tweets. A commercial
extension of classical binary sentiment analysis is reputation management systems. In
this paper, we examined several methods for filtering relevant tweet by a given
entity and for classifying them by their sentiments for reputation management.
We proposed special features which characterize the polarity and other aspects
of tweets and we concluded that due to the informality (slang, spelling
mistakes, etc.) of the messages it is crucial to normalize them properly. Our system
achieved outstanding ranks in both filtering and polarity tasks of the RepLab
2013 challenge.</p>
      <p>In the future, we plan to investigate the utility of relations between Twitter
users and between their tweets. Furthermore we would like to examine several
domain adaption methods in such way that we use messages for training from
other sources than Twitter.</p>
    </sec>
    <sec id="sec-5">
      <title>Acknowledgments</title>
      <p>This work was supported in part by the European Union and the European Social
Fund through project FuturICT.hu (grant no.:
TÁMOP-4.2.2.C-11/1/KONV2012-0013).</p>
    </sec>
  </body>
  <back>
    <ref-list>
      <ref id="ref1">
        <mixed-citation>
          1.
          <string-name>
            <surname>Agarwal</surname>
            ,
            <given-names>A.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Xie</surname>
            ,
            <given-names>B.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Vovsha</surname>
            ,
            <given-names>I.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Rambow</surname>
            ,
            <given-names>O.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Passonneau</surname>
          </string-name>
          , R.:
          <article-title>Sentiment Analysis of Twitter Data</article-title>
          .
          <source>In: Proceedings of the Workshop on Language in Social Media (LSM</source>
          <year>2011</year>
          ).
          <source>(June</source>
          <year>2011</year>
          )
          <fpage>30</fpage>
          -
          <lpage>38</lpage>
        </mixed-citation>
      </ref>
      <ref id="ref2">
        <mixed-citation>
          2.
          <string-name>
            <surname>Amigó</surname>
            ,
            <given-names>E.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Corujo</surname>
            ,
            <given-names>A.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Gonzalo</surname>
            ,
            <given-names>J.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Meij</surname>
            ,
            <given-names>E.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Rijke</surname>
            ,
            <given-names>M.</given-names>
          </string-name>
          : Overview of replab 2012:
          <article-title>Evaluating online reputation management systems</article-title>
          .
          <source>In: CLEF 2012 Labs and Workshop</source>
          Notebook Papers. (
          <year>2012</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref3">
        <mixed-citation>
          3.
          <string-name>
            <surname>AmigÃş</surname>
          </string-name>
          , E., Carrillo de Albornoz, J.,
          <string-name>
            <surname>Chugur</surname>
            ,
            <given-names>I.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Corujo</surname>
            ,
            <given-names>A.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Gonzalo</surname>
            ,
            <given-names>J.</given-names>
          </string-name>
          , MartÃŋn, T.,
          <string-name>
            <surname>Meij</surname>
            , E., de Rijke,
            <given-names>M.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Spina</surname>
            ,
            <given-names>D.</given-names>
          </string-name>
          <article-title>In: Fourth International Conference of the CLEF initiative</article-title>
          ,
          <source>CLEF</source>
          <year>2013</year>
          ,
          <article-title>Valencia, Spain</article-title>
          . Proceedings. Springer LNCS, location =
        </mixed-citation>
      </ref>
      <ref id="ref4">
        <mixed-citation>
          4.
          <string-name>
            <surname>Baccianella</surname>
            ,
            <given-names>S.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Esuli</surname>
            ,
            <given-names>A.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Sebastiani</surname>
            ,
            <given-names>F.</given-names>
          </string-name>
          :
          <article-title>SentiWordNet 3.0: An Enhanced Lexical Resource for Sentiment Analysis and Opinion Mining</article-title>
          . In Chair),
          <string-name>
            <given-names>N.C.C.</given-names>
            ,
            <surname>Choukri</surname>
          </string-name>
          ,
          <string-name>
            <given-names>K.</given-names>
            ,
            <surname>Maegaard</surname>
          </string-name>
          ,
          <string-name>
            <given-names>B.</given-names>
            ,
            <surname>Mariani</surname>
          </string-name>
          ,
          <string-name>
            <given-names>J.</given-names>
            ,
            <surname>Odijk</surname>
          </string-name>
          ,
          <string-name>
            <given-names>J.</given-names>
            ,
            <surname>Piperidis</surname>
          </string-name>
          ,
          <string-name>
            <given-names>S.</given-names>
            ,
            <surname>Rosner</surname>
          </string-name>
          ,
          <string-name>
            <given-names>M.</given-names>
            ,
            <surname>Tapias</surname>
          </string-name>
          , D., eds.
          <source>: Proceedings of the Seventh International Conference on Language Resources and Evaluation (LREC'10)</source>
          , Valletta, Malta, European Language Resources Association (ELRA) (May
          <year>2010</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref5">
        <mixed-citation>
          5.
          <string-name>
            <surname>Barbosa</surname>
            ,
            <given-names>L.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Feng</surname>
          </string-name>
          , J.:
          <article-title>Robust Sentiment Detection on Twitter from Biased and Noisy Data</article-title>
          . In: Poster volume.
          <source>Coling</source>
          <year>2010</year>
          (
          <year>August 2010</year>
          )
          <fpage>36</fpage>
          -
          <lpage>44</lpage>
        </mixed-citation>
      </ref>
      <ref id="ref6">
        <mixed-citation>
          6.
          <string-name>
            <surname>Bifet</surname>
            ,
            <given-names>A.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Frank</surname>
          </string-name>
          , E.:
          <article-title>Sentiment Knowledge Discovery in Twitter Streaming Data</article-title>
          . (
          <year>2010</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref7">
        <mixed-citation>
          7.
          <string-name>
            <surname>Blei</surname>
            ,
            <given-names>D.M.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Ng</surname>
            ,
            <given-names>A.Y.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Jordan</surname>
            ,
            <given-names>M.I.</given-names>
          </string-name>
          :
          <article-title>Latent dirichlet allocation</article-title>
          .
          <source>the Journal of machine Learning research 3</source>
          (
          <year>2003</year>
          )
          <fpage>993</fpage>
          -
          <lpage>1022</lpage>
        </mixed-citation>
      </ref>
      <ref id="ref8">
        <mixed-citation>
          8.
          <string-name>
            <surname>Hangya</surname>
            ,
            <given-names>V.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Berend</surname>
            ,
            <given-names>G.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Farkas</surname>
          </string-name>
          , R.:
          <article-title>Szte-nlp: Sentiment detection on twitter messages</article-title>
          .
          <source>In: Second Joint Conference on Lexical and Computational Semantics (*SEM)</source>
          , Volume
          <volume>2</volume>
          :
          <source>Proceedings of the Seventh International Workshop on Semantic Evaluation (SemEval</source>
          <year>2013</year>
          ), Atlanta, Georgia, USA, Association for Computational Linguistics (
          <year>June 2013</year>
          )
          <fpage>549</fpage>
          -
          <lpage>553</lpage>
        </mixed-citation>
      </ref>
      <ref id="ref9">
        <mixed-citation>
          9.
          <string-name>
            <surname>Jansen</surname>
            ,
            <given-names>B.J.</given-names>
          </string-name>
          , Zhang,
          <string-name>
            <given-names>M.</given-names>
            ,
            <surname>Sobel</surname>
          </string-name>
          ,
          <string-name>
            <given-names>K.</given-names>
            ,
            <surname>Chowdury</surname>
          </string-name>
          ,
          <string-name>
            <surname>A.</surname>
          </string-name>
          :
          <article-title>Twitter Power: Tweets as Electronic Word of Mouth</article-title>
          . In:
          <article-title>Journal of the American society for information science and technology</article-title>
          . (
          <year>2009</year>
          )
          <fpage>2169</fpage>
          -
          <lpage>2188</lpage>
        </mixed-citation>
      </ref>
      <ref id="ref10">
        <mixed-citation>
          10.
          <string-name>
            <surname>Jiang</surname>
            ,
            <given-names>L.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Yu</surname>
            ,
            <given-names>M.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Zhou</surname>
            ,
            <given-names>M.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Liu</surname>
            ,
            <given-names>X.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Zhao</surname>
            ,
            <given-names>T.</given-names>
          </string-name>
          :
          <article-title>Target-dependent Twitter Sentiment Classification</article-title>
          .
          <source>In: Proceedings of the 49th Annual Meeting of the Association for Computational Linguistics</source>
          . (
          <year>June 2011</year>
          )
          <fpage>151</fpage>
          -
          <lpage>160</lpage>
        </mixed-citation>
      </ref>
      <ref id="ref11">
        <mixed-citation>
          11.
          <string-name>
            <surname>Liu</surname>
            ,
            <given-names>B.</given-names>
          </string-name>
          :
          <article-title>Sentiment Analysis and Subjectivity</article-title>
          . In Indurkhya, N.,
          <string-name>
            <surname>Damerau</surname>
          </string-name>
          , F.J., eds.:
          <source>Handbook of Natural Language Processing</source>
          . (
          <year>2010</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref12">
        <mixed-citation>
          12.
          <string-name>
            <surname>McCallum</surname>
            ,
            <given-names>A.K.</given-names>
          </string-name>
          :
          <article-title>Mallet: A machine learning for language toolkit</article-title>
          . http://mallet.cs.umass.edu (
          <year>2002</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref13">
        <mixed-citation>
          13.
          <string-name>
            <given-names>O</given-names>
            <surname>'Connor</surname>
          </string-name>
          ,
          <string-name>
            <given-names>B.</given-names>
            ,
            <surname>Balasubramanyan</surname>
          </string-name>
          ,
          <string-name>
            <given-names>R.</given-names>
            ,
            <surname>Routledge</surname>
          </string-name>
          ,
          <string-name>
            <given-names>B.R.</given-names>
            ,
            <surname>Smith</surname>
          </string-name>
          ,
          <string-name>
            <surname>N.A.</surname>
          </string-name>
          :
          <article-title>From Tweets to Polls: Linking Text Sentiment to Public Opinion Time Series</article-title>
          .
          <source>In: Proceedings of the International AAAI Conference on Weblogs and Social Media</source>
          . (May
          <year>2010</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref14">
        <mixed-citation>
          14.
          <string-name>
            <surname>Sang</surname>
            ,
            <given-names>E.T.K.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Bos</surname>
            ,
            <given-names>J.: Predicting</given-names>
          </string-name>
          <article-title>the 2011 Dutch Senate Election Results with Twitter</article-title>
          .
          <source>In: Proceedings of the 13th Conference of the European Chapter of the Association for Computational Linguistics. (April</source>
          <year>2012</year>
          )
          <fpage>53</fpage>
          -
          <lpage>60</lpage>
        </mixed-citation>
      </ref>
      <ref id="ref15">
        <mixed-citation>
          15.
          <string-name>
            <surname>Vu</surname>
            ,
            <given-names>T.T.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Chang</surname>
            ,
            <given-names>S.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Ha</surname>
            ,
            <given-names>Q.T.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Collier</surname>
            ,
            <given-names>N.:</given-names>
          </string-name>
          <article-title>An experiment in integrating sentiment features for tech stock prediction in twitter</article-title>
          .
          <source>In: 24th International Conference on Computational Linguistics</source>
          . (
          <year>2012</year>
          )
          <fpage>23</fpage>
        </mixed-citation>
      </ref>
    </ref-list>
  </back>
</article>