<!DOCTYPE article PUBLIC "-//NLM//DTD JATS (Z39.96) Journal Archiving and Interchange DTD v1.0 20120330//EN" "JATS-archivearticle1.dtd">
<article xmlns:xlink="http://www.w3.org/1999/xlink">
  <front>
    <journal-meta />
    <article-meta>
      <title-group>
        <article-title>Building Word-Emotion Mapping Dictionary for Online News</article-title>
      </title-group>
      <contrib-group>
        <contrib contrib-type="author">
          <string-name>Yanghui Rao</string-name>
          <xref ref-type="aff" rid="aff0">0</xref>
        </contrib>
        <contrib contrib-type="author">
          <string-name>Xiaojun Quan</string-name>
          <xref ref-type="aff" rid="aff0">0</xref>
        </contrib>
        <contrib contrib-type="author">
          <string-name>Liu Wenyin</string-name>
          <xref ref-type="aff" rid="aff0">0</xref>
        </contrib>
        <contrib contrib-type="author">
          <string-name>Qing Li</string-name>
          <xref ref-type="aff" rid="aff0">0</xref>
        </contrib>
        <contrib contrib-type="author">
          <string-name>Mingliang Chen</string-name>
          <xref ref-type="aff" rid="aff0">0</xref>
        </contrib>
        <aff id="aff0">
          <label>0</label>
          <institution>Department of Computer Science, City University of Hong Kong</institution>
        </aff>
      </contrib-group>
      <fpage>28</fpage>
      <lpage>39</lpage>
      <abstract>
        <p>Sentiment analysis of online documents such as news articles, blogs and microblogs has received increasing attention. We propose an efficient method of automatically building the word-emotion mapping dictionary for social emotion detection. In the dictionary, each word is associated with the distribution on a series of human emotions. In addition, three different pruning strategies are proposed to refine the dictionary. Experiment on the real-world data sets has validated the effectiveness and reliability of the method. Compared with other lexicons, the dictionary generated using our approach is more adaptive for personalized data set, language-independent, fine-grained, and volume-unlimited. The generated dictionary has a wide range of applications, including predicting the emotional distribution of news articles and tracking the change of social emotions on certain events over time.</p>
      </abstract>
      <kwd-group>
        <kwd>Social emotion detection</kwd>
        <kwd>emotion dictionary</kwd>
        <kwd>maximum likelihood estimation</kwd>
      </kwd-group>
    </article-meta>
  </front>
  <body>
    <sec id="sec-1">
      <title>-</title>
      <p>In the traditional society, when we make a decision, opinions and emotions of
others have always been important information for reference. Knowing the
answer of “What others think and feel” is usually very necessary for general people,
marketers, public relations officials, politicians and managers.</p>
      <p>
        Nowadays, everyone can express their opinions and emotions easily through
news portals, blogs and microblogs, and they become both the listeners and
speakers. Facing the vast amount of data, tasks of automatically detecting
public emotions evoked by online documents is emerging recently [
        <xref ref-type="bibr" rid="ref1">1</xref>
        ], such as the
SemEval task 14. This task is treated as a classification problem according to
the polarity (positive, neutral or negative) or multiple emotion categories such
as joy, sadness, anger, fear, disgust and surprise. However, due to the limited
information in the news titles, annotating news headlines for emotions is a hard
task. It is usually intractable to annotate headlines consistently even for human
[
        <xref ref-type="bibr" rid="ref2">2</xref>
        ]. As a result, we mainly focus on annotating news bodies for emotions, and
building word-emotion mapping dictionaries in this paper.
      </p>
      <p>
        In previous works, emotions are mostly annotated based on the existing
emotional lexicons [
        <xref ref-type="bibr" rid="ref1">1</xref>
        ] [
        <xref ref-type="bibr" rid="ref3">3</xref>
        ], e.g., Subjectivity Wordlist [
        <xref ref-type="bibr" rid="ref4">4</xref>
        ], WordNet-Affect [
        <xref ref-type="bibr" rid="ref5">5</xref>
        ] and
SentiWordNet [
        <xref ref-type="bibr" rid="ref6">6</xref>
        ]. Emotion classification or opinion mining based on these
existing lexicons have their limited utility, because 1) the lexicons are mainly for
public use in general domains, some resulting classifications of words can appear
incorrect, and need to be adjusted to fit the personalized data set. 2) Most of
the lexicons are available only for bits of languages, such as English, and the
volume of words annotated is restricted, which limits the applicability of these
methods. 3) Some of the lexicons label words on coarse-grained dimensions
(positivity, negativity and neutrality), which are insufficient to individuate the whole
spectrum of emotional concepts [
        <xref ref-type="bibr" rid="ref5">5</xref>
        ].
      </p>
      <p>Unlike the above methods, we focus on building emotional dictionary
automatically, in which each item is scored along a number of predefined emotions.
Then, the emotion distributions of current news article are estimated accurately
based on the emotional dictionary. The main contributions are as follows:
– A method of building the word-emotion mapping dictionary is proposed,
which is efficient, precise and automatic, no human resource is needed.
– Three kinds of parameter-free pruning algorithms are presented to refine the
dictionary, and to improve the performance.
– Compared with the existing emotional lexicons, the emotional dictionary
constructed in this paper is more adaptively for personalized data set,
languageindependent, fine-grained, and can be updated constantly.</p>
      <p>Related works are given in Section 2. The problem definition, the method of
building the word-emotion mapping dictionary, pruning algorithms and
potential applications of the dictionary are presented in Section 3. The experimental
data sets, evaluation metrics, results and discussions are illustrated in Section
4. Finally, we draw conclusions and discuss future work in Section 5.
2</p>
    </sec>
    <sec id="sec-2">
      <title>Related Work</title>
      <p>Most of the previous works focus on constructing the emotional lexicons for
reviews, which is different with ours for news articles. The main features of
reviews and news articles are as follows:</p>
      <p>For the former data set, people usually explicitly express their opinions and
emotions in the reviews, which results in the subjective text; while for the latter
data set, news editor normally present the events objectively in the news
reports, and their opinions and emotions are transmitted implicitly. In other words,
the former data set mainly contains subjective sentences, which express some
personal feelings, opinions, views, emotions, or beliefs; while the latter data set
mainly contains objective sentences, which present some factual information.
Besides, for the former data set, as there exist fraudulent reviews or rumors, the
emotional dictionary maybe incorrect or biased; while for the latter data set, the
news reports are mainly objective and do not trigger the same problem.</p>
      <p>
        Works of sentiment analysis for reviews rose from the year 2001 or so. Das
and Chen [
        <xref ref-type="bibr" rid="ref7">7</xref>
        ] utilized classification algorithm to extract market emotions from
stock message boards, which was further used for decision on whether to buy or
sell a stock. However, the performance heavily depended on certain words. For
instance, the sentence “It is not a bear market” means a bull market actually,
because negation words such as “no”, “not” are much more important and serve
to reverse meaning. Turney [
        <xref ref-type="bibr" rid="ref8">8</xref>
        ] applied an unsupervised learning technique to
classify the emotional orientation of users’ reviews (such as reviews of movies,
travel destinations, automobiles and banks), in which the mutual information
differences between each phrase and the words “excellent” and “poor” were
calculated firstly. Then, the average emotional orientation of the phrases in the
review was used to classify the review as recommended or not recommended.
      </p>
      <p>
        During this incipient stage of research on sentiment analysis from reviews,
some of them focus on using linguistic heuristics or a set of seed words
preselected, to classify the emotional orientation of words or phrases [
        <xref ref-type="bibr" rid="ref9">9</xref>
        ]. Other
works focus on emotional categorization of entire documents, which are based
on the construction of discriminate-word dictionaries manually or semi-manually
[
        <xref ref-type="bibr" rid="ref7">7</xref>
        ]. However, previous experiments shown that the intuition of selecting
discriminating words may not always be the best for humans [
        <xref ref-type="bibr" rid="ref10">10</xref>
        ]. Besides classifying
emotions to positive or negative, predicting the rating scores of reviews has also
been done by researchers [
        <xref ref-type="bibr" rid="ref11">11</xref>
        ] [
        <xref ref-type="bibr" rid="ref12">12</xref>
        ]. As the rating scores are ordinal (e.g., 1-5
stars), the problem is tackled by regression. These previous works of sentiment
analysis from reviews are often performed on document, sentence, entity, and
feature/aspect level. Emotion classification at both the document and sentence
levels is useful, but it cannot find what aspects people liked or disliked.
Aspectbased emotional analysis is proposed to tackle such problem, but it is hard to
perform on news articles, in which aspects of entity are unknown.
      </p>
      <p>
        Works of emotion classification for news began from the SemEval tasks in
2007. Chaumartin [
        <xref ref-type="bibr" rid="ref1">1</xref>
        ] utilized a linguistic and rule-based approach to tag news
headlines for predefined emotions, which includes joy, sadness, anger, fear,
disgust and surprise, and for polarity, i.e. positive or negative. The algorithm was
based on existing emotional dictionaries, like WordNet-Affect and SentiWordNet.
Kolya et al. [
        <xref ref-type="bibr" rid="ref3">3</xref>
        ] identified event and emotional expressions at word level from the
sentences of TempEval-2010 corpus, in which the emotional expressions are also
identified simply based on the sentiment lexicons, e.g., Subjectivity Wordlist,
WordNet-Affect and SentiWordNet.
      </p>
      <p>
        These approaches based on public emotional dictionaries needed extra effort
of preprocessing and post-processing on individual words, because some resulting
classifications of words can appear incorrect, and need to be adjusted to fit the
personalized data set. Katz et al. [
        <xref ref-type="bibr" rid="ref2">2</xref>
        ] scored the emotions of each word as the
average of the emotions of every news headline, in which that word appears, all
non-content words were ignored. However, as the limited words in the news titles,
it faced the problem of the small number of words available for the analysis.
      </p>
      <p>In this paper, we mainly focus on annotating news bodies for emotions,
and building emotional dictionary automatically. The emotion expressions are
fine-grained (such as moving, sympathy, boring, angry and funny), rather than
coarse-grained (positive, negative and neutral). The dictionary can be used to
classify the emotional distributions of previous unseen news articles.</p>
    </sec>
    <sec id="sec-3">
      <title>Word-Emotion Mapping Dictionary Construction</title>
      <p>In this section, we will firstly define our research problem. Then, we introduce
the generation method of the word-emotion mapping dictionary, as well as the
pruning algorithms of the generated dictionary. Finally, we discuss the potential
applications of the dictionary.
3.1</p>
      <sec id="sec-3-1">
        <title>Problem Definition</title>
        <p>The research problem is defined as follows.</p>
        <p>Given N training news articles, a word-emotion mapping dictionary is
generated. The dictionary is a W × E matrix, and the (j, k ) item in this matrix is
the score (probability) of emotion ek conditioned on word wj .</p>
        <p>For each document di(i = 1, 2, ..., N ), the news content, the publication date
(timestamp), and the distribution of ratings of emotions in the predefined list
(see Fig. 1 as an example) are available. From these news contents, a vocabulary
is obtained as the source of the word-emotion mapping dictionary. The j -th word
in the vocabulary is denoted by wj (j = 1, 2, ..., W ), all the emotions is denoted
by e = (e1, e2, ..., eE ), the normalization form of ratings of di over e is denoted
by ri. ri = (ri1, ri2, ..., riE ), and |ri| = 1.
In this section, we introduce the method of generating word-emotion mapping
dictionary based on maximum likelihood estimation and the Jensen’s inequality.</p>
        <p>For each document di, the probability of ri conditioned on di can be
modeled as:</p>
        <p>W
P (ri|di) = X P (wj |di)P (ri|wj ) .</p>
        <p>j=1
(1)</p>
        <p>Where, the probability of ri conditioned on wj is a multinomial distribution,
and P (ri|wj ) = kE=1 P (ek|wj )rik . Then,</p>
        <p>W E
P (ri|di) = X P (wj|di) Y P (ek|wj)rik .</p>
        <p>j=1 k=1
In the above, words in document di are assumed to be independent.</p>
        <p>Let σij = P (wj|di) and θjk = P (ek|wj), the log-likelihood over all the N
documents can be defined as:</p>
        <p>N W E N W E
log l = log(Y(X σij Y θjrkik )) = X log(X σij Y θjrkik ) .</p>
        <p>i=1 j=1 k=1 i=1 j=1 k=1
According to Jensen’s inequality, we reconstruct the log-likelihood as follows:</p>
        <p>N W E
log l ≥ X X σij X rik log θjk .</p>
        <p>i=1 j=1 k=1</p>
        <p>Since PkE=1 θjk = 1, we add a Lagrange multiplier to the log-likelihood
equation as follows:</p>
        <p>N W E E
ˆl = X X σij X rik log θjk + λ(X θjk − 1) .</p>
        <p>i=1 j=1 k=1 k=1</p>
        <p>Then, we maximize the likelihood by calculating the first-order partial
derivative of θjk,</p>
        <p>Thus,
i.e.,</p>
        <p>∂ˆl
∂θjk</p>
        <p>i=1 θjk
= XN σijrik + λ = PiN=1 σijrik + λ = 0 .</p>
        <p>θjk
Since PkE=1 θjk = 1, we have
Then, substitute formula (8) into formula (7) and get
θjk = −</p>
        <p>PiN=1 σijrik .</p>
        <p>λ</p>
        <p>E N
λ = − X X σijrik .</p>
        <p>k=1 i=1
θjk =</p>
        <p>PiN=1 σijrik
PkE=1 PiN=1 σijrik</p>
        <p>.</p>
        <p>P (ek|wj) =</p>
        <p>PiN=1 P (wj|di)rik
PkE=1 PiN=1 P (wj|di)rik</p>
        <p>.
(2)
(3)
(4)
(5)
(6)
(7)
(8)
(9)
(10)</p>
        <p>In the above, P (ek|wj ) is the probability of emotion ek conditioned on
word wj from which we can generate the word-emotion mapping dictionary.
rik is the distribution of ratings of document di on emotion ek, P (wj |di) is the
probability of word wj conditioned on document di which can be calculated by
relative term frequency. The relative term frequency is the number of occurrences
of the term wj in di divide by the total number of occurrences of all the terms
in di.
3.3</p>
      </sec>
      <sec id="sec-3-2">
        <title>Pruning Algorithm</title>
        <p>As the size of the training data set increases, the scale of the dictionary extends,
making it hard for us to maintain and utilize. Thus, pruning operation is
necessary for such lexicons. We will give the definition of background word firstly,
and then illustrate how it can be used to prune the dictionary.</p>
        <p>Definition: Background word is the word that appears in most of the
documents in the training data set, it is general for specific domains and topics of
the training set, which is quite different with stop words for general domains.</p>
        <p>In the context of emotional annotation, the background words are general
words that contain little emotional information actually and will disturb the
effect of utilizing the dictionary. In contrast to other useful emotional tagging
words, the probability of a word being to background words, which is denoted by
P (B|w), is larger than the probability of the word being to emotions, which is
denoted by P (E|w). According to the definition, the probability P (B|w) can
be represented as follows:</p>
        <p>P (B|w) = dfw . (11)</p>
        <p>N</p>
        <p>In the above, dfw is the document frequency of word w, N is the total number
of documents in the training set. The proportion of documents that contains the
word w is larger, the probability of w being to background words is higher.</p>
        <p>As there are multiple emotions tagged for each word according to formula
(10), the latter probability P (E|w) has three forms, which are the maximum,
average and minimum of all values of P (ek|w), k is from 1 to E (the total
number of types of emotions). Then, the words are pruned from the dictionary
if P (B|w) is larger than P (E|w).</p>
        <p>When the pruning algorithm above is performed, the word-emotion mapping
dictionary is constructed to the end, which can be used to predict the emotions
of given news articles as follows:</p>
        <p>Pˆ(e|d) =</p>
        <p>X p(w|d)p(e|w) .
w∈W
(12)</p>
        <p>In the above, Pˆ(e|d) is the probability of social users having emotions e on
document d, P (w|d) is the distribution of new document d on word w, which can
be calculated by relative term frequency, P (e|w) is the probability of emotions e
conditioned on word w, which can be looked up from the word-emotion mapping
dictionary generated with formula (10).</p>
      </sec>
    </sec>
    <sec id="sec-4">
      <title>Experiments</title>
      <p>In this section, experiments are conducted on one Chinese data set and one
English data set, so as to test the effect of the word-emotion mapping dictionary
on sentiment analysis. The good performances and multilingual data sets reflect
the method’s effectiveness, reliability, and language-independent of building the
dictionary.
4.1</p>
      <sec id="sec-4-1">
        <title>Data Sets</title>
        <p>To test the adaptiveness, effectiveness and language-independent of our method
of building the word-emotion mapping dictionary, large-scale and multilingual
data sets are needed. Two kinds of data sets are employed in the experiment.</p>
        <p>Sina. This is a large-scale Chinese data set scrawled from Sina society, which
is one of the most popular news sites in China.1 The attributes include the URL
address of the news article, the news headline (title), the publish date (from
29 July, 2005 to 9 Sep, 2011), the news body (content), the user ratings on
emotions of touched, empathy, boredom, anger, amusement, sadness, surprise
and warmness. The data set contains 32,493 valid news articles with the total
number of ratings on the 8 emotions larger than 0. We use x (x = 90%, 80% ,
10%) of the data set for training and the remaining (1-x ) for testing, to evaluate
the scalability and stability of the method.</p>
        <p>SemEval. This is an English data set used in the 14th task of the 4th
International Workshop on Semantic Evaluations (SemEval-2007).2 The attributes
include the news headline, the score of emotions of anger, disgust, fear, joy, sad
and surprise normalizing from 0 to 100. The data set contains 1,246 valid news
headlines with the total score of the 6 emotions larger than 0. We use the 1,000
in the test-set (80% of the data set) for training and the 246 in the trial-set
(20%) for test.
4.2</p>
      </sec>
      <sec id="sec-4-2">
        <title>Evaluation Metrics</title>
        <p>Classifying and predicting the emotions of given news articles are efficient ways
to validate the effectiveness of the generated word-emotion mapping dictionary.</p>
        <p>The Pearson’s correlation coefficient is employed to measure the accuracy of
emotion prediction, which indicates the linear dependence between two variables.
A value closer to 1 indicates the predicted and the actual emotional distribution
fit better, and is reasonable to assert that the trend of ratings on emotions is
predicted well by the word-emotion mapping dictionary.</p>
        <p>We denote the Pearson’s correlation coefficient between the predicted and
the actual emotion distributions of the i -th article by pri, and the average value
of the Pearson’s correlation coefficient of all articles by r average, which is used
as the first metric.
1 http://news.sina.com.cn/society/.
2 http://nlp.cs.swarthmore.edu/semeval/tasks/.</p>
        <p>N
r average = X
i=1 N
pri
.</p>
        <p>(13)</p>
        <p>Besides the average value of the Pearson’s correlation coefficient, there is
another interesting metrics to evaluate the quality of emotion prediction. In
practice, when we predicting the multiple emotional distributions, the dominate
one with the maximum predicted rating value is attractive.</p>
        <p>m
p max = . (14)</p>
        <p>N</p>
        <p>In the above, m is the number of articles that the predicted and the actual
dominate emotion matched. N is the total number of articles for training or test.
4.3</p>
      </sec>
      <sec id="sec-4-3">
        <title>Results and Analysis</title>
        <p>In generating the word-emotion mapping dictionary, the probability of word
conditioned on document is calculated by relative term frequency according to
formula (10). We denote it by rtf. In pruning algorithms, maximum, average and
minimum are used to refine the dictionary generated by rtf (see section 3.3). We
denote these three algorithms by rtf-max, rtf-ave and rtf-min.</p>
        <p>Results of Sina For different scales of training data set in Sina, the number
of words pruned by rtf-max, rtf-ave and rtf-min are presented in Table 1. The
number of the original words in the dictionary ranges from 39,278 to 72,773,
within which 45.4% to 31.2% words are pruned using rtf-min, the words being
pruned are quite less using rtf-max and rtf-ave.
observed. The first one lies in that, our dictionary fits the training set well
even when the available emotional tagged data is limited. Second, although it is
harder to fit the training set when the scale is larger, our dictionary is robust
for the stability performance on large training sets. For pruning algorithms,
the performances after pruning by rtf-max, rtf-ave and rtf-min are better than
others without pruning, among which rtf-min performs the best, which shows
the significance of our pruning algorithm on refining the dictionary.</p>
        <p>For the test set, as the number of test articles increases, the quality of
emotion prediction remains stable mostly, except when the size of test articles is
12,997. This indicates the reliability and stability of the dictionary on predicting
emotions of previously unseen articles. For pruning algorithms, the performances
by rtf-max, rtf-ave and rtf-min are better than others without pruning, among
which rtf-min performs the best.</p>
        <p>Although rtf-min yields the best results for both training and test data sets,
and the improvement over benchmark is remarkable, we also refine the dictionary
by deleting the same proportion of words as rtf-min randomly, and perform t
hypothesis testing on pairwise methods, so as to verify the significant improvement
of our pruning algorithm on performances statistically.</p>
        <p>The results are depicted in Table 2. For the dictionary after pruning randomly
(prune-random) and the dictionary without pruning (rtf ), all of the significance
values are much larger than the conventional significance level 0.05, which
indicates the dictionary after pruning randomly is no significant different with the
dictionary without pruning. In fact, the quality of prune-random on the training
data set is worse than rtf when the size of training documents is 9,748, 12,997 and
19,496, and the quality between them is approximate for other scales of training
documents. These findings are similar on the test data set. On the other hand,
for the dictionary after pruning by rtf-min and rtf, or rtf-min and prune-random,
all of the significance values are below the conventional significance level 0.05,
which indicates the dictionary after pruning by our method is significant
different with others. In our case, we can infer that the dictionary after pruning by
rtf-min achieves significant performance improvement on both training and test
data sets, while pruning randomly does not get such improvement statistically.</p>
        <p>Above all, the word-emotion mapping dictionary is effective on emotion
classification and prediction. One of the most interesting observations is that when
the dictionary is pruned by rtf-min, more than 30% words are deleted, while the
performances are much better than others.</p>
        <p>Results of SemEval Despite that our focus is mainly on annotating emotions
for news bodies with long text, it would be very interesting to evaluate the
method and pruning algorithms on emotion prediction for news headlines.</p>
        <p>The first observation is that when building the word-emotion mapping
dictionary based on the short text, as the sparse of the vector, the prune operation
maybe unnecessary. For the 1,000 English news headlines used for training here,
the vocabulary size is only 2,380 after stemming while retaining the stop words.
When the pruning algorithm is applied, the number of pruned words is 0 for
rtf-max and rtf-ave, which means the pruning operation by maximum and
average is unnecessary for the data set. The ratio of pruned words is 68.66% for
rtf-min, which makes the size of the dictionary even smaller, and 7.30% of the
training headlines have no word exists in the dictionary, the ratio is 11.38% for
test headlines. As a result, pruning by minimize is unsuitable for the SemEval
data set, which contains quite limited words.</p>
        <p>The second observation is that our method of generating the dictionary works
well on fitting the training set for news headlines. The average correlation
coefficient of all training articles is 0.86 using the relative term frequency, which
shows a strong positive correlation between the predicted and actual emotion
distribution. However, the average correlation coefficient of all test articles is 0.36
using the relative term frequency, which means the precision of the dictionary on
predicting the emotion distribution of previous unseen documents is relatively
low. The reason is that the volume of the word-emotion mapping dictionary is
quite small for the limited information of news headlines.
5</p>
      </sec>
    </sec>
    <sec id="sec-5">
      <title>Conclusion</title>
      <p>Emotion and opinion mining is useful and meaningful from political, economical,
commercial, social and psychological perspectives, the word-emotion mapping
dictionary constructed in this paper is the first step to meet the needs. Different
from previous methods, our method of building the dictionary is adaptive for
personalized data set, volume-unlimited, automatically, language-independent,
and fine-grained. The main conclusions are as follows:</p>
      <p>First of all, the pruning algorithm is effective in refining the dictionary, and
improving the performances of emotion prediction. For three forms of removing
background words, which are maximum, average and minimum, the last one
achieves the largest improvement on the performances, and the improvement is
statistically significant under hypothesis testing.</p>
      <p>Secondly, as the number of training articles increases, the quality of emotion
prediction on training data sets decreases firstly and then remains stable. This
indicates that our dictionary fits the training data set well even when the
available tagged data is limited. Although it is harder to fit the training data set when
the scale is larger, our dictionary is robust for the stability performance on large
training sets. As the number of test articles increases, the quality of emotion
prediction on test data sets remains stable mostly. This indicates the reliable
of the word-emotion mapping dictionary on predicting emotions of previously
unseen articles.</p>
      <p>Last but not least, for annotating emotions of news headlines, it is
unnecessary to prune the dictionary, due to the limited vocabulary in the short text.
Thus, researches on emotional annotation for both long and short text are our
future focuses.</p>
      <p>Acknowledgments. The work described in this paper has been supported by
the NSFC Overseas, Hong Kong &amp; Macao Scholars Collaborated Researching
Fund (61028003)</p>
    </sec>
  </body>
  <back>
    <ref-list>
      <ref id="ref1">
        <mixed-citation>
          1.
          <string-name>
            <surname>Chaumartin</surname>
            ,
            <given-names>F.R.:</given-names>
          </string-name>
          <article-title>Upar7: A knowledge-based system for headline sentiment tagging</article-title>
          .
          <source>In: 4th International Workshop on Semantic Evaluations</source>
          , pp.
          <fpage>422</fpage>
          -
          <lpage>425</lpage>
          . Association for Computational Linguistics, Prague (
          <year>2007</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref2">
        <mixed-citation>
          2.
          <string-name>
            <surname>Katz</surname>
            ,
            <given-names>P.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Singleton</surname>
            ,
            <given-names>M.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Wicentowski</surname>
          </string-name>
          , R.:
          <article-title>Swat-mp: The semeval-2007 systems for task 5 and task 14</article-title>
          . In: 4th International Workshop on Semantic Evaluations, pp.
          <fpage>308</fpage>
          -
          <lpage>313</lpage>
          . Association for Computational Linguistics, Prague (
          <year>2007</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref3">
        <mixed-citation>
          3.
          <string-name>
            <surname>Kolya</surname>
            ,
            <given-names>A.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Das</surname>
            ,
            <given-names>D.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Ekbal</surname>
            ,
            <given-names>A.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Bandyopadhyay</surname>
            ,
            <given-names>S.</given-names>
          </string-name>
          :
          <article-title>Identifying Event-Sentiment Association using Lexical Equivalence and Co-reference Approaches</article-title>
          .
          <source>In: Workshop on Relational Models of Semantics Collocated with ACL</source>
          , pp.
          <fpage>19</fpage>
          -
          <lpage>27</lpage>
          .
          <string-name>
            <surname>Portland</surname>
          </string-name>
          (
          <year>2011</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref4">
        <mixed-citation>
          4.
          <string-name>
            <surname>Banea</surname>
            ,
            <given-names>C.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Mihalcea</surname>
            <given-names>R.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Wiebe</surname>
            <given-names>J.:</given-names>
          </string-name>
          <article-title>A Bootstrapping Method for Building Subjectivity Lexicons for Languages with Scarce Resources</article-title>
          .
          <source>In: 6th International Conference on Language Resources and Evaluation</source>
          ,
          <string-name>
            <surname>Marrakech</surname>
          </string-name>
          (
          <year>2008</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref5">
        <mixed-citation>
          5.
          <string-name>
            <surname>Strapparava</surname>
            ,
            <given-names>C.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Valitutti</surname>
            ,
            <given-names>A.</given-names>
          </string-name>
          :
          <article-title>Wordnet-affect: an affective extension of wordnet</article-title>
          .
          <source>In: 4th International Conference on Language Resources and Evaluation</source>
          , pp.
          <fpage>1083</fpage>
          -
          <lpage>1086</lpage>
          . Lisbon (
          <year>2004</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref6">
        <mixed-citation>
          6.
          <string-name>
            <surname>Baccianella</surname>
            <given-names>S.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Esuli</surname>
            <given-names>A.</given-names>
          </string-name>
          , Sebastiani F.:
          <article-title>SentiWordNet 3.0: An Enhanced Lexical Resource for Sentiment Analysis and Opinion Mining</article-title>
          .
          <source>In: 7th Conference on Language Resources and Evaluation</source>
          , pp.
          <fpage>2200</fpage>
          -
          <lpage>2204</lpage>
          .
          <string-name>
            <surname>Valletta</surname>
          </string-name>
          (
          <year>2010</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref7">
        <mixed-citation>
          7.
          <string-name>
            <surname>Das</surname>
            ,
            <given-names>S.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Chen</surname>
            ,
            <given-names>M.</given-names>
          </string-name>
          :
          <article-title>Yahoo! for Amazon: Extracting market sentiment from stock message boards</article-title>
          .
          <source>In: 8th Asia Pacific Finance Association Annual Conference</source>
          , Shanghai (
          <year>2001</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref8">
        <mixed-citation>
          8.
          <string-name>
            <surname>Turney</surname>
          </string-name>
          , P.D.:
          <article-title>Thumbs up or thumbs down? Semantic orientation applied to unsupervised classification of reviews. In: 40th annual meeting of the Association for Computational Linguistics</article-title>
          , pp.
          <fpage>417</fpage>
          -
          <lpage>424</lpage>
          . Philadelphia (
          <year>2002</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref9">
        <mixed-citation>
          9.
          <string-name>
            <surname>Turney</surname>
            ,
            <given-names>P.D.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Littman</surname>
            ,
            <given-names>M.L.</given-names>
          </string-name>
          :
          <article-title>Unsupervised learning of semantic orientation from a hundred-billion-word corpus</article-title>
          .
          <source>Technical Report EGB-1094</source>
          , National Research Council Canada (
          <year>2002</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref10">
        <mixed-citation>
          10.
          <string-name>
            <surname>Pang</surname>
            ,
            <given-names>B.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Lee</surname>
            ,
            <given-names>L.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Vaithyanathan</surname>
            ,
            <given-names>S.</given-names>
          </string-name>
          :
          <article-title>Thumbs up? Sentiment classification using machine learning techniques</article-title>
          .
          <source>In: Empirical Methods in Natural Language Processing</source>
          , pp.
          <fpage>79</fpage>
          -
          <lpage>86</lpage>
          , Philadelphia (
          <year>2002</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref11">
        <mixed-citation>
          11.
          <string-name>
            <surname>Liu</surname>
            ,
            <given-names>J.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Seneff</surname>
            ,
            <given-names>S.</given-names>
          </string-name>
          :
          <article-title>Review sentiment scoring via a parse-and-paraphrase paradigm</article-title>
          .
          <source>In: Empirical Methods in Natural Language Processing</source>
          ,
          <string-name>
            <given-names>ACL</given-names>
            ,
            <surname>Suntec</surname>
          </string-name>
          (
          <year>2009</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref12">
        <mixed-citation>
          12.
          <string-name>
            <surname>Ifrim</surname>
            ,
            <given-names>G.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Weikum</surname>
            ,
            <given-names>G.</given-names>
          </string-name>
          :
          <article-title>The Bag-of-Opinions Method for Review Rating Prediction from Sparse Text Patterns</article-title>
          .
          <source>In: Coling</source>
          <year>2010</year>
          , Beijing (
          <year>2010</year>
          )
        </mixed-citation>
      </ref>
    </ref-list>
  </back>
</article>