<!DOCTYPE article PUBLIC "-//NLM//DTD JATS (Z39.96) Journal Archiving and Interchange DTD v1.0 20120330//EN" "JATS-archivearticle1.dtd">
<article xmlns:xlink="http://www.w3.org/1999/xlink">
  <front>
    <journal-meta>
      <journal-title-group>
        <journal-title>IICST</journal-title>
      </journal-title-group>
    </journal-meta>
    <article-meta>
      <title-group>
        <article-title>TEXT MINING FOR ASPECT BASED SENTIMENT ANALYSIS ON CUSTOMER REVIEW: A CASE STUDY IN THE HOTEL INDUSTRY</article-title>
      </title-group>
      <contrib-group>
        <contrib contrib-type="author">
          <string-name>Fitra A. Bachtiar</string-name>
          <email>fitra.bachtiar@ub.ac.id</email>
          <xref ref-type="aff" rid="aff0">0</xref>
        </contrib>
        <contrib contrib-type="author">
          <string-name>Wirdhayanti Paulina</string-name>
          <xref ref-type="aff" rid="aff0">0</xref>
        </contrib>
        <contrib contrib-type="author">
          <string-name>Alfi Nur Rusydi</string-name>
          <xref ref-type="aff" rid="aff0">0</xref>
        </contrib>
        <aff id="aff0">
          <label>0</label>
          <institution>Faculty of Computer Science, Brawijaya University</institution>
        </aff>
      </contrib-group>
      <pub-date>
        <year>2020</year>
      </pub-date>
      <volume>5</volume>
      <fpage>105</fpage>
      <lpage>112</lpage>
      <abstract>
        <p>The development of the role of the OTA (Online Travel Agent) site has become one of the E-WOM (Electronic Word of Mouth) media in addition to its main function as a platform for ticket reservations to encourage stakeholders in the hotel industry to utilize E-WOM for business continuity. One of the guest houses in Malang realized the importance of E-WOM because 90 percent of the booking process originated from the OTA website. However, the process of processing customer reviews only focuses on physical reviews, namely Guest Reviews. Meanwhile, information from online sources can have a more significant impact on E-WOM. One of the techniques of text mining is sentiment analysis which can be used to process and group text reviews. Sentiment analysis can be done to determine the sentiment of opinions on customer reviews to determine customer satisfaction with guest house services that aim to produce a positive E-WOM. Sentiment analysis is carried out at the aspect level using aspects of location, room, food, price, and service. The text of the review used in Indonesian originates from the sites Agoda.com, Expedia, Pegi-Pegi, Booking.Com, TripAdvisor and has a timeline from 2012 to 2019. This research yields findings in the form of customer satisfaction analysis of the five aspects where food aspects have urgency to be addressed and corrected immediately. Evaluation of the classification results also proves the effectiveness of the SVM method from Naïve Bayes</p>
      </abstract>
      <kwd-group>
        <kwd>Guest House</kwd>
        <kwd>E-WOM</kwd>
        <kwd>Sentiment Analysis</kwd>
      </kwd-group>
    </article-meta>
  </front>
  <body>
    <sec id="sec-1">
      <title>1. INTRODUCTION</title>
      <p>
        The rapid growth of technology has encouraged the development of the role of the OTA (Online Travel Agent)
site as one of the E-WOM (Electronic Word Of Mouth) media in addition to its main function as a platform for
ticket reservations.
        <xref ref-type="bibr" rid="ref12">Westbrook (1987)</xref>
        states that all informal communication aimed at consumers through
internetbased technology related to the use or characteristics of a product, service, or provider is called E-WOM. Through
these sites, customers are expected to provide reviews / reviews about what they feel and experience after a visit
at the place they are going to either a hotel, restaurant, amusement vehicle and so on. Positive e-WOM can be
generated through good customer reviews, while good reviews can be generated from satisfying customer
experience for the services and accommodations produced.
      </p>
      <p>One of the guest houses in Malang realized the importance of E-WOM in its business continuity because 90
percent of the booking process originated from the OTA website. Currently, the guest house has been listed on
several OTA sites, namely TripAdvisor, Booking.com, Expedia, Agoda and Pegi-Pegi. However, the process of
processing customer reviews only focuses on physical reviews, namely Guest Reviews. Meanwhile, information
from online sources can have a more significant impact on E-WOM. The process of processing customer reviews
becomes ineffective because it only focuses on one source, while the evaluation of guest house management
services needs to be on target. In addition to these problems, the wide range of hotel attributes makes it difficult
for stakeholders to determine aspects that have urgency to be addressed immediately.</p>
      <p>
        A method that can be used to process and group text reviews is a sentiment analysis. Sentiment analysis or
opinion mining is a computational study of people's opinions, sentiments, and emotions through entities and
attributes that are expressed in text form
        <xref ref-type="bibr" rid="ref8">(Liu, 2012)</xref>
        . This sentiment analysis can classify the polarity of the text
in sentences or documents to find out whether the opinions on the sentence or document are positive or negative.
Sentiment analysis can be done to determine the sentiment of opinions on customer reviews to determine customer
satisfaction with guest house services that aim to produce a positive E-WOM.
      </p>
      <p>
        <xref ref-type="bibr" rid="ref4">Ekawati and Khodra (2017)</xref>
        uses sentiment analysis on restaurant reviews to help restaurant owners improve
the quality of their products and services. Sentiment analysis is carried out at the aspect level using aspects of food,
service, price, and place.
        <xref ref-type="bibr" rid="ref5">El-Jawad et al. (2018)</xref>
        describes several stages in sentiment analysis namely the data
collection phase, the preprocessing phase, the term weighting phase, the classification phase, and the evaluation
phase. Sentiment analysis will be carried out in 5 stages using the SVM method based on research
        <xref ref-type="bibr" rid="ref2">Bhavitha et al.
(2017)</xref>
        who analyzes comparisons of the techniques used in sentiment analysis. Researchers compared the lexicon
based approach with the machine learning approach. Lexicon based has an average accuracy of 70% where
      </p>
      <p>
        Machine Learning has an average accuracy that is much better above 80% and among the machine learning
classifier results Support Vector Machine has the best accuracy compared to other classifiers.
        <xref ref-type="bibr" rid="ref10">Miao et al. (2018)</xref>
        compared 3 machine learning methods, namely SVM, KNN, and Naïve Bayes for making a Chinese text news
classification system. SVM has advantages in the value of Recall, Precision, and Recall although it takes longer
than both methods because of the iteration process. Meanwhile, KNN and Naïve Bayes produce values that are
not much different. Finally,
        <xref ref-type="bibr" rid="ref11">Shi and Li (2011)</xref>
        uses the SVM method to compare TF-IDF with Frequency in the
Term Weighting process in Sentiment Analysis. TF-IDF proved to be more effective with a Recall value of 89.2%,
Precision 85.2%, and F1-Score 87.2%.
      </p>
      <p>
        This study will discuss the use of sentiment analysis will be carried out at the selected aspect level to group
reviews into 5 aspects and determine the sentiment of customer reviews by applying stages in text mining using
machine learning classification methods. These five aspects were chosen based on research by
        <xref ref-type="bibr" rid="ref3">Dolnicar and Otter
(2003)</xref>
        . The aspects used are the location, room, price, food and services chosen according to the needs of the
organization. The results of this study are in the form of findings that can help stakeholders understand what are
the customer complaints so that the decision making process to determine the services that need to be repaired and
addressed becomes more effective and targeted. This paper is divided into several sections, namely Introduction,
Literature Review, Methodology, Experiment, Analysis, and Conclusion.
      </p>
    </sec>
    <sec id="sec-2">
      <title>2. LITERATUR REVIEW 2.1</title>
    </sec>
    <sec id="sec-3">
      <title>Sentiment Analysis</title>
      <p>
        Sentiment analysis or opinion mining is an umbrella of branches of study such as opinion extraction, sentiment
mining, subjectivity analysis, affect analysis, emotion analysis, mining review, etc. Sentiment analysis is a field
of study that analyzes opinions, praise sentiments, one's emotions towards entities such as products, services,
organizations, events, problems, and attributes of entities
        <xref ref-type="bibr" rid="ref8">(Liu, 2012)</xref>
        . Sentiment analysis is divided into 3 levels,
namely: Document Level, Sentence Level, and Entity Level (aspect).
2.2
      </p>
    </sec>
    <sec id="sec-4">
      <title>Support Vector Machine</title>
      <p>
        The SVM algorithm aims to find Maximum Marginal Hyperplane (MMH) using support vectors and margins.
MMH is the best hyperplane with the largest margin distance used to separate data maximally and accurately for
each class. Margin can be defined as the shortest distance of a hyperplane to one side of the margin is the same as
the hyperplane distance to the other side of the margin, provided that both margins are in a parallel position with
the hyperplane
        <xref ref-type="bibr" rid="ref7">(Han et al., 2012)</xref>
        .
      </p>
      <p>If there is a dataset in the form (X1, y1), (X2, y2), (X3, y3), ... , (Xi, yi) where Xi is tuple training and yi is the
class label with  = 1 ....   ∈  and  ∈ {−1,1}. Every yi can choose one of two values either +1 or -1 SVM
will form a classifier as shown in the Equation (1) as follows:</p>
      <p>(!) = {("##,,%%!!"" )&amp;'' (1)
In SVM a hyperplane will be described in the following equation:</p>
      <p>.  +  = 0 (2)</p>
      <p>Based on equations (2), W is scalar Weight, n is attribute, b is scalar value or called bias, and X is training data
set or training tuples.
106
2.3</p>
    </sec>
    <sec id="sec-5">
      <title>Naïve Bayes</title>
      <p>
        Naive Bayes is an algorithm put forward by a British scientist Thomas Bayes which is a classification algorithm
using probability and statistical methods. This algorithm predicts opportunities based on past experience to be used
in the future so it is known as the Bayes Theorem
        <xref ref-type="bibr" rid="ref1">(Berrar, 2018)</xref>
        . This study uses a multinomial model that takes
into account the frequency of each word that appears in the document
        <xref ref-type="bibr" rid="ref9">(Manning et al., 2008)</xref>
        . For example there
are documents d and class c. To calculate the class of document d, it can be calculated using the formula:
(|  ) = ()  (|)  (|)  (|)  …  (|) (3)
Based on Equation (3), P (c) is prior probability of class c, tn is word document d nth, P (c | term document d)
is the probability of a document including class c and P (tn | c) = N-word probability with class c. The probability
of prior class c is determined by the formula:
() =  (4)
      </p>
      <p>Based on Equation (4), Nc is number of class c in all documents, N is number of all documents. The n-word
probability is determined using the laplacian smoothing technique:</p>
      <p>( | ) = ((),))|)| (5)</p>
      <p>In Equation (5), count (tn, c) is number of terms tn found in all training data with category c, count (c) is
number of terms in all training data with category c, V is number of all templates in the training data
2.4</p>
      <p>
        TF-IDF
TF-IDF Term Weighting is a weighting that is often used and is a combination of Term Frequency and Inverse
Document Frequency. TF-IDF consists of frequency terms and inverse documents obtained from dividing the total
number of documents to the number of documents that have these terms
        <xref ref-type="bibr" rid="ref6">(Feldman and Sanger, 2007)</xref>
        .
      </p>
      <p>!() = ! ∗ log /! (6)
Based on Equation (6), !() is the weight of the term t in the document d, ! is frequency of occurrence of
8 is inverse document frequency value of the term t.
the term t in document d, and log 9:!
2.5</p>
    </sec>
    <sec id="sec-6">
      <title>Confusion Matrix</title>
      <p>Confusion Matrix contains information about the performance of a classification system that is evaluated using
the data or metrics contained in the Confusion matrix. Confusion Matrix analyzes how well the classification has
been done on the actual class and the predicted class.</p>
      <p>Positive</p>
      <p>
        Negative
Source:
        <xref ref-type="bibr" rid="ref7">Han et al. (2012)</xref>
        .
      </p>
      <p>= ;&lt;) &gt;;&lt;&lt;)) ;;==) &gt;=
Precision is the proportion of correctly identified labeling (Completeness), the formula for finding Precision:
 = ;&lt;;)&lt;;&lt; (8)
Recall is the proportion of information that can be found from a label, the search formula for Recall:
 = ;&lt;;)&lt;&gt;=</p>
      <p>Precision and Recall can be used to get the proportion of other measurements namely F1-Score. F1-Score is
the harmonic mean of the calculation of Precision and Recall, the formula to find F1-Score:</p>
    </sec>
    <sec id="sec-7">
      <title>3. METHODOLOGY</title>
    </sec>
    <sec id="sec-8">
      <title>Research Methodology</title>
      <p>The research phase begins by identifying the problem through an interview process with one of the stakeholders
to determine the aspects needed in the Sentiment Analysis. Next, the data collection stage is carried out on the
OTA (Online Travel Agent) website using webscraping tools to extract customer review data that will be used in
the Sentiment Analysis process. Then the data design is done which includes the process of categorizing data and
labeling data manually. The next stage is preprocessing text which aims to prepare data before entering the term
weighting stage by using the NLTK and Literature modules in Python. Some of the stages in the preprocessing
text are formalization and translation, cleansing, tokenizing, case folding, stemming, and stopword removal.
Furthermore, the term weighting phase or term weighting will produce a review data that has word weight using
the Sckit-learn module in Python. The sentiment classification stage requires two data: review data which already
has labels and word weights. This stage is divided into two parts, where the first stage aims to choose the best
model that will be used in the second stage, namely the classification of sentiments for each aspect. Model selection
is done by applying Stratified k-Fold Cross Validation and classification is done by classifier machine learning.
Finally, the analysis and evaluation of the classification results are carried out using F1-Score, Precision, and
Recall metrics and the Accuracy calculation. Evaluation is done by utilizing the Scikit-learn module in the Python
Programming Language. See Figure 2 for details.</p>
    </sec>
    <sec id="sec-9">
      <title>4. EXPERIMENT 4.1</title>
    </sec>
    <sec id="sec-10">
      <title>Data Collecting and Labelling</title>
      <p>
        This study uses a webscraping method with webcraper.io tools to extract variables containing information for
sentiment analysis. The data used is the text of customer reviews on the TripAdvisor, Booking.com, Expedia,
Agoda and Pegi-Pegi sites. Customer reviews that are used in Indonesian and have a period of 2012-2019. Total
data collected was 1,561 consisting of 435 texts from the Agoda.com site, 37 texts from the Expedia site, 27 texts
from the Pegi-Pegi site, 235 texts from the Booking.Com site, and 827 texts from the TripAdvisor site. The
example of customer review can be seen in Figure 3. After that, the data that has been collected will be labeled
manually based on the five aspects that have been selected namely location, room, food, price, and service as well
as the polarity of sentiment aspects that are positive and negative. Data labeling is based on guidelines in
        <xref ref-type="bibr" rid="ref3">Dolnicar
and Otter (2003)</xref>
        and adjusted back to the Kertanegara to eliminate subjectivity. Data that already has a label is
ready to enter the next stage, namely preprocessing. Table 2 shows the polarity of sentiment aspect guidance
labeling.
108
      </p>
      <p>Negative
hotel does not meet customer expectations
regarding the services they have
hotel cannot set prices according to the value
obtained
hotel cannot fulfill matters related to hotel
room conditions
hotel cannot meet the customer's expectations
regarding the food and drinks served
hotel cannot meet the customer's expectations
regarding the condition of the location of the
hotel and its surroundings
Text Preprocessing is a step taken to prepare data before being analyzed in the sentiment classification process.
The implementation of Text Preprocessing uses text variables of customer reviews on Kertanegara Premium Guest
House through the NLTK and Sastrawi modules in Python. Formalization is the stage of changing words into
standard forms in accordance with KBBI and Translation is the stage for translating words from foreign languages
into Indonesian. Both stages are carried out manually by adjusting the KBBI Standard Form. Cleansing aims to
eliminate elements that are not needed in the sentiment analysis process. These elements consist of punctuation,
numbers, html tags. Tokenizing aims to separate the text or sentence review into pieces of words. Case Folding
aims to change the words produced in the tokenizing process into lowercase characters. Stemming aims to
eliminate word affixes in the customer review text so that the basic words are obtained. Stopword Removal aims
to eliminate words that have little effect on sentiment analysis or words contained in a stopword list. Furthermore,
the data enters into the Term Weighting phase which aims to provide weights that measure how important the
value of a term is to the document. This study uses TF-IDF (Term Frequency-Inverse Document Frequency) which
will display the word weight which increases according to the appearance of the word and IDF calculation produces
a weight corresponding to the level of uniqueness of the word in the word index. This stage utilizes the
Scikitlearn module in Python. Table 3 shows a representation customer reviews before and after the text preprocessing
process.
Sentiment classification uses review text that has labels and weights. There are 2 stages in the Classification phase,
namely the stage of model selection and the stage of classification model implementation. The model selection
phase aims to choose the model that will produce the best accuracy. This stage uses K-Fold Cross Validation to
limit the problem of overfitting data in compiling data into training and testing, there is a single variance or overlap
in the distribution of training data and testing data so that the model does not lose significant data for the modeling
and testing process. The k value used in this stage is 6 so the dataset is divided into 6 parts. Next, the model
implementation stage uses SVM and Naïve Bayes classification algorithms. All stages in classification utilize the
Scikit-learn module at Python.</p>
    </sec>
    <sec id="sec-11">
      <title>5. ANALYSIS</title>
      <p>The highest number of reviews is in the aspect of rooms with a value of 1070 as shown in Table 4. This shows that
the customer is very concerned about the condition and quality of rooms provided by the Guest House. Room
cleanliness, amenities, bathroom conditions are things that are often discussed by customers during the stay.
Comparison of positive and negative sentiment can be seen in Figure 4 and 5, respectively. The food aspect has
the biggest ratio between positive and negative reviews. This shows that this aspect has urgency to be addressed
immediately. The results of a trend analysis of the five aspects also show that food aspects have not been addressed
optimally with increasing graphs in negative sentiment in 2019. Meanwhile, the other four aspects experienced an
increase in the number of positive reviews in 2019. In 2013 and 2017 the highest average for negative sentiment.
This shows that stakeholders have reformed and improved the management of the guest house so that the graph of
positive sentiment increases rapidly in 2019 except in the aspect of food. The variety and taste of food that does
not meet the standards of a guest house is often complained by customers.</p>
      <p>The study also produced a comparison between the two classification algorithms, SVM and Naïve Bayes,
which uses TF-IDF term weighting. The classification algorithm testing utilizes the classification report function
in the Python Sckit-learn. Table 5 shows that overall SVM has a better Accuracy value than Naïve Bayes on all
five aspects. This proves that SVM has better effectiveness than Naïve Bayes. The value of Precision, Recall and
F1-Score almost reached 70 percent in all aspects except the food aspect. This can be caused by the process of
labeling data review manually or an unbalanced dataset comparison.</p>
    </sec>
    <sec id="sec-12">
      <title>6. CONCLUSIONS</title>
      <p>This study discusses the use of sentiment analysis at the aspect level on customer reviews as a guest house that
aims to assist stakeholders in improving and improving management evaluations. Sentiment analysis conducted
on the five aspects resulted in the finding that the food aspect had an urgency to be immediately addressed by the
stakeholders. Evaluation of sentiment classification results in the SVM class of Accuracy, Precision, Recall, and
F1-Score better than Naive Bayes which proves that SVM has better effectiveness. From this study it can be
concluded that the guest house customers are satisfied with the management and accommodation of the guest
house, but it is good for the guest house to continue to monitor and improve its services to all aspects, especially
food aspects to produce a positive E-WOM.</p>
    </sec>
  </body>
  <back>
    <ref-list>
      <ref id="ref1">
        <mixed-citation>
          <string-name>
            <surname>Berrar</surname>
            ,
            <given-names>D.</given-names>
          </string-name>
          (
          <year>2018</year>
          ).
          <article-title>Bayes' Theorem and Naive Bayes Classifier</article-title>
          .
          <source>Encyclopedia of Bioinformatics and Computational Biology</source>
          ,
          <volume>1</volume>
          ,
          <fpage>403</fpage>
          -
          <lpage>412</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref2">
        <mixed-citation>
          <string-name>
            <surname>Bhavitha</surname>
            ,
            <given-names>B.K.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Rodrigues</surname>
            ,
            <given-names>A.P.</given-names>
          </string-name>
          , and
          <string-name>
            <surname>Chiplunkar</surname>
            ,
            <given-names>N.N.</given-names>
          </string-name>
          (
          <year>2017</year>
          ).
          <article-title>Comparative Study of Machine Learning Techniques in Sentimental Analysis</article-title>
          ,
          <source>In: International Conference on Inventive Communication and Computational Technologies (ICICCT</source>
          <year>2017</year>
          ),
          <fpage>216</fpage>
          -
          <lpage>221</lpage>
          . IEEE: New York NY.
        </mixed-citation>
      </ref>
      <ref id="ref3">
        <mixed-citation>
          <string-name>
            <surname>Dolnicar</surname>
            ,
            <given-names>S.</given-names>
          </string-name>
          , and
          <string-name>
            <surname>Otter</surname>
            ,
            <given-names>T.</given-names>
          </string-name>
          (
          <year>2003</year>
          ).
          <article-title>Which Hotel attributes Matter? A review of previous and a framework for future research</article-title>
          ,
          <source>In: Proceedings of the 9th Annual Conference of the Asia Pacific Tourism Association (APTA)</source>
          , Griffin,
          <string-name>
            <given-names>T.</given-names>
            , and
            <surname>Harris</surname>
          </string-name>
          , R. (Eds.),
          <volume>1</volume>
          ,
          <fpage>176</fpage>
          -
          <lpage>188</lpage>
          . APTA: Busan, South Korea.
        </mixed-citation>
      </ref>
      <ref id="ref4">
        <mixed-citation>
          <string-name>
            <surname>Ekawati</surname>
            ,
            <given-names>D.</given-names>
          </string-name>
          , and
          <string-name>
            <surname>Khodra</surname>
            ,
            <given-names>M.L.</given-names>
          </string-name>
          (
          <year>2017</year>
          ).
          <article-title>Aspect-based Sentiment Analysis for Indonesian Restaurant Reviews</article-title>
          , In: 2017 International Conference on Advanced Informatics, Concepts, Theory, and
          <string-name>
            <surname>Applications</surname>
          </string-name>
          (ICAICTA),
          <fpage>1</fpage>
          -
          <lpage>6</lpage>
          . Curan Associates: Red Hook NY.
        </mixed-citation>
      </ref>
      <ref id="ref5">
        <mixed-citation>
          <string-name>
            <surname>El-Jawad</surname>
            ,
            <given-names>M.H.A.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Hodhod</surname>
            ,
            <given-names>R.</given-names>
          </string-name>
          , and
          <string-name>
            <surname>Omar</surname>
            ,
            <given-names>Y.M.</given-names>
          </string-name>
          (
          <year>2018</year>
          ).
          <article-title>Sentiment Analysis of Social Media Networks Using Machine Learning</article-title>
          , In: 2018 14th International Computer Engineering Conference (ICENCO),
          <fpage>174</fpage>
          -
          <lpage>176</lpage>
          . IEEE: New York NY.
        </mixed-citation>
      </ref>
      <ref id="ref6">
        <mixed-citation>
          <string-name>
            <surname>Feldman</surname>
            ,
            <given-names>R.</given-names>
          </string-name>
          , and
          <string-name>
            <surname>Sanger</surname>
            <given-names>J.</given-names>
          </string-name>
          (
          <year>2007</year>
          ).
          <article-title>The Text Mining Handbook: Advanced Approaches in Analyzing Unstructured Data</article-title>
          . New York: Cambridge University Press.
        </mixed-citation>
      </ref>
      <ref id="ref7">
        <mixed-citation>
          <string-name>
            <surname>Han</surname>
            ,
            <given-names>J</given-names>
          </string-name>
          .,
          <string-name>
            <surname>Kamber</surname>
            ,
            <given-names>M.</given-names>
          </string-name>
          , and
          <string-name>
            <surname>Pei</surname>
            ,
            <given-names>J.</given-names>
          </string-name>
          (
          <year>2012</year>
          ).
          <article-title>Data Mining: concepts and techniques</article-title>
          . Elsevier/Morgan: Amsterdam.
        </mixed-citation>
      </ref>
      <ref id="ref8">
        <mixed-citation>
          <string-name>
            <surname>Liu</surname>
            ,
            <given-names>B.</given-names>
          </string-name>
          (
          <year>2012</year>
          ).
          <article-title>Sentiment Analysis and Opinion Mining</article-title>
          . Morgan &amp; Claypool Publishers: Williston VT.
        </mixed-citation>
      </ref>
      <ref id="ref9">
        <mixed-citation>
          <string-name>
            <surname>Manning</surname>
            ,
            <given-names>C.D.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Raghavan</surname>
            ,
            <given-names>P.</given-names>
          </string-name>
          , and
          <string-name>
            <surname>Schütze</surname>
          </string-name>
          , H. (Eds.) (
          <year>2008</year>
          ).
          <article-title>An Introduction to Information Retrieval</article-title>
          . Cambrigde University Press: Cambridge.
        </mixed-citation>
      </ref>
      <ref id="ref10">
        <mixed-citation>
          <string-name>
            <surname>Miao</surname>
            ,
            <given-names>F.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Zhang</surname>
            ,
            <given-names>P.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Jin</surname>
            ,
            <given-names>L.</given-names>
          </string-name>
          and
          <string-name>
            <surname>Wu</surname>
            ,
            <given-names>H.</given-names>
          </string-name>
          (
          <year>2018</year>
          ).
          <article-title>Chinese News Text Classification Based on Machine learning algorithm</article-title>
          ,
          <source>In: 2018 10th International Conference on Intelligent Human-Machine Systems and Cybernetics (IHMSC)</source>
          ,
          <fpage>48</fpage>
          -
          <lpage>51</lpage>
          . IEEE: New York NY.
        </mixed-citation>
      </ref>
      <ref id="ref11">
        <mixed-citation>
          <string-name>
            <surname>Shi</surname>
            ,
            <given-names>H.X.</given-names>
          </string-name>
          and
          <string-name>
            <surname>Li</surname>
            ,
            <given-names>X.J.</given-names>
          </string-name>
          (
          <year>2011</year>
          ).
          <article-title>A Sentiment Analysis Model for Hotel Reviews Based On Supervised Learning</article-title>
          ,
          <source>In: Proceedings of the 2011 International Conference on Machine Learning and Cybernetics</source>
          , Guilin,
          <fpage>10</fpage>
          -
          <issue>13</issue>
          <year>July</year>
          ,
          <year>2011</year>
          ,
          <fpage>10</fpage>
          -
          <lpage>13</lpage>
          . IEEE: New York NY.
        </mixed-citation>
      </ref>
      <ref id="ref12">
        <mixed-citation>
          <string-name>
            <surname>Westbrook</surname>
            ,
            <given-names>R.A.</given-names>
          </string-name>
          (
          <year>1987</year>
          ).
          <article-title>Product/Consumption-Based Affective Responses and Postpurchase Processes</article-title>
          .
          <source>Journal of Marketing Research</source>
          ,
          <volume>24</volume>
          (
          <issue>3</issue>
          ),
          <fpage>258</fpage>
          -
          <lpage>270</lpage>
          .
        </mixed-citation>
      </ref>
    </ref-list>
  </back>
</article>