<!DOCTYPE article PUBLIC "-//NLM//DTD JATS (Z39.96) Journal Archiving and Interchange DTD v1.0 20120330//EN" "JATS-archivearticle1.dtd">
<article xmlns:xlink="http://www.w3.org/1999/xlink">
  <front>
    <journal-meta />
    <article-meta>
      <title-group>
        <article-title>Behavioural Analytics using Process Mining in On-line Advertising</article-title>
      </title-group>
      <contrib-group>
        <contrib contrib-type="author">
          <string-name>Maria Diapouli</string-name>
          <email>M.Diapouli@brighton.ac.uk</email>
          <xref ref-type="aff" rid="aff0">0</xref>
        </contrib>
        <contrib contrib-type="author">
          <string-name>Stelios Kapetanakis</string-name>
          <email>S.Kapetanakis@brighton.ac.uk</email>
          <xref ref-type="aff" rid="aff0">0</xref>
        </contrib>
        <contrib contrib-type="author">
          <string-name>Miltos Petridis</string-name>
          <email>2M.Petridis@mdx.ac.ukuk</email>
        </contrib>
        <contrib contrib-type="author">
          <string-name>Roger Evans</string-name>
          <email>R.P.Evans@brighton.ac.uk</email>
          <xref ref-type="aff" rid="aff0">0</xref>
        </contrib>
        <aff id="aff0">
          <label>0</label>
          <institution>School of Computing, Engineering and Mathematical Sciences, University of Brighton</institution>
          ,
          <addr-line>Watts Building, Lewes Road, Brighton, BN2 4GJ</addr-line>
          ,
          <country country="UK">UK</country>
        </aff>
      </contrib-group>
      <fpage>147</fpage>
      <lpage>156</lpage>
      <abstract>
        <p>Online behavioural targeting is one of the most popular business strategies on the display advertising today. It is based primarily on analysing web user behavioural data with the usage of machine learning techniques with the aim to optimise web advertising. Being able to identify “unknown” and “first time seen” customers is of high importance in online advertising since a successful guess could identify “possible prospects” who would be more likely to purchase an advertisement's product. By identifying prospective customers, online advertisers may be able to optimise campaign performance, maximise their revenue as well as deliver advertisements tailored to a variety of user interests. This work presents a hybrid approach benchmarking machine-learning algorithms and attribute preprocessing techniques in the context of behavioural targeting in process oriented environments. The performance of our suggested methodology is evaluated using the key performance metric in online advertising which is the predicted conversion rate. Our experimental results indicate that the presented process mining framework can significantly identify prospect customers in most cases. Our results seem promising, indicating that there is a need for further workflow research in online display advertising.</p>
      </abstract>
      <kwd-group>
        <kwd>Process mining</kwd>
        <kwd>Process Oriented Workflows</kwd>
        <kwd>Classification</kwd>
        <kwd>Online Display Advertising</kwd>
      </kwd-group>
    </article-meta>
  </front>
  <body>
    <sec id="sec-1">
      <title>Introduction</title>
      <p>
        According to statistics published by the Internet Advertising Bureau online advertisers in
the UK have spent more than 8.6 billion UK-pounds in 2016 on behavioural targeted
advertising a figure which grew 16.4% compared to 2015. The estimate represents steady
growth rates of about 20% from 2010 through 2016 [
        <xref ref-type="bibr" rid="ref3">3</xref>
        ]. Behavioural targeting and
customer prospecting are both promising and challenging aspects in display advertising.
Promising since the more information of user behavioural activity exists the better targeted
advertisements could be delivered to end users and challenging since display advertising is
a rather complex ecosystem which involves multiple interested parties such as end users,
advertisers, publishers, and ad platforms. The size of data generated and collected from any
Copyright © 2017 for this paper by its authors. Copying permitted for private and
academic purpose. In Proceedings of the ICCBR 2017 Workshops. Trondheim, Norway
involved parties is significantly large: Billions of websites requests every day trigger
millions of advertisements that are finally displayed to millions of users.
      </p>
      <p>
        Digital advertisers attract increasing traffic on their websites aiming for certain user
marketing actions, more commonly, accomplishing an online purchase. This action is
recorded as a conversion. There are two ways for viewing an advert upon arrival on an
affiliate ad-friendly website. Firstly, by clicking on the advert and immediately buying
and/or by viewing an advert and waiting for a future return and a possible purchase. The
journey of a user throughout several websites can be represented as a series of events with
intermediate temporal durations. This can be interpreted into a “workflow” of variant length
which may or may not convert at its final stages. Petridis et al. [
        <xref ref-type="bibr" rid="ref19">19</xref>
        ] have shown that
workflow behaviours with such a distinct event-duration coupling can be formalised over a
general theory of time [
        <xref ref-type="bibr" rid="ref20">20</xref>
        ], be graph-represented, monitored [
        <xref ref-type="bibr" rid="ref21">21</xref>
        ] and explained [
        <xref ref-type="bibr" rid="ref22">22</xref>
        ]
effectively using Case-based Reasoning techniques [
        <xref ref-type="bibr" rid="ref23">23</xref>
        ].
      </p>
      <p>Our research questions on top of the online marketing business model are twofold – One:
which metric features in terms of evaluating an online campaign performance are mostly
important and -Two: based on the set of identified metrics what is the profile of an ad
viewer who is keen to make a purchase. In such way by analysing and classifying past
behavioural observations among ad viewers, could allow marketers to identify future
prospect customers more effectively.</p>
      <p>The work we present in this paper handles a challenging area in the online display
advertising marketplace, this of customer prospecting. Customer prospecting identifies web
users who are likely to purchase a product after seeing an advertisement. We developed a
process mining methodology based on an advertising campaign implemented by an ad
network provider. We collected and analysed campaign data that contained audience
demographic information and audience behavioural segments to predict whether a user who
had no previous seen an advert is likely to convert. The goal of this research was to increase
an individual advertising campaign performance by augmenting its CPA ratio.</p>
      <p>This paper is structured as follows: Section 2 presents the context of search engine
advertising and its online display landscape, section 3 will describe our adopted process
mining methodology. Additionally, the imbalanced problem of conversion rate will be
explained and our approach to the class imbalance problem will be analysed. Section 4 will
present a series of empirical experiments for selecting the best performing classification
algorithm. Finally, a discussion upon our experiment results will be presented in section 5.
2
2.1</p>
    </sec>
    <sec id="sec-2">
      <title>Related Work</title>
      <sec id="sec-2-1">
        <title>Search Engine Advertising</title>
        <p>
          The application of statistical algorithms and process mining methodologies is widely
applied in search engine advertising. Its outcomes could be observed from user-relevant
textual advertisements placed next to search results as they come from several search
engines. Choosing the most relevant ad for a user query and the optimal place in which it is
displayed could affect significantly the probability for a user to click on that chosen ad.
For any adverts with already known Click-Through Rate (CTR) historical information,
CTR could be estimated empirically by dividing the number of impressions over clicks. In
any other case when a new ad, with no historical information, is going to be displayed to a
random user; a major challenge emerges: This is mainly in terms of identifying a suitable
advert where a user would be “tempted to click”. Richardson et al. [
          <xref ref-type="bibr" rid="ref9">9</xref>
          ] answer this challenge
by predicting the CTR for a new ad with no prior historical information. Based on ad text
information only (title, body, search keywords, display URL, impressions, clicks, landing
URL) logistic regression can produce an accurate model ad CTR prediction. Research in
the area has shown that decision rules [
          <xref ref-type="bibr" rid="ref7">7</xref>
          ] can be produced for predicting the CTR for
unseen ads, from data that contain information regarding advertisements, query terms and
URLs. Clustering techniques could also be used to improve the keywords CTR for rare or
new keywords [
          <xref ref-type="bibr" rid="ref8">8</xref>
          ] based on generation of clusters of related keywords. This can be applied
by if different search keywords have a different likelihood of receiving a user click [
          <xref ref-type="bibr" rid="ref8">8</xref>
          ].
2.2
        </p>
      </sec>
      <sec id="sec-2-2">
        <title>Online display advertising landscape</title>
        <p>Online Display Advertising is a highly congested and convoluted environment involving
an extended range of vendors, services and high volumes of transactions. Its landscape
mainly comprises workflows of advertisers, ad agencies, web-users and publishers.
Advertisers set up product or service workflows in publisher web sites, also known as
inventories, with the aim to attract as many web-users as possible. This is achieved by
providing rich and engaging ad-context to the most receptive online audience. In return the
end users will click on the ad and will be redirected on the advertiser website to purchase
potential product(s).</p>
        <p>
          The complexity of achieving such desired actions has led to the development of new
display parties, these of Ad-exchanges and Ad-networks. Both could perform better on
accessing and controlling inventories. Ad-exchanges are online auction marketplaces (like
e-bay) trading in advertisements as their “products”. Ad-exchanges could provide three
main services: adverts allocation, prices determination, traffic control [
          <xref ref-type="bibr" rid="ref1">1</xref>
          ]. Ad networks
provide a variety of excluded services to advertisers such as ad serving, privacy
verification, targeting the most suitable audience and advertising campaign reporting. Ad
networks could access publishers directly as well as ad exchanges for identifying
inventories.
        </p>
        <p>Publisher</p>
        <p>Measuring the revenue and the effectiveness of online display advertising campaigns is
achieved through three prevalent pricing models. These are: Cost per impression (CPM),
Advertiser</p>
        <p>Ad Agency</p>
        <p>Ad Network
Ad Exchange</p>
        <p>
          Web User
Cost per click (CPC) and Cost per action (CPA). CPM ratio is popularly used for brand
recognition campaigns where a fixed cost is charged to advertisers based on the number of
displays of advertisements. CPC ratio was introduced to build the advertisers confidence
upon their return on investment(s). Advertisers pay publishers when a web user clicks on an
advertisement without considering the number of impression displays. However, since most
advertisers are retailers the actual advertising benefit derives from the commercial
transaction within their websites [
          <xref ref-type="bibr" rid="ref2">2</xref>
          ]. CPA metrics is used to serve this purpose. According
to the Interactive Advertising Bureau report (2016) in the US market, 65% of online
advertising transactions share was attributed on the customer performance models (CPC,
CPA), while the second on the list was the CPM model with 33%. Hybrids of impression
and performance models reside in 2% of online advertising transactions [
          <xref ref-type="bibr" rid="ref3">3</xref>
          ].
        </p>
        <p>
          According to Lewis and Reiley (2014) [
          <xref ref-type="bibr" rid="ref16">16</xref>
          ] the effect of online advertising on sales is
not fully associated with CTR. Their collaboration with Yahoo!Research and an eminent
retailer had reported that 78% of the lift in retailer sales was originated from users who had
viewed ads but had not clicked them, while only 22% was attributed to those who had
clicked. In our research work we found that the online advertising campaign had
substantial impact on the users who merely viewed the ads. Based on thorough analysis we
identified that impressions are more strongly correlated to conversions than clicks. Most
interestingly, clicks had a very trivial correlation (correlation = 0.00000115) with
conversion. These findings suggest that the most meaningful metric for evaluating
campaign performance is conversions instead of clicks.
3
        </p>
      </sec>
    </sec>
    <sec id="sec-3">
      <title>Methodology</title>
      <p>
        Customer prospecting is related with predictive modelling in process mining terms.
Predictive modelling, also referred as supervised data mining, aims to predict the probable
future event based on previous historical knowledge [
        <xref ref-type="bibr" rid="ref10">10</xref>
        ]. The appropriate selection of data
samples is important for effective analysis and prediction based on underlying patterns [
        <xref ref-type="bibr" rid="ref13">13</xref>
        ].
In this work decision trees and kNN were preferred over the commonly used logistic
regression and collaborative filtering classification methods. Decision trees have been
shown as effective in building profiles for the web users who have converted in the past and
then predict whether a new web-user is likely to convert [
        <xref ref-type="bibr" rid="ref17">17</xref>
        ]. Decision trees although
powerful in expressing continuous and categorical inputs, they seem to fall out when there
is a mixture of continuation and categorical type data. Thus, kNN has seemed more
appropriate since it performs better with continuous data [
        <xref ref-type="bibr" rid="ref18">18</xref>
        ]. For predictive modelling
decision trees seem most powerful and prevalent tools [
        <xref ref-type="bibr" rid="ref15">15</xref>
        ] compared to logistic regression
and collaborative filtering since:
      </p>
      <p>-Logistic regression could be used to examine the model's exponentials of the
coefficients to explore which user attributes affect the likelihood of conversion. In such
way, we would be able to explore the necessary coefficients but be unable to explore the
underlying ruleset which could indicate and predict online “target” users willing to convert.</p>
      <p>
        -Collaborative Filtering was also considered since it has been proven effective in finding
prospect customers based on past customer behaviours (training samples) [
        <xref ref-type="bibr" rid="ref14">14</xref>
        ]. In such
way, a successful model would be retrained on a regular basis to include recent user activity
information. However, in our investigated dataset such information was not available and
this research was not able benefit from “live ad-feeds” including: Ad ids, ad-width,
adheight, visibility time viewed, format, etc.
      </p>
      <p>Therefore, based on the limitations of the dataset and the model target audience, our
selection methodology was based on data-mined patterns for ruleset generation to
understand and predict successful (or not) online user-conversions.
3.1</p>
      <sec id="sec-3-1">
        <title>The Data</title>
        <p>The dataset used in this experiment was gathered as a day-campaign and it consisted of
3,425,119 impressions that were displayed on 3,407,293 users. Among them 8,082 users
clicked on the displayed advert (click response rate 0.24%) and 913 converted (convert
response rate 0.03%). Due to the very sparse number of response rate both for click and
conversion the data were highly skewed. To overcome this limitation of imbalanced data a
sampling technique was adopted and will be discussed in the next section. The size of the
data set had 3,407,293 observations. Each observation characterised a web user and was
described by 46 independent features and 1 dependent feature (sample of log data are
shown in Table 1). The independent features were both categorical and nominal. These
features were related to a user’s browsing history (URLs that user has visited in the past).
The dependent feature was a nominal one-named conversion rate which illustrated whether
a user had made a purchase in the past.
3.2</p>
      </sec>
      <sec id="sec-3-2">
        <title>Our Approach</title>
        <p>
          The process of Knowledge Discovery in Databases (KDD) was adopted as
methodological approach on this case study, where process mining was the gist in the
overall process [
          <xref ref-type="bibr" rid="ref11">11</xref>
          ]. The experiment went through all the steps of KDD. The data were
extracted from the DDMS based on the process mining project specifications to ensure
consistency and completeness. In the data transformation phase, categorical values where
transformed on numerical ones to adhere to the algorithm specifications. Additionally, since
the response rate for conversion was a small number, only 0.03% of the total dataset, a
method that modified data values and derived new fields from response rate for conversion
was adopted. A new field was created which contained two values for the conversion field:
zero (indicated that a user did not convert) and one (indicated a successful user conversion).
However, this new field did not overcome the very low percentage of conversion rate. This
was regarded a challenge since a high-performance classifier was required to have an
accurate model. In such: training data should be evenly distributed between conversion and
no-conversion values. In any other case the classifier could be biased since it would try to
achieve the overall classification accuracy by identifying mostly the majority class
(conversions) and would overlook the minority one (no-conversion). This would offer little
contribution to the model accuracy [
          <xref ref-type="bibr" rid="ref4">4</xref>
          ]. Our approach in balancing the classifier will be
described in section 4.2.
4
4.1
        </p>
      </sec>
    </sec>
    <sec id="sec-4">
      <title>Experiments and Evaluation</title>
      <sec id="sec-4-1">
        <title>Modelling and imbalanced data sets</title>
        <p>For this work, IBM SPSS Modeler 18.0 was used for all experiments. Different algorithms
were assessed to benchmark the most appropriate and accurate dataset for classification and
prediction. The data set was separated into training and testing set to build and evaluate our
decision tree model with a 70 - 30 split rate respectively. In the training phase, the model
was processed by using the training set and then tested to evaluate our model’s accuracy</p>
        <p>
          To overcome the problem of heavily imbalanced data two approaches were considered:
a) using over-sampling and b) using under-sampling. In the over-sampling approach, the
training set was populated with replicated data that belonged to the minority class until the
training set was balanced [
          <xref ref-type="bibr" rid="ref6">6</xref>
          ]. The information remains the same but the misclassification
cost of the minority class is increased.
        </p>
        <p>
          In the under-sampling approach data from the majority class were removed to balance
the training set [
          <xref ref-type="bibr" rid="ref5">5</xref>
          ]. For our experiments, the under-sampling approach was used where
cases from the dependent feature conversion rate were randomly eliminated.
4.2
        </p>
      </sec>
      <sec id="sec-4-2">
        <title>Comparing the performance of different algorithms</title>
        <p>Different decision tree algorithms have been selected for searching patterns in data as
well as kNN with k =3 for more accurate classification of continuous data attributes. This
process included deciding which algorithm provided the lowest average classification error.
The selected algorithms were: classification and regression tree (C&amp;RT), Chi-squared
automatic interaction detector (CHAID), C5.0 and 3NN. The performance results of these
algorithms as applied on the test set are shown in Graph 1.</p>
      </sec>
      <sec id="sec-4-3">
        <title>Graph 1: Summary of results</title>
        <p>Converter True Positives (Hit) Rate
Non-Converter True Positives (Hit) Rate
Accuracy
Sensitivity
Specificity</p>
        <p>
          Table 2 illustrates a higher likelihood (90.63%) for C5.0 to predict the event for someone
converting on an advertisement compared to the other baselines (88.25% for C&amp;RT,
89.34% for CHAID and 89.14% for 3NN). In Table 2 C5.0 is showing 99.02% sensitivity
(the portion of users that were correctly predicted to convert) and 90.62% specificity (the
portion of users that did not convert and were successfully predicted) which accounts for
the overall accuracy of 90.63%. The sensitivity and specificity measures are used to
ascertain the model validity and accuracy [
          <xref ref-type="bibr" rid="ref12">12</xref>
          ].
        </p>
        <p>The data for building decision trees with C5.0 algorithm models were re-sampled using a
bootstrap aggregation technique to form several pairs of training and testing data sets. A
decision tree model was developed for each pair of training and testing data sets.
Predictions from any individual decision trees were merged via a voting system which led
to the highest possible accuracy for the final model (ensemble) predictions. It was observed
that similar models were generated throughout ensemble learning. This was evidence that
the chosen algorithm was stable throughout the dataset.
5</p>
      </sec>
    </sec>
    <sec id="sec-5">
      <title>Conclusions</title>
      <p>In this paper, we demonstrated process mining as an effective tool for direct marketing
which can improve online marketing campaigns. Most the existing research in this area so
far focuses on computational and theoretical aspects of direct marketing though little efforts
have been put on technological aspects of applying process mining in the process of direct
marketing. The complexity of process mining models makes it difficult for marketers to use
it, hence we outlined a simplified framework to guide marketers and managers in making
use of process mining methods and focus their advertising and promotion on those
categories of people to reduce time and costs. We explained the steps and tasks that are
carried out at each stage of the process mining framework and showed some examples of
the type of predictive efficiency that can be achieved using the proposed approach. This has
shown that substantial gains can be achieved by adopting this pragmatic and exploratory
approach to predict user behaviour in on-line advertising.</p>
      <p>This work has shown capable of dealing with the uncertainty underlying within
behavioural data as on-line advertising experts have noted that user behaviours can vary
significantly. Our suggested approach seems capable of dealing with more complex online
advertising models and thus our future directions will focus on more complex, variant and
fuzzy attributes.</p>
      <p>The results obtained so far, are promising and encourage us to continue experimentation
with more sophisticated models or other algorithms to further improve the performance of
the system. It seems sensible to experiment with the following settings in future work:
– introducing the temporal dimension to our model to apply time series analysis
techniques to build the model
– combining the model with content-based approach
– additional category-based and continuous-based attributes specifying the times
spent on each of the categories, with possible division into work-days and week-days, for
example a different choice of categories.</p>
      <p>As future work, we will incorporate more user and publisher information obtained from
third party media providers into data hierarchies to improve model prediction.</p>
      <sec id="sec-5-1">
        <title>Acknowledgement</title>
        <p>This work was funded by Innovate UK (formerly known as Technology Strategy Board).</p>
      </sec>
    </sec>
  </body>
  <back>
    <ref-list>
      <ref id="ref1">
        <mixed-citation>
          1.
          <string-name>
            <surname>Balseiro</surname>
            ,
            <given-names>S.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Feldman</surname>
            ,
            <given-names>J.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Mirrokni</surname>
            ,
            <given-names>V.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Muthukrishnan</surname>
          </string-name>
          . S.,
          <article-title>Yield optimization of display advertising with ad exchange</article-title>
          .
          <source>In ACM Conf on E-Commerce</source>
          ,
          <article-title>(</article-title>
          <year>2011</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref2">
        <mixed-citation>
          2.
          <string-name>
            <surname>Moona</surname>
            ,
            <given-names>Y.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Kwonb</surname>
            ,
            <given-names>C.</given-names>
          </string-name>
          :
          <article-title>Online advertisement service pricing and an option contract</article-title>
          ,
          <source>Electronic Commerce Research and Applications</source>
          ,
          <string-name>
            <surname>Elsevier</surname>
          </string-name>
          (
          <year>2011</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref3">
        <mixed-citation>
          3.
          <string-name>
            <given-names>Interactive</given-names>
            <surname>Advertising</surname>
          </string-name>
          <article-title>Board (IAB)</article-title>
          .
          <source>Internet Advertising Revenue commissioned report, (Updated June</source>
          <year>2016</year>
          ), https://iabuk.net/, accessed May 2017
        </mixed-citation>
      </ref>
      <ref id="ref4">
        <mixed-citation>
          4.
          <string-name>
            <surname>Visa</surname>
            ,
            <given-names>S.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Ralescu</surname>
          </string-name>
          . A.,
          <article-title>Issues in mining imbalanced data sets - a review paper</article-title>
          .
          <source>Midwest AI and Cognitive Science Conference</source>
          ,
          <volume>67</volume>
          -
          <fpage>73</fpage>
          (
          <year>2005</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref5">
        <mixed-citation>
          5.
          <string-name>
            <surname>Kubat</surname>
            ,
            <given-names>M.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Matwin</surname>
            ,
            <given-names>S.</given-names>
          </string-name>
          :
          <article-title>Addressing the curse of imbalanced training sets: onesided selection</article-title>
          ,
          <source>International Conference on Machine Learning</source>
          ,
          <fpage>179</fpage>
          -
          <lpage>186</lpage>
          (
          <year>1997</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref6">
        <mixed-citation>
          6.
          <string-name>
            <surname>Ling</surname>
            ,
            <given-names>C.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Li</surname>
            ,
            <given-names>C.</given-names>
          </string-name>
          :
          <article-title>Data mining for Direct Marketing: Problems and Solutions Proc</article-title>
          .
          <article-title>Of the Fourth Inter</article-title>
          . Conf.
          <article-title>on Knowledge Discovery and Data mining</article-title>
          , AAAI Press,
          <fpage>73</fpage>
          -
          <lpage>79</lpage>
          , (
          <year>1998</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref7">
        <mixed-citation>
          7.
          <string-name>
            <surname>Kotlowski</surname>
            ,
            <given-names>W.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Debmbsczynski</surname>
            ,
            <given-names>K.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Weiss</surname>
          </string-name>
          , D.:
          <article-title>Predicting ads click through rate with decision rules</article-title>
          . In Workshop on Target and
          <article-title>Ranking for Online Advertising</article-title>
          , WWW
          <volume>08</volume>
          , (
          <year>2008</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref8">
        <mixed-citation>
          8.
          <string-name>
            <surname>Regelson</surname>
            ,
            <given-names>M.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Fain</surname>
            <given-names>D. C.</given-names>
          </string-name>
          :
          <article-title>Predicting click-through rate using keyword clusters</article-title>
          .
          <source>In Electronic Commerce (EC)</source>
          .
          <source>ACM</source>
          ,
          <year>2006</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref9">
        <mixed-citation>
          9.
          <string-name>
            <surname>Richardson</surname>
            ,
            <given-names>M.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Dominowska</surname>
            ,
            <given-names>E.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Ragno</surname>
          </string-name>
          , R.:
          <article-title>Predicting clicks: estimating the click-through rate for new ads</article-title>
          .
          <source>International conference on World Wide Web</source>
          ,
          <fpage>521</fpage>
          -
          <lpage>530</lpage>
          , (
          <year>2007</year>
          ).
        </mixed-citation>
      </ref>
      <ref id="ref10">
        <mixed-citation>
          10.
          <string-name>
            <surname>Apte</surname>
            ,
            <given-names>C. V.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Hong</surname>
            ,
            <given-names>S. J.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Natarajan</surname>
            ,
            <given-names>R.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Pednault</surname>
            ,
            <given-names>E. P. D.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Tipu</surname>
            ,
            <given-names>F. A.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Weiss</surname>
            ,
            <given-names>S. M.</given-names>
          </string-name>
          :
          <article-title>Data-intensive analytics for predictive modelling</article-title>
          ,
          <source>IBM journal of research and development</source>
          ,
          <volume>47</volume>
          (
          <issue>1</issue>
          ) (
          <year>2003</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref11">
        <mixed-citation>
          11.
          <string-name>
            <surname>Chapman</surname>
            ,
            <given-names>P.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Clinton</surname>
            ,
            <given-names>J.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Kerber</surname>
            ,
            <given-names>R.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Khabaza</surname>
            <given-names>T.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Reinartz</surname>
            ,
            <given-names>T.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Shearer</surname>
            ,
            <given-names>C.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Wirth</surname>
            ,
            <given-names>R.</given-names>
          </string-name>
          ,
          <year>1999</year>
          .
          <article-title>CRISP-DM 1.0 Step-by-step data mining guide</article-title>
          .
          <source>Crisp DM Consortium</source>
          , (
          <year>Updated 2010</year>
          ), https://www.the
          <article-title>-modeling-agency.com/crisp-dm.pdf, accessed April 2016</article-title>
        </mixed-citation>
      </ref>
      <ref id="ref12">
        <mixed-citation>
          12.
          <string-name>
            <surname>Han</surname>
            ,
            <given-names>J</given-names>
          </string-name>
          .,
          <string-name>
            <surname>Kamber</surname>
            ,
            <given-names>M.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Kaufmann</surname>
            ,
            <given-names>M.</given-names>
          </string-name>
          :
          <article-title>Data mining: Concepts and Techniques (2nd Edition) (</article-title>
          <year>2006</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref13">
        <mixed-citation>
          13.
          <string-name>
            <surname>Kim</surname>
            ,
            <given-names>Y.</given-names>
          </string-name>
          :
          <article-title>Toward a Successful CRM: Variable Selection, Sampling, and</article-title>
          <string-name>
            <surname>Ensemble</surname>
          </string-name>
          ,
          <source>Decision Support Systems</source>
          ,
          <volume>41</volume>
          (
          <issue>2</issue>
          ),
          <fpage>542</fpage>
          -
          <lpage>553</lpage>
          (
          <year>2006</year>
          ).
        </mixed-citation>
      </ref>
      <ref id="ref14">
        <mixed-citation>
          14.
          <string-name>
            <surname>Anastasakos</surname>
            ,
            <given-names>T.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Hillard</surname>
            ,
            <given-names>D.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Kshetramade</surname>
            ,
            <given-names>S.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Raghavan</surname>
          </string-name>
          , H.:
          <article-title>A collaborative filtering approach to ad recommendation using the query-ad click graph</article-title>
          ,
          <source>ACM conference on Information and knowledge management</source>
          ,
          <year>1927</year>
          -
          <fpage>1930</fpage>
          (
          <year>2009</year>
          ).
        </mixed-citation>
      </ref>
      <ref id="ref15">
        <mixed-citation>
          15.
          <string-name>
            <surname>Reddy</surname>
            ,
            <given-names>C.S.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Vasu</surname>
            ,
            <given-names>V.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Kumara Swamy</surname>
          </string-name>
          Achari B.:
          <article-title>Effective Decision Tree Learning</article-title>
          ,
          <source>International Journal of Computer Applications</source>
          ,
          <volume>82</volume>
          (
          <issue>9</issue>
          ), (
          <year>2013</year>
          ).
        </mixed-citation>
      </ref>
      <ref id="ref16">
        <mixed-citation>
          16.
          <string-name>
            <surname>Lewis</surname>
            ,
            <given-names>R.A.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Reiley</surname>
            ,
            <given-names>D.H.</given-names>
          </string-name>
          :
          <article-title>Online ads and offline sales: measuring the effect of retail advertising via a controlled experiment on Yahoo!</article-title>
          ,
          <source>Quantitative Marketing and Economics</source>
          ,
          <volume>12</volume>
          (
          <issue>3</issue>
          ),
          <fpage>235</fpage>
          -
          <lpage>266</lpage>
          (
          <year>2014</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref17">
        <mixed-citation>
          17.
          <string-name>
            <surname>Diapouli</surname>
            ,
            <given-names>M.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Petridis</surname>
            ,
            <given-names>M.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Evans</surname>
            ,
            <given-names>R.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Kapetanakis</surname>
            ,
            <given-names>S.:</given-names>
          </string-name>
          <article-title>An Industrial Application of Data Mining Techniques to Enhance the Effectiveness of On-line Advertising</article-title>
          .
          <source>In Proceeding of Thirty-sixth SGAI International Conference on Artificial Intelligence</source>
          ,
          <source>AI</source>
          <year>2016</year>
          , pp.
          <fpage>395</fpage>
          -
          <lpage>401</lpage>
          (
          <year>2016</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref18">
        <mixed-citation>
          18.
          <string-name>
            <surname>Shah</surname>
            ,
            <given-names>Z.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Mahmood</surname>
            ,
            <given-names>A. N.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Orgun</surname>
            ,
            <given-names>M. A.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Mashinchi</surname>
            ,
            <given-names>M. H.</given-names>
          </string-name>
          :
          <article-title>Subset Selection Classifier (SSC): A Training Set Reduction Method</article-title>
          .
          <source>In proceedings of the 16th International Conference in Computational Science and Engineering (CSE)</source>
          ,
          <year>2013</year>
          IEEE, pp.
          <fpage>862</fpage>
          -
          <lpage>869</lpage>
          (
          <year>2013</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref19">
        <mixed-citation>
          19.
          <string-name>
            <surname>Petridis</surname>
            ,
            <given-names>M.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Ma</surname>
          </string-name>
          , J.,
          <string-name>
            <surname>Knight</surname>
            ,
            <given-names>B.</given-names>
          </string-name>
          :
          <article-title>Temporal model for business process</article-title>
          .
          <source>Intelligent Decision Technologies</source>
          .
          <volume>5</volume>
          (
          <issue>4</issue>
          ), pp.
          <fpage>321</fpage>
          -
          <lpage>332</lpage>
          (
          <year>2011</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref20">
        <mixed-citation>
          20.
          <string-name>
            <surname>Ma</surname>
          </string-name>
          , J.,
          <string-name>
            <surname>Knight</surname>
            ,
            <given-names>B.: A General</given-names>
          </string-name>
          <string-name>
            <surname>Temporal</surname>
            <given-names>Theory</given-names>
          </string-name>
          , the
          <source>Computer Journal</source>
          ,
          <volume>37</volume>
          (
          <issue>2</issue>
          ),
          <fpage>114</fpage>
          -
          <lpage>123</lpage>
          (
          <year>1994</year>
          ).
        </mixed-citation>
      </ref>
      <ref id="ref21">
        <mixed-citation>
          21.
          <string-name>
            <surname>Kapetanakis</surname>
            ,
            <given-names>S.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Petridis</surname>
            ,
            <given-names>M.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Knight</surname>
            ,
            <given-names>B.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Ma</surname>
          </string-name>
          , J.,
          <string-name>
            <surname>Bacon</surname>
            ,
            <given-names>L.</given-names>
          </string-name>
          :
          <article-title>A Case Based Reasoning Approach for the Monitoring of Business Workflows</article-title>
          ,
          <source>18th International Conference on Case-Based Reasoning, ICCBR</source>
          <year>2010</year>
          , Alessandria, Italy,
          <string-name>
            <surname>LNAI</surname>
          </string-name>
          (
          <year>2010</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref22">
        <mixed-citation>
          22.
          <string-name>
            <surname>Kapetanakis</surname>
            ,
            <given-names>S.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Petridis</surname>
            ,
            <given-names>M.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Knight</surname>
            ,
            <given-names>B.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Ma</surname>
          </string-name>
          , J.,
          <string-name>
            <surname>Bacon</surname>
            ,
            <given-names>L.</given-names>
          </string-name>
          :
          <article-title>Providing Explanations for the Intelligent Monitoring of Business Workflows Using Case-Based Reasoning</article-title>
          , in workshop proceedings of ExACT-10
          <source>at ECAI</source>
          <year>2010</year>
          , Lisbon, Portugal (
          <year>2010</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref23">
        <mixed-citation>
          23.
          <string-name>
            <surname>Kapetanakis</surname>
            ,
            <given-names>S.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Petridis</surname>
            ,
            <given-names>M.</given-names>
          </string-name>
          :
          <article-title>Evaluating a Case-Based Reasoning Architecture for the Intelligent Monitoring of Business Workflows</article-title>
          ,
          <source>in Successful Case-based Reasoning</source>
          Applications-2,
          <string-name>
            <given-names>S.</given-names>
            <surname>Montani</surname>
          </string-name>
          and L.C. Jain, Editors.
          <year>2014</year>
          , Springer Berlin Heidelberg. p.
          <fpage>43</fpage>
          -
          <lpage>54</lpage>
          .
        </mixed-citation>
      </ref>
    </ref-list>
  </back>
</article>