<!DOCTYPE article PUBLIC "-//NLM//DTD JATS (Z39.96) Journal Archiving and Interchange DTD v1.0 20120330//EN" "JATS-archivearticle1.dtd">
<article xmlns:xlink="http://www.w3.org/1999/xlink">
  <front>
    <journal-meta>
      <journal-title-group>
        <journal-title>Workshop on Artificial Intelligence and Formal Verification, Logics, Automata and Synthesis (OVERLAY),
Rende, Italy, November</journal-title>
      </journal-title-group>
    </journal-meta>
    <article-meta>
      <title-group>
        <article-title>Towards Interval Temporal Logic Rule-Based Classification∗</article-title>
      </title-group>
      <contrib-group>
        <contrib contrib-type="author">
          <string-name>E. Lucena-Sánchez</string-name>
          <xref ref-type="aff" rid="aff2">2</xref>
          <xref ref-type="aff" rid="aff3">3</xref>
        </contrib>
        <contrib contrib-type="author">
          <string-name>E. Muñoz-Velasco</string-name>
          <xref ref-type="aff" rid="aff0">0</xref>
        </contrib>
        <contrib contrib-type="author">
          <string-name>G. Sciavicco</string-name>
          <xref ref-type="aff" rid="aff2">2</xref>
        </contrib>
        <contrib contrib-type="author">
          <string-name>I.E. Stan</string-name>
          <xref ref-type="aff" rid="aff1">1</xref>
          <xref ref-type="aff" rid="aff2">2</xref>
        </contrib>
        <contrib contrib-type="author">
          <string-name>A. Vaccari</string-name>
          <email>alessandr.vaccari@student.unife.it</email>
          <xref ref-type="aff" rid="aff2">2</xref>
        </contrib>
        <aff id="aff0">
          <label>0</label>
          <institution>Dept. of Applied Mathematics, University of Málaga</institution>
          ,
          <country country="ES">Spain</country>
        </aff>
        <aff id="aff1">
          <label>1</label>
          <institution>Dept. of Mathematical</institution>
          ,
          <addr-line>Physical, and Computer Sciences</addr-line>
          ,
          <institution>University of Parma</institution>
          ,
          <country country="IT">Italy</country>
        </aff>
        <aff id="aff2">
          <label>2</label>
          <institution>Dept. of Mathematics and Computer Science, University of Ferrara</institution>
          ,
          <country country="IT">Italy</country>
        </aff>
        <aff id="aff3">
          <label>3</label>
          <institution>Dept. of Physics</institution>
          ,
          <addr-line>Informatics, and Mathematics</addr-line>
          ,
          <institution>University of Modena and Reggio Emilia</institution>
          ,
          <country country="IT">Italy</country>
        </aff>
      </contrib-group>
      <pub-date>
        <year>2019</year>
      </pub-date>
      <volume>1</volume>
      <fpage>9</fpage>
      <lpage>20</lpage>
      <abstract>
        <p>Supervised classification is one of the main computational tasks of modern Artificial Intelligence, and it is used to automatically extract an underlying theory from a set of already classified instances. The available learning schemata are mostly limited to static instances, in which the temporal component of the information is absent, neglected, or abstracted into atemporal data, and purely, native temporal classification is still largely unexplored. In this paper, we propose a temporal rulebased classifier based on interval temporal logic, that is able to learn a classification model for multivariate classified (abstracted) time series, and we discuss some implementation issues.</p>
      </abstract>
    </article-meta>
  </front>
  <body>
    <sec id="sec-1">
      <title>Introduction</title>
      <p>
        Supervised classification, specifically, supervised classification model learning [
        <xref ref-type="bibr" rid="ref14 ref17">14, 17</xref>
        ] is one of the main
tasks associated with Artificial Intelligence. Roughly, given a data set of instances, each one of which
belongs to a known class, and each instance described by a finite set of attributes, supervised classification
model learning is the task of learning how the values of the attributes determine the class in the context of
the data set. Supervised classification models can be broadly divided into function-based, tree-based, and
rule-based, depending on how the model is represented. Function-based classification models range from
the very simple regression models, to neural network models (see, among others, [
        <xref ref-type="bibr" rid="ref10 ref11 ref8">8, 11, 10</xref>
        ]), and they are
characterized by the fact that the underlying theory is modelled as a function whose output value is used
to determine the class. Tree-based classification models are characterized by describing the underlying
theory as a tree, and they range from deterministic single-tree models, such as ID3 or C4.5, to random
forest models (see, e.g., [
        <xref ref-type="bibr" rid="ref18 ref19 ref3">3, 18, 19</xref>
        ]). Finally, rule-based classification models, which are of interest in this
work, describe an underlying theory by means of sets of rules, which are not necessarily dependent from
each other, and, in a sense, imitate the human reasoning (see, among others, [
        <xref ref-type="bibr" rid="ref12 ref13 ref16">12, 13, 16</xref>
        ]). Rule-based
classification generalizes tree-based classification, since a tree can be seen as a particular case set of rules,
but not the other way around. Moreover, rule-based models are generally interpretable, that is, they are
models that can be read (and explained) by a human, while function-based models (especially those based
on neural networks) are not.
      </p>
      <p>
        A common denominator to most classification models is the fact that they are atemporal. Classical
classification problems consider static instances, in which the temporal component is absent, or neglected.
In some cases, the temporal aspects of the information are abstracted (e.g., averaging the fever over all
the observation period of the patients), so that static algorithms can be used. Nevertheless, in some
applications, resorting to static models may not be adequate. Purely temporal instances such as in the
above example can be seen as multivariate time series, that is, collections of points in time in which
the interesting values are recorded: while there are several techniques devoted to single (univariate and
multivariate) time series explanation and prediction, classification models for time series are still in their
infancy. As shown in [
        <xref ref-type="bibr" rid="ref20">20</xref>
        ], time series can be abstracted into (interval-based) timelines in a rather general
way, and timelines, in turn, can be described in a very convenient way by means of interval temporal
logics, such as Halpern and Shoham’s logic of Allen’s relations (HS [
        <xref ref-type="bibr" rid="ref9">9</xref>
        ]). Static rule-based classification
produces models that can be described by means of propositional logic-like rules; here, we propose an
algorithm that is able to extract an interval temporal logic-like rule-based classifier, and in which finite
interval temporal model checking [
        <xref ref-type="bibr" rid="ref15">15</xref>
        ] plays a central role. A tree-based classification model algorithm
based on the same principle has been recently presented in [
        <xref ref-type="bibr" rid="ref5">5</xref>
        ], while a deterministic algorithm for HS
association rule extraction from timelines has been proposed in [
        <xref ref-type="bibr" rid="ref4">4</xref>
        ].
2
      </p>
    </sec>
    <sec id="sec-2">
      <title>The general context</title>
      <p>
        There are several application domains in which learning a temporal classification model may be useful.
Time information is commonly represented as a time series. The learning paradigm for the temporal
case can be orthogonally partitioned in univariate versus multivariate learning, and reasoning about
one instance versus a collection of instances. Time series forecasting [
        <xref ref-type="bibr" rid="ref2">2</xref>
        ] is the problem of predicting
what will happen in the next unseen observation of an instance, given the finite temporal history of one
variable. In the case of temporal multivariate learning with a single instance, one may try to describe the
behaviour of a variable based on the values of other variables at the same time instant and/or previous
ones. WEKA [
        <xref ref-type="bibr" rid="ref21">21</xref>
        ], among others, has packages for modelling such scenarios, where the time series are
transformed in some convenient way so that any classical propositional-like model (e.g., multivariate
regression) can be used to solve the problem. The cases when the problem is addressed to a collection of
instances require a different treatment. When the analysis is centered on a single temporal variable, the
typical solution proposed in the literature is to identify/mine frequent patterns across the instances to use
as driving discriminant characteristics.
      </p>
      <p>
        This work is part of a bigger project where the investigation is devoted to the more general problem
of multiple instances multivariate temporal classification. Given a set of labelled (or classified) instances
from a set of countable classes, the classification problem is to find a function/model that relates each
instance to its class in such a way to minimize the misclassification error. The goal, in our case, is to
treat time in an explicit way within the classification problem, by using the (well-studied) interval-based
temporal logic formalism to describe an interpretable theory which takes full advantage of the temporal
information, where the expressive power of the (interval) language plays a central role; in this sense,
this problem generalizes the classical (static) classification problem (because of the temporal component)
and the classical (temporal) forecasting problem (because it deals with multiple instances). A first step
towards solving the multiple instances multivariate temporal classification problem has been done by
devising a general algorithm to extract timelines (i.e., interval models) from time series, addressed in [
        <xref ref-type="bibr" rid="ref20">20</xref>
        ].
Such a representation allows an approach to abstract interval temporal learning, and highlights how the
interval-based approach is, in general, more natural than a point-based one.
      </p>
      <p>Inspired by the fact that interval temporal learning generalizes classical learning, it seems natural to
hAi
hLi
hBi
hEi
hDi
hOi</p>
      <p>Allen’s relations
[x, y]RA[x0, y0] ⇔ y = x0
[x, y]RL[x0, y0] ⇔ y &lt; x0
[x, y]RB[x0, y0] ⇔ x = x0, y0 &lt; y
[x, y]RE[x0, y0] ⇔ y = y0, x &lt; x0
[x, y]RD[x0, y0] ⇔ x &lt; x0, y0 &lt; y
[x, y]RO[x0, y0] ⇔ x &lt; x0 &lt; y &lt; y0</p>
      <p>Graphical representation
x y
x0</p>
      <p>y0
x0</p>
      <p>
        y0
approach this problem by generalizing classical (propositional-like) algorithms. A preliminary result in
this sense has been presented in [
        <xref ref-type="bibr" rid="ref5">5</xref>
        ], in the form of an interval-based generalization of the well-known
decision tree induction algorithm ID3. In this work, we consider the problem of extracting a temporal
classifier composed of a set of implicative rules following the same principle: we start from a temporal
data set of labelled timelines and we extract a temporal classifier (in which the left-hand side of rules is
written in interval temporal logic) by generalizing a classical extraction algorithm.
3
      </p>
    </sec>
    <sec id="sec-3">
      <title>Background</title>
      <p>
        Interval temporal logic. Let D = hD, &lt;i be a strict linear order. An interval over D is an ordered
pair [x, y], where x, y ∈ D and x &lt; y. We denote the set of all intervals over D with I(D). There are 13
different Allen’s relations between two intervals in a linear order [
        <xref ref-type="bibr" rid="ref1">1</xref>
        ]: the six relations RA (adjacent to),
RL (later than), RB (begins), RE (ends), RD (during), and RO (overlaps), depicted in Fig. 1, and their
inverses, for each X ∈ A, where A = {A, L, B, E, D, O}. Halpern and Shoham’s modal logic of temporal
intervals (HS) is defined from a set of propositional letters AP, and by associating a universal modality
[X] and an existential one hXi to each Allen’s relation RX . Formulas of HS are obtained by:
ϕ ::= p | ¬ϕ | ϕ ∨ ϕ | hXiϕ | hXiϕ | [=]ϕ,
where p ∈ AP and X ∈ A. The other Boolean connectives and the logical constants, as well as the
universal modalities [X], can be defined in the standard way. Observe that the operator [=] is redundant
and not included in the original language; it becomes important in the phase of inducting a formula, as we
shall see. The semantics of HS formulas is given in terms of timelines T = hI(D), V i, where V : AP → 2I(D)
is a valuation function which assigns to each atomic proposition p ∈ AP the set of intervals V (p) on which
p holds. The truth of a formula ϕ on a given interval [x, y] in an interval model T is defined by structural
induction on formulas as follows:
      </p>
      <p>
        T, [x, y]
T, [x, y]
T, [x, y]
T, [x, y]
T, [x, y]
T, [x, y]
p if [x, y] ∈ V (p), for p ∈ AP;
¬ψ if T, [x, y] 6 ψ;
ψ ∨ ξ if T, [x, y] ψ or T, [x, y] ξ;
hXiψ if there is [w, z] s.t [x, y]RX [w, z] and T, [w, z]
hXiψ if there is [w, z] s.t [x, y]RX [w, z] and T, [w, z]
[=]ψ if T, [x, y] ψ.
ψ,
ψ,
In [
        <xref ref-type="bibr" rid="ref20">20</xref>
        ], a general-purpose algorithm has been presented that is able to transform multivariate time series
into timelines (i.e., a temporal data set), which can be used to learn a classification model. Temporal
data sets have been first introduced in [
        <xref ref-type="bibr" rid="ref15">15</xref>
        ], where a polynomial time finite model checking algorithm for
formulas of HS has been proposed.
Learning a rule-based supervised classifier. Given a data set S of m instances I1, . . . , Im, each
represented by n distinct attributes, and each belonging to a predetermined class, the supervised classification
model extraction (or learning) is the task of inducing a model that explains how the values of the attributes
determine the value of the class [
        <xref ref-type="bibr" rid="ref14 ref17">14, 17</xref>
        ]. A rule-based classification model is characterized by the set of its
rules (e.g., if the patient is over 40, then he/she presents a recurrence). Even if static instances present
numeric attributes, classical classification models are considered propositional. A typical propositional
rule has the form:
where A(ρ) = p1 ∧ p2 ∧ . . . is its antecedent, and C(ρ) = c is its consequent, and a rule-based model has
the form:
      </p>
      <p>ρ : p1 ∧ p2 ∧ . . . → c,
 ρ1 : p11 ∧ p122 ∧ . . . ps11 → c1
Γ =  ρ2 : p21 ∧ p2 ∧ . . . ps22 → c2
 . . . k
 ρk : p1k ∧ p2 ∧ . . . pskk → ck
Each proposition plh above represents a condition (such as the patient being over 40 years); even if the
symbols cl do not belong to the object language, the rules can be considered as written in propositional
logic (in particular, A(ρ), that is, the antecedent of ρ, is a pure propositional formula). Rules have the
form of logical implications, but they are not interpreted as such: the notion of a rule being satisfied on
a data set S (composed by several instances) is statistical, rather than logical. In particular, given an
instance I ∈ S, we say that a rule ρ holds on I if A(ρ) is satisfied by the values of I, and we say that I is
classified as c by Γ (denoted Γ(I) = c) if, given the rules of Γ that hold on I, and given the firing policy
of Γ, the class c is assigned to I. There are standard ways to evaluate the performances of a supervised
classifier.
4</p>
    </sec>
    <sec id="sec-4">
      <title>An Optimization Based Classifiers Model for Inducing Temporal Logic Rule</title>
      <p>(1)
(2)
(3)
(4)
Let us consider a temporal data set T of timelines T1, T2, . . . , Tm. Each timeline is the abstraction of a
multivariate time series with n distinct variables, and it is (binary) classified. We define a generic temporal
classification rule of HS as an object ρ of the type:</p>
      <p>ρ : p1 ∧ Op(p2 ∧ Op(. . .) . . .) → c,
where each pi is either a propositional letter or the constant &gt;, and each Op is either hXi, [X], hXi, or
[X], where X ∈ A. While there are several possible alternative forms for temporal classification rules, an
object of the type of (3) is immediately interpretable, and allows rules to take the form of (temporal)
patterns, by generalizing (1). For us, a temporal classifier Γ is an object of the type:
 ρ1 : p11 ∧ Op(p12 ∧ Op(. . . ∧ ps11 ) . . .) → c1

Γ =  ρ2 : p21 ∧ Op(p22 ∧ Op(. . . ∧ ps22 ) . . .) → c2
 . . .</p>
      <p> ρk : p1k ∧ Op(p2k ∧ Op(. . . ∧ pskk ) . . .) → ck.</p>
      <p>
        Now let u¯ be a vector of decision variables (a candidate solution) that encodes a classifier, and let Γu¯ (resp.,
u¯Γ) be the classifier that corresponds to u¯ (resp., the encoding that corresponds to Γ). The classifier Γu¯
can be evaluated on the temporal data set T to obtain a measure P er(Γu¯) of its performance. Moreover,
let Dim(Γ) be a function measuring the total number of symbols used in Γ, md(Γ) a function measuring
the maximum modal depth of any antecedent A(ρ) for any ρ ∈ Γ, and |Γ| a function that measures the
number of rules in Γ. Then, the problem of inducing a temporal rule-based classifier for a given temporal
data set can be seen as the following multi-objective optimization problem (see, e.g. [
        <xref ref-type="bibr" rid="ref6">6</xref>
        ]):
      </p>
      <p>Algorithm 1: Rule-based classifier extraction algorithm.</p>
      <p>
        input : Γ, T , maxρ, and maxmd.
1 Π ← InitialP opulation(maxρ, maxmd, AP) // random population of candidate solutions
2 for i = 1 to maxGen do
3 foreach Γ ∈ Π do
4 P er(Γ) ← ComputeP erf ormance(Γ, T ) // using finite model checking [
        <xref ref-type="bibr" rid="ref15">15</xref>
        ]
5 Dim(Γ) ← ComputeDimension(Γ)
6
7
8
9 end
10 return Π
end
Π0 ← Selection(Π, P er(Γ), Dim(Γ))
Π ← N ewGeneration(Π0)
      </p>
      <p>
        // (Pareto) non-dominated solutions
// new population via evolutionary operators
 max P er(Γu¯)


 min Dim(Γu¯)
 subject to
 2 ≤ |Γu¯| ≤ maxρ

 md(Γu¯) ≤ maxmd,
where maxρ and maxmd are, respectively, the maximum number of rules in classifier and the maximum
modal depth for each rule. Multi-objective evolutionary algorithms (see, e.g., [
        <xref ref-type="bibr" rid="ref7">7</xref>
        ]) are known to be
particularly suitable to perform multi-objective optimization, as they search for multiple optimal solutions
in parallel. Algorithm 1 is the adaptation of the general schemata of a evolutionary algorithm to our case,
in which we abstract over the difference between a solution Γ and its internal representation u¯Γ. After
maxGen generations, the procedure stops. The solution of multi-objective optimization is a population of
candidates that tends towards the so-called Pareto front. This means that, in general, we shall have many
possible classifiers. Each classifier is optimal within the part of the search space that has been explored,
that is, during the execution, all classifiers that have been produced with worse performance and same
dimension, as well as all classifiers with greater dimension and same performance have been discarded.
From the returned population, a decision-making strategy must be applied to choose one classifier, and
that may depend from the intended application.
5
      </p>
    </sec>
    <sec id="sec-5">
      <title>Conclusions</title>
      <p>Supervised classification is one of the main computational tasks of modern Artificial Intelligence. In this
paper we defined the problem of supervised temporal classification, and then we proposed a solution based
on interval temporal logic, essentially defining the task of temporal classification model induction as an
optimization problem in which finite interval temporal model checking plays a central role in the design of
Algorithm 1.</p>
    </sec>
  </body>
  <back>
    <ref-list>
      <ref id="ref1">
        <mixed-citation>
          [1]
          <string-name>
            <given-names>J.</given-names>
            <surname>Allen</surname>
          </string-name>
          .
          <article-title>Maintaining knowledge about temporal intervals</article-title>
          .
          <source>Communications of the ACM</source>
          ,
          <volume>26</volume>
          (
          <issue>11</issue>
          ):
          <fpage>832</fpage>
          -
          <lpage>843</lpage>
          ,
          <year>1983</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref2">
        <mixed-citation>
          [2]
          <string-name>
            <given-names>G. E. P.</given-names>
            <surname>Box</surname>
          </string-name>
          ,
          <string-name>
            <given-names>G. M.</given-names>
            <surname>Jenkins</surname>
          </string-name>
          ,
          <string-name>
            <given-names>G. C.</given-names>
            <surname>Reinsel</surname>
          </string-name>
          , and
          <string-name>
            <given-names>G. M.</given-names>
            <surname>Ljung</surname>
          </string-name>
          .
          <source>Time Series Analysis: Forecasting and Control. Wiley, 5 edition</source>
          ,
          <year>2015</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref3">
        <mixed-citation>
          [3]
          <string-name>
            <given-names>L.</given-names>
            <surname>Breiman</surname>
          </string-name>
          .
          <article-title>Random forests</article-title>
          .
          <source>Machine Learning</source>
          ,
          <volume>1</volume>
          (
          <issue>45</issue>
          ):
          <fpage>5</fpage>
          -
          <lpage>32</lpage>
          ,
          <year>2001</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref4">
        <mixed-citation>
          [4]
          <string-name>
            <given-names>D.</given-names>
            <surname>Bresolin</surname>
          </string-name>
          ,
          <string-name>
            <given-names>E.</given-names>
            <surname>Cominato</surname>
          </string-name>
          ,
          <string-name>
            <given-names>S.</given-names>
            <surname>Gnani</surname>
          </string-name>
          ,
          <string-name>
            <given-names>E.</given-names>
            <surname>Muñoz-Velasco</surname>
          </string-name>
          , and
          <string-name>
            <given-names>G.</given-names>
            <surname>Sciavicco</surname>
          </string-name>
          .
          <article-title>Extracting interval temporal logic rules: A first approach</article-title>
          .
          <source>In Proc. of the 24th International Symposium on Temporal Representation and Reasoning (TIME)</source>
          ,
          <source>number 120 in Leibniz International Proceedings in Informatics</source>
          , pages
          <fpage>1</fpage>
          -
          <lpage>15</lpage>
          ,
          <year>2018</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref5">
        <mixed-citation>
          [5]
          <string-name>
            <given-names>A.</given-names>
            <surname>Brunello</surname>
          </string-name>
          , G. Sciavicco,
          <string-name>
            <given-names>and I.</given-names>
            <surname>Stan</surname>
          </string-name>
          .
          <article-title>Interval temporal logic decision tree learning</article-title>
          .
          <source>In Proc. of the 16th European Conference on Logics in Artificial Intelligence (JELIA</source>
          <year>2019</year>
          ),
          <source>number 11468 in Lecture Notes in Computer Science</source>
          , pages
          <fpage>778</fpage>
          -
          <lpage>793</lpage>
          ,
          <year>2019</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref6">
        <mixed-citation>
          [6]
          <string-name>
            <given-names>Y.</given-names>
            <surname>Collette</surname>
          </string-name>
          and
          <string-name>
            <given-names>P.</given-names>
            <surname>Siarry</surname>
          </string-name>
          .
          <source>Multiobjective Optimization: Principles and Case Studies</source>
          . Springer Berlin Heidelberg,
          <year>2004</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref7">
        <mixed-citation>
          [7]
          <string-name>
            <given-names>K.</given-names>
            <surname>Deb</surname>
          </string-name>
          <article-title>. Multi-objective optimization using evolutionary algorithms</article-title>
          . Wiley, London, UK,
          <year>2001</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref8">
        <mixed-citation>
          [8]
          <string-name>
            <given-names>I.</given-names>
            <surname>Goodfellow</surname>
          </string-name>
          ,
          <string-name>
            <given-names>Y.</given-names>
            <surname>Bengio</surname>
          </string-name>
          ,
          <article-title>and</article-title>
          <string-name>
            <given-names>A.</given-names>
            <surname>Courville</surname>
          </string-name>
          . Deep Learning. MIT Press,
          <year>2016</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref9">
        <mixed-citation>
          [9]
          <string-name>
            <given-names>J.</given-names>
            <surname>Halpern</surname>
          </string-name>
          and
          <string-name>
            <given-names>Y.</given-names>
            <surname>Shoham</surname>
          </string-name>
          .
          <article-title>A propositional modal logic of time intervals</article-title>
          .
          <source>Journal of the ACM</source>
          ,
          <volume>38</volume>
          (
          <issue>4</issue>
          ):
          <fpage>935</fpage>
          -
          <lpage>962</lpage>
          ,
          <year>1991</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref10">
        <mixed-citation>
          [10]
          <string-name>
            <given-names>S.</given-names>
            <surname>Haykin</surname>
          </string-name>
          .
          <article-title>Neural networks: A comprehensive foundation</article-title>
          .
          <source>Prentice Hall</source>
          ,
          <year>1999</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref11">
        <mixed-citation>
          [11]
          <string-name>
            <given-names>D.</given-names>
            <surname>Hosmer</surname>
          </string-name>
          and
          <string-name>
            <given-names>S.</given-names>
            <surname>Lemeshow</surname>
          </string-name>
          . Applied Logistic Regression (2nd ed.). Wiley,
          <year>2000</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref12">
        <mixed-citation>
          [12]
          <string-name>
            <given-names>J.</given-names>
            <surname>Ji</surname>
          </string-name>
          ,
          <string-name>
            <given-names>N.</given-names>
            <surname>Zhang</surname>
          </string-name>
          , C. Liu, and
          <string-name>
            <given-names>N.</given-names>
            <surname>Zhong</surname>
          </string-name>
          .
          <article-title>An ant colony optimization algorithm for learning classification rules</article-title>
          .
          <source>In Proc. of the 2006 IEEE/WIC/ACM International Conference on Web Intelligence</source>
          , pages
          <fpage>1034</fpage>
          -
          <lpage>1037</lpage>
          ,
          <year>2006</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref13">
        <mixed-citation>
          [13]
          <string-name>
            <given-names>Y.</given-names>
            <surname>Liu</surname>
          </string-name>
          ,
          <string-name>
            <given-names>Q.</given-names>
            <surname>Zheng</surname>
          </string-name>
          ,
          <string-name>
            <given-names>Z.</given-names>
            <surname>Shi</surname>
          </string-name>
          ,
          <string-name>
            <given-names>and J.</given-names>
            <surname>Chen</surname>
          </string-name>
          .
          <article-title>Rule discovery with particle swarm optimization</article-title>
          .
          <source>In Proc. of the 2004 Conference on Advanced Workshop on Content Computing (AWCC)</source>
          , volume
          <volume>3309</volume>
          of Lecture Notes in Computer Science, pages
          <fpage>291</fpage>
          -
          <lpage>296</lpage>
          ,
          <year>2004</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref14">
        <mixed-citation>
          [14]
          <string-name>
            <given-names>T.</given-names>
            <surname>Mitchell</surname>
          </string-name>
          . Machine Learning.
          <source>McGraw-Hill</source>
          , New York, NY, USA,
          <year>1997</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref15">
        <mixed-citation>
          [15]
          <string-name>
            <surname>D. D. Monica</surname>
          </string-name>
          , D. de
          <string-name>
            <surname>Frutos-Escrig</surname>
            ,
            <given-names>A.</given-names>
          </string-name>
          <string-name>
            <surname>Montanari</surname>
            ,
            <given-names>A.</given-names>
          </string-name>
          <string-name>
            <surname>Murano</surname>
            , and
            <given-names>G.</given-names>
          </string-name>
          <string-name>
            <surname>Sciavicco</surname>
          </string-name>
          .
          <article-title>Evaluation of temporal datasets via interval temporal logic model checking</article-title>
          .
          <source>In Proc. of the 24th International Symposium on Temporal Representation and Reasoning (TIME)</source>
          , pages
          <fpage>1</fpage>
          -
          <lpage>18</lpage>
          ,
          <year>2017</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref16">
        <mixed-citation>
          [16]
          <string-name>
            <given-names>M.</given-names>
            <surname>Muntean</surname>
          </string-name>
          ,
          <string-name>
            <given-names>C.</given-names>
            <surname>Rotar</surname>
          </string-name>
          ,
          <string-name>
            <surname>I. Ileană</surname>
          </string-name>
          , and
          <string-name>
            <given-names>H.</given-names>
            <surname>Vălean</surname>
          </string-name>
          .
          <article-title>Learning classification rules with genetic algorithm</article-title>
          .
          <source>In Proc. of the 8th International Conference on Communications</source>
          , pages
          <fpage>213</fpage>
          -
          <lpage>216</lpage>
          ,
          <year>2010</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref17">
        <mixed-citation>
          [17]
          <string-name>
            <given-names>K.</given-names>
            <surname>Murphy</surname>
          </string-name>
          .
          <article-title>Machine learning: a probabilistic perspective</article-title>
          . MIT press,
          <year>2012</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref18">
        <mixed-citation>
          [18]
          <string-name>
            <given-names>J.</given-names>
            <surname>Quinlan</surname>
          </string-name>
          .
          <article-title>Induction of decision trees</article-title>
          .
          <source>Machine Learning</source>
          ,
          <volume>1</volume>
          :
          <fpage>81</fpage>
          -
          <lpage>106</lpage>
          ,
          <year>1986</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref19">
        <mixed-citation>
          [19]
          <string-name>
            <given-names>J.</given-names>
            <surname>Quinlan</surname>
          </string-name>
          .
          <article-title>Simplifying decision trees</article-title>
          .
          <source>International Journal of Human-Computer Studies</source>
          ,
          <volume>51</volume>
          (
          <issue>2</issue>
          ):
          <fpage>497</fpage>
          -
          <lpage>510</lpage>
          ,
          <year>1999</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref20">
        <mixed-citation>
          [20]
          <string-name>
            <given-names>G.</given-names>
            <surname>Sciavicco</surname>
          </string-name>
          ,
          <string-name>
            <surname>I.</surname>
          </string-name>
          <article-title>Stan, and</article-title>
          <string-name>
            <given-names>A.</given-names>
            <surname>Vaccari</surname>
          </string-name>
          .
          <article-title>Towards a general method for logical rule extraction from time series</article-title>
          .
          <source>In Proc. of the 8th International Work-conference on the Interplay between Natural and Artificial Computation (IWINAC)</source>
          ,
          <source>number 11487 in Lecture Notes in Computer Science</source>
          , pages
          <fpage>3</fpage>
          -
          <lpage>12</lpage>
          ,
          <year>2019</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref21">
        <mixed-citation>
          [21]
          <string-name>
            <given-names>I. H.</given-names>
            <surname>Witten</surname>
          </string-name>
          , E. Frank, and
          <string-name>
            <given-names>M. A.</given-names>
            <surname>Hall</surname>
          </string-name>
          .
          <source>Data Mining: Practical Machine Learning Tools and Techniques. Wiley, 3 edition</source>
          ,
          <year>2015</year>
          .
        </mixed-citation>
      </ref>
    </ref-list>
  </back>
</article>