<!DOCTYPE article PUBLIC "-//NLM//DTD JATS (Z39.96) Journal Archiving and Interchange DTD v1.0 20120330//EN" "JATS-archivearticle1.dtd">
<article xmlns:xlink="http://www.w3.org/1999/xlink">
  <front>
    <journal-meta />
    <article-meta>
      <title-group>
        <article-title>Interpretable Narrative Explanation for ML Predictors with LP: A Case Study for XAI</article-title>
      </title-group>
      <contrib-group>
        <aff id="aff0">
          <label>0</label>
          <institution>Roberta Calegari Giovanni Ciatto Jason Dellaluce Andrea Omicini Dipartimento di Informatica - Scienza e Ingegneria (DISI) A</institution>
        </aff>
      </contrib-group>
      <pub-date>
        <year>2019</year>
      </pub-date>
      <fpage>105</fpage>
      <lpage>112</lpage>
      <abstract>
        <p>-In the era of digital revolution, individual lives are going to cross and interconnect ubiquitous online domains and offline reality based on smart technologies-discovering, storing, processing, learning, analysing, and predicting from huge amounts of environment-collected data. Sub-symbolic techniques, such as deep learning, play a key role there, yet they are often built as black boxes, which are not inspectable, interpretable, explainable. New research efforts towards explainable artificial intelligence (XAI) are trying to address those issues, with the final purpose of building understandable, accountable, and trustable AI systems-still, seemingly with a long way to go. Generally speaking, while we fully understand and appreciate the power of sub-symbolic approaches, we believe that symbolic approaches to machine intelligence, once properly combined with sub-symbolic ones, have a critical role to play in order to achieve key properties of XAI such as observability, interpretability, explainability, accountability, and trustability. In this paper we describe an example of integration of symbolic and sub-symbolic techniques. First, we sketch a general framework where symbolic and sub-symbolic approaches could fruitfully combine to produce intelligent behaviour in AI applications. Then, we focus in particular on the goal of building a narrative explanation for ML predictors: to this end, we exploit the logical knowledge obtained translating decision tree predictors into logical programs. Index Terms-XAI, logic programming, machine learning, symbolic vs. sub-symbolic</p>
      </abstract>
    </article-meta>
  </front>
  <body>
    <sec id="sec-1">
      <title>-</title>
      <p>I. INTRODUCTION</p>
      <p>
        Artificial intelligence (AI), machine learning (ML), and
deep learning (DL) are nowadays intertwined with a growing
number of aspects of people’s every day life [
        <xref ref-type="bibr" rid="ref1">1</xref>
        ], [
        <xref ref-type="bibr" rid="ref2">2</xref>
        ]. In fact,
more and more decisions are delegated by humans to software
agents whose intelligent behaviour is not the result of some
skilled developer endowing it with some clever code, but rather
the consequence the agents’ capability of learning, planning,
or inferring what to do from data—or, roughly speaking, their
artificial intelligence.
      </p>
      <p>For instance, banks and insurance companies have adopted
ML and statistical methods since decades, in order to decide
whether or not to grant a loan to a given customer, or to
estimate the most profitable insurance plan for her. Similarly,
ML has been employed in order to help doctors with their
diagnoses, provided that a set of symptoms has been properly
identified for a given patient; whereas statistical and
probabilistic inference have been employed to test drugs, in order
to prove them effective or safe. Furthermore, virtually any
person, as a consumer of services and goods, lets a number
of ML-trained agents decide or suggest what to buy, like, or
read—as any consumer is likely to be profiled by most of the
companies and organisations he/she has interacted.</p>
      <p>In spite of the large adoption, intelligent agents whose
behaviour is the result of automatic synthesis / learning
procedures are difficult to trust for most people—in particular when
people are not expert in the fields of computer or data sciences,
AI, statistics. This is especially true for agents leveraging on
machine or deep learning based techniques, often producing
models whose internal behaviour is opaque and hard to explain
for their developers too.</p>
      <p>There, agents often tend to accumulate their knowledge into
black-box predictive models which are trained through ML or
DL. Broadly speaking, the “black-box” expression is used to
refer to models where knowledge is not explicitly represented
– such as in neural networks, support vector machines, or
Hidden Markov Chains –, and it is therefore difficult, for
humans, to understand what a black-box actually knows, or
what leads to a particular decision.</p>
      <p>
        Such difficulty in understanding black-boxes content and
functioning is what prevents people from fully trusting –
and thus accepting – them. In several contexts, such as the
medical or financial ones, it is not sufficient for intelligent
agents to output bare decisions, since, for instance, ethical
and legal issues may arise. An explanation for each decision
is therefore often desirable, preferable, or even required. For
instance, applications dealing with personal data need to face
the challenges of achieving valid consent for data use and
protecting confidentiality, and addressing threats to privacy,
data protection, and copyright. Those issues are particularly
challenging in critical application scenarios such as
healthcare, often involving the use of image (i.e., identifiable) data
from children. While issues of data ownership, data security,
and data access are important, other ethical issues may arise:
since the diagnostic accuracy and value of the result is
determined by the amount and quality of data used in model
training, the first potential concern is to avoid algorithmic
bias, which may lead to social discrimination and result in
inequitable access to healthcare, just related to the provenience
of the collected data [
        <xref ref-type="bibr" rid="ref1">1</xref>
        ], [
        <xref ref-type="bibr" rid="ref3">3</xref>
        ].
      </p>
      <p>
        Furthermore, it may happen that black-boxes silently learn
something wrong (e.g., Google image recognition software
that classified black people as gorillas [
        <xref ref-type="bibr" rid="ref4">4</xref>
        ], [
        <xref ref-type="bibr" rid="ref5">5</xref>
        ]), or something
right, but in a biased way (like the “background bias” problem,
causing for instance husky images to be recognised only
because of their snowy background [
        <xref ref-type="bibr" rid="ref6">6</xref>
        ]). In such situations,
explanations are expected to provide useful insights for
blackbox developers.
      </p>
      <p>
        To tackle such trust issues, the eXplainable Artificial
Intelligence (XAI) research field has recently emerged, and
a comprehensive research road map has been proposed by
DARPA [
        <xref ref-type="bibr" rid="ref7">7</xref>
        ], targeting the themes of explainability and
interpretability in AI – and in particular ML – as a challenge
of paramount importance in a world where AI is becoming
more and more pervasively adopted. There, DARPA reviews
the main approaches to make AI either more interpretable or a
posteriori explainable, it categorise the many currently
available techniques aimed at building meaningful interpretations
or explanations for black-box models, it summarises the open
problems and challenges, and it provides a successful reference
framework for the researchers interested in the field.
      </p>
      <p>
        The main idea behind XAI is to employ explanators [
        <xref ref-type="bibr" rid="ref8">8</xref>
        ]
to provide easy to understand insights for a given black-box
and its particular decisions. An explanator is any procedure
producing a meaningful explanation for some human observer,
by leveraging on any combination of (i) the black-box, (ii) its
input data, or (iii) its decisions or predictions. To this end,
we believe that symbolic approaches to machine intelligence
– properly integrated with sub-symbolic approaches – may
have a role to play in order to achieve key properties such
as interpretability, observability, explainability, accountability,
and trustability.
      </p>
      <p>
        In this paper we focus on the specific problem of building a
narrative explanation of ML techniques—thus positioning our
contribution into the specific Narrative Generation DARPA
category [
        <xref ref-type="bibr" rid="ref7">7</xref>
        ]. In particular, we first show a general framework
where symbolic and sub-symbolic techniques are fruitfully
combined to produce intelligent behaviour in AI applications.
Then, we focus on the translation of ML predictors into logical
knowledge with the aim to (i) infer new knowledge, (ii) reason
and act accordingly, and (iii) build the narrative explanation
of a decision output (or prediction).
      </p>
      <p>To this end, we propose an automatic procedure aimed at
translating a ML predictor – here in particular we consider the
case of decision trees (DT) – into logical knowledge. We argue
that, when the source DT has been trained over a set of real
data in order to produce a predictor, the corresponding logic
program may be employed to produce a narrative explanation
for any given prediction.</p>
      <p>
        Despite being mostly focused on DT, our proposal represent
a first step towards a more general approach. In fact, DT have
been proposed as a general means for explaining the behaviour
of virtually any black-box model [
        <xref ref-type="bibr" rid="ref9">9</xref>
        ], [
        <xref ref-type="bibr" rid="ref10">10</xref>
        ].
      </p>
      <p>Accordingly, the reminder of this paper is organised as
follows. Section II briefly recalls the ML concepts and
terminology used in the paper as well as the main research efforts in
the field. Then Section III introduces our vision of a framework
for the integration of symbolic and sub-symbolic techniques.
Finally, Section IV discusses early experiments alongside the
prototype implementation.</p>
    </sec>
    <sec id="sec-2">
      <title>II. CONTEXT</title>
    </sec>
    <sec id="sec-3">
      <title>Machine learning often produces black-box predictors based</title>
      <p>on opaque models, thus hiding their internal logic to the user.
This hinders explainability, and represents both a practical and
an ethical issue for ML. As a result, many research approaches
in the XAI field aim at overcoming that crucial weakness,
sometimes at the cost of trading off accuracy against
interpretability. So, we first (Subsection II-B) summarise the state
of the art as well as the goal of XAI, then (Subsection II-A)
introduce some background notions to define the terminology
adopted.</p>
      <p>A. Background</p>
    </sec>
    <sec id="sec-4">
      <title>Since several practical AI problems – such as image recog</title>
      <p>
        nition, financial and medical decision support systems – can
be reduced to supervised ML – which can be further grouped
in terms of either classification or regression problems [
        <xref ref-type="bibr" rid="ref11">11</xref>
        ],
[
        <xref ref-type="bibr" rid="ref12">12</xref>
        ] –, in the reminder of this paper we focus on this set of
ML problems.
      </p>
      <p>In those cases, a learning algorithm is commonly exploited
to estimate the specific nature and shape of an unknown
prediction function (or predictor) p∗ : X → Y, mapping
each input vector x from a given input space X into a
prediction from a given output space Y. To do so, the learning
algorithm takes into account a number N of examples in the
form (xi, yi) such that xi ∈ X ⊂ X , yi ∈ Y ⊂ Y, and
|X| ≡ |Y | ≡ N . There, each xi represents an instance of the
input data for which the expected output value yi is known
or has already been estimated. Such sorts of ML problems
are said to be “supervised” because the expected targets
Y are available, whereas they are said to be “regression”
problems if Y consists of continuous or numerable values, or
“classification” problems if Y consists of categorical values.</p>
      <p>The learning algorithm usually assumes p∗ ∈ P, for a given
family P of predictors—meaning that the unknown prediction
function exists, and it is from P. The algorithm then trains a
predictor pˆ ∈ P such that the value of a given loss function
λ : Y × Y → R – computing the discrepancy among predicted
and expected outputs – is minimal or reasonably low—i.e.:
pˆ = argmin nPiN=1 λ(yi, p(xi))o.</p>
      <p>p∈P</p>
      <p>
        Depending on the predictor family P of choice, the nature
of the learning algorithm and the admissible shapes of pˆ may
vary dramatically, as well as the their interpretability. Even if
the interpretability of predictor families is not a well-defined
feature, most authors agree on the fact that some predictor
families are more interpretable than others [
        <xref ref-type="bibr" rid="ref13">13</xref>
        ]—in the sense
that it is easier for humans to understand the functioning and
the predictions of the former ones. For instance, it is widely
acknowledged that generalized linear models (GLM) are more
interpretable than neural networks (NN), whereas decision
trees (DT) [
        <xref ref-type="bibr" rid="ref14">14</xref>
        ] are among the most interpretable families
[
        <xref ref-type="bibr" rid="ref8">8</xref>
        ]. DT can be considered more interpretable due to their
construction: that is, recursively partitioning the input space
X through a number of splits or decisions based on the input
data X, in such a way that the prediction in each partition
is constant, and the loss w.r.t. Y is low, while keeping the In spite of the many approaches proposed to explain black
amount of partitions low as well. Without affecting generality, boxes, some important scientific questions still remain
unanwe focus on the case of mono-dimensional classification – swered. One of the most important open problems is that,
thus we write y instead of y –, since other cases can be easily until now, there is no agreement on what an explanation is.
reduced to this one. We further assume the input space X is Indeed, some approaches adopt as explanation a set of rules,
N -dimensional, and let nj be the meta-variable representing others a decision tree, others rely on visualisation techniques
the name of the jth dimension of X . [
        <xref ref-type="bibr" rid="ref8">8</xref>
        ]. Moreover, recent works highlight the importance for an
      </p>
      <p>
        Under such hypotheses, a DT predictor pT ∈ Pdt assumes explanation to guarantee some properties, e.g., soundness,
a binary tree T exists such that each node is either completeness, and compactness [
        <xref ref-type="bibr" rid="ref8">8</xref>
        ].
      </p>
      <p>
        • a leaf, carrying and representing a prediction, i.e. and This is why our proposal aims at integrating sub-symbolic
assignment for y, approaches with symbolic ones. To this end, DT can be
• an internal node, carrying and representing a decision, i.e. exploited as an effective bridge between the symbolic and
a formula in the form (nj ≤ c)—where c is a constant sub-symbolic realms. In fact, DT can be easily (i) built from
threshold chosen by the learning algorithm. an existing sub-symbolic predictor, and (ii) translated into
symbolic knowledge – as it is shown in the reminder of this
Each node ν inherits a partition Xν ⊆ X of the original input paper – thanks to their rule-based nature.
data, from its parent. Since the root node ν0 has no parent, it Decision trees are an interpretable family of predictors that
is assigned to the whole set of input data—i.e. Xν0 ≡ X. The have been proposed as a global means for explaining other,
decision carried by each internal node splits its Xν into two less interpretable, sorts of black-box predictors [
        <xref ref-type="bibr" rid="ref9">9</xref>
        ], [
        <xref ref-type="bibr" rid="ref10">10</xref>
        ]—
disjoint parts – XνL and XνR – along the jth dimension of X . such as neural networks [
        <xref ref-type="bibr" rid="ref18">19</xref>
        ]. The main idea behind such an
In particular, XνL contains all the residual xi ∈ Xν such that approach is to build a DT approximating the behaviour of a
(xij ≤ cν ) – which are inherited by ν left child –, whereas XνR given predictor, possibly, by only considering its inputs and its
contains all the residual xi ∈ Xν such that xij &gt; cν —which outputs. Such approximation essentially trades off predictive
are inherited by by ν right child. A leaf node l is created performance with interpretability. In fact, the structure of such
whenever a sequence of splits (i.e., a path from the tree root a DT would then be used to provide useful insights concerning
to the leaf parent) leads to a partition Xl which is (almost) the original predictor inner functioning.
pure—roughly, meaning that Xl (mostly) contains input data Describing the particular means for extracting DT from
xi for which the expected output is the same yl. In this case, black-boxes is outside the scope of this paper. Given the vast
we say that the prediction carried by l is yl. Assuming such a literature on the topic – e.g., consider reading [
        <xref ref-type="bibr" rid="ref8">8</xref>
        ], [
        <xref ref-type="bibr" rid="ref19">20</xref>
        ] for
tree T exists, in order to classify some input data x ∈ X , the an overview or [
        <xref ref-type="bibr" rid="ref18">19</xref>
        ], [
        <xref ref-type="bibr" rid="ref20">21</xref>
        ], [
        <xref ref-type="bibr" rid="ref21">22</xref>
        ] for a practical examples – we
predictor pT simply navigates the path P = (ν0, ν1, ν2, . . . , l) simply assume an extracted DT is available and it has an high
of T such that all decisions νk are matched by x, then it fidelity—meaning that the loss in terms of predictive
perforoutputs yl. mance is low, w.r.t. the original black-box. In fact, whereas
B. XAI: The need for explanation and interpretable models tohuetreofexbilsatcske-vbeorxalpwreodrikcstofros,cunsosinagtteonntiohnowis topasidynttohemsiseergDinTg
      </p>
      <p>
        Since the adoption of interpretable predictors usually comes them with symbolic approaches, which can play a key role in
at cost of a lower potential in terms of predictive performance, enhancing the interpretability and explainability of the system.
explanations are the newly preferred way for providing under- In this paper we focus on such a matter.
standable predictions without necessarily sacrificing accuracy. We believe that a logical representation of DT may be
The idea, and the main goal of XAI is to create intelligible interesting and enabling for further research directions. For
and understandable explanations for uninterpretable predictors instance, as far as explainability is concerned, we show how
without replacing or modifying them. Thus explanations are logic-translated DT can be used to both navigate the
knowlbuilt through a number of heterogeneous techniques, broadly edge stored within the corresponding predictors – thus acting
referred to as explanators [
        <xref ref-type="bibr" rid="ref8">8</xref>
        ]—just to cite some, decision rules as global explanators –, and produce narrative explanations
[15], feature importance [
        <xref ref-type="bibr" rid="ref15">16</xref>
        ], saliency masks [
        <xref ref-type="bibr" rid="ref16">17</xref>
        ], sensitivity for their predictions—thus acting as local explanators. Note
analysis [
        <xref ref-type="bibr" rid="ref17">18</xref>
        ], etc. that the restriction on the DT representation makes it easy to
      </p>
      <p>
        The state of the art for explainability currently recognises map DT onto logical clauses, since DT are finite and with a
two main sorts of explanators, namely, either local or global. limited expressivity (if / else conditions).
While local explanators attempt to provide an explanation for
each particular prediction of a given predictor p, the global III. VISION
ones attempt to provide an explanation for the predictor p as Many approaches to ML nowadays are increasingly
foa whole. In other words, local explanators provide an answer cussing on sub-symbolic approaches – such as deep learning
to the question “why does p predict y for the input x?” – with neural networks [
        <xref ref-type="bibr" rid="ref22">23</xref>
        ] – and on how to make them
such as the LIME technique presented in [
        <xref ref-type="bibr" rid="ref6">6</xref>
        ] –, whereas global work on the large scale. As promising as this may look –
explanators provide an answer to the question “how does p with the premise of potentially minimizing the engineering
build its predictions?”—such as decision rules. efforts needed – it is increasingly acknowledged that those
approaches do not cope well with the socio-technical nature of
the systems they are exploited in, which often demand a degree
of interpretability, observability, explainability, accountability,
and trustability they just cannot deliver.
      </p>
      <p>
        To this end, since logic-based approaches already have
a well-understood role in building intelligent (multi-agent)
systems [
        <xref ref-type="bibr" rid="ref23">24</xref>
        ], declarative, logic-based approaches have the
potential to represent an alternative way of delivering
symbolic intelligence, complementary to the one pursued by
sub-symbolic approaches. In fact, declarative and logic-based
technologies much better address the aforementioned
sociotechnical issues, in particular when exploiting their inferential
capabilities—e.g., [
        <xref ref-type="bibr" rid="ref24">25</xref>
        ].
      </p>
      <p>
        The potential of logic-based models and their extensions is
first of all related to their declarativeness as well as to explicit
knowledge representation, enabling knowledge sharing at the
most adequate level of abstraction, while supporting
modularity and separation of concerns [
        <xref ref-type="bibr" rid="ref25">26</xref>
        ]—which are especially
valuable in open and dynamic distributed systems. As a further
element, LP sound and complete semantics straightforwardly
enables intelligent agents to reason and infer new information
in a sound and complete way.
      </p>
      <p>
        Another relevant point is that LP has been already proven to
work well both as a knowledge representation language and as
an inference platform for rational agents [
        <xref ref-type="bibr" rid="ref26">27</xref>
        ], [
        <xref ref-type="bibr" rid="ref27">28</xref>
        ]. The latter
usually may interact with an external environment by means
of a suitably defined observe–think–act cycle.
      </p>
      <p>Accordingly to this vision, here we propose an integrated
framework of hybrid reasoning – where symbolic and
subsymbolic techniques fruitfully combine to produce intelligent
behaviour.</p>
      <p>Indeed, looking in depth at pervasive socio-technical
systems, it turns out that agents (either human or software)
effortlessly undertake a complex decision making process in
almost all situations, which seamlessly integrates perceptions
(and actions) at two different scales—the macro and the micro:
• at the macro scale, by considering the knowledge of the
global system, rules of general validity and concerning
the most likely situation;
• at the micro scale, we modulate such decision by
considering all the contingencies arising during the precise
situation – such as, for instance, a last minute
inconvenient, etc. As a consequence, we adapt the original plan
to the local perceptions we gather while enacting it.</p>
      <p>
        In order to better illustrate the above remarks, one may
consider as a concrete example the case of a disease diagnosis
in a hospital, where the notions of micro and macro scale w.r.t.
to the nature of algorithms and techniques can be declined as
follows:
• at the macro level, the main concerns regard a mid/long
term horizon and focus the issue of analysis of
highdimensional and multimodal biomedical data train
algorithms to recognize cancerous tissue at a level comparable
to trained physicians—there including, for instance,
representation and recognition of patterns and sequences in
the input data. With such a sort of goals to pursue, it
is not surprising that most IT tools supporting decision
making are based on sub-symbolic approaches such as
deep learning, Bayesian networks, machine vision, latent
Dirichlet analysis, and in general any kind of statistical
approach to ML [
        <xref ref-type="bibr" rid="ref28">29</xref>
        ], [
        <xref ref-type="bibr" rid="ref29">30</xref>
        ], [
        <xref ref-type="bibr" rid="ref30">31</xref>
        ]
• at the micro level, the main concerns regard instead the
short term horizon, and mostly focus on the specific
problem of the patient, there including a few
highlyintertwined sub-problems—e.g. specific symptom or
situation, ongoing epidemic in that hospital or place that
carries the same symptoms. Although sub-symbolic
approaches can still be used, symbolic ones such as fuzzy
logic, specialized level (white box) learning instead of
higher-level learning, symbolic time series are most
common [
        <xref ref-type="bibr" rid="ref28">29</xref>
        ], [
        <xref ref-type="bibr" rid="ref31">32</xref>
        ], [
        <xref ref-type="bibr" rid="ref32">33</xref>
        ]
      </p>
    </sec>
    <sec id="sec-5">
      <title>Generally speaking, we believe the computational intelligence accounts for this two kind of rules: general rules whose validity is essentially unconstrained (speed limits, right of way,</title>
      <p>etc.) which represent the commonsense knowledge necessary
to inhabit the environment and specific rules, with a validity
bound in space and time (school hours and days, open-air
market hours and days, unpredictable events such as incoming
emergency vehicles the need to gather at an evacuation
assembly point), which represent the contextual or expert knowledge
necessary to deal with transient, unforeseen, and unpredictable
situations.</p>
      <p>That is why in the framework envisioned here we plan to
combine sub-symbolic techniques with symbolic ones (LP in
particular): sub-symbolic techniques are exploited for training
the system and learn new rules (commonsense knowledge),
rules are translated into logical knowledge (contextual / expert
knowledge), and the two approaches interact and interleave
to share knowledge and learn from each other in a coherent
framework.</p>
      <p>The framework architecture, depicted in Fig. 1, shows the
embodiment of the vision discussed above: sensor data and
dataset are translated into the logic knowledge base. In
particular the Machine Learning Interface allows for the interaction
of different kinds of ML algorithms with the framework: a
standard interface is proposed in order to combine the specific
features of each algorithm in a coherent manner. ML to Prolog
is the core of the translation into logical knowledge, while the
Prolog to ML returns insights of the logical KB to the ML
predictor—for instance, new inferred rules, or rules learned
by a specific situation. The blocks on the left (Knowledge
Base, Demonstration) reflect the standard architecture of a
Prolog engine. Overall, the framework looks general enough
to account for the variety of ML techniques and algorithms,
and also to ensure the consistency between symbolic and
sub-symbolic approaches. Finally, the block Prolog to ML
currently expresses our vision, and is obviously subject of
future research.</p>
    </sec>
    <sec id="sec-6">
      <title>The first prototype we design and implement enables the</title>
      <p>construction of a narrative explanation of the prediction
generated exploiting the ML technique, thus achieving
interpretability and making a step towards explainability.</p>
      <p>
        With respect to Fig. 1, we experiment the predictor
translation into logical rules, provided by the ML to Prolog. The
experimental results refer to the case in which the predictor
corresponds to a decision tree or to the corresponding crisp
rules [
        <xref ref-type="bibr" rid="ref33">34</xref>
        ]. The conversion generates a Prolog predicate for ✞diagnosis(temperatureOfPatient(T), occurrenceOfNausea(N),
each decision taken by the predictor: inside the predicate, a lumbarPain(L), urinePushing(U), micturitionPains(M),
term for each input/output attribute is instantiated with the bDuercniisnigoOn)f,Urceotnhfriad(eBUn)c,e(nCe)p)hr:i-tiBsoOdfyR.enalPelvisOrigin(
values of the leaf of the decision tree. A rule is generated ✡✝
for each leaf in the tree: between the other advantages, this where the Body body consists of check and computation on
allows for a very compact representation, easy to handle and the variables of the Head terms. For instance, considering the
interoperate with. above tree of Fig. 2, the first generated rule is
      </p>
      <p>
        For a concrete example, let us consider the “Acute in- ✞
flammations data set”1 [
        <xref ref-type="bibr" rid="ref34">35</xref>
        ] supplying data to perform the
presumptive diagnosis of two diseases of urinary system: the
acute inflammations of urinary bladder and acute nephritises.
      </p>
      <p>Input parameters collect all the patient symptoms, each
instance represents a potential patient. The data was created by
a medical expert as a data set to test the expert system, which
performs the presumptive diagnosis of two diseases of urinary
system. The dataset considered is summarised in TABLE I and
TABLE II.</p>
      <p>Starting from the general form Head ← Body for a logical
clause, a predicate in the Head is generated for the decision
of the predictor—in the example, the diagnosis predicate.</p>
      <p>Inside the predicate, a term for each input/output attribute is
instantiated with the value of the decision tree (leaf).</p>
      <p>In our example, the following predicate is generated:
diagnosis(temperatureOfPatient(T), occurrenceOfNausea(N),
lumbarPain(L), urinePushing(U), micturitionPains(M),
burningOfUrethra(BU), nephritis(no), confidence(1.00))
:- T =&lt; 37.95.</p>
    </sec>
    <sec id="sec-7">
      <title>1http://archive.ics.uci.edu/ml/datasets/acute+inflammations</title>
      <p>✆
✆
✝
✡
representing the fact that if the temperature of patient is lesser inferring hidden knowledge in the rules. It is worth noticing
or equal of 37.9, it is unlikely the patient presents nephritis that similar results (emphasising the relations between decision
of renal pelvis; the answer contains a degree of confidence output) can be obtained manipulating the dataset a priori—
based on the case of the dataset that confirm the rule—in the i.e. before the ML algorithm training (a common operation
case 1.00 stands that all the patients in the dataset that have but not always applicable). The manipulation of the above
a temperature lower that 37.9 do not present the disease. dataset, for instance, can build a unique decision output</p>
      <p>To improve readability, the rule above could be written as Result that combines the two different diseases and their
✞diagnosis(temperatureOfPatient(T), _, _, _, _, _, nephritis symptoms. In such a case the dataset is enriched with the
(no), confidence(1.00)) :- T =&lt; 37.95. Result attribute containing the complete diagnosis, i.e., it can
✡✝ ✆assume the values Healthy, Inflammation, Nephritis, Both. The
by omitting the undefined variables, i.e., highlighting the input corresponding decision tree and LP knowledge is depicted in
attribute that are effectively to be considered as influencer. Fig. 3.</p>
      <p>Fig. 2 (left) depicts the whole picture: the decision trees d) Interpretable narrative explanation: LP makes it
posgenerated as output of the example dataset when we run sible to generate a narration for each answer of the predictor.
the basic classification tree algorithm2 and the corresponding The inference Prolog tree becomes inspectable, tracking the
translation into LP rules. With respect to Fig. 1, the decision path for obtaining the answer. For instance, w.r.t. the KB of
trees are the output of the Machine Learning Interface block Fig. 3 – including all diseases –, the diagnosis in the case of
and become the input for the ML to Prolog block. the following symptoms:</p>
      <p>
        Fig. 2 represents experiments of running the ML algorithm ✞
with no manipulation of the dataset: so, since the ML algo- diagnosis(
rithm allows only one decision output to be considered for tleummpbearraPtauirne(yOefsPa)t,ieunrti(n3e6P.u5s)h,inogc(cnuor)r,enceOfNausea(yes),
producing the corresponding decision tree, the information and micturitionPains(yes), burningOfUrethra(yes), _, _).
the related knowledge is fragmented into two different trees – ✡✝
the first obtained running the algorithm with decision output would produce the corresponding narration:
nephritis and the second with decision output inflammation of ✞
urinary bladder. By running the ML to Prolog block of Fig. 1 The diagnosis is healthy, with a full confidence because
we translate the two DT in LP rules as depicted in Fig. 2 the patient has no fever.
(right). %%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%
a) Interpretabilty: The LP program provides an inter- fIonllpoawritnigcuplaatrh:the solution has been built across the
pretable explanation of virtually any predictor. At a glance, Solution: result(healthy) with confidence(1.00).
the user can identify which attributes are meaningful and con- For the proof, the following clauses are considered:
sidered for response and which are not. In case of nephritis, the [
        <xref ref-type="bibr" rid="ref1">1</xref>
        ] diagnosis(temperatureOfPatient(T), _, _, urinePushing(
only significant input attributes are the temperature of patient 7n.o9)5,._, _, result(healthy), confidence(1.00)) :- T =&lt; 3
and the presence or absence of lumbar pain. The same is for [
        <xref ref-type="bibr" rid="ref2">2</xref>
        ] X =&lt; Y that is verified if ’
inflammation of urinary bladder, where the only discriminative expression_less_or_equal_than’(X, Y)
attributes are presence of urine pushing, micturition pains and In the query the temperature T is of 36.5.
lumbar pain. abnedcaubseecauosferoufle[2[]1]’ex3p6.r5es=s&lt;io3n6_.l9eshsa_sort_oeqbuealv_etrhiafni’e(d36.5,
b) Interoperability: The adoption of a standard AI lan- 36.9) has to be verified
guage (LP), in spite of the plethora of different specific ML s%o%%%r%u%l%e%s%[%%1]%%%%a%n%d%%[2%]%%a%r%e%%v%e%r%i%f%i%e%d%.%%%%%%%%%%%%%%%%%%%%%%%%
toolkits, paves the way towards an interoperable explanation ✡✝
where LP is exploited as sort of lingua franca that goes beyond
the technical implementation of each ML framework.
      </p>
      <p>c) Relations between outputs: As emphasised by Fig. 2,
relations between outputs are lost, and possible links between
the diseases are not clearly highlighted having two different
decision trees. Instead, once obtained a LP representation, it
is easy to run simple queries on it in order to get much more
information with respect to the two different decision tree.</p>
      <p>For instance, we can learn that in case of fever (temperature
of patient &gt; 37.95) not presenting nephritis (i.e. no lumbar
pain detected), the only case in which inflammation of
urinary bladder is present is when urine pushing is detected in
absence of symptoms of micturition pains. With the logical
representation, relations between output can be recovered by</p>
    </sec>
    <sec id="sec-8">
      <title>Despite its simplicity, the narration allows for a reconstruction of the decision track, showing the path to the decision. With a large amount of nested rules this could result very effective.</title>
      <p>e) Exploitation of LP extension / abduction on the KB:
Moreover, we believe that exploiting abduction techniques we
could pave the way to hypothetical reasoning with incomplete
knowledge, i.e., learning new possible hypotheses that can
be assumed to hold, provided that they are consistent with
the given knowledge base. The idea, to be explored in future
research, is to provide the most likely solution given a set of
evidence. The conclusion would leave a degree of uncertainty
while highlighting a plausible answer based on the collected
information. In the healthcare field, for instance, it could be
represented by having the collection of symptoms (although
incomplete) and finding the most likely disease for them.</p>
    </sec>
    <sec id="sec-9">
      <title>2We exploit two different implementations: C45 [36] weka J48 for the Java</title>
      <p>
        translator and SciKit-Learn CART [
        <xref ref-type="bibr" rid="ref14">14</xref>
        ] for the Phyton one
✆
✆
      </p>
      <p>Output Decision:
Inflammation of urinary bladder {yes, no}
Output Decision:</p>
      <p>Nephritis of renal pelvis origin {yes, no}</p>
      <p>V. CONCLUSION tive policing. Nevertheless, concerns about the intentional
and unintentional negative consequences of AI systems are</p>
      <p>AI systems nowadays synthesise large amounts of data, legitimate, as well as ethical and legal concerns, mostly related
learning from experience and making predictions with the to darkness and opaqueness of AI decision algorithm. For that
goal of taking autonomous decisions—applications range from reason, recent work on interpretability in machine learning and
clinical decision support to autonomous driving and
predic✞
diagnosis(_, _, _, urinePushing(no), _, _, inflammation(no),</p>
      <p>confidence(1.00)).
diagnosis(_, _, lumbarPain(yes), urinePushing(yes),</p>
      <p>micturitionPains(no), _, inflammation(no), confidence(1.00)).
diagnosis(_, _, lumbarPain(no), urinePushing(yes), micturitionPains</p>
      <p>(no), _, inflammation(yes), confidence(1.00)).
diagnosis(_, _, _, urinePushing(yes), micturitionPains(yes), _,
inflammation, confidence(1.00).
diagnosis(temperatureOfPatient(T), _, lumbarPain(no), _,</p>
      <p>_, _, result(healthy), confidence(1.00)) :- T &gt; 37.95.
diagnosis(temperatureOfPatient(T), _, lumbarPain(yes), _,
micturitionPains(no), _, result(nephritis),
confidence(1.00)) :- T &gt; 37.95.
diagnosis(temperatureOfPatient(T), _, lumbarPain(yes), _,
micturitionPains(yes), _, result(both),
confidence(0.66)) :- T &gt; 37.95.
✞
✝
✡
✞
✝
✡
✝
✡
✝
✡
✆
✆
✆
✆
✆
✆</p>
    </sec>
    <sec id="sec-10">
      <title>AI has focused on simplified models that approximate the true</title>
      <p>criteria used to make decisions.</p>
      <p>In this paper we focus on building a narrative explanation
of the machine learning techniques: we first translate a ML
predictor into logical knowledge, then inspect the proof tree
leading to a solution. The narration is built tracking the path
(i.e., the rules) that leads from the query to the answer.</p>
      <p>Along this line, we foresee a broader vision that involves
the design of a consistent framework where symbolic and
sub-symbolic techniques are fruitfully combined to produce
intelligent behaviour in AI applications while exploiting the
benefits of each approach—like, in the case of symbolic ones,
interpretability, observability, explainability, and
accountability.</p>
      <p>The results presented here represent just a preliminary
exploration of the potential benefits of merging symbolic and
sub-symbolic approaches—where, of course, many critical
issues are still unexplored and will be subject of future work.
However, despite its simplicity, the case study already allows
us to point out the feasibility and the potential benefits of the
exploitation of symbolic techniques towards XAI.</p>
    </sec>
  </body>
  <back>
    <ref-list>
      <ref id="ref1">
        <mixed-citation>
          [1]
          <string-name>
            <given-names>D.</given-names>
            <surname>Helbing</surname>
          </string-name>
          , “
          <article-title>Societal, economic, ethical and legal challenges of the digital revolution: From big data to deep learning, artificial intelligence, and manipulative technologies,” in Towards Digital Enlightenment</article-title>
          . Springer,
          <year>2019</year>
          , pp.
          <fpage>47</fpage>
          -
          <lpage>72</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref2">
        <mixed-citation>
          [2]
          <string-name>
            <given-names>A.</given-names>
            <surname>Elliott</surname>
          </string-name>
          ,
          <article-title>The Culture of AI: Everyday Life and the Digital Revolution</article-title>
          . Routledge,
          <year>2019</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref3">
        <mixed-citation>
          [3]
          <string-name>
            <given-names>S.</given-names>
            <surname>Bird</surname>
          </string-name>
          ,
          <string-name>
            <given-names>K.</given-names>
            <surname>Kenthapadi</surname>
          </string-name>
          , E. Kiciman, and M. Mitchell, “
          <article-title>Fairnessaware machine learning: Practical challenges and lessons learned,” in 12th ACM International Conference on Web Search and Data Mining (WSDM'19)</article-title>
          . ACM,
          <year>2019</year>
          , pp.
          <fpage>834</fpage>
          -
          <lpage>835</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref4">
        <mixed-citation>
          [4]
          <string-name>
            <given-names>M.</given-names>
            <surname>Fourcade</surname>
          </string-name>
          and
          <string-name>
            <given-names>K.</given-names>
            <surname>Healy</surname>
          </string-name>
          , “
          <article-title>Categories all the way down</article-title>
          ,” Historical Social Research/Historische Sozialforschung, pp.
          <fpage>286</fpage>
          -
          <lpage>296</lpage>
          ,
          <year>2017</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref5">
        <mixed-citation>
          [5]
          <string-name>
            <given-names>K.</given-names>
            <surname>Crawford</surname>
          </string-name>
          , “
          <article-title>Artificial intelligence's white guy problem,” The New York Times</article-title>
          , vol.
          <volume>25</volume>
          ,
          <year>2016</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref6">
        <mixed-citation>
          [6]
          <string-name>
            <given-names>M. T.</given-names>
            <surname>Ribeiro</surname>
          </string-name>
          ,
          <string-name>
            <given-names>S.</given-names>
            <surname>Singh</surname>
          </string-name>
          , and
          <string-name>
            <given-names>C.</given-names>
            <surname>Guestrin</surname>
          </string-name>
          , “
          <article-title>Why should I trust you? Explaining the predictions of any classifier,” CoRR</article-title>
          , vol.
          <source>abs/1602.04938</source>
          ,
          <year>2016</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref7">
        <mixed-citation>
          [7]
          <string-name>
            <given-names>D.</given-names>
            <surname>Gunning</surname>
          </string-name>
          , “
          <article-title>Explainable artificial intelligence (XAI),”</article-title>
          <string-name>
            <surname>DARPA</surname>
          </string-name>
          ,
          <string-name>
            <surname>Funding Program</surname>
          </string-name>
          DARPA-BAA-
          <volume>16</volume>
          -53,
          <year>2016</year>
          . [Online]. Available: http://www.darpa.mil/program/explainable-artificial-intelligence
        </mixed-citation>
      </ref>
      <ref id="ref8">
        <mixed-citation>
          [8]
          <string-name>
            <given-names>R.</given-names>
            <surname>Guidotti</surname>
          </string-name>
          ,
          <string-name>
            <given-names>A.</given-names>
            <surname>Monreale</surname>
          </string-name>
          ,
          <string-name>
            <given-names>F.</given-names>
            <surname>Turini</surname>
          </string-name>
          ,
          <string-name>
            <given-names>D.</given-names>
            <surname>Pedreschi</surname>
          </string-name>
          , and
          <string-name>
            <given-names>F.</given-names>
            <surname>Giannotti</surname>
          </string-name>
          , “
          <article-title>A survey of methods for explaining black box models,” CoRR</article-title>
          , vol. abs/
          <year>1802</year>
          .
          <year>01933</year>
          ,
          <year>2018</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref9">
        <mixed-citation>
          [9]
          <string-name>
            <given-names>F.</given-names>
            <surname>Di Castro</surname>
          </string-name>
          and E. Bertini, “
          <article-title>Surrogate decision tree visualization,” in Joint Proceedings of the ACM IUI 2019 Workshops (ACMIUI-WS 2019), ser</article-title>
          .
          <source>CEUR Workshop Proceedings</source>
          , vol.
          <volume>2327</volume>
          ,
          <string-name>
            <surname>Mar</surname>
          </string-name>
          .
          <year>2019</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref10">
        <mixed-citation>
          [10]
          <string-name>
            <given-names>O.</given-names>
            <surname>Bastani</surname>
          </string-name>
          ,
          <string-name>
            <given-names>C.</given-names>
            <surname>Kim</surname>
          </string-name>
          , and
          <string-name>
            <given-names>H.</given-names>
            <surname>Bastani</surname>
          </string-name>
          , “
          <article-title>Interpreting blackbox models via model extraction</article-title>
          ,
          <source>” CoRR</source>
          , vol.
          <source>abs/1705.08504</source>
          ,
          <year>2017</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref11">
        <mixed-citation>
          [11]
          <string-name>
            <given-names>B.</given-names>
            <surname>Twala</surname>
          </string-name>
          , “
          <article-title>Multiple classifier application to credit risk assessment,” Expert Systems with Applications</article-title>
          , vol.
          <volume>37</volume>
          , no.
          <issue>4</issue>
          , pp.
          <fpage>3326</fpage>
          -
          <lpage>3336</lpage>
          ,
          <year>2010</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref12">
        <mixed-citation>
          [12]
          <string-name>
            <given-names>S.</given-names>
            <surname>Kotsiantis</surname>
          </string-name>
          , “
          <article-title>Supervised machine learning: A review of classification techniques,” in Emerging Artificial Intelligence Applications in Computer Engineering, ser</article-title>
          .
          <source>Frontiers in Artificial Intelligence and Applications</source>
          . IOS Press, Oct.
          <year>2007</year>
          , vol.
          <volume>160</volume>
          , pp.
          <fpage>3</fpage>
          -
          <lpage>24</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref13">
        <mixed-citation>
          [13]
          <string-name>
            <given-names>Z. C.</given-names>
            <surname>Lipton</surname>
          </string-name>
          , “
          <article-title>The mythos of model interpretability,” CoRR</article-title>
          , vol.
          <source>abs/1606.03490</source>
          ,
          <year>2016</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref14">
        <mixed-citation>
          [14]
          <string-name>
            <given-names>L.</given-names>
            <surname>Breiman</surname>
          </string-name>
          ,
          <string-name>
            <given-names>J. H.</given-names>
            <surname>Friedman</surname>
          </string-name>
          ,
          <string-name>
            <given-names>R. A.</given-names>
            <surname>Olshen</surname>
          </string-name>
          , and
          <string-name>
            <given-names>C. J.</given-names>
            <surname>Stone</surname>
          </string-name>
          , Classification and
          <string-name>
            <given-names>Regression</given-names>
            <surname>Trees</surname>
          </string-name>
          .
          <source>Chapman &amp; Hall/CRC</source>
          ,
          <year>1984</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref15">
        <mixed-citation>
          [16]
          <string-name>
            <given-names>G.</given-names>
            <surname>Tolomei</surname>
          </string-name>
          ,
          <string-name>
            <given-names>F.</given-names>
            <surname>Silvestri</surname>
          </string-name>
          ,
          <string-name>
            <given-names>A.</given-names>
            <surname>Haines</surname>
          </string-name>
          , and
          <string-name>
            <given-names>M.</given-names>
            <surname>Lalmas</surname>
          </string-name>
          , “
          <article-title>Interpretable predictions of tree-based ensembles via actionable feature tweaking,” in 23rd ACM SIGKDD International Conference on Knowledge Discovery and Data Mining</article-title>
          . ACM,
          <year>2017</year>
          , pp.
          <fpage>465</fpage>
          -
          <lpage>474</lpage>
          . [Online]. Available: http://dl.acm.org/citation.cfm?id=
          <volume>3098039</volume>
          [15]
          <string-name>
            <given-names>M. G.</given-names>
            <surname>Augasta</surname>
          </string-name>
          and
          <string-name>
            <given-names>T.</given-names>
            <surname>Kathirvalavakumar</surname>
          </string-name>
          , “
          <article-title>Reverse engineering the neural networks for rule extraction in classification problems,”</article-title>
          <source>Neural Processing Letters</source>
          , vol.
          <volume>35</volume>
          , no.
          <issue>2</issue>
          , pp.
          <fpage>131</fpage>
          -
          <lpage>150</lpage>
          , Apr.
          <year>2012</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref16">
        <mixed-citation>
          [17]
          <string-name>
            <given-names>R.</given-names>
            <surname>Fong</surname>
          </string-name>
          and
          <string-name>
            <given-names>A.</given-names>
            <surname>Vedaldi</surname>
          </string-name>
          , “
          <article-title>Interpretable explanations of black boxes by meaningful perturbation,” CoRR</article-title>
          , vol.
          <source>abs/1704.03296</source>
          ,
          <year>2017</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref17">
        <mixed-citation>
          [18]
          <string-name>
            <given-names>M.</given-names>
            <surname>Sundararajan</surname>
          </string-name>
          ,
          <string-name>
            <given-names>A.</given-names>
            <surname>Taly</surname>
          </string-name>
          , and
          <string-name>
            <given-names>Q.</given-names>
            <surname>Yan</surname>
          </string-name>
          , “
          <article-title>Axiomatic attribution for deep networks</article-title>
          ,
          <source>” CoRR</source>
          , vol.
          <source>abs/1703.01365</source>
          ,
          <year>2017</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref18">
        <mixed-citation>
          [19]
          <string-name>
            <given-names>M. W.</given-names>
            <surname>Craven</surname>
          </string-name>
          and
          <string-name>
            <given-names>J. W.</given-names>
            <surname>Shavlik</surname>
          </string-name>
          , “
          <article-title>Extracting tree-structured representations of trained networks</article-title>
          ,
          <source>” in 8th International Conference on Neural Information Processing Systems (NIPS'95)</source>
          . MIT Press,
          <year>1995</year>
          , pp.
          <fpage>24</fpage>
          -
          <lpage>30</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref19">
        <mixed-citation>
          [20]
          <string-name>
            <given-names>R.</given-names>
            <surname>Andrews</surname>
          </string-name>
          ,
          <string-name>
            <given-names>J.</given-names>
            <surname>Diederich</surname>
          </string-name>
          ,
          <article-title>and</article-title>
          <string-name>
            <given-names>A. B.</given-names>
            <surname>Tickle</surname>
          </string-name>
          , “
          <article-title>Survey and critique of techniques for extracting rules from trained artificial neural networks,” Knowledge-Based Systems</article-title>
          , vol.
          <volume>8</volume>
          , no.
          <issue>6</issue>
          , pp.
          <fpage>373</fpage>
          -
          <lpage>389</lpage>
          , Dec.
          <year>1995</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref20">
        <mixed-citation>
          [21]
          <string-name>
            <given-names>U.</given-names>
            <surname>Johansson</surname>
          </string-name>
          and l. Niklasson, “
          <article-title>Evolving decision trees using oracle guides</article-title>
          ,” in
          <source>2009 IEEE Symposium on Computational Intelligence and Data Mining, Mar</source>
          .
          <year>2009</year>
          , pp.
          <fpage>238</fpage>
          -
          <lpage>244</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref21">
        <mixed-citation>
          [22]
          <string-name>
            <given-names>N.</given-names>
            <surname>Frosst</surname>
          </string-name>
          and
          <string-name>
            <given-names>G. E.</given-names>
            <surname>Hinton</surname>
          </string-name>
          , “
          <article-title>Distilling a neural network into a soft decision tree,” in CEX 2017 Comprehensibility and Explanation in AI and ML 2017 (CEX 2017), ser</article-title>
          .
          <source>CEUR Workshop Proceedings</source>
          , vol.
          <year>2071</year>
          , Nov.
          <year>2017</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref22">
        <mixed-citation>
          [23]
          <string-name>
            <given-names>D.</given-names>
            <surname>Silver</surname>
          </string-name>
          et al.,
          <article-title>“Mastering the game of Go with deep neural networks and tree search</article-title>
          ,
          <source>” Nature</source>
          , vol.
          <volume>529</volume>
          , pp.
          <fpage>484</fpage>
          -
          <lpage>489</lpage>
          , Jan.
          <year>2016</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref23">
        <mixed-citation>
          [24]
          <string-name>
            <given-names>A.</given-names>
            <surname>Omicini</surname>
          </string-name>
          and
          <string-name>
            <given-names>F.</given-names>
            <surname>Zambonelli</surname>
          </string-name>
          , “
          <article-title>MAS as complex systems: A view on the role of declarative approaches,” in Declarative Agent Languages and Technologies, ser</article-title>
          .
          <source>Lecture Notes in Computer Science</source>
          . Springer, May
          <year>2004</year>
          , vol.
          <volume>2990</volume>
          , pp.
          <fpage>1</fpage>
          -
          <lpage>17</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref24">
        <mixed-citation>
          [25]
          <string-name>
            <given-names>F.</given-names>
            <surname>Idelberger</surname>
          </string-name>
          , G. Governatori,
          <string-name>
            <given-names>R.</given-names>
            <surname>Riveret</surname>
          </string-name>
          , and G. Sartor, “
          <article-title>Evaluation of logic-based smart contracts for blockchain systems</article-title>
          ,” in Rule Technologies. Research, Tools, and Applications,
          <source>ser. Lecture Notes in Computer Science</source>
          , vol.
          <volume>9718</volume>
          . Springer,
          <year>2016</year>
          , pp.
          <fpage>167</fpage>
          -
          <lpage>183</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref25">
        <mixed-citation>
          [26]
          <string-name>
            <given-names>M.</given-names>
            <surname>Oliya and H. K. Pung</surname>
          </string-name>
          , “
          <article-title>Towards incremental reasoning for context aware systems,” in Advances in Computing and Communications, ser</article-title>
          .
          <source>Communications in Computer and Information Science</source>
          . Springer,
          <year>2011</year>
          , vol.
          <volume>190</volume>
          , pp.
          <fpage>232</fpage>
          -
          <lpage>241</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref26">
        <mixed-citation>
          [27]
          <string-name>
            <given-names>G.</given-names>
            <surname>Sotnik</surname>
          </string-name>
          , “
          <article-title>The SOSIEL platform: Knowledge-based, cognitive, and multi-agent,” Biologically Inspired Cognitive Architectures</article-title>
          , vol.
          <volume>26</volume>
          , pp.
          <fpage>103</fpage>
          -
          <lpage>117</lpage>
          , Oct.
          <year>2018</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref27">
        <mixed-citation>
          [28]
          <string-name>
            <given-names>R.</given-names>
            <surname>Kowalski</surname>
          </string-name>
          and
          <string-name>
            <given-names>F.</given-names>
            <surname>Sadri</surname>
          </string-name>
          , “
          <article-title>From logic programming towards multi-agent systems</article-title>
          ,
          <source>” Annals of Mathematics and Artificial Intelligence</source>
          , vol.
          <volume>25</volume>
          , no.
          <issue>3</issue>
          , pp.
          <fpage>391</fpage>
          -
          <lpage>419</lpage>
          , Nov.
          <year>1999</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref28">
        <mixed-citation>
          [29]
          <string-name>
            <surname>M. D. Pandya</surname>
            ,
            <given-names>P. D.</given-names>
          </string-name>
          <string-name>
            <surname>Shah</surname>
            , and
            <given-names>S.</given-names>
          </string-name>
          <string-name>
            <surname>Jardosh</surname>
          </string-name>
          , “
          <article-title>Medical image diagnosis for disease detection: A deep learning approach,” in U-Healthcare Monitoring Systems, ser</article-title>
          .
          <source>Advances in Ubiquitous Sensing Applications for Healthcare</source>
          . Academic Press,
          <year>2019</year>
          , vol.
          <volume>1</volume>
          : Design and Applications,
          <source>ch. 3</source>
          , pp.
          <fpage>37</fpage>
          -
          <lpage>60</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref29">
        <mixed-citation>
          [30]
          <string-name>
            <given-names>S.</given-names>
            <surname>Kuwayama</surname>
          </string-name>
          ,
          <string-name>
            <given-names>Y.</given-names>
            <surname>Ayatsuka</surname>
          </string-name>
          ,
          <string-name>
            <given-names>D.</given-names>
            <surname>Yanagisono</surname>
          </string-name>
          ,
          <string-name>
            <given-names>T.</given-names>
            <surname>Uta</surname>
          </string-name>
          ,
          <string-name>
            <given-names>H.</given-names>
            <surname>Usui</surname>
          </string-name>
          ,
          <string-name>
            <given-names>A.</given-names>
            <surname>Kato</surname>
          </string-name>
          ,
          <string-name>
            <given-names>N.</given-names>
            <surname>Takase</surname>
          </string-name>
          ,
          <string-name>
            <given-names>Y.</given-names>
            <surname>Ogura</surname>
          </string-name>
          , and T. Yasukawa, “
          <article-title>Automated detection of macular diseases by optical coherence tomography and artificial intelligence machine learning of optical coherence tomography images</article-title>
          ,
          <source>” Journal of Ophthalmology</source>
          , vol.
          <year>2019</year>
          , p.
          <fpage>7</fpage>
          ,
          <year>2019</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref30">
        <mixed-citation>
          [31]
          <string-name>
            <given-names>P.</given-names>
            <surname>Sajda</surname>
          </string-name>
          , “
          <article-title>Machine learning for detection and diagnosis of disease,” Annual Review of Biomedical Engineering</article-title>
          , vol.
          <volume>8</volume>
          , pp.
          <fpage>537</fpage>
          -
          <lpage>565</lpage>
          , Aug.
          <year>2006</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref31">
        <mixed-citation>
          [32]
          <string-name>
            <given-names>C.</given-names>
            <surname>Zhang</surname>
          </string-name>
          ,
          <string-name>
            <given-names>Y.</given-names>
            <surname>Chen</surname>
          </string-name>
          ,
          <string-name>
            <given-names>A.</given-names>
            <surname>Yin</surname>
          </string-name>
          , and
          <string-name>
            <given-names>X.</given-names>
            <surname>Wang</surname>
          </string-name>
          , “
          <article-title>Anomaly detection in ECG based on trend symbolic aggregate approximation</article-title>
          ,
          <source>” Mathematical Biosciences and Engineering</source>
          , vol.
          <volume>16</volume>
          , no.
          <issue>4</issue>
          , pp.
          <fpage>2154</fpage>
          -
          <lpage>2167</lpage>
          ,
          <year>2019</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref32">
        <mixed-citation>
          [33]
          <string-name>
            <given-names>A.</given-names>
            <surname>Rastogi</surname>
          </string-name>
          ,
          <string-name>
            <given-names>R.</given-names>
            <surname>Arora</surname>
          </string-name>
          , and
          <string-name>
            <given-names>S.</given-names>
            <surname>Sharma</surname>
          </string-name>
          , “
          <article-title>Leaf disease detection and grading using computer vision technology &amp; fuzzy logic</article-title>
          ,
          <source>” in 2nd International Conference on Signal Processing and Integrated Networks (SPIN</source>
          <year>2015</year>
          ). IEEE,
          <year>2015</year>
          , pp.
          <fpage>500</fpage>
          -
          <lpage>505</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref33">
        <mixed-citation>
          [34]
          <string-name>
            <given-names>A.</given-names>
            <surname>Lozowski</surname>
          </string-name>
          ,
          <string-name>
            <given-names>T. J.</given-names>
            <surname>Cholewo</surname>
          </string-name>
          , and
          <string-name>
            <given-names>J. M.</given-names>
            <surname>Zurada</surname>
          </string-name>
          , “
          <article-title>Crisp rule extraction from perceptron network classifiers,”</article-title>
          <source>in IEEE International Conference on Neural Networks (ICNN</source>
          <year>1996</year>
          ), vol. Plenary, Panel and Special Sessions, Jun.
          <year>1996</year>
          , pp.
          <fpage>94</fpage>
          -
          <lpage>99</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref34">
        <mixed-citation>
          [35]
          <string-name>
            <given-names>J.</given-names>
            <surname>Czerniak</surname>
          </string-name>
          and
          <string-name>
            <given-names>H.</given-names>
            <surname>Zarzycki</surname>
          </string-name>
          , “
          <article-title>Application of rough sets in the presumptive diagnosis of urinary system diseases,” in Artificial Intelligence and Security in Computing Systems, ser</article-title>
          . The Springer International Series in Engineering and Computer Science. Springer,
          <year>2003</year>
          , vol.
          <volume>752</volume>
          , pp.
          <fpage>41</fpage>
          -
          <lpage>51</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref35">
        <mixed-citation>
          [36]
          <string-name>
            <given-names>J. R.</given-names>
            <surname>Quinlan</surname>
          </string-name>
          ,
          <year>C4</year>
          .
          <article-title>5: Programs for Machine Learning</article-title>
          . San Francisco, CA, USA: Morgan Kaufmann Publishers Inc.,
          <year>1993</year>
          .
        </mixed-citation>
      </ref>
    </ref-list>
  </back>
</article>