<!DOCTYPE article PUBLIC "-//NLM//DTD JATS (Z39.96) Journal Archiving and Interchange DTD v1.0 20120330//EN" "JATS-archivearticle1.dtd">
<article xmlns:xlink="http://www.w3.org/1999/xlink">
  <front>
    <journal-meta />
    <article-meta>
      <title-group>
        <article-title>Argumentative Explanations from Causal Models</article-title>
      </title-group>
      <contrib-group>
        <contrib contrib-type="author">
          <string-name>Antonio Rago</string-name>
          <email>antonio@imperial.ac.uk</email>
          <xref ref-type="aff" rid="aff0">0</xref>
        </contrib>
        <contrib contrib-type="author">
          <string-name>Fabrizio Russo</string-name>
          <email>fabrizio@imperial.ac.uk</email>
          <xref ref-type="aff" rid="aff0">0</xref>
        </contrib>
        <contrib contrib-type="author">
          <string-name>Emanuele Albini</string-name>
          <email>emanuele@imperial.ac.uk</email>
          <xref ref-type="aff" rid="aff0">0</xref>
        </contrib>
        <contrib contrib-type="author">
          <string-name>Pietro Baroni</string-name>
          <email>pietro.baroni@unibs.it</email>
          <xref ref-type="aff" rid="aff1">1</xref>
        </contrib>
        <aff id="aff0">
          <label>0</label>
          <institution>Department of Computing, Imperial College London</institution>
          ,
          <country country="UK">UK</country>
        </aff>
        <aff id="aff1">
          <label>1</label>
          <institution>Dipartimento di Ingegneria dell'Informazione, Università degli Studi di Brescia</institution>
          ,
          <country country="IT">Italy</country>
        </aff>
      </contrib-group>
      <fpage>0000</fpage>
      <lpage>0001</lpage>
      <abstract>
        <p>We introduce a conceptualisation for generating argumentation frameworks (AFs) from causal models for the purpose of forging explanations for models' outputs. The conceptualisation is based on reinterpreting properties of semantics of AFs as explanation moulds, which are means for characterising argumentative relations. We demonstrate our methodology by reinterpreting the property of bi-variate reinforcement in bipolar AFs, showing how the extracted bipolar AFs may be used as relation-based explanations for the outputs of causal models.</p>
      </abstract>
      <kwd-group>
        <kwd>Explainable AI</kwd>
        <kwd>Argumentation frameworks</kwd>
        <kwd>Causal models</kwd>
      </kwd-group>
    </article-meta>
  </front>
  <body>
    <sec id="sec-1">
      <title>1. Introduction</title>
      <p>
        The field of explainable AI (XAI) has in recent years become a major focal point of the
eforts of researchers, with a wide variety of models for explanation being proposed (see
[
        <xref ref-type="bibr" rid="ref1">1</xref>
        ] for an overview). More recently, incorporating a causal perspective into explanations
has been explored by some, e.g. [
        <xref ref-type="bibr" rid="ref2 ref3 ref4">2, 3, 4</xref>
        ]. The link between causes and explanations has
long been studied [
        <xref ref-type="bibr" rid="ref5">5</xref>
        ]; indeed, the two have even been equated (under a broad sense of the
concept of “cause”) [
        <xref ref-type="bibr" rid="ref6">6</xref>
        ]. Causal reasoning is, in fact, how humans explain to one another
[
        <xref ref-type="bibr" rid="ref7">7</xref>
        ], and so mimicking such a trend lends credence to the hypothesis that machines should
do likewise. Further, research from the social sciences [
        <xref ref-type="bibr" rid="ref8">8</xref>
        ] has indicated the value of
causal links, particularly in the form of counterfactual reasoning, within explanations,
and that the importance of such information surpasses that of probabilities or statistical
relationships for users.
      </p>
      <p>
        Despite these findings, many of the approaches for generating explanations for AI
models have, nevertheless, neglected causality as a potential drive for explainability. Some
of the most popular methods are heuristic and model-agnostic [
        <xref ref-type="bibr" rid="ref10 ref9">9, 10</xref>
        ], and, although
LGOBE
      </p>
      <p>
        https://www.doc.ic.ac.uk/~afr114/ (A. Rago); https://briziorusso.github.io/ (F. Russo);
https://www.imperial.ac.uk/people/f.toni (F. Toni)
they are useful, particularly with regards to their wide-ranging applicability, they neglect
how models are determining their outputs and therefore the underlying causes therein.
This has arguably left a chasm between how explanations are provided by models at the
forefront of XAI technology and what users actually require from explanations [
        <xref ref-type="bibr" rid="ref11">11</xref>
        ].
      </p>
      <p>
        Meanwhile, computational argumentation (see [
        <xref ref-type="bibr" rid="ref12 ref13">12, 13</xref>
        ] for recent overviews) has received
increasing interest in recent years as a means for providing explanations of the outputs of
a number of AI models, e.g. recommender systems [
        <xref ref-type="bibr" rid="ref14">14</xref>
        ], classifiers [
        <xref ref-type="bibr" rid="ref15">15</xref>
        ], Bayesian networks
[
        <xref ref-type="bibr" rid="ref16">16</xref>
        ] and PageRank [
        <xref ref-type="bibr" rid="ref17">17</xref>
        ]. Argumentative explanations have also been advocated in the
social sciences [
        <xref ref-type="bibr" rid="ref18 ref8">18, 8</xref>
        ], and several works focus on the power of argumentation to provide
a bridge between explained models and users, validated by user studies [
        <xref ref-type="bibr" rid="ref19 ref20">19, 20</xref>
        ]. While
argumentative explanations are wide-ranging in their application (see [
        <xref ref-type="bibr" rid="ref21 ref22">21, 22</xref>
        ] for recent
surveys), the links between causal models and argumentative explanations have remained
largely unexplored to date.
      </p>
      <p>
        In this paper, we introduce a conceptualisation for generating argumentation
frameworks (AFs) with any number of dialectical relations as envisaged in [
        <xref ref-type="bibr" rid="ref23 ref24">23, 24</xref>
        ], from causal
models for the purpose of forging explanations for the models’ outputs. Like [
        <xref ref-type="bibr" rid="ref25">25</xref>
        ], we focus
not on explaining by features, but instead by relations, hence the use of argumentation as
the underpinning explanatory mechanism. After giving the necessary background (§2), we
show how properties of argumentation semantics from the literature can be reinterpreted
to serve as explanation moulds, i.e. means for characterising argumentative relations (§3).
In (§4) we propose a way to define explanation moulds based on inverting properties
of argumentation semantics. Briefly, the idea is to detect, inside a causal model, the
satisfaction of the conditions specified by some semantics property: if these conditions
are satisfied by some influence in the causal model, then the influence can be assigned an
explanatory role by casting it as a dialectical relation, whose type is in correspondence
with the detected property. The identified dialectical relations compose, altogether,
an argumentation framework. We demonstrate our methodology by reinterpreting the
property of bi-variate reinforcement [
        <xref ref-type="bibr" rid="ref26">26</xref>
        ] from bipolar AFs [
        <xref ref-type="bibr" rid="ref27">27</xref>
        ] and then showing in (§5)
how the extracted bipolar AFs may be used as counterfactual explanations for the outputs
of causal models representing diferent classification methods. Finally, we discuss related
work (§6) before concluding, indicating potentially fruitful future work (§7).
      </p>
      <p>Overall, we make the following main contributions:
• We propose a novel concept for defining relation-based explanations for causal
models by inverting properties of argumentation semantics.
• We use this concept to define a novel form of reinforcement explanation (RX) for
causal models.
• We show deployability of RXs with two machine-learning models, from which causal
models are drawn.</p>
    </sec>
    <sec id="sec-2">
      <title>2. Background</title>
      <p>Our method relies upon causal models and some notions from computational
argumentation. We provide core background for both.</p>
      <p>
        Causal models. A causal model [
        <xref ref-type="bibr" rid="ref28">28</xref>
        ] is a triple ⟨ ,  , ⟩
, where:
•  is a (finite) set of exogenous variables, i.e. variables whose values are determined
by external factors (outside the causal model);
•  is a (finite) set of endogenous variables, i.e. variables whose values are determined
by internal factors, namely by (the values of some of the) variables in  ∪  ;
• each variable may take any values in its associated domain; we refer to the domain
of   ∈  ∪ 
as  (
      </p>
      <p>);
•  is a (finite) set of structural equations that, for each endogenous variable   ∈  ,
define   ’s values as a function   
of the values of   ’ parents  (
 ) ⊆  ∪  ⧵ {
Example 1. Let us consider a simple causal model ⟨ ,  , ⟩
comprising  = {
 = { 1,  2} and for all   ∈  ∪ 
,  (</p>
      <p>) = {⊤, ⊥}. Figure 1i (we ignore Figure 1ii for the
moment: this will be discussed later in §4) visualises the variables’ parents, and Table 1
group chooses to enter the pizzeria.
gives the combinations of values for the variables resulting from the structural equations
 . This may represent a group’s decision on whether or not to enter a restaurant, with
variables  1: “margherita” is spelt correctly on the menu, not like the drink;  2: there is
pineapple on the pizzas;  1: the pizzeria seems to be legitimately Italian; and  2: the
 }.
1,  2},
the assignment to exogenous variables u ∈  such that   1[u] = ⊤ and   2[u] = ⊤.
example { 1,  2} = (</p>
      <p>1), i.e.  1 and  2 are the parents of  1). (ii) SAF explanation (see §3) for
Given a causal model ⟨ ,  , ⟩
where  = {
1, … ,   }, we denote with  =  (
1)×…× (

)
the a set of all possible combinations of values of the exogenous variables (realisations).
With an abuse of notation, we refer to the value of any variable   ∈  ∪ 
given u ∈ 
as    [u]: if   is an exogenous variable,    [u] will be its assigned value in u; if   is
 1
⊤ margherita
⊤ margherita
⊥ margarita
⊥ margarita</p>
      <p>2
⊤ pineapple
⊥ ∼pineapple
⊤ pineapple
⊥ ∼pineapple</p>
      <p>1
⊥ ∼Italian
⊤ Italian
⊥ ∼Italian
⊥ ∼Italian</p>
      <p>2
⊥ ∼enter
⊤ enter
⊥ ∼enter
⊥ ∼enter
an endogenous variable, it will be the value dictated by the structural equations in the
causal model.</p>
      <p>
        We use the do operator [
        <xref ref-type="bibr" rid="ref29">29</xref>
        ] to indicate interventions, i.e., for any variable   ∈  and
value thereof   ∈  (  ), ( =   ) implies that the function    is replaced by the constant
function   , and for any variable   ∈  and value thereof   ∈  (  ), (  =   ) implies
that   is assigned   .
      </p>
      <p>
        Argumentation. In general, an argumentation framework (AF) is any tuple ⟨ , ℛ 1, … , ℛ ⟩,
with  a set (of arguments),  &gt; 0 and ℛ ⊆  ×  , for  ∈ {1, … , } , (binary and directed)
dialectical relations between arguments [
        <xref ref-type="bibr" rid="ref23 ref24">23, 24</xref>
        ]. In the abstract argumentation [
        <xref ref-type="bibr" rid="ref30">30</xref>
        ]
tradition, arguments in these AFs are unspecified abstract entities that can be
instantiated diferently to suit diferent settings of deployment. Several specific choices of
dialectical relations can be made, giving rise to specific AFs instantiating the above
general definition, including abstract AFs (AAFs) [
        <xref ref-type="bibr" rid="ref30">30</xref>
        ], with  = 1 (and ℛ1 a dialectical
relation of attack, referred to later as ℛ−), support AFs (SAFs) [
        <xref ref-type="bibr" rid="ref31">31</xref>
        ], with  = 1 (and
ℛ1 a dialectical relation of support, referred to later as ℛ+), and bipolar AFs (BAFs)
[
        <xref ref-type="bibr" rid="ref27">27</xref>
        ], with  = 2 (and ℛ1 and ℛ2 dialectical relations of attack and support, respectively,
referred to later as ℛ− and ℛ+).
      </p>
      <p>
        The meaning of AFs (including the intended dialectical role of the relations) may be
given in terms of gradual semantics (e.g. see [
        <xref ref-type="bibr" rid="ref24 ref32">24, 32</xref>
        ] for BAFs), defined, for AFs with
arguments  , by means of mappings  ∶  →  , with  a given set of values of interest
for evaluating arguments.
      </p>
      <p>
        The choice of gradual semantics for AFs may be guided by properties that the mappings
 should satisfy (e.g. as in [
        <xref ref-type="bibr" rid="ref26 ref32">26, 32</xref>
        ]). We will utilise, in §4, a variant of the property of
bi-variate reinforcement for BAFs from [
        <xref ref-type="bibr" rid="ref26">26</xref>
        ].
      </p>
    </sec>
    <sec id="sec-3">
      <title>3. From Causal Models to Explanation Moulds and</title>
    </sec>
    <sec id="sec-4">
      <title>Argumentative Explanations</title>
      <p>In this section we see the task of obtaining explanations for causal models’ assignments
of values to variables as a two-step process: first we define moulds characterising the core
ingredients of explanations; then we use these moulds to obtain, automatically, (instances
of) AFs as argumentative explanations. Moulds and explanations are defined in terms
of influences between variables in the causal model, focusing on those from parents to
children given by the causal structure underpinning the model, as follows.</p>
      <p>be a causal model. The influence graph corresponding to
Definition 1. Let  = ⟨ ,  , ⟩
 is the pair ⟨ , ℐ ⟩ with:
•  =  ∪ 
• ℐ ⊆  × 
influences).</p>
      <p>is the set of all (exogenous and endogenous) variables;
is defined as
ℐ = {( 1,  2)| 1 ∈  (</p>
      <sec id="sec-4-1">
        <title>2)} (referred to as the set of</title>
        <p>
          Note that, while straightforward, the concept of influence graph (closely related to the
notion of causal diagram [
          <xref ref-type="bibr" rid="ref33">33</xref>
          ]) is useful as it underpins much of what follows.
        </p>
        <p>Next, the idea underlying explanation moulds is that, typically, inside the causal model,
some variables afect others in a way that may not be directly understandable or even
cognitively manageable by a user. The influence graph synthetically expresses which
variables afect which others but does not give an account of how the influences actually
occur in the context (namely, the values given to the exogenous variables) that a user
may be interested in. Thus, the perspective we take is that each influence can be assigned
an explanatory role, indicating how that influence is actually working in that context.
The explanatory roles ascribable to influences can be regarded as a form of explanatory
knowledge which is user specific: diferent users may be willing (and/or able) to accept
explanations built using diferent sets of explanatory roles as they correspond to their
understanding of how variables may afect each other. We assume that each explanatory
role is specified by a relation characterisation, i.e. a Boolean logical requirement, which
can be used to mould the explanations to be presented to the users by indicating which
relations play a role in the explanations.</p>
        <p>Definition 2. Given a causal model ⟨ ,  , ⟩
an explanation mould is a non-empty set:</p>
        <p>and its corresponding influence graph ⟨ , ℐ ⟩ ,
{ 1, … ,   }
where for all  ∈ {1, … , } ,   ∶  × ℐ → {⊤, ⊥} is a relation characterisation, in the form
of a Boolean condition expressed in some formal language. Given some u ∈  and
( 1,  2) ∈ ℐ, if   (u, ( 1,  2)) = ⊤ we say that the influence ( 1,  2) satisfies   for u.</p>
        <p>Note that we are not prescribing any formal language for specifying relation
characterisations, as several such languages may be suitable.</p>
        <p>Given an assignment u to the exogenous variables, based on an explanation mould, we
can obtain an AF including, as (diferent) dialectical relations, the influences satisfying
the (diferent) relation characterisations for the given u. Thus, the choice of relation
characterisations is to a large extent dictated by the specific form of AF the intended
users expect. Before defining argumentative explanations formally, we give an illustration.
Example 1 (Cont.). Let us imagine a situation where one would like to explain the
behaviour of the causal model from Figure 1i and Table 1 with a SAF (see §2). We thus
require one single form of relation (i.e. support) to be extracted from the corresponding
influence graph ⟨{ 1,  2,  1,  2}, {( 1,  1), ( 2,  1), ( 1,  2)}⟩. In order to define the
explanation mould for such a situation, we note that the behaviour defining this relation could
be characterised as changing the state of rejected arguments that it supports to accepted
when the supporting argument’s state is accepted. In our simple causal model, accepted
arguments may amount to variables assigned to value ⊤ and rejected arguments may
amount to variables assigned to value ⊥. Thus, the intended behaviour can be captured
by a relation characterisation   such that, given u ∈  and ( 1,  2) ∈ ℐ:
  (u, ( 1,  2)) = ⊤ if
(  1[u] = ⊤∧  2[u] = ⊤ ∧  2[u, (
(  1[u] = ⊥ ∧  2[u] = ⊥ ∧   2[u, (
1 = ⊥)] = ⊥)∨
1 = ⊤)] = ⊤).</p>
        <p>Then, for the assignment to exogenous variables u ∈  such that   1[u] = ⊤ and   2[u] = ⊥,
we may obtain the SAF in Figure 1ii (visualised as a graph with nodes as arguments and
edges indicating elements of the support relation). For illustration, consider ( 1,  1) ∈ ℐ
for this u. We can see from Table 1 that   1[u] = ⊤ and also that   1[u, ( 1 = ⊥)] = ⊥
and thus from the above it is clear that   (u, ( 1,  1)) = ⊤ and thus the influence is of the
type of support that   characterises. Meanwhile, consider ( 2,  1) ∈ ℐ for the same u: the
fact that   2[u] = ⊥ and   1[u] = ⊤ means that   (u, ( 2,  1)) = ⊥ and thus the influence is
not cast as a support. Indeed, if we consider the first and second rows of Table 1, we
can see that  2 being true actually causes  1 to be false, thus it is no surprise that the
influence is not cast as a support and plays no role in the resulting SAF. If we wanted for
this influence to play a role, we could, for example, choose to incorporate an additional
relation of attack into the explanation mould, to generate instead BAFs (see §2) as
argumentative explanations. This example thus shows how explanation moulds must be
designed to fit causal models depending on external explanatory requirements dictated
by users. It should be noted also that some explanation moulds may be unsuitable to
some causal models, e.g. the explanation mould with the earlier   would not be directly
applicable to causal models with variables with non-binary or continuous domains.</p>
        <p>In general, AFs serving as argumentative explanations can be generated as follows.
Definition 3. Given a causal model ⟨ ,  , ⟩ , its corresponding influence graph ⟨ , ℐ ⟩ ,
some u ∈  and an explanation mould { 1, … ,   }, an argumentative explanation is an AF
⟨ , ℛ 1, … ℛ ⟩, where
•  ⊆</p>
        <p>, and
• ℛ1, … , ℛ ⊆ ℐ ∩ ( ×  )
 )|  (u, ( 1,  2)) = ⊤}.</p>
        <p>such that, for any  = 1 …  , ℛ = {( 1,  2) ∈ ℐ ∩ ( ×</p>
        <p>Note that we have left open the choice of  (as a generic, possibly non-strict subset of
 ). In practice,  may be the full  , but we envisage that users may prefer to restrict
attention to some variables of interest (for example, excluding variables not “involved” in
any influence satisfying the relation characterisations).</p>
        <p>Example 1 (Cont.). The behaviour of the causal model from Figure 1i and Table 1 for u
such that   1[u] = ⊤ and   2[u] = ⊤, using the explanation mould {  } given earlier, can
be captured by either of the two SAFs (argumentative explanations) below, depending
on the choice of  :
• the SAF in Figure 1ii, where every variable is an argument;
• the SAF with the same support relation but  2 excluded from  , as not “involved”
and thus not contributing to the explanation.</p>
        <p>Both SAFs explain that   1[u] = ⊤ is supported by   1[u] = ⊤, in turn supporting
  2[u] = ⊤ . Namely, the causal model recommends that the group should enter the
pizzeria because the pizzeria seems legitimately Italian, given that “margherita” is spelt
correctly on the menu. Note that the pineapple not being on the pizza could also be seen
as a support towards the pizzeria being legitimately Italian, the inclusion of which could
be achieved with a slightly more complex explanation mould.</p>
      </sec>
    </sec>
    <sec id="sec-5">
      <title>4. Inverting Properties of Argumentation Semantics:</title>
    </sec>
    <sec id="sec-6">
      <title>Reinforcement Explanations</title>
      <p>The choice (number and form) of relation characterisations in explanation moulds is
crucial for the generation of explanations concerning the value assignments to
endogenous variables in the causal models. Even after having decided which argumentative
relations to include in the AF/argumentative explanation, the definition of the relation
characterisations is non-trivial, in general. In this section we demonstrate a novel
concept for utilising properties of gradual semantics for AFs for the definition of relation
characterisations and the consequent extraction of argumentative explanations.</p>
      <p>
        The common usage of these properties in computational argumentation can be roughly
equated to: if a semantics, given an AF, satisfies some desirable properties, then the
semantics is itself desirable (for the intended context, where those properties matter).
We propose a form of inversion of this notion for use in our XAI setting, namely: if some
desirable properties are identified for the gradual semantics of (still unspecified) AFs,
then these properties can guide the definition of the dialectical relations underpinning
the AFs. For this inversion to work, we need to identify first and foremost a suitable
notion of gradual semantics for the AFs we extract from causal models. Given that, with
our AFs, we are trying to explain the results obtained from underlying causal models, we
cannot impose just any gradual semantics from the literature, but need to make sure that
we capture, with the chosen semantics, the behaviour of the causal model itself. This is
similar, in spirit, to recent work to extract (weighted) BAFs from multi-layer perceptrons
(MLPs) [
        <xref ref-type="bibr" rid="ref34">34</xref>
        ], using the underlying computation of the MLPs as a gradual semantics, and
to the proposals to explain recommender systems (RSs) via tripolar AFs [
        <xref ref-type="bibr" rid="ref35">35</xref>
        ] or BAFs
[
        <xref ref-type="bibr" rid="ref20">20</xref>
        ], using the underlying predicted ratings by the RSs as a gradual semantics.
      </p>
      <p>A natural semantic choice for causal models, since we are trying to explain why
endogenous variables are assigned specific values in their domains given assignments to
the exogenous variables, is to use the assignments themselves as a gradual semantics.
Then, the idea of inverting properties of semantics to obtain dialectical relations in AFs
can be recast to obtain relation characterisations in explanation moulds as follows: given
an influence graph and a selected value assignment to exogenous variables, if an influence
satisfies a given, desirable property, then the influence can be cast as part of a dialectical
relation in the resulting AF.</p>
      <p>Naturally, for this inversion to be useful, we need to identify useful properties from
an explanatory viewpoint.</p>
      <p>
        We will illustrate this concept with the property of
bivariate reinforcement for BAFs [
        <xref ref-type="bibr" rid="ref26">26</xref>
        ], which we posit is generally intuitive in the realm of
explanations. Bi-variate reinforcement is defined when the set of values  for evaluating
arguments is equipped with a pre-order &lt;. Intuitively, bi-variate reinforcement states that1
strengthening an attacker (a supporter) cannot strengthen (cannot weaken, respectively)
an argument it attacks (supports, respectively), where strengthening an argument amounts
to increasing its value from  1 ∈  to  2 ∈  such that  2 &gt;  1 (whereas weakening an
argument amounts to decreasing its value from such  2 to  1). In our formulation of
this property, we require that increasing the value of variables represented as attackers
(supporters) can only decrease (increase, respectively) the values of variables they attack
(support, respectively).
      </p>
      <p>Property 1. Given a causal model ⟨ ,  , ⟩
such that, for each   ∈  ∪ , the domain  (

)
is equipped with a pre-order &lt;,2 and given its corresponding influence graph
⟨ , ℐ ⟩ , an
argumentative explanation ⟨ , ℛ</p>
      <p>−, ℛ+⟩ for u ∈ 
 + ∈  (</p>
      <p>1) such that  + &gt;  1:
any ( 1,  2) ∈ ℐ where  1 =   1[u], for any  − ∈  (
satisfies causal reinforcement if for
1) such that  − &lt;  1, and for any
• if ( 1,  2) ∈ ℛ−, then   2[u, (
1 =  +)] ≤   2[u] and   2[u, (
1 =  −)] ≥
• if ( 1,  2) ∈ ℛ+, then   2[u, (
1 =  +)] ≥   2[u] and   2[u, (
1 =  −)] ≤
  2[u];
  2[u].</p>
      <p>
        We can then invert this property to obtain an explanation mould. In doing so, we
introduce slightly stricter conditions to ensure that influencing variables that have no efect
on influenced variables do not constitute both an attack and a support, a phenomenon
which we believe would be counter-intuitive from an explanation viewpoint.
1Here, we ignore the intrinsic basic strength of arguments used in the formal definition in [
        <xref ref-type="bibr" rid="ref26">26</xref>
        ].
2With an abuse of notation we use the same symbol for all pre-orders.
      </p>
      <p>Definition 4. Given a causal model ⟨ ,  , ⟩ such that, for each   ∈  ∪  , the domain
 (  ) is equipped with a pre-order &lt;, and given its corresponding influence graph ⟨ , ℐ ⟩ ,
a reinforcement explanation mould is an explanation mould { −,  +} such that, given some
u ∈  and ( 1,  2) ∈ ℐ, letting  1 =   1[u]:
•  −(u, ( 1,  2)) = ⊤ if:
•  +(u, ( 1,  2)) = ⊤ if:
1. ∀ + ∈  (
2. ∀ − ∈  (
1) such that  + &gt;  1, it holds that   2[u, (
1) such that  − &lt;  1, it holds that   2[u, (
1 =  +)] ≤   2[u];
1 =  −)] ≥   2[u];
3. ∃≥1 + ∈  ( 1) or ∃≥1 − ∈  (
in points 1 and 2 above.</p>
      <p>1) satisfying strictly the inequality conditions
1. ∀ + ∈  (
2. ∀ − ∈  (
1) such that  + &gt;  1, it holds that   2[u, (
1) such that  − &lt;  1, it holds that   2[u, (
1 =  +)] ≥   2[u];
1 =  −)] ≤   2[u];
3. ∃≥1 + ∈  ( 1) or ∃≥1 − ∈  (
in points 1 and 2 above.</p>
      <sec id="sec-6-1">
        <title>1) satisfying strictly the inequality conditions</title>
        <p>We call any argumentative explanation resulting from the explanation mould { −,  +} a
reinforcement explanation (RX).</p>
        <p>Note that, as for generic argumentative explanations, we do not commit in general to
any choice of  in RXs.</p>
        <p>Proposition 1. Any RX satisfies causal reinforcement.</p>
        <p>Proof. Follows directly from the definition of Property 1 and Definition 4.</p>
        <p>The satisfaction of the property of causal reinforcement indicates how RXs could
be used counterfactually, given that the results of changes to the variables’ values on
influenced variables are guaranteed. For example, if a user is looking to increase an
influenced variable’s value, supporters (attackers) indicate variables whose values should
be increased (decreased, respectively). In the following sections, we will explore the
potential of this capability when causal models provide abstractions of classifiers whose
output needs explaining.</p>
      </sec>
    </sec>
    <sec id="sec-7">
      <title>5. Reinforcement Explanations for Classification</title>
      <p>In this section, we instantiate causal models for two diferent AI models commonly used
for classification in the literature and preliminarily discuss a potential use of RXs in this
context.</p>
      <p>The two classification models that we use to instantiate causal models are Bayesian
network classifiers (BCs) and classifiers built from feed-forward neural networks (NNs).
Given some assignments to input variables I (from the variables’ domains), these classifiers
can be seen as determining the most likely value for classification variables, which, in
this paper, we assume to be binary, in a given set C. Thus, the classification task may
be seen as a mapping ℳ(x) returning, for assignment x to input variables, either 1 or
0 (for the classification variables in C) depending on whether the probability exceeds a
given threshold  . We summarise the classification process in Figure 2. Note that the
choice of threshold is crucial to guarantee that a single value   is determined by ℳ for
each classification variable   : if  is too high, then no value may be computed, whereas
if  is too low, the probability of both values may exceed it. Note also that, in the case
of NNs, the probabilities may result from using, e.g., a softmax activation for the output
layer. Furthermore, note that for the purposes of this paper, the underpinning details of
these classifiers and how they can be obtained are irrelevant and will be ignored. In other
words, we treat the classifier as a black-box, as standard in much of the XAI literature,
and explain its outputs in terms of its inputs.</p>
      <p>We represent the classification task by a (naive) BC or by a NN with the following
causal model:
Definition 5. A causal model for a naive BC or classifier built from a NN is a causal
model ⟨  ,   ,   ⟩, where:
•   consists of the input variables I of the classifier, with their respective domains;
•   = C such that, for each   ∈ C,  (  ) = { 1,  0};
•   corresponds to the computation of the probability values  (  =  1)) by the
classifier (see Figure 2).</p>
      <p>ℐ =   ×   represents the influences in the causal model for the classifier; these are
such that the exogenous variables   are densely connected to the endogenous variables
  .</p>
      <p>In line with our assumptions for RXs, we assume that the variables’ domains are
equipped with a pre-order.</p>
      <p>On this basis, we envisage the following use of RXs as actionable explanations in
contexts where the user has a classification goal to reach and has control on (some of)
the input variables.</p>
      <p>Given a situation with an undesired classification outcome (e.g. a rejected loan
application) and an explanation indicating the relevant attackers and supporters, if a
user would like to decrease the probability of the current classification, s/he would look
to increase (decrease) the value of the corresponding variable’s attackers (supporters,
respectively), in line with Property 1.</p>
      <p>A broader investigation of the possible uses of RXs is left to future work.</p>
    </sec>
    <sec id="sec-8">
      <title>6. Related Work</title>
      <p>
        The role of causality within explanations for AI models has received increasing attention
of late. [
        <xref ref-type="bibr" rid="ref2">2</xref>
        ] define a framework for determining the causal efects between features
and predictions using a variational autoencoder. The detection of causal relations and
explanations between arguments within text has also proven efective within NLP [
        <xref ref-type="bibr" rid="ref36">36</xref>
        ].
[
        <xref ref-type="bibr" rid="ref3">3</xref>
        ] give causal explanations for NNs in that they train a separate NN by masking features
to determine causal relations (in the original NN) from the features to the classifications.
Generative causal explanations of black box classifiers [
        <xref ref-type="bibr" rid="ref37">37</xref>
        ] are built by learning the latent
factors involved in a classification, which are then included in a causal model. [
        <xref ref-type="bibr" rid="ref38">38</xref>
        ] take
a diferent approach, proposing a general framework for constructing structural causal
models with deep learning components, allowing tractable counterfactual inference. Other
approaches towards explaining NNs, e.g., [
        <xref ref-type="bibr" rid="ref39 ref40">39, 40</xref>
        ], take into account causal relations when
calculating features’ attribution values for explanation. Meanwhile, [
        <xref ref-type="bibr" rid="ref4">4</xref>
        ] introduce causal
explanations for reinforcement learning models based on [
        <xref ref-type="bibr" rid="ref5">5</xref>
        ].
      </p>
      <p>
        Computational argumentation has been widely used in the literature as a mechanism for
explaining AI models, from data-driven explanations of classifiers’ outputs [
        <xref ref-type="bibr" rid="ref41">41</xref>
        ], powered
by AA-CBR [
        <xref ref-type="bibr" rid="ref42">42</xref>
        ], to the explanation of the PageRank algorithm [
        <xref ref-type="bibr" rid="ref43">43</xref>
        ] via bipolar AFs
[
        <xref ref-type="bibr" rid="ref17">17</xref>
        ]. The outputs of Bayesian networks have been explained by SAFs [
        <xref ref-type="bibr" rid="ref16">16</xref>
        ], while
decisionmaking [44] and scheduling [45] have also been targeted. Property-driven explanations
based on bipolar [
        <xref ref-type="bibr" rid="ref20">20</xref>
        ] and tripolar [
        <xref ref-type="bibr" rid="ref35">35</xref>
        ] AFs have been extracted for recommendations,
where the properties driving the extraction are defined in the orthodox manner (with
respect to the resulting frameworks), rather than inversions thereof, as we propose.
Other forms of argumentation have also proven efective in providing explanations for
recommender systems [
        <xref ref-type="bibr" rid="ref14">14</xref>
        ], decision making [46] and planning [47].
      </p>
      <p>Various works have explored the links between causality and argumentation. [48] shows
that a propositional argumentation system in a full classical language is equivalent to a
causal reasoning system, while [49] develops a formal theory combining “causal stories”
and evidential arguments. Somewhat similarly to us, [50] present a method for extracting
argumentative explanations for the outputs of causal models. However, their method
requires more information than the causal model alone, namely, ontological links, and
the argumentation supplements the rule-based explanations, rather than being the main
constituent, as is the case in our approach.</p>
    </sec>
    <sec id="sec-9">
      <title>7. Conclusions</title>
      <p>We have introduced a novel approach for extracting AFs from causal models in order
to explain the latter’s outputs. We have shown how explanation moulds can be defined
for particular explanatory requirements in order to generate argumentative explanations.
We focused, in particular, on inverting the existing property of argumentation semantics
of bi-variate reinforcement to create an explanation mould, before demonstrating how
the resulting reinforcement explanations (RXs) can be used to explain causal models
representing diferent machine-learning-based classifiers.</p>
      <p>One of the most promising aspects of this preliminary work is the vast array of directions
for future investigation it suggests. First, an experimental validation of the proposed
ideas and a comparison with other explanation approaches on a suficiently large variety
of case studies is needed.</p>
      <p>Clearly, the wide-ranging applicability of causal models broadens the scope of
explanation moulds and argumentative explanations well beyond machine learning models, and
we plan to undertake an investigation into other contexts in which they may be useful,
for example for decision support in healthcare.</p>
      <p>
        We also plan to study inversions of diferent properties of argumentation semantics and
diferent forms of AFs to understand their potential, e.g. counting for AAFs [ 51]. Within
the context of explaining machine learning models, we plan to assess RXs’ suitability
for diferent data structures and diferent classifiers, considering in particular deeper
explanations, e.g. including influences amongst input variables and/or intermediate, in
addition to input and output, variables, in the spirit of [
        <xref ref-type="bibr" rid="ref25">52, 25</xref>
        ]. This may be aided by
the deployment of methods for the extraction of more sophisticated causal models from
classifiers, e.g., [ 53] for NNs.
      </p>
      <p>Finally, while we posit that, when properly defined, the meaning and explanatory role
of the dialectical relations can be rather intuitive at a general level, providing efective
explanations to users through AFs will require the investigation of proper presentation
and visualization methods, possibly tailored to users’ competences and goals and to
diferent application domains.</p>
    </sec>
    <sec id="sec-10">
      <title>Acknowledgments</title>
      <p>Toni was partially funded by the European Research Council (ERC) under the
European Union’s Horizon 2020 research and innovation programme (grant agreement No.
101020934). Further, Russo was supported by UK Research and Innovation [grant number
EP/S023356/1], in the UKRI Centre for Doctoral Training in Safe and Trusted Artificial
Intelligence (www.safeandtrustedai.org). Finally, Rago and Toni were partially funded by
J.P. Morgan and by the Royal Academy of Engineering under the Research Chairs and
Senior Research Fellowships scheme. Any views or opinions expressed herein are solely
those of the authors listed, and may difer, in particular, from the views and opinions
expressed by J.P. Morgan or its afiliates. This material is not a product of the Research
Department of J.P. Morgan Securities LLC. This material should not be construed as an
individual recommendation for any particular client and is not intended as a
recommendation of particular securities, financial instruments or strategies for a particular client.
This material does not constitute a solicitation or ofer in any jurisdiction.</p>
      <p>Bringing Order to the Web, WWW: Internet and Web Inf. Syst. 54 (1998) 1–17.
[44] L. Amgoud, H. Prade, Using arguments for making and explaining decisions,</p>
      <p>Artificial Intelligence 173 (2009) 413–436.
[45] K. Cyras, D. Letsios, R. Misener, F. Toni, Argumentation for explainable scheduling,
in: Proc. AAAI, 2019, pp. 2752–2759.
[46] Q. Zhong, X. Fan, X. Luo, F. Toni, An explainable multi-attribute decision model
based on argumentation, Exp. Syst. Appl. 117 (2019) 42–61.
[47] N. Oren, K. van Deemter, W. W. Vasconcelos, Argument-based plan explanation,
in: Knowledge Engineering Tools and Techniques for AI Planning, Springer, 2020,
pp. 173–188.
[48] A. Bochman, Propositional argumentation and causal reasoning, in: Proc. IJCAI,
2005, pp. 388–393.
[49] F. Bex, An integrated theory of causal stories and evidential arguments, in: Proc.</p>
      <p>ICAIL, 2015, pp. 13–22.
[50] P. Besnard, M. Cordier, Y. Moinard, Arguments using ontological and causal
knowledge, in: Proc. FoIKS, 2014, pp. 79–96.
[51] L. Amgoud, J. Ben-Naim, Axiomatic foundations of acceptability semantics, in:</p>
      <p>Proc. KR, 2016, pp. 2–11.
[52] C. Olah, A. Satyanarayan, I. Johnson, S. Carter, L. Schubert, K. Ye, A. Mordvintsev,</p>
      <p>The building blocks of interpretability, Distill 3 (2018) e10.
[53] T. Kyono, Y. Zhang, M. van der Schaar, CASTLE: regularization via auxiliary
causal graph discovery, in: Proc. NeurIPS, 2020.</p>
    </sec>
  </body>
  <back>
    <ref-list>
      <ref id="ref1">
        <mixed-citation>
          [1]
          <string-name>
            <given-names>R.</given-names>
            <surname>Guidotti</surname>
          </string-name>
          ,
          <string-name>
            <given-names>A.</given-names>
            <surname>Monreale</surname>
          </string-name>
          ,
          <string-name>
            <given-names>S.</given-names>
            <surname>Ruggieri</surname>
          </string-name>
          ,
          <string-name>
            <given-names>F.</given-names>
            <surname>Turini</surname>
          </string-name>
          ,
          <string-name>
            <given-names>F.</given-names>
            <surname>Giannotti</surname>
          </string-name>
          ,
          <string-name>
            <given-names>D.</given-names>
            <surname>Pedreschi</surname>
          </string-name>
          ,
          <article-title>A survey of methods for explaining black box models</article-title>
          ,
          <source>ACM Computing Surveys</source>
          <volume>51</volume>
          (
          <year>2019</year>
          )
          <volume>93</volume>
          :
          <fpage>1</fpage>
          -
          <lpage>93</lpage>
          :
          <fpage>42</fpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref2">
        <mixed-citation>
          [2]
          <string-name>
            <given-names>D.</given-names>
            <surname>Alvarez-Melis</surname>
          </string-name>
          ,
          <string-name>
            <given-names>T. S.</given-names>
            <surname>Jaakkola</surname>
          </string-name>
          ,
          <article-title>A causal framework for explaining the predictions of black-box sequence-to-sequence models</article-title>
          ,
          <source>in: Proc. EMNLP</source>
          ,
          <year>2017</year>
          , pp.
          <fpage>412</fpage>
          -
          <lpage>421</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref3">
        <mixed-citation>
          [3]
          <string-name>
            <given-names>P.</given-names>
            <surname>Schwab</surname>
          </string-name>
          , W. Karlen,
          <article-title>CXPlain: Causal explanations for model interpretation under uncertainty</article-title>
          ,
          <source>in: Proc. NeurIPS</source>
          ,
          <year>2019</year>
          , pp.
          <fpage>10220</fpage>
          -
          <lpage>10230</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref4">
        <mixed-citation>
          [4]
          <string-name>
            <given-names>P.</given-names>
            <surname>Madumal</surname>
          </string-name>
          ,
          <string-name>
            <given-names>T.</given-names>
            <surname>Miller</surname>
          </string-name>
          ,
          <string-name>
            <given-names>L.</given-names>
            <surname>Sonenberg</surname>
          </string-name>
          ,
          <string-name>
            <given-names>F.</given-names>
            <surname>Vetere</surname>
          </string-name>
          ,
          <article-title>Explainable reinforcement learning through a causal lens</article-title>
          ,
          <source>in: Proc. AAAI</source>
          ,
          <year>2020</year>
          , pp.
          <fpage>2493</fpage>
          -
          <lpage>2500</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref5">
        <mixed-citation>
          [5]
          <string-name>
            <given-names>J. Y.</given-names>
            <surname>Halpern</surname>
          </string-name>
          ,
          <string-name>
            <given-names>J.</given-names>
            <surname>Pearl</surname>
          </string-name>
          ,
          <article-title>Causes and explanations: A structural-model approach: Part 1: Causes</article-title>
          , in: UAI,
          <year>2001</year>
          , pp.
          <fpage>194</fpage>
          -
          <lpage>202</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref6">
        <mixed-citation>
          [6]
          <string-name>
            <given-names>J.</given-names>
            <surname>Woodward</surname>
          </string-name>
          , Explanation, invariance, and intervention,
          <source>Philosophy of Science</source>
          <volume>64</volume>
          (
          <year>1997</year>
          )
          <fpage>S26</fpage>
          -
          <lpage>S41</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref7">
        <mixed-citation>
          [7]
          <string-name>
            <surname>M. M. A. de Graaf</surname>
            ,
            <given-names>B. F.</given-names>
          </string-name>
          <string-name>
            <surname>Malle</surname>
          </string-name>
          ,
          <article-title>How people explain action (and autonomous intelligent systems should too)</article-title>
          ,
          <source>in: Proc. AAAI Fall Symposia</source>
          ,
          <year>2017</year>
          , pp.
          <fpage>19</fpage>
          -
          <lpage>26</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref8">
        <mixed-citation>
          [8]
          <string-name>
            <given-names>T.</given-names>
            <surname>Miller</surname>
          </string-name>
          ,
          <article-title>Explanation in artificial intelligence: Insights from the social sciences</article-title>
          ,
          <source>Artificial Intelligence</source>
          <volume>267</volume>
          (
          <year>2019</year>
          )
          <fpage>1</fpage>
          -
          <lpage>38</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref9">
        <mixed-citation>
          [9]
          <string-name>
            <given-names>M. T.</given-names>
            <surname>Ribeiro</surname>
          </string-name>
          ,
          <string-name>
            <given-names>S.</given-names>
            <surname>Singh</surname>
          </string-name>
          ,
          <string-name>
            <given-names>C.</given-names>
            <surname>Guestrin</surname>
          </string-name>
          , ”
          <article-title>why should I trust you?”: Explaining the predictions of any classifier</article-title>
          ,
          <source>in: Proc. ACM SIGKDD</source>
          ,
          <year>2016</year>
          , pp.
          <fpage>1135</fpage>
          -
          <lpage>1144</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref10">
        <mixed-citation>
          [10]
          <string-name>
            <given-names>S. M.</given-names>
            <surname>Lundberg</surname>
          </string-name>
          ,
          <string-name>
            <given-names>S.</given-names>
            <surname>Lee</surname>
          </string-name>
          ,
          <article-title>A unified approach to interpreting model predictions</article-title>
          ,
          <source>in: Proc. NeurIPS</source>
          ,
          <year>2017</year>
          , pp.
          <fpage>4765</fpage>
          -
          <lpage>4774</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref11">
        <mixed-citation>
          [11]
          <string-name>
            <given-names>A.</given-names>
            <surname>Ignatiev</surname>
          </string-name>
          ,
          <article-title>Towards trustable explainable AI</article-title>
          ,
          <source>in: Proc. IJCAI</source>
          ,
          <year>2020</year>
          , pp.
          <fpage>5154</fpage>
          -
          <lpage>5158</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref12">
        <mixed-citation>
          [12]
          <string-name>
            <given-names>K.</given-names>
            <surname>Atkinson</surname>
          </string-name>
          ,
          <string-name>
            <given-names>P.</given-names>
            <surname>Baroni</surname>
          </string-name>
          ,
          <string-name>
            <given-names>M.</given-names>
            <surname>Giacomin</surname>
          </string-name>
          ,
          <string-name>
            <given-names>A.</given-names>
            <surname>Hunter</surname>
          </string-name>
          ,
          <string-name>
            <given-names>H.</given-names>
            <surname>Prakken</surname>
          </string-name>
          ,
          <string-name>
            <given-names>C.</given-names>
            <surname>Reed</surname>
          </string-name>
          ,
          <string-name>
            <given-names>G. R.</given-names>
            <surname>Simari</surname>
          </string-name>
          ,
          <string-name>
            <given-names>M.</given-names>
            <surname>Thimm</surname>
          </string-name>
          ,
          <string-name>
            <given-names>S.</given-names>
            <surname>Villata</surname>
          </string-name>
          , Towards artificial argumentation,
          <source>AI</source>
          Magazine
          <volume>38</volume>
          (
          <year>2017</year>
          )
          <fpage>25</fpage>
          -
          <lpage>36</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref13">
        <mixed-citation>
          [13]
          <string-name>
            <given-names>P.</given-names>
            <surname>Baroni</surname>
          </string-name>
          ,
          <string-name>
            <given-names>D.</given-names>
            <surname>Gabbay</surname>
          </string-name>
          ,
          <string-name>
            <given-names>M.</given-names>
            <surname>Giacomin</surname>
          </string-name>
          , L. van der Torre (Eds.), Handbook of Formal Argumentation, College Publications,
          <year>2018</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref14">
        <mixed-citation>
          [14]
          <string-name>
            <given-names>J. C.</given-names>
            <surname>Teze</surname>
          </string-name>
          ,
          <string-name>
            <given-names>L.</given-names>
            <surname>Godo</surname>
          </string-name>
          ,
          <string-name>
            <given-names>G. R.</given-names>
            <surname>Simari</surname>
          </string-name>
          ,
          <article-title>An argumentative recommendation approach based on contextual aspects</article-title>
          ,
          <source>in: Proc. SUM</source>
          ,
          <year>2018</year>
          , pp.
          <fpage>405</fpage>
          -
          <lpage>412</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref15">
        <mixed-citation>
          [15]
          <string-name>
            <given-names>A.</given-names>
            <surname>Dejl</surname>
          </string-name>
          ,
          <string-name>
            <given-names>P.</given-names>
            <surname>He</surname>
          </string-name>
          ,
          <string-name>
            <given-names>P.</given-names>
            <surname>Mangal</surname>
          </string-name>
          ,
          <string-name>
            <given-names>H.</given-names>
            <surname>Mohsin</surname>
          </string-name>
          ,
          <string-name>
            <given-names>B.</given-names>
            <surname>Surdu</surname>
          </string-name>
          ,
          <string-name>
            <given-names>E.</given-names>
            <surname>Voinea</surname>
          </string-name>
          , E. Albini,
          <string-name>
            <given-names>P.</given-names>
            <surname>Lertvittayakumjorn</surname>
          </string-name>
          ,
          <string-name>
            <given-names>A.</given-names>
            <surname>Rago</surname>
          </string-name>
          ,
          <string-name>
            <given-names>F.</given-names>
            <surname>Toni</surname>
          </string-name>
          ,
          <article-title>Argflow: A toolkit for deep argumentative explanations for neural networks</article-title>
          ,
          <source>in: Proc. AAMAS</source>
          ,
          <year>2021</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref16">
        <mixed-citation>
          [16]
          <string-name>
            <given-names>S. T.</given-names>
            <surname>Timmer</surname>
          </string-name>
          , J. C. Meyer, H. Prakken,
          <string-name>
            <given-names>S.</given-names>
            <surname>Renooij</surname>
          </string-name>
          ,
          <string-name>
            <given-names>B.</given-names>
            <surname>Verheij</surname>
          </string-name>
          ,
          <article-title>Explaining Bayesian networks using argumentation</article-title>
          ,
          <source>in: Proc. ECSQARU</source>
          ,
          <year>2015</year>
          , pp.
          <fpage>83</fpage>
          -
          <lpage>92</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref17">
        <mixed-citation>
          [17]
          <string-name>
            <given-names>E.</given-names>
            <surname>Albini</surname>
          </string-name>
          ,
          <string-name>
            <given-names>P.</given-names>
            <surname>Baroni</surname>
          </string-name>
          ,
          <string-name>
            <given-names>A.</given-names>
            <surname>Rago</surname>
          </string-name>
          ,
          <string-name>
            <given-names>F.</given-names>
            <surname>Toni</surname>
          </string-name>
          ,
          <article-title>PageRank as an argumentation semantics</article-title>
          ,
          <source>in: Proc. COMMA</source>
          ,
          <year>2020</year>
          , pp.
          <fpage>55</fpage>
          -
          <lpage>66</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref18">
        <mixed-citation>
          [18]
          <string-name>
            <given-names>C.</given-names>
            <surname>Antaki</surname>
          </string-name>
          ,
          <string-name>
            <surname>I. Leudar</surname>
          </string-name>
          ,
          <article-title>Explaining in conversation: Towards an argument model</article-title>
          ,
          <source>Europ. J. of Social Psychology</source>
          <volume>22</volume>
          (
          <year>1992</year>
          )
          <fpage>181</fpage>
          -
          <lpage>194</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref19">
        <mixed-citation>
          [19]
          <string-name>
            <given-names>P.</given-names>
            <surname>Madumal</surname>
          </string-name>
          ,
          <string-name>
            <given-names>T.</given-names>
            <surname>Miller</surname>
          </string-name>
          ,
          <string-name>
            <given-names>L.</given-names>
            <surname>Sonenberg</surname>
          </string-name>
          ,
          <string-name>
            <given-names>F.</given-names>
            <surname>Vetere</surname>
          </string-name>
          ,
          <article-title>A grounded interaction protocol for explainable artificial intelligence</article-title>
          ,
          <source>in: Proc. AAMAS</source>
          ,
          <year>2019</year>
          , pp.
          <fpage>1033</fpage>
          -
          <lpage>1041</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref20">
        <mixed-citation>
          [20]
          <string-name>
            <given-names>A.</given-names>
            <surname>Rago</surname>
          </string-name>
          ,
          <string-name>
            <given-names>O.</given-names>
            <surname>Cocarascu</surname>
          </string-name>
          ,
          <string-name>
            <given-names>C.</given-names>
            <surname>Bechlivanidis</surname>
          </string-name>
          ,
          <string-name>
            <given-names>F.</given-names>
            <surname>Toni</surname>
          </string-name>
          ,
          <article-title>Argumentation as a framework for interactive explanations for recommendations</article-title>
          ,
          <source>in: Proc. KR</source>
          ,
          <year>2020</year>
          , pp.
          <fpage>805</fpage>
          -
          <lpage>815</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref21">
        <mixed-citation>
          [21]
          <string-name>
            <given-names>K.</given-names>
            <surname>Cyras</surname>
          </string-name>
          ,
          <string-name>
            <given-names>A.</given-names>
            <surname>Rago</surname>
          </string-name>
          , E. Albini,
          <string-name>
            <given-names>P.</given-names>
            <surname>Baroni</surname>
          </string-name>
          ,
          <string-name>
            <given-names>F.</given-names>
            <surname>Toni</surname>
          </string-name>
          ,
          <string-name>
            <surname>Argumentative</surname>
            <given-names>XAI</given-names>
          </string-name>
          :
          <article-title>A survey</article-title>
          ,
          <source>in: Proc. IJCAI</source>
          ,
          <year>2021</year>
          , pp.
          <fpage>4392</fpage>
          -
          <lpage>4399</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref22">
        <mixed-citation>
          [22]
          <string-name>
            <given-names>A.</given-names>
            <surname>Vassiliades</surname>
          </string-name>
          ,
          <string-name>
            <given-names>N.</given-names>
            <surname>Bassiliades</surname>
          </string-name>
          , T. Patkos,
          <source>Argumentation and Explainable Artificial Intelligence: A Survey</source>
          ,
          <article-title>Knowledge Eng</article-title>
          .
          <source>Rev</source>
          .
          <volume>36</volume>
          (
          <year>2021</year>
          ).
        </mixed-citation>
      </ref>
      <ref id="ref23">
        <mixed-citation>
          [23]
          <string-name>
            <surname>D. M. Gabbay</surname>
          </string-name>
          ,
          <article-title>Logical foundations for bipolar and tripolar argumentation networks: preliminary results</article-title>
          ,
          <source>J. Log. Comput</source>
          .
          <volume>26</volume>
          (
          <year>2016</year>
          )
          <fpage>247</fpage>
          -
          <lpage>292</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref24">
        <mixed-citation>
          [24]
          <string-name>
            <given-names>P.</given-names>
            <surname>Baroni</surname>
          </string-name>
          ,
          <string-name>
            <given-names>G.</given-names>
            <surname>Comini</surname>
          </string-name>
          ,
          <string-name>
            <given-names>A.</given-names>
            <surname>Rago</surname>
          </string-name>
          ,
          <string-name>
            <given-names>F.</given-names>
            <surname>Toni</surname>
          </string-name>
          ,
          <article-title>Abstract games of argumentation strategy and game-theoretical argument strength</article-title>
          ,
          <source>in: Proc. PRIMA</source>
          ,
          <year>2017</year>
          , pp.
          <fpage>403</fpage>
          -
          <lpage>419</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref25">
        <mixed-citation>
          [25]
          <string-name>
            <given-names>E.</given-names>
            <surname>Albini</surname>
          </string-name>
          ,
          <string-name>
            <given-names>A.</given-names>
            <surname>Rago</surname>
          </string-name>
          ,
          <string-name>
            <given-names>P.</given-names>
            <surname>Baroni</surname>
          </string-name>
          ,
          <string-name>
            <given-names>F.</given-names>
            <surname>Toni</surname>
          </string-name>
          ,
          <article-title>Relation-based counterfactual explanations for bayesian network classifiers</article-title>
          ,
          <source>in: Proc. IJCAI</source>
          ,
          <year>2020</year>
          , pp.
          <fpage>451</fpage>
          -
          <lpage>457</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref26">
        <mixed-citation>
          [26]
          <string-name>
            <given-names>L.</given-names>
            <surname>Amgoud</surname>
          </string-name>
          ,
          <string-name>
            <given-names>J.</given-names>
            <surname>Ben-Naim</surname>
          </string-name>
          ,
          <article-title>Weighted bipolar argumentation graphs: Axioms and semantics</article-title>
          ,
          <source>in: Proc. IJCAI</source>
          ,
          <year>2018</year>
          , pp.
          <fpage>5194</fpage>
          -
          <lpage>5198</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref27">
        <mixed-citation>
          [27]
          <string-name>
            <given-names>C.</given-names>
            <surname>Cayrol</surname>
          </string-name>
          ,
          <string-name>
            <surname>M.-C.</surname>
          </string-name>
          Lagasquie-Schiex,
          <article-title>On the acceptability of arguments in bipolar argumentation frameworks</article-title>
          ,
          <source>in: Proc. ECSQARU</source>
          ,
          <year>2005</year>
          , pp.
          <fpage>378</fpage>
          -
          <lpage>389</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref28">
        <mixed-citation>
          [28]
          <string-name>
            <given-names>J.</given-names>
            <surname>Pearl</surname>
          </string-name>
          ,
          <article-title>Reasoning with cause and efect</article-title>
          ,
          <source>in: Proc. IJCAI</source>
          ,
          <year>1999</year>
          , pp.
          <fpage>1437</fpage>
          -
          <lpage>1449</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref29">
        <mixed-citation>
          [29]
          <string-name>
            <given-names>J.</given-names>
            <surname>Pearl</surname>
          </string-name>
          ,
          <article-title>The do-calculus revisited</article-title>
          ,
          <source>in: Proc. UAI</source>
          ,
          <year>2012</year>
          , pp.
          <fpage>3</fpage>
          -
          <lpage>11</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref30">
        <mixed-citation>
          [30]
          <string-name>
            <surname>P. M. Dung</surname>
          </string-name>
          ,
          <article-title>On the Acceptability of Arguments and its Fundamental Role in Nonmonotonic Reasoning, Logic Programming</article-title>
          and
          <string-name>
            <surname>n-Person</surname>
            <given-names>Games</given-names>
          </string-name>
          ,
          <source>Artificial Intelligence</source>
          <volume>77</volume>
          (
          <year>1995</year>
          )
          <fpage>321</fpage>
          -
          <lpage>358</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref31">
        <mixed-citation>
          [31]
          <string-name>
            <given-names>L.</given-names>
            <surname>Amgoud</surname>
          </string-name>
          ,
          <string-name>
            <given-names>J.</given-names>
            <surname>Ben-Naim</surname>
          </string-name>
          ,
          <article-title>Evaluation of arguments from support relations: Axioms and semantics</article-title>
          ,
          <source>in: Proc. IJCAI</source>
          ,
          <year>2016</year>
          , pp.
          <fpage>900</fpage>
          -
          <lpage>906</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref32">
        <mixed-citation>
          [32]
          <string-name>
            <given-names>P.</given-names>
            <surname>Baroni</surname>
          </string-name>
          ,
          <string-name>
            <given-names>A.</given-names>
            <surname>Rago</surname>
          </string-name>
          ,
          <string-name>
            <given-names>F.</given-names>
            <surname>Toni</surname>
          </string-name>
          ,
          <article-title>How many properties do we need for gradual argumentation?</article-title>
          ,
          <source>in: Proc. AAAI</source>
          ,
          <year>2018</year>
          , pp.
          <fpage>1736</fpage>
          -
          <lpage>1743</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref33">
        <mixed-citation>
          [33]
          <string-name>
            <given-names>J.</given-names>
            <surname>Pearl</surname>
          </string-name>
          ,
          <article-title>Causal diagrams for empirical research</article-title>
          ,
          <source>Biometrika</source>
          <volume>82</volume>
          (
          <year>1995</year>
          )
          <fpage>669</fpage>
          -
          <lpage>710</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref34">
        <mixed-citation>
          [34]
          <string-name>
            <given-names>N.</given-names>
            <surname>Potyka</surname>
          </string-name>
          ,
          <article-title>Interpreting neural networks as quantitative argumentation frameworks</article-title>
          ,
          <source>in: Proc. AAAI</source>
          ,
          <year>2021</year>
          , pp.
          <fpage>6463</fpage>
          -
          <lpage>6470</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref35">
        <mixed-citation>
          [35]
          <string-name>
            <given-names>A.</given-names>
            <surname>Rago</surname>
          </string-name>
          ,
          <string-name>
            <given-names>O.</given-names>
            <surname>Cocarascu</surname>
          </string-name>
          ,
          <string-name>
            <given-names>F.</given-names>
            <surname>Toni</surname>
          </string-name>
          ,
          <article-title>Argumentation-based recommendations: Fantastic explanations and how to find them</article-title>
          , ????
        </mixed-citation>
      </ref>
      <ref id="ref36">
        <mixed-citation>
          [36]
          <string-name>
            <given-names>Y.</given-names>
            <surname>Son</surname>
          </string-name>
          ,
          <string-name>
            <given-names>N.</given-names>
            <surname>Bayas</surname>
          </string-name>
          ,
          <string-name>
            <given-names>H. A.</given-names>
            <surname>Schwartz</surname>
          </string-name>
          ,
          <article-title>Causal explanation analysis on social media</article-title>
          ,
          <source>in: Proc. EMNLP</source>
          ,
          <year>2018</year>
          , pp.
          <fpage>3350</fpage>
          -
          <lpage>3359</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref37">
        <mixed-citation>
          [37]
          <string-name>
            <surname>M. R. O'Shaughnessy</surname>
            ,
            <given-names>G.</given-names>
          </string-name>
          <string-name>
            <surname>Canal</surname>
            ,
            <given-names>M.</given-names>
          </string-name>
          <string-name>
            <surname>Connor</surname>
            ,
            <given-names>C.</given-names>
          </string-name>
          <string-name>
            <surname>Rozell</surname>
            ,
            <given-names>M. A.</given-names>
          </string-name>
          <string-name>
            <surname>Davenport</surname>
          </string-name>
          ,
          <article-title>Generative causal explanations of black-box classifiers</article-title>
          ,
          <source>in: Proc. NeurIPS</source>
          ,
          <year>2020</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref38">
        <mixed-citation>
          [38]
          <string-name>
            <given-names>N.</given-names>
            <surname>Pawlowski</surname>
          </string-name>
          , D. C. de Castro,
          <string-name>
            <given-names>B.</given-names>
            <surname>Glocker</surname>
          </string-name>
          ,
          <article-title>Deep structural causal models for tractable counterfactual inference</article-title>
          ,
          <source>in: Proc. NeurIPS</source>
          ,
          <year>2020</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref39">
        <mixed-citation>
          [39]
          <string-name>
            <given-names>A.</given-names>
            <surname>Chattopadhyay</surname>
          </string-name>
          ,
          <string-name>
            <given-names>P.</given-names>
            <surname>Manupriya</surname>
          </string-name>
          ,
          <string-name>
            <given-names>A.</given-names>
            <surname>Sarkar</surname>
          </string-name>
          ,
          <string-name>
            <given-names>V. N.</given-names>
            <surname>Balasubramanian</surname>
          </string-name>
          ,
          <article-title>Neural network attributions: A causal perspective</article-title>
          ,
          <source>in: Proc. ICML</source>
          ,
          <year>2019</year>
          , pp.
          <fpage>981</fpage>
          -
          <lpage>990</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref40">
        <mixed-citation>
          [40]
          <string-name>
            <given-names>T.</given-names>
            <surname>Heskes</surname>
          </string-name>
          ,
          <string-name>
            <given-names>E.</given-names>
            <surname>Sijben</surname>
          </string-name>
          ,
          <string-name>
            <given-names>I. G.</given-names>
            <surname>Bucur</surname>
          </string-name>
          , T. Claassen,
          <article-title>Causal Shapley values: Exploiting causal knowledge to explain individual predictions of complex models</article-title>
          ,
          <source>in: Proc. NeurIPS</source>
          ,
          <year>2020</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref41">
        <mixed-citation>
          [41]
          <string-name>
            <given-names>O.</given-names>
            <surname>Cocarascu</surname>
          </string-name>
          ,
          <string-name>
            <given-names>A.</given-names>
            <surname>Stylianou</surname>
          </string-name>
          ,
          <string-name>
            <given-names>K.</given-names>
            <surname>Cyras</surname>
          </string-name>
          ,
          <string-name>
            <given-names>F.</given-names>
            <surname>Toni</surname>
          </string-name>
          ,
          <article-title>Data-empowered argumentation for dialectically explainable predictions</article-title>
          ,
          <source>in: Proc. ECAI</source>
          ,
          <year>2020</year>
          , pp.
          <fpage>2449</fpage>
          -
          <lpage>2456</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref42">
        <mixed-citation>
          [42]
          <string-name>
            <given-names>K.</given-names>
            <surname>Cyras</surname>
          </string-name>
          ,
          <string-name>
            <given-names>K.</given-names>
            <surname>Satoh</surname>
          </string-name>
          ,
          <string-name>
            <given-names>F.</given-names>
            <surname>Toni</surname>
          </string-name>
          ,
          <article-title>Abstract argumentation for case-based reasoning</article-title>
          ,
          <source>in: Proc. KR</source>
          ,
          <year>2016</year>
          , pp.
          <fpage>549</fpage>
          -
          <lpage>552</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref43">
        <mixed-citation>
          [43]
          <string-name>
            <given-names>L.</given-names>
            <surname>Page</surname>
          </string-name>
          ,
          <string-name>
            <given-names>S.</given-names>
            <surname>Brin</surname>
          </string-name>
          ,
          <string-name>
            <given-names>R.</given-names>
            <surname>Motwani</surname>
          </string-name>
          , T. Winograd, The PageRank Citation Ranking:
        </mixed-citation>
      </ref>
    </ref-list>
  </back>
</article>