<!DOCTYPE article PUBLIC "-//NLM//DTD JATS (Z39.96) Journal Archiving and Interchange DTD v1.0 20120330//EN" "JATS-archivearticle1.dtd">
<article xmlns:xlink="http://www.w3.org/1999/xlink">
  <front>
    <journal-meta>
      <journal-title-group>
        <journal-title>Italian Workshop on Explainable Artificial Intelligence
" andrea.apicella@unina.it (A. Apicella)</journal-title>
      </journal-title-group>
    </journal-meta>
    <article-meta>
      <title-group>
        <article-title>Explanations in terms of Hierarchically organised Middle Level Features</article-title>
      </title-group>
      <contrib-group>
        <contrib contrib-type="author">
          <string-name>Andrea Apicella</string-name>
          <xref ref-type="aff" rid="aff0">0</xref>
          <xref ref-type="aff" rid="aff2">2</xref>
          <xref ref-type="aff" rid="aff3">3</xref>
        </contrib>
        <contrib contrib-type="author">
          <string-name>Salvatore Giugliano</string-name>
          <xref ref-type="aff" rid="aff0">0</xref>
          <xref ref-type="aff" rid="aff2">2</xref>
          <xref ref-type="aff" rid="aff3">3</xref>
        </contrib>
        <contrib contrib-type="author">
          <string-name>Francesco Isgrò</string-name>
          <xref ref-type="aff" rid="aff0">0</xref>
          <xref ref-type="aff" rid="aff2">2</xref>
          <xref ref-type="aff" rid="aff3">3</xref>
        </contrib>
        <contrib contrib-type="author">
          <string-name>Roberto Prevete</string-name>
          <xref ref-type="aff" rid="aff0">0</xref>
          <xref ref-type="aff" rid="aff1">1</xref>
          <xref ref-type="aff" rid="aff2">2</xref>
          <xref ref-type="aff" rid="aff3">3</xref>
        </contrib>
        <aff id="aff0">
          <label>0</label>
          <institution>Department of Electrical Engineering and Information Technology, Università degli Studi di Napoli Federico II</institution>
          ,
          <addr-line>Naples</addr-line>
          ,
          <country country="IT">Italy</country>
        </aff>
        <aff id="aff1">
          <label>1</label>
          <institution>Interdepartmental Center for Research on Management and Innovation in Healthcare (CIRMIS), Università degli Studi di Napoli Federico II</institution>
          ,
          <addr-line>Naples</addr-line>
          ,
          <country country="IT">Italy</country>
        </aff>
        <aff id="aff2">
          <label>2</label>
          <institution>Laboratory of Artificial Intelligence, Privacy &amp; Applications (AIPA Lab), Università degli Studi di Napoli Federico II</institution>
          ,
          <addr-line>Naples</addr-line>
          ,
          <country country="IT">Italy</country>
        </aff>
        <aff id="aff3">
          <label>3</label>
          <institution>Laboratory of Augmented Reality for Health Monitoring (ARHeMLab), Università degli Studi di Napoli Federico II</institution>
          ,
          <addr-line>Naples</addr-line>
          ,
          <country country="IT">Italy</country>
        </aff>
      </contrib-group>
      <pub-date>
        <year>2021</year>
      </pub-date>
      <volume>000</volume>
      <fpage>0</fpage>
      <lpage>0002</lpage>
      <abstract>
        <p>The rapidly growing research area of eXplainable Artificial Intelligence (XAI) focuses on making Machine Learning systems' decisions more transparent and humanly understandable. One of the most successful XAI strategies is to provide explanations in terms of visualisations and, more specifically, low-level input features such as relevance scores or heat maps of the input, like sensitivity analysis or layer-wise relevance propagation methods. The main problem with such methods is that starting from the relevance of low-level features, the human user needs to identify the overall input properties that are salient. Thus, a current line of XAI research attempts to alleviate this weakness of low-level approaches, constructing explanations in terms of input features that represent more salient and understandable input properties for a user, which we call here Middle-Level input Features (MLF). In addition, another interesting and very recent approach is that of considering hierarchically organised explanations. Thus, in this paper, we investigate the possibility to combine both MLFs and hierarchical organisations. The potential advantages of providing explanations in terms of hierarchically organised MLFs are grounded on the possibility of exhibiting explanations to a diferent granularity of MLFs interacting with each other. We experimentally tested our approach on 300 Birds Species and Cars dataset. The results seem encouraging.</p>
      </abstract>
      <kwd-group>
        <kwd>eol&gt;XAI</kwd>
        <kwd>Explainable AI</kwd>
        <kwd>Hierarchical</kwd>
        <kwd>middle-level</kwd>
        <kwd>Interpretable models</kwd>
      </kwd-group>
    </article-meta>
  </front>
  <body>
    <sec id="sec-1">
      <title>1. Introduction</title>
      <p>
        A large part of current successfully Machine Learning(ML) techniques can be considered as
black-box systems insofar as they give responses whose relationships with the input are often
challenging to understand. As ML systems are frequently being used in more and more domains
and, so, by a more varied audience, there is the need of making them understandable and trusting
to general users [
        <xref ref-type="bibr" rid="ref1 ref2">1, 2</xref>
        ], leaving unaltered, or even improving, their performance. The rapidly
growing research area of eXplainable Artificial Intelligence (XAI) is focused on this challenge.
In the XAI context many approaches have been proposed to overcome the opaqueness of ML
systems [
        <xref ref-type="bibr" rid="ref2 ref3 ref4 ref5 ref6">3, 4, 5, 2, 6</xref>
        ]. We note that in the literature, one of the most successful strategies is to
provide explanations in terms of “visualisations” [
        <xref ref-type="bibr" rid="ref1 ref7">1, 7</xref>
        ], and, more specifically, in terms of
lowlevel input features such as relevance scores or heat maps of the input, like sensitivity analysis
[
        <xref ref-type="bibr" rid="ref8">8</xref>
        ] or Layer-wise Relevance Propagation (LRP) [
        <xref ref-type="bibr" rid="ref3">3</xref>
        ] methods. For example, LRP associates a
relevance value to each input element (to each pixel in case of images) to explain the ML model
answer.
      </p>
      <p>
        The main problem with such methods is that human users are left with a significant
interpretive burden. Starting from the relevance of each low-level feature, the human user needs to
identify the overall input properties that are perceptually and cognitively salient to him [
        <xref ref-type="bibr" rid="ref6 ref9">9, 6</xref>
        ].
Thus, there is a current line of XAI research that attempts to alleviate this weakness of low-level
approaches and overcome their limitations. The explanations are obtained in terms of input
features that represent more salient and understandable input properties for a user [
        <xref ref-type="bibr" rid="ref6 ref9">9, 10, 11, 6</xref>
        ],
which we call Middle-Level input Features (MLFs) (see, as an example, Figure 1).
      </p>
      <p>Another interesting and very recent approach is that of considering hierarchically organised
explanations [12, 13, 14, 13]. Thus, in this paper, we investigate the possibility to combine both
MLFs and hierarchical organisations. The potential advantages of providing explanations in
terms of hierarchically organised MLFs are grounded on the possibility of exhibiting explanations
to a diferent granularity of MLFs interacting with each other. For example, natural images can
be described in terms of the objects they show at various levels of granularity, and their relations
[15, 16, 17]. In this paper, in particular, we take advantage of using a general framework that we
recently proposed in a paper under review [18]. The paper is organised as follows: in Section
2 we discuss diferences and advantages of our approach with respect to similar approaches
presented in the literature; Section 3 describes in detail the proposed approach; experiments and
results are discussed in Section 4 and 5; the concluding Section summarises the main high-level
features of the proposed explanation framework and outlines some future developments.</p>
    </sec>
    <sec id="sec-2">
      <title>2. Background</title>
      <p>
        Building explanations in terms of Middle-Level input Features (MLFs) is a growing research area
in the XAI community. For example, in [
        <xref ref-type="bibr" rid="ref9">9</xref>
        ] Concept Activation Vectors (CAV) are introduced
as a way to visually represent human-understandable concepts. These concepts are extracted
by an external labelled dataset. The authors proposed a method to quantify the influence
of each concept on the classifier output. CAV-based explanations are expressed in terms of
high-level visual concepts which, unlike our approach, do not necessarily belong to the ML
system input. In [10], the CAVs are automatically extracted without using external labelled
data expressing human-friendly concepts, but to produce explanations related to an entire class
(global explanation) instead of the single ML system input (local explanation). Thus, again,
human users are left with a significant interpretive load: starting from external high-level visual
concepts, the human user needs to identify the input properties perceptually and cognitively
related to these concepts. On the contrary, in our approach MLFs are expressed in terms of
elements belonging to the input itself. In [19] LIME is proposed. Especially in the context of
images, LIME is one of the predominant XAI methods discussed in the literature [20, 21]. It
provides local explanations in terms of relevant image regions of the input that the classifier
receives. In [11] the authors use the concept of “fault lines”, defined as “high-level semantic
aspects of reality that humans zoom in on when imagining an alternative to it”. The proposed
explanations are given in terms of images representing semantic aspects which should help
the user to understand why the classifier output was a given label instead of another one. The
semantic aspects are constructed using feature maps of the pre-trained Convolutional Neural
network. Thus, a critical point is that high-level or middle-level user-friendly concepts are
computed on the basis of the neural network classifier to be explained. In this way, an unsafe
short-circuit can be created in which the visual concepts used to explain the classifier are
closely related to the classifier itself. This fact could lead to the creation of false human-friendly
visual concepts if the classifier is not reliable. By contrast, in our approach, MLFs are extracted
independently from the classifier.
      </p>
      <p>A crucial aspect that distinguishes our proposal from the above-discussed research line is
grounded on the fact that we propose explanations in terms of hierarchically organised MLFs.</p>
      <p>From the point of view of the hierarchical properties, a number of research works have already
tried to give hierarchical explanations in XAI domain. In [13] a method to build hierarchical
saliency map exploiting the relations between the input features is proposed. In [12] the Mahé
framework is described. As in LIME, Mahé perturbs the input data and construct a proxy model
using the outputs of the original model on the perturbed data. Diferently from LIME, Mahé
uses a Neural Network as model approximator instead of a linear model. As in [13], the possible
interactions between variables are captured to build hierarchical explanations. However, none
of these two methods focus on input features that represent more salient and understandable
input properties such as MLFs. In [14] another method for hierarchical explanation is proposed,
but limited to text classification problems.</p>
    </sec>
    <sec id="sec-3">
      <title>3. Proposal</title>
      <sec id="sec-3-1">
        <title>3.1. General description</title>
        <p>Given an ML classification model  which receives an input x ∈  and outputs y ∈ , we
want to produce an explanation of the output y in terms of MLFs organised in a hierarchical
way. Our XAI method is based on a more general approach which we have proposed in a paper
currently under review [18]. Our method can be divided into two consecutive steps.</p>
        <p>In the first step, we build an auto-encoder  ≡ (, ) such that the input x can be
represented by the encoder  in a hierarchical structure. For instance, considering x as an
input image, we generate an encoder which returns a set of image partitions organised in
a hierarchical way, i.e. each coarser detail partition can be obtained by merging partitions
elements representing finer details. As discussed in [ 22], this target can be achieved by a
segmentation algorithm which ensures the following two conditions: 1) if a contour is present
at a given scale, the same contour has to be present at any finer scale ( causality principle of the
multiscale analysis [23]) and 2) even when the number of regions decreases, the contours remain
stable (location principle). If a segmentation algorithm ensures these conditions, then it can be
{︃1 if all the pixels in vj+1
considered a hierarchical segmentation algorithm. More formally, given a set of  diferent
partitions {1, 2, . . . ,  } of the image x ∈  produced by a segmentation algorithm sorted
from the coarse to the finer detail level, each region v ∈  can be expressed as a linear
v+1, where   =
combination of the elements of the finer level +1, i.e. vik = ∑︀   
∈ vi for each  ∈ {1, 2, . . . , }, and each vik is a candidate
0 otherwise,
MLF to be included in the explanation. In this way, it is straightforward to build a
feedforward neural network as an autoencoder which outputs a reconstruction of the image x
exploiting a hierarchical segmentation of x in its own internal layers. In a nutshell, given a
fully-connected feed-forward neural network of  + 1 layers having || inputs and |+1|
outputs for  ∈ . . . 1, 2, . . . , , and  inputs and  outputs for the final  + 1 layer, we
can set the identity as activation functions, biases equal to 0 and each weights  =   for
 ∈ {1, 2, . . . , }. For the last layer, we can consider the image x as the trivial partition where
each partition elements represents a single image pixel, i.e. +1 = {v1+1, v2+1, . . . , v+1}
{︃+1 =  if  = 
with +1 =  . The weights of the  + 1 layer can be set equal
0 otherwise.
to (v+1)=1. The resulting network, fed with the 1 vector, outputs the original image x. It
can be viewed as a decoder  that decodes the sequence of all 1s as x. Being a simple feed
forward neural network, once stacked on the top of the model  , each relevance propagation
method can be used to give a relevance score to each MLF vik, hierarchically organised by . In
Algorithm 1 our approach is reported in pseudo-code, while in Fig. 2 a graphical representation
of the proposed method is shown.</p>
      </sec>
    </sec>
    <sec id="sec-4">
      <title>4. Experimental setup</title>
      <p>In this section the setup used to validate the proposed method is described. In Section 5 a set
of possible explanations of the classifier outputs on image sampled from 300 Birds Species [24]
and Cars [25] dataset are shown and discussed. As classifier, a VGG16 [ 26] architecture trained
on Imagenet was used. We used as MLFs the set of segmentations produced by [22]. This
algorithm ensures the causality and the location principle discussed in Section 3, so the returned
segmentations can be considered hierarchically related. However, any segmentation algorithm
which respects the principles described in Section 3 can be used. Firstly, ℎ sets of segments
ℎ
{}=1 related between them in a hierarchical way are generated for each test image to explain,
going from the coarsest ( = 1) to the finest (  = ℎ) segmentation level. These segments are the
candidate MLFs for providing an hierarchically organised explanation. Next, the weights of a
ℎ-layers fully-connected neural network are set using the obtained segmentations as described
in Section 3. Finally, the network is stacked on the top of VGG16 model and the LRP algorithm
is applied to the whole network fed with the "1"s vector. An LRP heatmaps for each hierarchical
layer is plotted and the most relevant MLFs are highlighted.</p>
      <p>Algorithm 1: Hierarchical segmention-based explanation Generator</p>
      <p>Input: data point x ∈ , hierarchical segmentation procedure  and its parameters
 = ( 1,  2, . . . ,   ), a relevance propagation algorithm  which outputs the
relevances
Output: A hierarchical explanation exp of the model  on the input x. exp will be
composed of  relevance vectors indicating the importance feature scores for
each layer of the hierarchy.
1 {1, 2, . . . ,  } ← (x,  );
2 let +1 ← ∅ ;
3 for  ∈ x do
4 let v+1 ∈ {0};
+1 ←  ;
+1 ← +1 ∪ {v+1};</p>
      <p>←
end
end
 ←
   (weights = { }=+11,
biases = {b}=+11,
activation fun = );
21
22
23
24
25 end</p>
      <p>1 |1|;
define  : x ↦→  ∈ { }
stack together (,  );
exp ←  (x, , (,  ))
return exp1,..., ;</p>
    </sec>
    <sec id="sec-5">
      <title>5. Results</title>
      <p>
        In this section, an evaluation of our approach is reported and discussed with input images taken
from 300 Birds Species and Cars datasets using ℎ ∈ {2, 3} hierarchy layers. The evaluation is
made in both qualitative and quantitative terms. The former showing the explanations produced
by the proposed method on several inputs of 300 Birds Species and Cars dataset fed to VGG16
model, the latter showing the MoRF (Most Relevant First) curves [
        <xref ref-type="bibr" rid="ref3">3, 27</xref>
        ]. For all the inputs, a
comparison with the explanations generated by LIME method is made.
      </p>
      <sec id="sec-5-1">
        <title>5.1. Qualitative evaluation</title>
        <sec id="sec-5-1-1">
          <title>5.1.1. Dataset 300 Birds Species</title>
          <p>In Fig. 3 an example of a two-layer hierarchical explanation of the class bald eagle, correctly
assigned to an input image, compared with LIME explanation is reported. One can observe that
the first hierarchical level highlights the head as the most significant MLF for the returned class.
Going deeper into the hierarchy, it is possible to see that not all the head parts contributed at
the same way to the returned output. Instead, the head plumage and the beak seem to have a
greater importance in the final classification. By contrast, one can note that LIME produces
“flat” explanations, without any relation between diferent segments. In other words, given a
selected input partition, LIME outputs similar results with respect our approach, but it is unable
to relate segments belonging to diferent input partitions.</p>
          <p>Similar results are shown in Figure 4, 5, 6 and 7. In Fig. 4 the first explanation level highlights
that the swan body and the beak have been relevant for the final classification. Going deeper
into the hierarchy, is possible to see that the neck and the head result to be determinant for the
obtained output.</p>
          <p>In Fig. 5 in the first hierarchical layer is highlighted how the upper body part is particularly
relevant for the classification. Next into the hierarchy, one can note that the neck and the wing
result to be particularly incisive for the model output. In Fig. 6 the first hierarchical level, the
most relevant middle-level-feature is the flamingo body. Interestingly, the sky has also an high
weight for the classification (2nd position). In the next hierarchy level, it is highlighted which
body parts contributed most to the final result, that is its distinctive S-shaped neck. Similar
considerations can be done for the Fig 7 where the most relevant part of the image is almost the
entire pelican. The relevance of the sky has great incidence for the classification (2nd position).
Going deeper along the hierarchy, the most important parts of the whole body result to be the
wing and the beak.</p>
          <p>(a) Proposed approach
(b) LIME</p>
        </sec>
        <sec id="sec-5-1-2">
          <title>5.1.2. Dataset Cars</title>
          <p>In Fig. 8, an image containing a sports car is fed to VGG16 classifier. The proposed explanation
method returns the car bodywork as the most relevant one for the classification output (left side,
ifrst row). Exploiting the second layer of the hierarchy (left side, second row), the diferent parts
of the bodywork sorted by relevance can be highlighted helping the user to better understand
the output. Interestingly, LIME returns the same most relevant segment (right side, first row).
However, with a finer segmentation (right side, second row), no relation with the coarser one is
taken into account. Similar consideration can be done for Fig. 9, 10, and 11.</p>
        </sec>
      </sec>
      <sec id="sec-5-2">
        <title>5.2. Quantitative evaluation</title>
        <p>
          The MoRF (Most Relevant First) curve analysis is a method to evaluate the explanation produced
by an XAI method proposed and discussed in [
          <xref ref-type="bibr" rid="ref3">3, 27</xref>
          ]. In a nutshell, features of the input are
iteratively replaced by random noise and fed to the classifier, in the descending order with
respect to the relevance values given by the explanation to evaluate. Finally, all the probability
values of the original class can be plotted generating a curve. We expect that the better the
explanation method is, the steeper the curve the better the explanation. Since the proposed
approach relies on Middle Level Features, the MoRF curves are computed following the region
lfipping as perturbation schema, a generalisation of the pixel-flipping measure proposed in [
          <xref ref-type="bibr" rid="ref3">3</xref>
          ].
Furthermore, MLFs were removed from the inputs, exploiting the hierarchy in a topological
sort depth-first search based on the descending order’s relevances. Therefore, the MLFs of the
ifnest hierarchical layer were considered.
        </p>
        <p>As baseline, MoRF obtained with LIME are produced. To make the results comparable, the
same segmentation algorithm was used both with the proposed work and LIME, instead of
the segmentation algorithm used in the original paper. Comparing the MoRF curves related to
the explanation obtained by the proposed method and LIME, it results that in most cases the
proposed approach returns comparable explanations or better, as shown in Fig. 12 and 13.</p>
      </sec>
    </sec>
    <sec id="sec-6">
      <title>6. Conclusion</title>
      <p>We proposed in this paper to combine two recent approaches in the XAI research literature:
explanations in terms of input features that represent more salient and understandable input
properties for a user, Middle-Level input Feature (MLF), and their hierarchical organisation. To
this aim, a new XAI method has been proposed, built on a more general approach proposed
in a paper [18] freely available on Arxiv. We have experimentally evaluated the advantages
of providing explanations in terms of hierarchically organised MLFs. In particular, we have
highlighted the possibility of exhibiting explanations to a diferent granularity of MLFs
interacting with each other. The experiments were conducted using 300 Birds Species and Cars
as dataset, and VGG16 as classifier. Our approach was evaluated in both qualitatively and
quantitatively manner and compared with LIME. The preliminary results seem encouraging.
When it is possible to find MLF relationships at diferent degrees of input granularity, we can
obtain explanations easier to understand.</p>
    </sec>
    <sec id="sec-7">
      <title>Acknowledgments</title>
      <p>This work was carried out under the initiative “Departments of Excellence” (Italian Budget
Law no. 232/2016), through an excellence grant awarded to the Department of Information
Technology and Electrical Engineering of the University of Naples Federico II, Naples, Italy.
[10] A. Ghorbani, J. Wexler, J. Zou, B. Kim, Towards automatic concept-based explanations,
arXiv preprint arXiv:1902.03129 (2019).
[11] A. Akula, S. Wang, S.-C. Zhu, Cocox: Generating conceptual and counterfactual
explanations via fault-lines, in: Proceedings of the AAAI Conference on Artificial Intelligence,
volume 34, 2020, pp. 2594–2601.
[12] M. Tsang, Y. Sun, D. Ren, Y. Liu, Can i trust you more? model-agnostic hierarchical
explanations, arXiv preprint arXiv:1812.04801 (2018).
[13] C. Singh, W. J. Murdoch, B. Yu, Hierarchical interpretations for neural network predictions,
in: International Conference on Learning Representations, 2019.
[14] H. Chen, G. Zheng, Y. Ji, Generating hierarchical explanations on text classification via
feature interaction detection, arXiv preprint arXiv:2004.02015 (2020).
[15] M. Tschannen, O. Bachem, M. Lucic, Recent advances in autoencoder-based representation
learning, arXiv preprint arXiv:1812.05069 (2018).
[16] C. K. Sønderby, T. Raiko, L. Maaløe, S. K. Sønderby, O. Winther, Ladder variational
autoencoders, in: Proceedings of the 30th International Conference on Neural Information
Processing Systems, 2016, pp. 3745–3753.
[17] S. Zhao, J. Song, S. Ermon, Learning hierarchical features from deep generative models,
in: International Conference on Machine Learning, PMLR, 2017, pp. 4091–4099.
[18] A. Apicella, F. Isgrò, R. Prevete, A general approach for explanations in terms of middle
level features, arXiv preprint arXiv:2106.05037 (2021).
[19] M. T. Ribeiro, S. Singh, C. Guestrin, "why should i trust you?": Explaining the predictions
of any classifier, in: Proceedings of the 22Nd ACM SIGKDD International Conference on
Knowledge Discovery and Data Mining, KDD ’16, ACM, New York, NY, USA, 2016, pp.
1135–1144.
[20] J. Dieber, S. Kirrane, Why model why? assessing the strengths and limitations of lime,
arXiv preprint arXiv:2012.00093 (2020).
[21] X. Zhao, X. Huang, V. Robu, D. Flynn, Baylime: Bayesian local interpretable model-agnostic
explanations, arXiv preprint arXiv:2012.03058 (2020).
[22] S. J. F. Guimarães, J. Cousty, Y. Kenmochi, L. Najman, A hierarchical image segmentation
algorithm based on an observation scale, in: Joint IAPR International Workshops on
Statistical Techniques in Pattern Recognition (SPR) and Structural and Syntactic Pattern
Recognition (SSPR), Springer, 2012, pp. 116–125.
[23] L. Guigues, J. P. Cocquerez, H. Le Men, Scale-sets image analysis, International Journal of</p>
      <p>Computer Vision 68 (2006) 289–317.
[24] G. Piosenka, 300 bird species dataset, 2014. Data retrieved from 300 Bird Species Dataset,
https://www.kaggle.com/gpiosenka/100-bird-species.
[25] J. Krause, M. Stark, J. Deng, L. Fei-Fei, 3d object representations for fine-grained
categorization, in: 4th International IEEE Workshop on 3D Representation and Recognition
(3dRR-13), Sydney, Australia, 2013.
[26] K. Simonyan, A. Zisserman, Very deep convolutional networks for large-scale image
recognition, in: International Conference on Learning Representations, 2015.
[27] W. Samek, A. Binder, G. Montavon, S. Lapuschkin, K.-R. Müller, Evaluating the visualization
of what a deep neural network has learned, IEEE transactions on neural networks and
learning systems 28 (2016) 2660–2673.</p>
    </sec>
  </body>
  <back>
    <ref-list>
      <ref id="ref1">
        <mixed-citation>
          [1]
          <string-name>
            <given-names>M.</given-names>
            <surname>Ribera</surname>
          </string-name>
          ,
          <string-name>
            <given-names>A.</given-names>
            <surname>Lapedriza</surname>
          </string-name>
          ,
          <article-title>Can we do better explanations? a proposal of user-centered explainable ai</article-title>
          .,
          <source>in: IUI Workshops</source>
          ,
          <year>2019</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref2">
        <mixed-citation>
          [2]
          <string-name>
            <given-names>A. B.</given-names>
            <surname>Arrieta</surname>
          </string-name>
          ,
          <string-name>
            <given-names>N.</given-names>
            <surname>Díaz-Rodríguez</surname>
          </string-name>
          ,
          <string-name>
            <given-names>J. Del</given-names>
            <surname>Ser</surname>
          </string-name>
          ,
          <string-name>
            <given-names>A.</given-names>
            <surname>Bennetot</surname>
          </string-name>
          ,
          <string-name>
            <given-names>S.</given-names>
            <surname>Tabik</surname>
          </string-name>
          ,
          <string-name>
            <given-names>A.</given-names>
            <surname>Barbado</surname>
          </string-name>
          ,
          <string-name>
            <given-names>S.</given-names>
            <surname>García</surname>
          </string-name>
          ,
          <string-name>
            <given-names>S.</given-names>
            <surname>Gil-López</surname>
          </string-name>
          ,
          <string-name>
            <given-names>D.</given-names>
            <surname>Molina</surname>
          </string-name>
          ,
          <string-name>
            <given-names>R.</given-names>
            <surname>Benjamins</surname>
          </string-name>
          , et al.,
          <article-title>Explainable artificial intelligence (xai): Concepts, taxonomies, opportunities and challenges toward responsible ai</article-title>
          ,
          <source>Information Fusion</source>
          <volume>58</volume>
          (
          <year>2020</year>
          )
          <fpage>82</fpage>
          -
          <lpage>115</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref3">
        <mixed-citation>
          [3]
          <string-name>
            <given-names>S.</given-names>
            <surname>Bach</surname>
          </string-name>
          ,
          <string-name>
            <given-names>A.</given-names>
            <surname>Binder</surname>
          </string-name>
          ,
          <string-name>
            <given-names>G.</given-names>
            <surname>Montavon</surname>
          </string-name>
          ,
          <string-name>
            <given-names>F.</given-names>
            <surname>Klauschen</surname>
          </string-name>
          ,
          <string-name>
            <surname>K.-R. Müller</surname>
          </string-name>
          , W. Samek,
          <article-title>On pixel-wise explanations for non-linear classifier decisions by layer-wise relevance propagation</article-title>
          ,
          <source>PloS one 10</source>
          (
          <year>2015</year>
          )
          <article-title>e0130140</article-title>
          .
        </mixed-citation>
      </ref>
      <ref id="ref4">
        <mixed-citation>
          [4]
          <string-name>
            <given-names>A.</given-names>
            <surname>Nguyen</surname>
          </string-name>
          ,
          <string-name>
            <given-names>J.</given-names>
            <surname>Yosinski</surname>
          </string-name>
          ,
          <string-name>
            <given-names>J.</given-names>
            <surname>Clune</surname>
          </string-name>
          , Multifaceted Feature Visualization:
          <article-title>Uncovering the Diferent Types of Features Learned By Each Neuron in Deep Neural Networks</article-title>
          , ArXiv e-prints (
          <year>2016</year>
          ).
        </mixed-citation>
      </ref>
      <ref id="ref5">
        <mixed-citation>
          [5]
          <string-name>
            <given-names>D.</given-names>
            <surname>Doran</surname>
          </string-name>
          ,
          <string-name>
            <given-names>S.</given-names>
            <surname>Schulz</surname>
          </string-name>
          ,
          <string-name>
            <given-names>T. R.</given-names>
            <surname>Besold</surname>
          </string-name>
          ,
          <article-title>What does explainable AI really mean? A new conceptualization of perspectives</article-title>
          ,
          <source>CoRR abs/1710</source>
          .00794 (
          <year>2017</year>
          ).
        </mixed-citation>
      </ref>
      <ref id="ref6">
        <mixed-citation>
          [6]
          <string-name>
            <given-names>A.</given-names>
            <surname>Apicella</surname>
          </string-name>
          ,
          <string-name>
            <given-names>F.</given-names>
            <surname>Isgro</surname>
          </string-name>
          ,
          <string-name>
            <given-names>R.</given-names>
            <surname>Prevete</surname>
          </string-name>
          , G. Tamburrini,
          <article-title>Middle-level features for the explanation of classification systems by sparse dictionary methods</article-title>
          ,
          <source>International Journal of Neural Systems</source>
          <volume>30</volume>
          (
          <year>2020</year>
          )
          <fpage>2050040</fpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref7">
        <mixed-citation>
          [7]
          <string-name>
            <given-names>Q.</given-names>
            <surname>Zhang</surname>
          </string-name>
          , S. Zhu,
          <article-title>Visual interpretability for deep learning: a survey</article-title>
          ,
          <source>Frontiers of Information Technology &amp; Electronic Engineering</source>
          <volume>19</volume>
          (
          <year>2018</year>
          )
          <fpage>27</fpage>
          -
          <lpage>39</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref8">
        <mixed-citation>
          [8]
          <string-name>
            <given-names>K.</given-names>
            <surname>Simonyan</surname>
          </string-name>
          ,
          <string-name>
            <given-names>A.</given-names>
            <surname>Vedaldi</surname>
          </string-name>
          ,
          <string-name>
            <given-names>A.</given-names>
            <surname>Zisserman</surname>
          </string-name>
          ,
          <article-title>Deep inside convolutional networks: Visualising image classification models and saliency maps</article-title>
          ,
          <source>in: 2nd International Conference on Learning Representations, Workshop Track Proceedings, Banf, Canada</source>
          ,
          <year>2014</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref9">
        <mixed-citation>
          [9]
          <string-name>
            <given-names>B.</given-names>
            <surname>Kim</surname>
          </string-name>
          ,
          <string-name>
            <given-names>M.</given-names>
            <surname>Wattenberg</surname>
          </string-name>
          ,
          <string-name>
            <given-names>J.</given-names>
            <surname>Gilmer</surname>
          </string-name>
          ,
          <string-name>
            <given-names>C.</given-names>
            <surname>Cai</surname>
          </string-name>
          ,
          <string-name>
            <given-names>J.</given-names>
            <surname>Wexler</surname>
          </string-name>
          ,
          <string-name>
            <given-names>F.</given-names>
            <surname>Viegas</surname>
          </string-name>
          , et al.,
          <article-title>Interpretability beyond feature attribution: Quantitative testing with concept activation vectors (tcav)</article-title>
          ,
          <source>in: International conference on machine learning, PMLR</source>
          ,
          <year>2018</year>
          , pp.
          <fpage>2668</fpage>
          -
          <lpage>2677</lpage>
          .
        </mixed-citation>
      </ref>
    </ref-list>
  </back>
</article>