<!DOCTYPE article PUBLIC "-//NLM//DTD JATS (Z39.96) Journal Archiving and Interchange DTD v1.0 20120330//EN" "JATS-archivearticle1.dtd">
<article xmlns:xlink="http://www.w3.org/1999/xlink">
  <front>
    <journal-meta>
      <journal-title-group>
        <journal-title>Workshops, Los Angeles, USA, March</journal-title>
      </journal-title-group>
    </journal-meta>
    <article-meta>
      <title-group>
        <article-title>Detecting Changes in User Behavior to Understand Interaction Provenance during Visual Data Analysis</article-title>
      </title-group>
      <contrib-group>
        <contrib contrib-type="author">
          <string-name>Alyssa Peña</string-name>
          <email>alyspena@gmail.com</email>
          <xref ref-type="aff" rid="aff0">0</xref>
        </contrib>
        <contrib contrib-type="author">
          <string-name>Ehsanul Haque Nirjhar</string-name>
          <email>nirjhar71@tamu.edu</email>
          <xref ref-type="aff" rid="aff0">0</xref>
        </contrib>
        <contrib contrib-type="author">
          <string-name>Andrew Pachuilo</string-name>
          <xref ref-type="aff" rid="aff0">0</xref>
        </contrib>
        <contrib contrib-type="author">
          <string-name>Theodora Chaspari</string-name>
          <email>chaspari@tamu.edu</email>
          <xref ref-type="aff" rid="aff0">0</xref>
        </contrib>
        <contrib contrib-type="author">
          <string-name>Eric D. Ragan</string-name>
          <email>eragan@ufl.edu</email>
          <xref ref-type="aff" rid="aff1">1</xref>
        </contrib>
        <contrib contrib-type="author">
          <string-name>Visual Analytics, Modeling User Behavior, Explainable AI, Sense-</string-name>
          <xref ref-type="aff" rid="aff2">2</xref>
        </contrib>
        <aff id="aff0">
          <label>0</label>
          <institution>Texas A&amp;M University</institution>
          ,
          <country country="US">USA</country>
        </aff>
        <aff id="aff1">
          <label>1</label>
          <institution>University of Florida</institution>
          ,
          <country country="US">USA</country>
        </aff>
        <aff id="aff2">
          <label>2</label>
          <institution>making</institution>
        </aff>
      </contrib-group>
      <pub-date>
        <year>2019</year>
      </pub-date>
      <volume>20</volume>
      <issue>2019</issue>
      <abstract>
        <p>Analysts can make better informed decisions with “Explainable AI”. Visualizations are often used to understand, diagnose, and refine AI models. Yet, it is unclear what type of interactions are appropriate for a given model and how the visualizations are perceived. Furthering research into sensemaking for visual analytics may be useful in understanding how user's are interacting with visualizations for AI and in developing a naturalistic model of explanation. Conventional approaches consist of human experts applying theoretical sensemaking models to identify changes in information processing or utilizing recorded rationale provided by the users. However, these approaches can be ineficient and inaccurate since they heavily rely on subjective human reports. In this research, we aim to understand how data-driven techniques can automatically identify changes in user behavior (inflection points) based on user interaction logs collected from eye tracking and mouse interactions. We relay the results of a supervised classification system using Hidden Markov Models to predict changes in a visual data analysis of a cyber security scenario. Preliminary results indicate a 70% accuracy in identifying inflection points. These preliminary results suggest the feasibility of data-driven approaches furthering our understanding of sensemaking processes and interaction provenance.</p>
      </abstract>
    </article-meta>
  </front>
  <body>
    <sec id="sec-1">
      <title>CCS CONCEPTS</title>
      <p>• Human-centered computing → Collaborative and social
computing theory, concepts and paradigms; Visual
analytics; • Computing methodologies → Knowledge representation
and reasoning; Machine learning approaches.
IUI Workshops’19, March 20, 2019, Los Angeles, USA
© 2019 Copyright ©2019 for the individual papers by the papers’ authors. Copying
permitted for private and academic purposes. This volume is published and copyrighted
by its editors.
1</p>
    </sec>
    <sec id="sec-2">
      <title>INTRODUCTION</title>
      <p>
        The efectiveness of AI applications is limited by the disconnect
between experts knowledgeable in AI and subject matter experts.
Using “Explainable AI” (XAI), analysts can make better informed
decisions [
        <xref ref-type="bibr" rid="ref17">17</xref>
        ]. In particular, visualizations are used to understand,
diagnose, and refine AI models. However, it is unclear what type of
interactions are appropriate for a given model and how the
visualization is perceived [
        <xref ref-type="bibr" rid="ref23">23</xref>
        ]. Sensemaking can aid in the development of
a naturalistic model of explanation and visualization tools [
        <xref ref-type="bibr" rid="ref17 ref19">17, 19</xref>
        ].
In visual analytics, understanding the user’s sensemaking process
comprises an important milestone for improving tools, supporting
collaboration among analysts, and training new workers [
        <xref ref-type="bibr" rid="ref30">30</xref>
        ].
Furthering existing research in sensemaking for visual analytics may
in turn further research in XAI.
      </p>
      <p>
        Due to the complexity and ad-hoc nature of sensemaking
during exploratory data analysis, it is dificult to comprehensively
characterize sensemaking using solely theoretical models
without expertise and assumptions to map analyst behavior to the
model [
        <xref ref-type="bibr" rid="ref20 ref29 ref32">20, 29, 32</xref>
        ]. Capturing the provenance of sensemaking
often requires explicit feedback or annotations from the analyst.
Researchers often investigate “thinkaloud protocols” [
        <xref ref-type="bibr" rid="ref11">11</xref>
        ] to study
analytic processes by asking analysts to verbally describe what
they are doing or thinking over time (e.g., [
        <xref ref-type="bibr" rid="ref31">31</xref>
        ]). Unfortunately,
collecting explicit user feedback—whether from annotations or verbal
comments—can be distracting or burdensome for the analyst.
However, research has demonstrated the benefits of using interaction
behavior to learn about the analysis and sensemaking processes
(e.g., [
        <xref ref-type="bibr" rid="ref26 ref5 ref8">5, 8, 26</xref>
        ]). In this way, information about sensemaking can be
inferred automatically without requiring explicit user feedback.
      </p>
      <p>In our research, we explore summarizing analysis activity from
collected visual analytics interaction logs. Our post-hoc approach
focuses on techniques to summarize diferent periods of analysis
time based on changes in the analyst’s goals and behaviors, referred
to as inflection points. Inflection points can be indicative of changes
in a analyst’s thought process, but qualitatively identifying these
changes can be a laborious process. Finding alternative ways to
detect inflection points could help quantitatively model such
behaviors and provide additional insights to analytic provenance. The
problem is formulated as a supervised classification task,
according to which the goal is to determine the presence or absence of
inflection points from a time-sequence of eye-tracking and mouse
interaction logs over a pre-determined length of time. We capture
interaction logs and think-aloud comments from human
participants during a data analysis activity with a sample visual analytics
application. We employ a simple heuristic rationale to label
inflection points during the analysis session and extract features from
the interaction logs. Two supervised Hidden Markov Models were
trained with the features and labels to predict inflection points.</p>
      <p>We present this research as preliminary work from an analysis
of sample records collected from a seven-participant user study. We
share our results on identifying changes in analyst behavior along
with insights learned regarding the challenges of using features
from diferent interaction logs to understand sequence segments.
The results of our study can be used in future research aimed
towards explaining the training/model building process of AI.
2</p>
    </sec>
    <sec id="sec-3">
      <title>RELATED LITERATURE</title>
      <p>
        Research of analytic provenance in visual analytics broadly consists
of integrated workflow management, post processing of captured
interactions, and visualization of captured tool states and/or
interactions using a combination of analytics and sensemaking theory [
        <xref ref-type="bibr" rid="ref30">30</xref>
        ].
Some visual analytics tools utilize node-based workflows by
allowing users access to a view showing previous states and operations
used to manipulate the data flow (e.g., [
        <xref ref-type="bibr" rid="ref6 ref9">6, 9</xref>
        ]). Alternatively, the
overall visual analysis task can be composed of subtasks derived
from user interactions [
        <xref ref-type="bibr" rid="ref15 ref4">4, 15</xref>
        ]. Dou et al. [
        <xref ref-type="bibr" rid="ref8">8</xref>
        ] compared human-coded
interpretations of interaction logs to their ground truth of
thinka-loud transcriptions based on analysis findings, strategies, and
methods. While these results were based on human interpretation
of the reasoning process, they showed that there was a correlation
between interactions and cognitive reasoning. Research has also
been made in summarizing the analysis task through meta data
visualization using topic modelling for breaks in time [
        <xref ref-type="bibr" rid="ref22 ref25">22, 25</xref>
        ].
However, there is debate in determining parameters for segmentation
and deciding which interactions are most meaningful to indicate
changes in the analysis process.
      </p>
      <p>
        Given feature rich information from interaction logs, it is
possible to use machine learning techniques to gain insights on the user’s
sensemaking process. Work by Harrison et al. [
        <xref ref-type="bibr" rid="ref18">18</xref>
        ] employed
interaction features such as mouse clicks in conjunction with Hidden
Markov Models (HMMs) to predict transitions in user frustration.
Research also shows that mouse and interface interactions can be
used to predict task performance and infer user personality traits
[
        <xref ref-type="bibr" rid="ref5">5</xref>
        ]. Kodagoda et al. [
        <xref ref-type="bibr" rid="ref21">21</xref>
        ] compared machine learning techniques to
infer reasoning provenance. However, they relied on using experts
to code user’s interactions as processes based on Klein’s data frame
theory of sensemaking [
        <xref ref-type="bibr" rid="ref20">20</xref>
        ] and were not successful in identifying
all processes. Similarly, Aboufoul et. al explore using HMMs to
determine the cognitive states of users by mapping hidden states
to sensemaking processes defined by Piroli et. al. to validate their
models [
        <xref ref-type="bibr" rid="ref1">1</xref>
        ].
      </p>
      <p>
        Recent techniques introduce objective signal-based indices of
eye-tracking and mouse log data that provide an alternative view
of the interaction between the user and the visualization. The work
by Coltekin et al. [
        <xref ref-type="bibr" rid="ref7">7</xref>
        ] provide a deeper understanding of how people
make inferences and decisions when using visualizations tools and
complex displays by comparing a hypothetical sequence of
strategies to recorded sequences of participants. Eye tracking-derived
metrics can also be highly beneficial for understanding visualization
usage and analysis behavior. Eye tracking metrics often highlight
the importance of areas of interests (AOIs) by including factors such
as number of fixations within areas of the display and sequential
transitions among areas [
        <xref ref-type="bibr" rid="ref13 ref2 ref3">2, 3, 13</xref>
        ]. AOIs are specified regions of an
interface such as a particular view, button, or piece of content.
      </p>
      <p>
        In our own research, instead of employing elaborate coding
schemes that might be time-consuming and hard to replicate, we
relied solely on inflection points observed when a user changed
tactics in tool usage as labels. We included eye tracking data in addition
to mouse interaction features and tested for which features
produced the highest model precision. The extracted features required
less dependency on the structure of the visual interface compared
to previous approaches that relied on a variety of task-dependent
indices (e.g., features related to keyword highlighting, new search).
Our automated system was designed to be implemented on a variety
of visual analytics tasks, including those towards XAI (e.g., feature
analysis [
        <xref ref-type="bibr" rid="ref17">17</xref>
        ]). Having an understanding of when inflections occur
can later aid in semantic interactions [
        <xref ref-type="bibr" rid="ref10">10</xref>
        ] or developing adaptive
UIs [
        <xref ref-type="bibr" rid="ref28">28</xref>
        ].
3
      </p>
    </sec>
    <sec id="sec-4">
      <title>METHODOLOGY</title>
      <p>In our approach we attempt to identify inflection points using
mouse and eye tracking interaction logs as features to a
supervised classification model. We discuss a prior study that collected
interaction logs from a visual analysis scenario and the
associated visualization tool. A human experimenter labelled perceived
changes in behavior during analysis to create a “ground truth” set
of labels. Processed features and labels were used to train two HMM
models.
3.1</p>
    </sec>
    <sec id="sec-5">
      <title>Analysis Scenario and Interaction Logs</title>
      <p>
        Our research makes use of prior data collected from a visual
analysis scenario, available on a public repository of analytic
provenance records and interaction logs [
        <xref ref-type="bibr" rid="ref24">24</xref>
        ]. The analysis scenario
implemented used the Badge and Network Trafic dataset from the
VAST 2009 mini-challenge [
        <xref ref-type="bibr" rid="ref16">16</xref>
        ]. The challenge had two high-level
tasks regarding an embassy employee suspected of sending data
to a malicious corporation within the network. The tasks were
ifnding which computer was used to send confidential data and
characterizing patterns of suspicious computer use. Information
such as employee identification numbers, proximity log card uses,
and network trafic logs were provided for a period of one month.
      </p>
      <p>
        As the basis for data collection, the visual analytics tool was
designed to be similar to common designs of visualization tools
with multiple coordinated views used for network analysis (e.g., [
        <xref ref-type="bibr" rid="ref14 ref27">14,
27</xref>
        ]). Using this tool, test participants conducted multi-dimensional
data analysis. The tool included six views in total that supported
interactive exploration of the cyber security data. Throughout the
paper, we refer to these views as areas of interest (AOI). Five AOIs
provided interactive tabular and visual presentations of the network
data, and one AOI provided static information relevant to network
trafic (see Figure 1). They are summarized as follows:
• Detail Histogram: An interactive histogram for filtering the
period of time being analyzed by the user.
• Overview Histogram: An interactive histogram view of the
data at a more macro scale.
• Network Graph: A network graph visualization showing
network trafic between IP addresses. Mousing over a node
dimmed unconnected nodes to facilitate visual inspection.
• Ofices: A graphical layout showing employee desks and
whether or not employees were in the ofice during the time
selected via the histograms.
• Info: A static textual view with basic information about
network trafic and common usage of socket numbers.
• Table: An interactive table containing network trafic
between IP addresses. Users could click to sort the table by
attribute, select particular exchanges for further inquiry, or
iflter to a specific time.
3.2
      </p>
    </sec>
    <sec id="sec-6">
      <title>User Study for Data Collection</title>
      <p>Seven participants engaged in an analysis session using our tool
for a 90 minute duration. All participants majored in computer
science and had taken at least one university level class in computer
networking. The median age was 23 years. Participants were briefed
on the visual analysis tool and tasked with finding the computer
being used to send confidential data to malicious corporations. They
were also asked to “think out loud” by verbalizing their thoughts
throughout the process.</p>
      <p>The tool was a standalone web application viewed on a 27-inch
monitor with mouse input. Mouse interactions such as the position
of the cursor, click events, hovering, and filtering were recorded
via the application. The screen was also recorded using screen
capture software. Eye tracking data was collected using a Tobii
EyeX (70 Hz) eye tracking equipment attached to the bottom of the
monitor. A standard microphone was employed to record audio of
the think-aloud data.
3.3</p>
    </sec>
    <sec id="sec-7">
      <title>Extracting Features from User Interactions</title>
      <p>
        We categorized the interactions from the raw eye tracking x-y
coordinates and mouse events to lay out a foundation for feature
extraction. Following Brehmer et al. [
        <xref ref-type="bibr" rid="ref4">4</xref>
        ], we identified abstract analysis
tasks by methods of interaction: select, navigate, arrange, change,
iflter , and aggregate. For our study, navigation solely relied on the
x-y coordinates of the eye tracking data as a means of navigating
between AOIs. The remaining methods were used to categorize the
mouse interactions.
      </p>
      <p>To summarize the navigation data, we mapped the x-y
coordinates of the eye tracking data to the six AOIs. By examining the
ratio of fixations per participant across the AOIs (see Figure 2), we
observed that the ofices and table views had the highest
occurrence of fixations, while the network graph and detail histogram
had the least. The interactions were heavily skewed towards select
interactions (Table 1), wherein a participant either hovered over
an item for more information or selected an item to single it out
visually. This might have been a result of it being the easiest and
most natural way of viewing information while using the tool.</p>
      <p>We derived three feature sets from the described interaction
methods and AOIs. Due to the high sampling rate of the
eyetracking data in comparison with mouse interactions, we binned
interactions into analysis frames of pre-determined length
(Figure 3). Within each analysis frame we computed the number of
ifxations of the eye-tracking data per AOI, the number of mouse
interactions methods, as well as the total number of eye tracking
transitions that originated from each AOI. The maximum number
of features per frame was seventeen (six for the fixation counts
in each AOI, six for transitions between AOIs, and five for mouse
interaction methods)
Our system was designed to solve a binary classification problem
of whether or not a sequence of interactions contained an
inflection point. The inflection points were manually labeled by the
experimenter based on the participant’s think out louds and
corresponding interactions with the visualization tool as the study
was conducted. The experimenter marked an inflection when he
observed the participant changing tactics in tool usage. While the
use of one experimenter is not ideal, the experimenter took notes of
the entire session, was well versed in the tool’s usage, and listened
for verbal cues. An example of an inflection would be a participant
shifting from a period of time using the detail histogram and
scanning ofices for changes to a new strategy of looking through the
table of network trafic. There was a mean of 15 inflections per
participant, with counts ranging from 9 to 22.</p>
      <p>
        We implemented a supervised learning algorithm based on two
Hidden Markov Models (HMM): one modeling interaction sequences
that included inflections and one with sequences that did not [
        <xref ref-type="bibr" rid="ref33">33</xref>
        ].
Approximate inflection points were classified by using the forward
algorithm to determine the log-likelihood score of a sequence
occurring from each model and selecting the model with the higher score.
Specifically, we marked the time at the beginning of a sequence
as an inflection point if the inflection model had a higher score.
Sequences were formed using a sliding window of frames on the
90 minute session. The duration of the frames was always under
one minute and the number of features ranged between 5 and 17.
Figure 3 provides an example of the input sequence, that consists
of successive analysis frames. For training, if an inflection occurred
within the first half of the sequence length, the sequence would
be used to estimate the model parameters of the inflection model.
Remaining sequences were used on the non-inflection model. The
number of hidden states for each model was either 4 or 5.
4
      </p>
    </sec>
    <sec id="sec-8">
      <title>EVALUATION</title>
      <p>To further understand the impact of interactions in determining
the occurrence of inflections, experiments were performed using a
leave-one-user-out cross-validation. One participant was held out
in the test set for each fold, while the remaining data was used in
the training set. The number of folds was equivalent to the
number of users. Evaluation metrics were averaged over all folds. We
experimented with the number of hidden states N for the
inflection and non-inflection models ( N ∈ {4, 5}), frame length K (K ∈
{0.1, 0.5, 1.0}minutes), sequence length L (L ∈ {4, 8, 12}frames),
and all possible combinations of the feature sets.</p>
      <p>
        Evaluation metrics include the recall for the inflection and non
inflection classes. In addition, unweighted accuracy is computed as
the average of the separate recalls per class. Unweighted measures
are usually employed to yield unbiased estimators of the system
accuracy respective of the number of samples per class. They are
common measures in applications with unbalanced data, since they
give the same weight to each class, regardless of how many samples
of the dataset it contains [
        <xref ref-type="bibr" rid="ref12">12</xref>
        ]. Table 2 shows the results from the
highest scoring models (overall and for each feature set).
      </p>
      <p>Results indicate that sequences of 4 one-minute frames generally
had the best scores. Intuitively, this can be justified by the fact
that reducing granularity in time provides more room for error,
resulting in improved accuracy. Counts of eye tracking fixations
per AOI and mouse interactions were the most helpful feature
sets. These were the only sets used in the highest scoring system
and individually scored better than the AOI transition feature set.
The highest unweighted accuracy of our system was 70.8% with
an average of 23 classified inflection points. While the achieved
accuracy is greater than chance (50%), additional experimentation
with a greater amount of participants is required to more-reliably
classify inflection points.
5</p>
    </sec>
    <sec id="sec-9">
      <title>DISCUSSION AND CONCLUSION</title>
      <p>Visualizations used to understand, diagnose, and refine models are
a subset of “Explainable AI” (XAI). However, it is unclear what
types of interactions are appropriate for a given model and how
the visualizations are perceived. By developing methods to
understand the user’s sensemaking process in visual analytics we may
be able to gain insight on how visualizations for AI are being used
and in development of a naturalistic model of explanation. The
study of sensemaking processes in visual analytics typically
relies on having elaborate coding schemes, where human experts
with domain specific knowledge provide annotations. However,
this requires time-consuming reporting from analysts who are
burdened with explaining their rationale. Our research explores the
use of supervised classification to detect inflection points and its
features derived solely from interaction logs. The proposed
classification system of two HMM models reached an unweighted accuracy
of 70% when detecting inflection points, suggesting feasibility of
inflection-point classification.</p>
      <p>Best results were obtained using a reduced time granularity of
4minute time precision within the total 90-minute user session. The
count of AOIs from eye tracking fixations and mouse interactions
proved to be slightly more useful feature sets in comparison to
transitions between AOIs. As part of our future work, we will
experiment with time trajectory approaches that will not bin
eyetracking and mouse interaction features over an analysis frame,
as this might remove valuable sequential information. We will
also consider more detailed feature extraction techniques such as
saccade information, raw x-y coordinates, or specific parts of the
AOI to increase granularity.</p>
      <p>Going forward, we seek to increase the number of participants
for more generalizable models and to group participants based on
interaction sequences. The analysis activity per participant consisted
of highly diverse combinations of tasks and subtasks due to the
nature of complex task and unstructured data exploration. Cleaner
analysis tasks with more clearly defined subtasks may help our
understanding of participants’ inflection points and guide insights
on individual sensemaking. While we do not expect to be able to
perfectly capture the human analytic process through interaction
logs alone, the results of this work are promising for further
analysis such as comparing sequences between inflection points amongst
users or segmenting the analysis session to augment visualizations
of meta data.</p>
    </sec>
    <sec id="sec-10">
      <title>ACKNOWLEDGMENTS</title>
      <p>This material is based on work supported by NSF 1565725 and the
DARPA XAI program N66001-17-2-4031.</p>
    </sec>
  </body>
  <back>
    <ref-list>
      <ref id="ref1">
        <mixed-citation>
          [1]
          <string-name>
            <given-names>Mohamad</given-names>
            <surname>Aboufoul</surname>
          </string-name>
          , Ryan Wesslen, Isaac Cho, Wenwen Dou, and
          <string-name>
            <given-names>Samira</given-names>
            <surname>Shaikh</surname>
          </string-name>
          .
          <year>2018</year>
          .
          <article-title>Using Hidden Markov Models to Determine Cognitive States of Visual Analytic Users</article-title>
          . (
          <year>2018</year>
          ).
        </mixed-citation>
      </ref>
      <ref id="ref2">
        <mixed-citation>
          [2]
          <string-name>
            <given-names>Tanja</given-names>
            <surname>Blascheck</surname>
          </string-name>
          ,
          <string-name>
            <surname>Markus John</surname>
          </string-name>
          , Kuno Kurzhals, Stefen Koch, and
          <string-name>
            <given-names>Thomas</given-names>
            <surname>Ertl</surname>
          </string-name>
          .
          <year>2016</year>
          .
          <article-title>VA 2: a visual analytics approach for evaluating visual analytics applications</article-title>
          .
          <source>IEEE transactions on visualization and computer graphics 22</source>
          ,
          <issue>1</issue>
          (
          <year>2016</year>
          ),
          <fpage>61</fpage>
          -
          <lpage>70</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref3">
        <mixed-citation>
          [3]
          <string-name>
            <given-names>Tanja</given-names>
            <surname>Blascheck</surname>
          </string-name>
          , Kuno Kurzhals,
          <string-name>
            <given-names>Michael</given-names>
            <surname>Raschke</surname>
          </string-name>
          , Michael Burch, Daniel Weiskopf, and
          <string-name>
            <given-names>Thomas</given-names>
            <surname>Ertl</surname>
          </string-name>
          .
          <year>2014</year>
          .
          <article-title>State-of-the-art of visualization for eye tracking data</article-title>
          .
          <source>In Proceedings of EuroVis</source>
          , Vol.
          <year>2014</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref4">
        <mixed-citation>
          [4]
          <string-name>
            <given-names>Matthew</given-names>
            <surname>Brehmer</surname>
          </string-name>
          and
          <string-name>
            <given-names>Tamara</given-names>
            <surname>Munzner</surname>
          </string-name>
          .
          <year>2013</year>
          .
          <article-title>A multi-level typology of abstract visualization tasks</article-title>
          .
          <source>IEEE Transactions on Visualization and Computer Graphics</source>
          <volume>19</volume>
          ,
          <issue>12</issue>
          (
          <year>2013</year>
          ),
          <fpage>2376</fpage>
          -
          <lpage>2385</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref5">
        <mixed-citation>
          [5]
          <string-name>
            <surname>Eli</surname>
            <given-names>T Brown</given-names>
          </string-name>
          , Alvitta Ottley,
          <string-name>
            <given-names>Helen</given-names>
            <surname>Zhao</surname>
          </string-name>
          , Quan Lin, Richard Souvenir, Alex Endert, and
          <string-name>
            <given-names>Remco</given-names>
            <surname>Chang</surname>
          </string-name>
          .
          <year>2014</year>
          .
          <article-title>Finding waldo: Learning about users from their interactions</article-title>
          .
          <source>IEEE Transactions on visualization and computer graphics 20</source>
          ,
          <issue>12</issue>
          (
          <year>2014</year>
          ),
          <fpage>1663</fpage>
          -
          <lpage>1672</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref6">
        <mixed-citation>
          [6]
          <string-name>
            <surname>Steven</surname>
            <given-names>P Callahan</given-names>
          </string-name>
          , Juliana Freire, Emanuele Santos, Carlos E Scheidegger,
          <string-name>
            <surname>Cláudio T Silva</surname>
          </string-name>
          , and
          <string-name>
            <surname>Huy T Vo</surname>
          </string-name>
          .
          <year>2006</year>
          .
          <article-title>VisTrails: visualization meets data management</article-title>
          .
          <source>In Proceedings of the 2006 ACM SIGMOD international conference on Management of data. ACM</source>
          ,
          <volume>745</volume>
          -
          <fpage>747</fpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref7">
        <mixed-citation>
          [7]
          <string-name>
            <given-names>Arzu</given-names>
            <surname>Çöltekin</surname>
          </string-name>
          , Sara Irina Fabrikant, and
          <string-name>
            <given-names>Martin</given-names>
            <surname>Lacayo</surname>
          </string-name>
          .
          <year>2010</year>
          .
          <article-title>Exploring the eficiency of users' visual analytics strategies based on sequence analysis of eye movement recordings</article-title>
          .
          <source>International Journal of Geographical Information Science</source>
          <volume>24</volume>
          ,
          <issue>10</issue>
          (
          <year>2010</year>
          ),
          <fpage>1559</fpage>
          -
          <lpage>1575</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref8">
        <mixed-citation>
          [8]
          <string-name>
            <given-names>Wenwen</given-names>
            <surname>Dou</surname>
          </string-name>
          , Dong Hyun Jeong, Felesia Stukes,
          <string-name>
            <given-names>William</given-names>
            <surname>Ribarsky</surname>
          </string-name>
          , Heather Richter Lipford, and
          <string-name>
            <given-names>Remco</given-names>
            <surname>Chang</surname>
          </string-name>
          .
          <year>2009</year>
          .
          <article-title>Recovering reasoning processes from user interactions</article-title>
          .
          <source>IEEE Computer Graphics and Applications</source>
          <volume>29</volume>
          ,
          <issue>3</issue>
          (
          <year>2009</year>
          ).
        </mixed-citation>
      </ref>
      <ref id="ref9">
        <mixed-citation>
          [9]
          <string-name>
            <given-names>Cody</given-names>
            <surname>Dunne</surname>
          </string-name>
          , Nathalie Henry Riche,
          <string-name>
            <given-names>Bongshin</given-names>
            <surname>Lee</surname>
          </string-name>
          ,
          <string-name>
            <given-names>Ronald</given-names>
            <surname>Metoyer</surname>
          </string-name>
          , and
          <string-name>
            <given-names>George</given-names>
            <surname>Robertson</surname>
          </string-name>
          .
          <year>2012</year>
          .
          <article-title>GraphTrail: Analyzing large multivariate, heterogeneous networks while supporting exploration history</article-title>
          .
          <source>In Proceedings of the SIGCHI conference on human factors in computing systems. ACM</source>
          ,
          <volume>1663</volume>
          -
          <fpage>1672</fpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref10">
        <mixed-citation>
          [10]
          <string-name>
            <surname>Alex</surname>
            <given-names>Endert</given-names>
          </string-name>
          , Patrick Fiaux, and Chris North.
          <year>2012</year>
          .
          <article-title>Semantic interaction for sensemaking: inferring analytical reasoning for model steering</article-title>
          .
          <source>IEEE Transactions on Visualization and Computer Graphics</source>
          <volume>18</volume>
          ,
          <issue>12</issue>
          (
          <year>2012</year>
          ),
          <fpage>2879</fpage>
          -
          <lpage>2888</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref11">
        <mixed-citation>
          [11]
          <string-name>
            <given-names>K Anders</given-names>
            <surname>Ericsson and Herbert A Simon</surname>
          </string-name>
          .
          <year>1998</year>
          .
          <article-title>How to study thinking in everyday life: Contrasting think-aloud protocols with descriptions and explanations of thinking</article-title>
          .
          <source>Mind, Culture, and Activity 5</source>
          ,
          <issue>3</issue>
          (
          <year>1998</year>
          ),
          <fpage>178</fpage>
          -
          <lpage>186</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref12">
        <mixed-citation>
          [12]
          <string-name>
            <given-names>Michael</given-names>
            <surname>Glodek</surname>
          </string-name>
          , Stephan Tschechne, Georg Layher, Martin Schels, Tobias Brosch, Stefan Scherer, Markus Kächele, Miriam Schmidt, Heiko Neumann,
          <string-name>
            <given-names>Günther</given-names>
            <surname>Palm</surname>
          </string-name>
          , et al.
          <year>2011</year>
          .
          <article-title>Multiple classifier systems for the classification of audio-visual emotional states</article-title>
          .
          <source>In Afective Computing and Intelligent Interaction</source>
          . Springer,
          <fpage>359</fpage>
          -
          <lpage>368</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref13">
        <mixed-citation>
          [13]
          <string-name>
            <surname>Joseph</surname>
            <given-names>H</given-names>
          </string-name>
          <string-name>
            <surname>Goldberg and Jonathan I Helfman</surname>
          </string-name>
          .
          <year>2010</year>
          .
          <article-title>Comparing information graphics: a critical look at eye tracking</article-title>
          .
          <source>In Proceedings of the 3rd BELIV'10 Workshop</source>
          :
          <article-title>BEyond time and errors: novel evaLuation methods for Information Visualization</article-title>
          . ACM,
          <volume>71</volume>
          -
          <fpage>78</fpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref14">
        <mixed-citation>
          [14]
          <string-name>
            <surname>John</surname>
            <given-names>R Goodall</given-names>
          </string-name>
          ,
          <string-name>
            <given-names>Eric D</given-names>
            <surname>Ragan</surname>
          </string-name>
          , Chad A Steed, Joel W Reed, G David Richardson,
          <source>Kelly MT Hufer, Robert A Bridges, and Jason A Laska</source>
          .
          <year>2018</year>
          .
          <article-title>Situ: Identifying and explaining suspicious behavior in networks</article-title>
          .
          <source>IEEE transactions on visualization and computer graphics</source>
          (
          <year>2018</year>
          ).
        </mixed-citation>
      </ref>
      <ref id="ref15">
        <mixed-citation>
          [15]
          <string-name>
            <given-names>David</given-names>
            <surname>Gotz</surname>
          </string-name>
          and
          <string-name>
            <surname>Michelle X Zhou</surname>
          </string-name>
          .
          <year>2009</year>
          .
          <article-title>Characterizing users' visual analytic activity for insight provenance</article-title>
          .
          <source>Information Visualization 8</source>
          ,
          <issue>1</issue>
          (
          <year>2009</year>
          ),
          <fpage>42</fpage>
          -
          <lpage>55</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref16">
        <mixed-citation>
          [16]
          <string-name>
            <surname>Georges</surname>
            <given-names>G Grinstein</given-names>
          </string-name>
          ,
          <article-title>Jean Scholtz, Mark A Whiting,</article-title>
          and
          <string-name>
            <given-names>Catherine</given-names>
            <surname>Plaisant</surname>
          </string-name>
          .
          <year>2009</year>
          .
          <article-title>VAST 2009 challenge: An insider threat</article-title>
          ..
          <source>In IEEE VAST</source>
          .
          <volume>243</volume>
          -
          <fpage>244</fpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref17">
        <mixed-citation>
          [17]
          <string-name>
            <given-names>David</given-names>
            <surname>Gunning</surname>
          </string-name>
          .
          <year>2017</year>
          .
          <article-title>Explainable artificial intelligence (xai)</article-title>
          .
          <source>Defense Advanced Research Projects Agency (DARPA)</source>
          , nd
          <string-name>
            <surname>Web</surname>
          </string-name>
          (
          <year>2017</year>
          ).
        </mixed-citation>
      </ref>
      <ref id="ref18">
        <mixed-citation>
          [18]
          <string-name>
            <surname>Lane</surname>
            <given-names>Harrison</given-names>
          </string-name>
          , Wenwen Dou, Aidong Lu,
          <string-name>
            <given-names>William</given-names>
            <surname>Ribarsky</surname>
          </string-name>
          , and
          <string-name>
            <given-names>Xiaoyu</given-names>
            <surname>Wang</surname>
          </string-name>
          .
          <year>2011</year>
          .
          <article-title>Analysts aren't machines: Inferring frustration through visualization interaction</article-title>
          .
          <source>In Visual Analytics Science and Technology (VAST)</source>
          ,
          <source>2011 IEEE Conference on. IEEE</source>
          ,
          <fpage>279</fpage>
          -
          <lpage>280</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref19">
        <mixed-citation>
          [19]
          <string-name>
            <surname>Robert</surname>
            <given-names>R Hofman</given-names>
          </string-name>
          , Gary Klein, and
          <string-name>
            <surname>Shane T Mueller</surname>
          </string-name>
          .
          <year>2018</year>
          .
          <article-title>Explaining Explanation For “Explainable AI”</article-title>
          .
          <source>In Proceedings of the Human Factors and Ergonomics Society Annual Meeting</source>
          , Vol.
          <volume>62</volume>
          . SAGE Publications Sage CA: Los Angeles, CA,
          <fpage>197</fpage>
          -
          <lpage>201</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref20">
        <mixed-citation>
          [20]
          <string-name>
            <surname>Gary</surname>
            <given-names>Klein</given-names>
          </string-name>
          , Jennifer K Phillips, Erica L Rall, and Deborah A Peluso.
          <year>2007</year>
          .
          <article-title>A data-frame theory of sensemaking. In Expertise out of context: Proceedings of the sixth international conference on naturalistic decision making</article-title>
          . New York, NY, USA: Lawrence Erlbaum,
          <fpage>113</fpage>
          -
          <lpage>155</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref21">
        <mixed-citation>
          [21]
          <string-name>
            <surname>Neesha</surname>
            <given-names>Kodagoda</given-names>
          </string-name>
          , Sheila Pontis, Donal Simmie, Simon Attfield, BL William Wong, Ann Blandford, and
          <string-name>
            <given-names>Chris</given-names>
            <surname>Hankin</surname>
          </string-name>
          .
          <year>2017</year>
          .
          <article-title>Using machine learning to infer reasoning provenance from user interaction log data: based on the data/frame theory of sensemaking</article-title>
          .
          <source>Journal of Cognitive Engineering and Decision Making</source>
          <volume>11</volume>
          ,
          <issue>1</issue>
          (
          <year>2017</year>
          ),
          <fpage>23</fpage>
          -
          <lpage>41</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref22">
        <mixed-citation>
          [22]
          <string-name>
            <surname>Rhema</surname>
            <given-names>Linder</given-names>
          </string-name>
          , Alyssa M Pena,
          <string-name>
            <given-names>Sampath</given-names>
            <surname>Jayarathna</surname>
          </string-name>
          , and
          <string-name>
            <given-names>Eric D</given-names>
            <surname>Ragan</surname>
          </string-name>
          .
          <year>2016</year>
          .
          <article-title>Results and challenges in visualizing analytic provenance of text analysis tasks using interaction logs</article-title>
          .
          <source>In Logging Interactive Visualizations and Visualizing Interaction Logs (LIVVIL) Workshop.</source>
        </mixed-citation>
      </ref>
      <ref id="ref23">
        <mixed-citation>
          [23]
          <string-name>
            <surname>Shixia</surname>
            <given-names>Liu</given-names>
          </string-name>
          , Xiting Wang, Mengchen Liu, and
          <string-name>
            <given-names>Jun</given-names>
            <surname>Zhu</surname>
          </string-name>
          .
          <year>2017</year>
          .
          <article-title>Towards better analysis of machine learning models: A visual analytics perspective</article-title>
          .
          <source>Visual Informatics</source>
          <volume>1</volume>
          ,
          <issue>1</issue>
          (
          <year>2017</year>
          ),
          <fpage>48</fpage>
          -
          <lpage>56</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref24">
        <mixed-citation>
          [24]
          <string-name>
            <surname>Sina</surname>
            <given-names>Mohseni</given-names>
          </string-name>
          , Andrew Pachuilo, Ehsanul Haque Nirjhar, Rhema Linder, Alyssa Pena, and
          <string-name>
            <given-names>Eric D</given-names>
            <surname>Ragan</surname>
          </string-name>
          .
          <year>2018</year>
          .
          <article-title>Analytic Provenance Datasets: A Data Repository of Human Analysis Activity and Interaction Logs</article-title>
          . arXiv preprint arXiv:
          <year>1801</year>
          .
          <volume>05076</volume>
          (
          <year>2018</year>
          ).
        </mixed-citation>
      </ref>
      <ref id="ref25">
        <mixed-citation>
          [25]
          <string-name>
            <surname>Sina</surname>
            <given-names>Mohseni</given-names>
          </string-name>
          , Alyssa Pena, and
          <string-name>
            <given-names>Eric D</given-names>
            <surname>Ragan</surname>
          </string-name>
          .
          <year>2018</year>
          .
          <article-title>ProvThreads: Analytic Provenance Visualization and Segmentation</article-title>
          . arXiv preprint arXiv:
          <year>1801</year>
          .
          <volume>05469</volume>
          (
          <year>2018</year>
          ).
        </mixed-citation>
      </ref>
      <ref id="ref26">
        <mixed-citation>
          [26]
          <string-name>
            <surname>Phong</surname>
            <given-names>H Nguyen</given-names>
          </string-name>
          , Kai Xu, Ashley Wheat, BL William Wong, Simon Attfield, and
          <string-name>
            <given-names>Bob</given-names>
            <surname>Fields</surname>
          </string-name>
          .
          <year>2016</year>
          .
          <article-title>SensePath: Understanding the sensemaking process through analytic provenance</article-title>
          .
          <source>IEEE transactions on visualization and computer graphics 22</source>
          ,
          <issue>1</issue>
          (
          <year>2016</year>
          ),
          <fpage>41</fpage>
          -
          <lpage>50</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref27">
        <mixed-citation>
          [27]
          <string-name>
            <surname>Steven</surname>
            <given-names>Noel</given-names>
          </string-name>
          , Michael Jacobs, Pramod Kalapa, and
          <string-name>
            <given-names>Sushil</given-names>
            <surname>Jajodia</surname>
          </string-name>
          .
          <year>2005</year>
          .
          <article-title>Multiple coordinated views for network attack graphs</article-title>
          .
          <source>In Visualization for Computer Security</source>
          ,
          <year>2005</year>
          .
          <article-title>(VizSEC 05)</article-title>
          . IEEE Workshop on. IEEE,
          <fpage>99</fpage>
          -
          <lpage>106</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref28">
        <mixed-citation>
          [28]
          <string-name>
            <given-names>Andrew</given-names>
            <surname>Pachuilo</surname>
          </string-name>
          ,
          <string-name>
            <surname>Eric D Ragan</surname>
            ,
            <given-names>and John R Goodall.</given-names>
          </string-name>
          <year>2016</year>
          .
          <article-title>Leveraging Interaction History for Intelligent Configuration of Multiple Coordinated Views in Visualization Tools</article-title>
          . LIVVIL:
          <article-title>Logging Interactive Visualizations &amp; Visualizing Interaction Logs (</article-title>
          <year>2016</year>
          ).
        </mixed-citation>
      </ref>
      <ref id="ref29">
        <mixed-citation>
          [29]
          <string-name>
            <given-names>Peter</given-names>
            <surname>Pirolli</surname>
          </string-name>
          and
          <string-name>
            <given-names>Stuart</given-names>
            <surname>Card</surname>
          </string-name>
          .
          <year>2005</year>
          .
          <article-title>The sensemaking process and leverage points for analyst technology as identified through cognitive task analysis</article-title>
          .
          <source>In Proceedings of international conference on intelligence analysis</source>
          , Vol.
          <volume>5</volume>
          .
          <string-name>
            <surname>McLean</surname>
            ,
            <given-names>VA</given-names>
          </string-name>
          , USA,
          <fpage>2</fpage>
          -
          <lpage>4</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref30">
        <mixed-citation>
          [30]
          <string-name>
            <surname>Eric</surname>
            <given-names>D</given-names>
          </string-name>
          <string-name>
            <surname>Ragan</surname>
            , Alex Endert, Jibonananda Sanyal, and
            <given-names>Jian</given-names>
          </string-name>
          <string-name>
            <surname>Chen</surname>
          </string-name>
          .
          <year>2016</year>
          .
          <article-title>Characterizing provenance in visualization and data analysis: an organizational framework of provenance types and purposes</article-title>
          .
          <source>IEEE transactions on visualization and computer graphics 22</source>
          ,
          <issue>1</issue>
          (
          <year>2016</year>
          ),
          <fpage>31</fpage>
          -
          <lpage>40</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref31">
        <mixed-citation>
          [31]
          <string-name>
            <surname>Eric</surname>
            <given-names>D</given-names>
          </string-name>
          <string-name>
            <surname>Ragan</surname>
          </string-name>
          and John R Goodall.
          <year>2014</year>
          .
          <article-title>Evaluation methodology for comparing memory and communication of analytic processes in visual analytics</article-title>
          .
          <source>In Proceedings of the Fifth Workshop on Beyond Time and Errors: Novel Evaluation Methods for Visualization. ACM</source>
          ,
          <volume>27</volume>
          -
          <fpage>34</fpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref32">
        <mixed-citation>
          [32]
          <string-name>
            <surname>Dominik</surname>
            <given-names>Sacha</given-names>
          </string-name>
          , Andreas Stofel, Florian Stofel, Bum Chul Kwon, Geofrey Ellis, and Daniel A Keim.
          <year>2014</year>
          .
          <article-title>Knowledge generation model for visual analytics</article-title>
          .
          <source>IEEE transactions on visualization and computer graphics 20</source>
          ,
          <issue>12</issue>
          (
          <year>2014</year>
          ),
          <fpage>1604</fpage>
          -
          <lpage>1613</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref33">
        <mixed-citation>
          [33]
          <string-name>
            <surname>Dario</surname>
            <given-names>D</given-names>
          </string-name>
          <string-name>
            <surname>Salvucci</surname>
          </string-name>
          and
          <string-name>
            <surname>Joseph H Goldberg</surname>
          </string-name>
          .
          <year>2000</year>
          .
          <article-title>Identifying fixations and saccades in eye-tracking protocols</article-title>
          .
          <source>In Proceedings of the 2000 symposium on Eye tracking research &amp; applications. ACM</source>
          ,
          <volume>71</volume>
          -
          <fpage>78</fpage>
          .
        </mixed-citation>
      </ref>
    </ref-list>
  </back>
</article>