<!DOCTYPE article PUBLIC "-//NLM//DTD JATS (Z39.96) Journal Archiving and Interchange DTD v1.0 20120330//EN" "JATS-archivearticle1.dtd">
<article xmlns:xlink="http://www.w3.org/1999/xlink">
  <front>
    <journal-meta />
    <article-meta>
      <title-group>
        <article-title>A New Good Listener, the Digital Human: A comparative research analysis of conversational virtual agents and robots</article-title>
      </title-group>
      <contrib-group>
        <contrib contrib-type="author">
          <string-name>Boeun Kwak</string-name>
          <email>kwakboeun@kookmin.ac.kr</email>
          <xref ref-type="aff" rid="aff0">0</xref>
        </contrib>
        <contrib contrib-type="author">
          <string-name>Jeongyun Heo</string-name>
          <xref ref-type="aff" rid="aff0">0</xref>
        </contrib>
        <contrib contrib-type="author">
          <string-name>Eunsoon You</string-name>
          <xref ref-type="aff" rid="aff0">0</xref>
        </contrib>
        <aff id="aff0">
          <label>0</label>
          <institution>Kookmin University</institution>
          ,
          <addr-line>Seoul</addr-line>
          ,
          <country>Republic of Korea</country>
        </aff>
      </contrib-group>
      <fpage>37</fpage>
      <lpage>48</lpage>
      <abstract>
        <p>This paper aims to discover the potential of the digital human to develop as a listener and the ability to generate appropriate non-verbal feedback. We look at what aspects of the current digital human are easier to interact with compared with older robots or traditional virtual agents. We examine comparative studies of conversational virtual agents and robots in various contexts and review previous studies investigating non-verbal expressions and characteristics. Based on the research results, four major listener response functions of digital humans are proposed.</p>
      </abstract>
      <kwd-group>
        <kwd>Digital Human</kwd>
        <kwd>Conversation</kwd>
        <kwd>Robot</kwd>
        <kwd>virtual agent</kwd>
        <kwd>Listener Feedback</kwd>
        <kwd>Non-verbal</kwd>
      </kwd-group>
    </article-meta>
  </front>
  <body>
    <sec id="sec-1">
      <title>-</title>
      <p>
        Since the advent of computers, the scope of the conversation partner in the human–
artificial agent dialogue system has been developed in various ways. The external
appearance of a digital human has reached a level high enough to be recognized as a real
human, and we often encounter them on the Internet, kiosks, and TV screens. In
addition to message information, human dialogue interactions include tone, pitch, and
nonverbal dialogue cues that constitute the context of speech intent and contain emotional
expressions.[
        <xref ref-type="bibr" rid="ref1">1</xref>
        ] Therefore, human–digital human interaction should be as natural and
intuitive as actual human interaction to avoid miscommunication. If a virtual agent’s
appearance is unnatural, it may offend the user’s feelings (e.g., the Uncanny Valley
Effect). In this sense, a digital human should resemble a real human being and produce
natural sounding/nonverbal responses to the human user. Nonverbal communication
makes up a large proportion of human-to-human communication, comprising about
two-thirds of human-to-human contact Nonverbal expressions are made up of the gaze,
facial expression, and gestures.[
        <xref ref-type="bibr" rid="ref2">2</xref>
        ] In addition to verbal expressions, nonverbal
expressions provide important communication functions such as providing information ahead
of spoken language in face-to-face interactions, controlling interactions, and expressing
intimacy. Therefore, the perception and social effects of digital human behavior should
be treated as important.[
        <xref ref-type="bibr" rid="ref3 ref4">3, 4</xref>
        ]
      </p>
      <p>
        Nonverbal expressions can sometimes be misleading in meaning transfer because
they are represented by symbolic symbols reflecting the culture of each country. People
also express their feelings unconsciously and instinctively. Because nonverbal
expressions appear as symbolic symbols reflecting the culture of each country, it can
sometimes cause misunderstandings in conveying meaning. Sometimes, the latent content of
communication can play a more decisive role through these unrecognized nonverbal
expressions. Information that is not conveyed through language gives the impression
of more than the cognitive activity required for language generation. It can help
improve the reliability of digital humans by giving an impression of the mind.[
        <xref ref-type="bibr" rid="ref5">5</xref>
        ] In
addition, related studies have found that nonverbal communication can improve the
likeability of, interest in, and satisfaction with virtual agents.[
        <xref ref-type="bibr" rid="ref6 ref7">6, 7</xref>
        ] Natural nonverbal
communication serves as a key function of relationship formation and is a means by which
communication can embody information beyond messages. In particular, since digital
humans have a higher degree of freedom of expression of emotions than other types of
conversational agents and have the characteristics of manipulation in digital space, it is
expected that the nonverbal expressions of digital humans will have a significant impact
on design.
      </p>
      <p>The purpose of this paper is to find out why digital humans as virtual agents have
greater potential as future conversation partners than robots through a case study
comparing virtual agents and robots. This paper consists of the following: Section 2. A
comparative study case analysis of virtual agents and robots. Section 3. Deriving the
strengths as a good listener of the digital human based on the results of research case
analysis. Section 4. Proposed the response four functions of a digital human as a listener
(good listener). Section 5. Discussion, conclusions, and future work.
2
2.1</p>
    </sec>
    <sec id="sec-2">
      <title>Related Work</title>
      <sec id="sec-2-1">
        <title>Concept and Application of Digital Human and Robot</title>
        <p>
          A digital human (or, virtual human) is an artificial agent with both a human-like body
(expression or natural body movement) and intelligent, cognitively-driven behavior.[
          <xref ref-type="bibr" rid="ref8">8</xref>
          ]
A set of joints is added to the 3D face for expression and movement. The 3D face has
features such as eyes, teeth, tongue, and skin. Current research on virtual agents is
largely conducted in four areas: environmental design, training, culture and education,
and medical care.[
          <xref ref-type="bibr" rid="ref9">9</xref>
          ] Recently, digital humans have helped humans by taking various
roles such as advertising models (e.g., Lil Miquela, Oh Rozy, Imma), virtual idols,
teachers, counselors, coaches, and bankers. Digital human production companies aim
to replace most corporate chatbot services with a digital human. Digital human
production companies aim to replace most corporate chatbot services with digital humans.
Digital humans cannot perform tasks beyond the environment outside the interface. The
representative virtual agent Greta can display hand gestures, but her lower body is
motionless, i.e., the activity space is limited to the virtual world. If digital humans can be
used to perform tasks or collaborate in real environments for humans, it is expected that
the scope of their contribution will be expanded further than now. Robots can have a
variety of capabilities, mimicking human emotional states, intention and behavior
recognition, interpretation of contextual information, communication, and contextual
behavior.[
          <xref ref-type="bibr" rid="ref10">10</xref>
          ] Current applications of robots include a variety of areas such as
counseling,[
          <xref ref-type="bibr" rid="ref11">11</xref>
          ] education and training,[
          <xref ref-type="bibr" rid="ref12">12</xref>
          ] security and rescue operations,[
          <xref ref-type="bibr" rid="ref13">13</xref>
          ] social services
and business,[
          <xref ref-type="bibr" rid="ref14">14</xref>
          ] entertainment,[
          <xref ref-type="bibr" rid="ref15">15</xref>
          ] and industrial assistance,[
          <xref ref-type="bibr" rid="ref16">16</xref>
          ] Robots are
becoming more and more advanced in human interaction with humans and compared to the
virtual agent, the biggest advantage is that they can exhibit a physical presence. They
are now designed to be supported in a personal environment, such as at home. However,
the application of robots is still limited in cooperating with humans or in carrying out
social tasks for human welfare. This is because robots lack both physical dexterity and
expressive ability to imitate simple expressions. In general, robots must meet space and
cost requirements and have less hand degree of freedom,[
          <xref ref-type="bibr" rid="ref17">17</xref>
          ] The expressions displayed
through most postures and facial expressions are also limited, and only a few systems
can respond to visual response demands. Physical limitations include the robot’s angle,
joint speed and torque limitations, awkward arm composition or trajectory, and
excessively fast movement. Performing tasks using these action systems can make humans
uncertain and anxious. Even state-of-the-art humanoid robots are still unnatural,
unhuman, and expensive.[
          <xref ref-type="bibr" rid="ref18">18</xref>
          ]
2.2
        </p>
      </sec>
      <sec id="sec-2-2">
        <title>The concept and function of the listener’s nonverbal response</title>
        <p>
          Listener feedback may be defined as a response by the listener to the content of the
speaker’s utterance. The listener can switch to the speaker’s position at any time and
can express what they think or feel. During the speaker’s turn, they can provide
feedback without interfering with the utterance.[
          <xref ref-type="bibr" rid="ref19">19</xref>
          ] The expression of the listener’s
reaction is collectively referred to by various terms such as listener feedback, listener
response, backchannel, and nonverbal communication strategy (NVCS). It is used as
feedback on receiving the communicative behavior of the interlocutor, and through
language and gestures, the listener can indicate the level of participation in the
conversation. For example, the speaker may stop the conversation or restructure the sentence if
the listener is not interested.[
          <xref ref-type="bibr" rid="ref20">20</xref>
          ] Among many feedback types, listener feedback is an
important feature in face-to-face interactions because it represents the willingness to
continue to hear or invites speakers to continue with the conversation.[
          <xref ref-type="bibr" rid="ref21">21</xref>
          ] It can also
be used to express evaluations such as surprise, interest, and sympathy. If there is no
feedback, the speaker may feel anxious about whether the communication is going well
and the listener may feel as if they are talking to a hard “machine,” so it should be
handled with interest in the communication process. The listener can switch to the
speaker’s position at any time, express what he or she thinks or feels, and can provide
feedback without interfering with the utterance during the presenter’s turn.[
          <xref ref-type="bibr" rid="ref19">19</xref>
          ] Fig. 1
reconstructed a Shannon-Weaver-based model to supplement our current
understanding of some basic concepts. The listener creates the meaning as code and then transmits
it through the message. The speaker receives the code and understands the meaning.
The speaker sends a code back to the listener’s area in response.[
          <xref ref-type="bibr" rid="ref22">22</xref>
          ] In this way, the
listener and the speaker are mutually cyclical, suggesting that if you become a good
listener, you can become a good speaker at the same time. This paper focuses on the
nonverbal listener feedback as the ultimate goal of digital human development as a
dialogue listener and the ability to generate appropriate nonverbal expressions.
        </p>
      </sec>
    </sec>
    <sec id="sec-3">
      <title>Method</title>
      <sec id="sec-3-1">
        <title>Cases of Comparative Studies of the Virtual Agents and Robots</title>
        <p>In human–robot interaction (HRI) studies, there have often been comparative studies
of physical and virtual implementations; therefore, we assume that the more
humanrelated social characteristics digital humans present, the more likely they will lead to
natural communication, so we would like to look at comparative studies of existing
virtual agents and robots. In the following, we list and explain the results of various
previous studies focusing on communication between robots and virtual agents.</p>
        <p>
          [
          <xref ref-type="bibr" rid="ref23">23</xref>
          ] that when giving recommendations to users in a color selection experiment,
robots were less convincing than virtual characters on the screen. Participants had to
choose one of four colored square names displayed on a computer monitor. Before
making a decision, a robot or virtual character recommended one option to the user,
noting that it was the option chosen by other users. As a result, participants followed
virtual character recommendations more than robots. In the post-questionnaire
response, the subjective “familiarity” factor was found to be different from the subject’s
behavior. The familiarity factor of the robot group was much stronger than that of the
agent group, and the subject accepted more recommendations from the agent.
        </p>
        <p>
          [
          <xref ref-type="bibr" rid="ref24">24</xref>
          ] compared people’s responses to robots, projected robots, and agents in health
interviews to help them understand differences in people’s social interactions with
agents and robots. The researchers hypothesized that robots would have more social
impact than agents simply because of their physical proximity. The results showed that
the robot had more social impact, but the participants who interacted with the agent
remembered more key information in the recall test than the participants who interacted
with the robot. studied whether humanoid robots in real life could elicit stronger
anthropomorphic interactions than software agents and whether physical presence
modulates this effect. The researchers predicted that subjects would anthropomorphize with
more anthropomorphic humanoid robots than less anthropomorphic agents. As a result,
the participant interacted more with the robot as a person than with the agent, and the
more anthropomorphic, the more subjects treated artificial agents as a person.
        </p>
        <p>
          [
          <xref ref-type="bibr" rid="ref25">25</xref>
          ] demonstrated the importance of non-functional aspects that can enhance the
level of enjoyment and social presence of older people. It was hypothesized that the
more natural and human the conversation with the conversational agent, the higher the
perceived pleasure and acceptance. The virtual agent used in this study is “Steffie,” and
Steffie is a virtual 3D agent that can use various facial expressions, hand/arm gestures,
lip-syncing, and voice repetition in the form of a woman. The robot used Philips
Electronics’ iCat. iCat can make a variety of facial expressions using lips, eyes, eyelids, and
eyebrows, has a female voice and is in the shape of a cat. Statistics show a stronger
relationship between intention and use of virtual agents than robots. M Heerink
revealed that the two agents could not explain why the virtual agent had a stronger
influence on intention and use than the robot because of the fundamental difference in the
appearance and action system of the two agents.
        </p>
        <p>
          [
          <xref ref-type="bibr" rid="ref26">26</xref>
          ] studied how the physical presence of a robot affects human judgments about a
robot as a social partner. Subjects participated in a simple book-moving task with either
a physically present robot or a humanoid robot displayed via live video. The Nico robot,
which was used in the experiment, was a humanoid robot in the upper body, wearing
children’s sportswear and a baseball cap. Nico’s head has a total of seven degrees of
freedom and six degrees of freedom (two on the shoulder, elbow and wrist) on each
arm. In the experiment, subjects easily approached Nico in video and augmented
conditions, while avoiding face-to-face encounters with the physically present Nico. The
researcher identified two causes for these results. First, physical robots can be perceived
as more expensive than monitors used in video display conditions, so robots may be
reluctant to come closer. Second, the granting of personal space between the robot and
the subject can be interpreted as a sign of respect. However, space was also created
between Nico in video and augmented conditions, indicating that the first case would
be more appropriate.
        </p>
        <p>
          [
          <xref ref-type="bibr" rid="ref13">13</xref>
          ] invited Brazilian subjects to interact with two types of receptionists with
different appearances (agent vs mechanical robot) and voices (human vs mechanical) to
investigate factors related to designing a receptionist robot for deployment in Brazil. In
the interaction experiment, participants interacted with two receptionists with different
characteristics (a conversational virtual agent and a humanoid robot) and voice
(humanlike vs robot). Two receptionists directed the participants to a specific room where the
assessment was conducted via questionnaire. The researchers found that when
comparing Ana and Kobiana through all categories of questions, they preferred Ana in both
groups of participants and that the main reason was its human appearance.
3.2
        </p>
      </sec>
      <sec id="sec-3-2">
        <title>Cases of Comparative Study of Virtual Agents and Robots</title>
        <p>Through the review of related prior studies, we derived the following.
1. Although it was shown that people had stronger behavioral and attitudinal responses
to physically existing agents, when both physically implemented agents and virtually
implemented agents were presented, each had different results depending on the
purpose of the study (Table 1). Depending on the appearance of the virtual agent and its
degree of freedom to express, it is assumed that each resulted from a different result.
2. The nonverbal expression of the virtual agent usually has a positive effect on users
in the experiment, but unnatural expression or repetition may give a feeling of dis-
traction or discomfort. Therefore, it is necessary to provide natural and appropriate
feedback.
3. Users expect a natural conversation response from these anthropomorphic agents.
Therefore, the more similar to a person the agent is, the more likely the user will be to
treat the agent as a person.
4. Finally, since the difference in the social reality given by the implementation
environment is greater than the difference in appearance and function, it is necessary to
consider how this sense of presence can be realized in digital humans.</p>
        <p>Based on the above four points, we felt that for a digital human to become a good
conversational partner, a new design unique to a digital human is needed that is
different from the existing robot design. The current nonverbal representations of virtual
Interactions with agents show a
stronger relationship between
intent and use than with robots.</p>
        <p>Overall, participants preferred the
robot but easily approached the
virtual Nico while avoiding
faceto-face interaction with the robot
Nico.</p>
        <p>Preferring virtual agents that
resemble humans to robots that do
not resemble humans.</p>
        <p>Philips iCat</p>
        <p>
          Steffie
[
          <xref ref-type="bibr" rid="ref27">27</xref>
          ]
        </p>
        <p>Humanoid Nico</p>
        <p>
          3D Modeling Nico
[
          <xref ref-type="bibr" rid="ref13">13</xref>
          ]
        </p>
        <p>Humanoid Kobiana Ana
agents are not as natural as the real appearance of the digital human. The above study
results are experiments that exclude realistic human-like virtual agents (digital
humans), and since they did not focus on subjective evaluation criteria or use
representative evaluation scales, there is a possibility that participants’ responses may be
different. What virtual agents and robots have in common is that they have a body. Whether
virtual or physical, due to differences in the physical specifications of the agent, the
nonverbal expression and implementation method of the two are different, and various
nonverbal expressions can be generated due to this transformation.</p>
      </sec>
      <sec id="sec-3-3">
        <title>3.3 Cases of Comparative Study of Virtual Agents and Robots</title>
        <p>
          Efforts to communicate with humans naturally are continuing as in previous studies and
several improvement methods have also been proposed.[
          <xref ref-type="bibr" rid="ref28 ref29">28, 29</xref>
          ] Allwood proposed four
feedback functions: contact, perception, understanding, and attitudinal reactions. In this
paper, we reconstruct the overlapping functions of the four feedback functional
elements of the preceding studies as “attention,” “understanding,” and “opinion.”[
          <xref ref-type="bibr" rid="ref30 ref31">30, 31</xref>
          ]
Here, we would like to examine the case of a listener’s reaction studies by additionally
using the element of “timing” (Fig. 2).
Attention is an expression of the listener’s willingness and ability to recognize a
message.[
          <xref ref-type="bibr" rid="ref30">30</xref>
          ] Attention can help users interact with agents to provide even more listener
feedback.[
          <xref ref-type="bibr" rid="ref32">32</xref>
          ] A representative expression of attention is staring at the speaker. Gaze
can signal usually that the speaker’s continuing encouragement of utterance and that
communication channels is open. Yoichi found in healthcare studies that patients pay
more attention to agents when returning listener feedback while the patient is
speaking.[
          <xref ref-type="bibr" rid="ref33">33</xref>
          ] Oh studied the degree of attention conveyed by nodding, audio and
audio-visual feedback.[
          <xref ref-type="bibr" rid="ref34">34</xref>
          ] The robot was evaluated more positively when it displayed hand and
arm gestures with words and asked participants to pay attention to the robot during the
interaction.[
          <xref ref-type="bibr" rid="ref35">35</xref>
          ]. Allwood and Cerrato found that nodding the head conveys that the
listener is paying attention and further triggers a sympathetic reaction.[
          <xref ref-type="bibr" rid="ref36">36</xref>
          ]
The listener understands the speaker’s intentions through the language information, and
the speaker monitors the listener to see if what the speaker wants to convey has been
achieved.[
          <xref ref-type="bibr" rid="ref32">32</xref>
          ] The listener can express understanding by nodding or staring.[
          <xref ref-type="bibr" rid="ref31 ref37">31, 37</xref>
          ]
Nakano et al. found that nonverbal cues perceived as positive evidence of
comprehension were context-dependent. They also found that staring at the speaker was
interpreted as evidence of incomprehension that provoked further explanation from the
speaker.[
          <xref ref-type="bibr" rid="ref31">31</xref>
          ]
The speaker checks how the partner receives the message. Listeners can express their
opinions (acceptance, consent, preference, etc.) to make communication livelier. The
expression of opinion may take an expression form similar to the above understanding
element. Understanding, however, is simply focused on the listener’s understanding of
the information, and expression of opinion is an implicit confirmation of the
understanding. Nodding proved to be very important because all participants responded
“agree” when displayed alone. Smiling, nodding and raising eyebrows also received
high marks as signs of consent.[
          <xref ref-type="bibr" rid="ref31">31</xref>
          ] When virtual agent Billie requests confirmation,
nodding is considered an acceptance and shaking the head is considered an expression
of rejection. On the other hand, if the user nodded while agent “Billie” presented the
information, nodding was interpreted as evidence of understanding.[
          <xref ref-type="bibr" rid="ref38">38</xref>
          ]
Timing needs to be considered a digital human listener feedback element because it can
provide an unnatural feeling and a sense that we are indeed talking. For human-like
communication, proper timing of the response to feedback is important.[
          <xref ref-type="bibr" rid="ref39 ref40">39, 40</xref>
          ] Even
a good expression of consent can cause misunderstanding in the process of conveying
meaning if it appears when it is not appropriate. Also, timing can be a signal of
turntaking and can contribute to creating a natural and realistic digital human.[
          <xref ref-type="bibr" rid="ref21">21</xref>
          ]
[
          <xref ref-type="bibr" rid="ref41">41</xref>
          ] scrutinized when and how the listener inserted responses in line with the speaker’s
context. They suggested that the speaker’s gaze mediates this cooperation. [
          <xref ref-type="bibr" rid="ref42">42</xref>
          ] The
“Rapport Agent” creates rapport by providing feedback to the person speaking about
the comics they have seen before. The camera analyzes the speaker and determines the
appropriate moment to provide feedback with head nods, head shakes, head rolls, and
gaze.[
          <xref ref-type="bibr" rid="ref39">39</xref>
          ] Previous investigation of nonverbal feedback from avatars or robots revealed
that cues or reactions, such as head turns, are effective when they occur at meaningful
times rather than at random times
4
        </p>
      </sec>
    </sec>
    <sec id="sec-4">
      <title>Discussion</title>
      <p>Existing studies have not explored the potential mediated effects of digital human
appearance as no work has been done using digital human beings to any degree.
Therefore, we felt the need for a new guide to digital humans in line with the rapid
commercialization trend of digital humans. For natural communication, we showed that people
had stronger behavioral and attitudinal responses to physically present agents as
opposed to a virtual presence, but when both the physically implemented agent and the
virtually implemented agent were presented, the potential for development as a good
listener compared to the robot was found respectively. It is presumed that different
results were derived depending on the appearance of the virtual agent and the degree of
freedom to express it. Also, depending on the appearance and degree of freedom of the
virtual agent, the response can be linked to the agent’s overall satisfaction. Unnatural
expressions or excessive repetition may cause discomfort, so it is necessary to consider
providing natural and appropriate feedback. Based on this, we defined four feedback
A New Good Listener, the Digital Human 45
functional factors for interactive agents to become good listeners and proposed a wide
range of concepts that can be applied. The four functional feedback elements presented
were summarized into three elements (attention, understanding, and expression) that
overlap or have the same meaning in previous studies, then redefined a total of four
elements by adding the “timing” elements that stand out in other nonverbal
communication studies.</p>
      <p>First, an expression that pays attention to the speaker is required. Second, whether the
understanding of the content of the ignition is successful or not. Third, it should be
possible to express the listener’s opinion about the content. Finally, the expression of
the listener’s attention, understanding, and opinion should be expressed in a timely
manner. Digital humans, which are currently commercialized, may be suitable as
lowcost personal assistants because they can be less expensive than robots and less
constrained by their environment of use. The digital human can be generally better than a
robot in that it can represent behavior, emotions, gestures, and expressions like humans.
All told, it suggests that the digital human has sufficient potential to be utilized as a
virtual listener. From the robots and agents used in this study alone, it is not clear to
what extent the results will be the same for each evaluation in different implementations
with different agent types. In future research, it will be necessary to check whether there
is an empirical effect as a good listener through the feedback presented.
5</p>
    </sec>
    <sec id="sec-5">
      <title>Conclusion</title>
      <p>This study is just the first to examine how listener feedback from digital humans can
affect human conversations. As the use of various agents increases, studies focusing on
specific contexts and the need for nonverbal representation design studies in digital
human conversation are shown. For several reasons, it was not possible to go deep into
the functional analysis of listener feedback as originally intended in this study. We have
just begun exploring listener feedback in digital humans, which should be combined
with a larger number of studies in the future. In the next study, we intend to verify the
four feedback factors proposed in this study through experiments to create a digital
human for conversation. Thus, our ultimate goal is to build a digital human that users
will want to talk to.</p>
      <p>A New Good Listener, the Digital Human 47</p>
    </sec>
  </body>
  <back>
    <ref-list>
      <ref id="ref1">
        <mixed-citation>
          1.
          <string-name>
            <surname>Kim</surname>
            ,
            <given-names>J.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Kim</surname>
            ,
            <given-names>W.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Nam</surname>
            ,
            <given-names>J.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Song</surname>
          </string-name>
          , H.:
          <article-title>“I Can Feel Your Empathic Voice”: Effects of Nonverbal Vocal Cues in Voice User Interface</article-title>
          .
          <source>Ext. Abstr</source>
          .
          <source>2020 CHI Conf. Hum. Factors Comput. Syst. 1-8</source>
          (
          <year>2020</year>
          ). https://doi.org/10.1145/3334480.3383075.
        </mixed-citation>
      </ref>
      <ref id="ref2">
        <mixed-citation>
          2.
          <string-name>
            <surname>Gobron</surname>
            ,
            <given-names>S.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Ahn</surname>
            ,
            <given-names>J.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Garcia</surname>
            ,
            <given-names>D.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Silvestre</surname>
            ,
            <given-names>Q.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Thalmann</surname>
            ,
            <given-names>D.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Boulic</surname>
            ,
            <given-names>R.:</given-names>
          </string-name>
          <article-title>An Event-Based Ar- chitecture to Manage Virtual Human Non-Verbal Communication in 3D Chatting Environment</article-title>
          . In: Perales,
          <string-name>
            <given-names>F.J.</given-names>
            , Fisher, R.B., and
            <surname>Moeslund</surname>
          </string-name>
          , T.B. (eds.)
          <source>Articulated Motion and Deformable Objects</source>
          . pp.
          <fpage>58</fpage>
          -
          <lpage>68</lpage>
          . Springer, Berlin, Heidelberg (
          <year>2012</year>
          ). https://doi.org/10.1007/978-3-
          <fpage>642</fpage>
          - 31567-
          <issue>1</issue>
          _
          <fpage>6</fpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref3">
        <mixed-citation>
          3.
          <string-name>
            <surname>Judee</surname>
            <given-names>K</given-names>
          </string-name>
          ,
          <string-name>
            <given-names>B.</given-names>
            ,
            <surname>Laura</surname>
          </string-name>
          <string-name>
            <surname>K</surname>
          </string-name>
          , L.K.G.,
          <article-title>Valerie Manusov: Nonverbal signals</article-title>
          .
          <source>SAGE Handb. Interpers. Commun</source>
          .
          <volume>239</volume>
          -
          <fpage>280</fpage>
          (
          <year>2011</year>
          ).
        </mixed-citation>
      </ref>
      <ref id="ref4">
        <mixed-citation>
          4.
          <string-name>
            <surname>Ekman</surname>
            ,
            <given-names>P.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Friesen</surname>
            ,
            <given-names>W.V.</given-names>
          </string-name>
          :
          <article-title>Nonverbal Leakage and Clues to Deception</article-title>
          .
          <source>Psychiatry</source>
          <volume>321</volume>
          .
          <fpage>88</fpage>
          -
          <lpage>106</lpage>
          (
          <year>1969</year>
          ).
        </mixed-citation>
      </ref>
      <ref id="ref5">
        <mixed-citation>
          5.
          <string-name>
            <surname>Cassell</surname>
          </string-name>
          , J.:
          <article-title>Nudge Nudge Wink Wink: Elements of Face-to-Face Conversation for Embodied Conversational Agents</article-title>
          . (
          <year>2000</year>
          ).
        </mixed-citation>
      </ref>
      <ref id="ref6">
        <mixed-citation>
          6.
          <string-name>
            <surname>Bergmann</surname>
            ,
            <given-names>K.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Kopp</surname>
            ,
            <given-names>S.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Eyssel</surname>
            ,
            <given-names>F.</given-names>
          </string-name>
          : Individualized Gesturing Outperforms Average Gesturing
        </mixed-citation>
      </ref>
      <ref id="ref7">
        <mixed-citation>
          7.
          <article-title>-Evaluating Gesture Production in Virtual Humans</article-title>
          . Intell. Virtual Agents.
          <fpage>104</fpage>
          -
          <lpage>117</lpage>
          (
          <year>2010</year>
          ). https://doi.org/10.1007/978-3-
          <fpage>642</fpage>
          -15892-6_
          <fpage>11</fpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref8">
        <mixed-citation>
          8.
          <string-name>
            <surname>Davis</surname>
            ,
            <given-names>R.O.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Wan</surname>
            ,
            <given-names>L.L.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Vincent</surname>
            ,
            <given-names>J.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Lee</surname>
            ,
            <given-names>Y.J.:</given-names>
          </string-name>
          <article-title>The effects of virtual human gesture frequency and reduced video speed on satisfaction and learning outcomes</article-title>
          .
          <source>Educ. Technol. Res. Dev</source>
          . (
          <year>2021</year>
          ). https://doi.org/10.1007/s11423-021-10010-x.
        </mixed-citation>
      </ref>
      <ref id="ref9">
        <mixed-citation>
          9.
          <string-name>
            <surname>Traum</surname>
            ,
            <given-names>D.</given-names>
          </string-name>
          :
          <article-title>Models of Culture for Virtual Human Conversation</article-title>
          . Univers. Access
          <string-name>
            <surname>Hum</surname>
          </string-name>
          .- Com- put.
          <source>Interact. Appl. Serv</source>
          .
          <volume>5616</volume>
          ,
          <fpage>434</fpage>
          -
          <lpage>440</lpage>
          (
          <year>2009</year>
          ). https://doi.org/10.1007/978-3-
          <fpage>642</fpage>
          - 02713-0_
          <fpage>46</fpage>
          . 9.
          <string-name>
            <surname>Carrozzino</surname>
            ,
            <given-names>M.A.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Galdieri</surname>
            ,
            <given-names>R.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Machidon</surname>
            ,
            <given-names>O.M.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Bergamasco</surname>
            ,
            <given-names>M.</given-names>
          </string-name>
          :
          <source>Do Virtual Humans Dream of Digital Sheep? IEEE Comput. Graph. Appl</source>
          .
          <volume>40</volume>
          ,
          <fpage>71</fpage>
          -
          <lpage>83</lpage>
          (
          <year>2020</year>
          ). https://doi.org/10.1109/
          <string-name>
            <surname>MCG</surname>
          </string-name>
          .
          <year>2020</year>
          .
          <volume>2993345</volume>
          .
        </mixed-citation>
      </ref>
      <ref id="ref10">
        <mixed-citation>
          10.
          <string-name>
            <surname>Leite</surname>
            ,
            <given-names>I.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Martinho</surname>
            ,
            <given-names>C.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Paiva</surname>
            ,
            <given-names>A.</given-names>
          </string-name>
          :
          <article-title>Social Robots for Long-Term Interaction: A Survey</article-title>
          .
          <source>Int. J. Soc. Robot</source>
          .
          <volume>5</volume>
          ,
          <fpage>291</fpage>
          -
          <lpage>308</lpage>
          (
          <year>2013</year>
          ). https://doi.org/10.1007/s12369-013-0178-y.
        </mixed-citation>
      </ref>
      <ref id="ref11">
        <mixed-citation>
          11.
          <string-name>
            <surname>David</surname>
            ,
            <given-names>D.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Matu</surname>
            ,
            <given-names>S.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>David</surname>
            ,
            <given-names>O.</given-names>
          </string-name>
          :
          <article-title>Robot-Based Psychotherapy: Concepts Development, State of the Art, and New Directions</article-title>
          .
          <source>Int. J. Cogn. Ther</source>
          .
          <volume>7</volume>
          ,
          <fpage>192</fpage>
          -
          <lpage>210</lpage>
          (
          <year>2014</year>
          ). https://doi.org/10.1521/ijct.
          <year>2014</year>
          .
          <volume>7</volume>
          .2.192.
        </mixed-citation>
      </ref>
      <ref id="ref12">
        <mixed-citation>
          12.
          <string-name>
            <surname>Forbrig</surname>
            ,
            <given-names>P.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Bundea</surname>
            ,
            <given-names>A.-N.</given-names>
          </string-name>
          :
          <article-title>Modelling the Collaboration of a Patient and an Assisting Human- oid Robot During Training Tasks</article-title>
          .
          <source>Hum.-Comput. Interact. Multimodal Nat. Interact</source>
          .
          <volume>592</volume>
          -
          <fpage>602</fpage>
          (
          <year>2020</year>
          ). https://doi.org/10.1007/978-3-
          <fpage>030</fpage>
          -49062-1_
          <fpage>40</fpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref13">
        <mixed-citation>
          13.
          <string-name>
            <surname>Trovato</surname>
            ,
            <given-names>G.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Lopez</surname>
            ,
            <given-names>A.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Paredes</surname>
            <given-names>Venero</given-names>
          </string-name>
          ,
          <string-name>
            <given-names>R.</given-names>
            ,
            <surname>Cuellar</surname>
          </string-name>
          ,
          <string-name>
            <surname>F.</surname>
          </string-name>
          :
          <article-title>Security and guidance: Two roles for a humanoid robot in an interaction experiment</article-title>
          . (
          <year>2017</year>
          ). https://doi.org/10.1109/ROMAN.
          <year>2017</year>
          .
          <volume>8172307</volume>
          .
        </mixed-citation>
      </ref>
      <ref id="ref14">
        <mixed-citation>
          14.
          <string-name>
            <surname>Nakanishi</surname>
          </string-name>
          , J.:
          <article-title>Can a Humanoid Robot Engage in Heartwarming Interaction Service at a Hotel?</article-title>
          <source>Proc. 6th Int. Conf. Hum</source>
          .-Agent
          <string-name>
            <surname>Interact</surname>
          </string-name>
          . (
          <year>2018</year>
          ).
        </mixed-citation>
      </ref>
      <ref id="ref15">
        <mixed-citation>
          15.
          <string-name>
            <surname>Johnson</surname>
            ,
            <given-names>D.O.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Cuijpers</surname>
            ,
            <given-names>R.H.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Kathrin</surname>
            ,
            <given-names>P.</given-names>
          </string-name>
          , J, van de V.
          <article-title>A.A.: Exploring the Entertainment Value of Playing Games with a Humanoid Robot</article-title>
          .
          <source>Int. J. Soc. Robot</source>
          .
          <volume>8</volume>
          ,
          <fpage>247</fpage>
          -
          <lpage>269</lpage>
          (
          <year>2016</year>
          ). http://dx.doi.org/10.1007/s12369-015-0331-x.
        </mixed-citation>
      </ref>
      <ref id="ref16">
        <mixed-citation>
          16.
          <string-name>
            <surname>Hasunuma</surname>
            ,
            <given-names>H.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Kobayashi</surname>
            ,
            <given-names>M.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Moriyama</surname>
            ,
            <given-names>H.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Itoko</surname>
            ,
            <given-names>T.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Yanagihara</surname>
            ,
            <given-names>Y.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Ueno</surname>
            ,
            <given-names>T.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Ohya</surname>
            ,
            <given-names>K.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Yokoil</surname>
            ,
            <given-names>K.</given-names>
          </string-name>
          :
          <article-title>A tele-operated humanoid robot drives a lift truck</article-title>
          .
          <source>Proc. 2002 IEEE Int. Conf. Robot. Autom. Cat No02CH37292</source>
          .
          <volume>3</volume>
          ,
          <fpage>2246</fpage>
          -
          <lpage>2252</lpage>
          vol.
          <volume>3</volume>
          (
          <year>2002</year>
          ). https://doi.org/10.1109/ROBOT.
          <year>2002</year>
          .
          <volume>1013566</volume>
          .
        </mixed-citation>
      </ref>
      <ref id="ref17">
        <mixed-citation>
          17.
          <string-name>
            <surname>Kose</surname>
            ,
            <given-names>H.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Yorganci</surname>
            ,
            <given-names>R.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Algan</surname>
            ,
            <given-names>E.H.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Syrdal</surname>
            ,
            <given-names>D.S.</given-names>
          </string-name>
          :
          <article-title>Evaluation of the Robot Assisted Sign Language Tutoring Using Video-Based Studies</article-title>
          .
          <source>Int. J. Soc. Robot</source>
          .
          <volume>4</volume>
          ,
          <fpage>273</fpage>
          -
          <lpage>283</lpage>
          (
          <year>2012</year>
          ). https://doi.org/10.1007/s12369-012-0142-2.
        </mixed-citation>
      </ref>
      <ref id="ref18">
        <mixed-citation>
          18.
          <string-name>
            <surname>Rahman</surname>
            ,
            <given-names>S.M.M.:</given-names>
          </string-name>
          <article-title>Generating human-like social motion in a human-looking humanoid robot: The biomimetic approach</article-title>
          .
          <source>2013 IEEE Int. Conf. Robot. Biomim. ROBIO</source>
          .
          <volume>1377</volume>
          -
          <fpage>1383</fpage>
          (
          <year>2013</year>
          ). https://doi.org/10.1109/ROBIO.
          <year>2013</year>
          .
          <volume>6739657</volume>
          .
        </mixed-citation>
      </ref>
      <ref id="ref19">
        <mixed-citation>
          19.
          <string-name>
            <surname>Schroder</surname>
            ,
            <given-names>M.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Heylen</surname>
            ,
            <given-names>D.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Poggi</surname>
            ,
            <given-names>I.</given-names>
          </string-name>
          :
          <article-title>Perception of Non-Verbal Emotional Listener Feedback</article-title>
          .
          <volume>4</volume>
          (
          <year>2006</year>
          ).
        </mixed-citation>
      </ref>
      <ref id="ref20">
        <mixed-citation>
          20.
          <string-name>
            <surname>Bevacqua</surname>
            ,
            <given-names>E.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Heylen</surname>
            ,
            <given-names>D.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Pelachaud</surname>
            ,
            <given-names>C.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Tellier</surname>
            ,
            <given-names>M.</given-names>
          </string-name>
          :
          <article-title>Facial Feedback Signals for ECAs</article-title>
          .
          <source>AISB</source>
          .
          <volume>328</volume>
          -
          <fpage>334</fpage>
          (
          <year>2007</year>
          ).
        </mixed-citation>
      </ref>
      <ref id="ref21">
        <mixed-citation>
          21.
          <string-name>
            <surname>Huang</surname>
            ,
            <given-names>L.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Morency</surname>
            ,
            <given-names>L.-P.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Gratch</surname>
          </string-name>
          , J.:
          <article-title>Learning Backchannel Prediction Model from Parasocial Consensus Sampling: A Subjective Evaluation</article-title>
          .
          <volume>6356</volume>
          ,
          <issue>172</issue>
          (
          <year>2010</year>
          ). https://doi.org/10.1007/978-3-
          <fpage>642</fpage>
          -15892-6_
          <fpage>17</fpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref22">
        <mixed-citation>
          22.
          <string-name>
            <surname>Al-Fedaghi</surname>
            ,
            <given-names>S.:</given-names>
          </string-name>
          <article-title>A Conceptual Foundation for the Shannon-Weaver Model of Communication</article-title>
          .
          <source>Int. J. Soft Comput</source>
          .
          <volume>7</volume>
          ,
          <fpage>12</fpage>
          -
          <lpage>19</lpage>
          (
          <year>2012</year>
          ). https://doi.org/10.3923/ijscomp.
          <year>2012</year>
          .
          <volume>12</volume>
          .19.
        </mixed-citation>
      </ref>
      <ref id="ref23">
        <mixed-citation>
          23.
          <string-name>
            <surname>Yamato</surname>
            ,
            <given-names>J.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Shinozawa</surname>
            ,
            <given-names>K.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Naya</surname>
            ,
            <given-names>F.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Kogure</surname>
            ,
            <given-names>K.</given-names>
          </string-name>
          :
          <article-title>Effects of Conversational Agent and Robot on User Decision. (</article-title>
          <year>2000</year>
          ).
        </mixed-citation>
      </ref>
      <ref id="ref24">
        <mixed-citation>
          24.
          <string-name>
            <surname>Powers</surname>
            ,
            <given-names>A.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Kiesler</surname>
            ,
            <given-names>S.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Fussell</surname>
            ,
            <given-names>S.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Torrey</surname>
            ,
            <given-names>C.</given-names>
          </string-name>
          :
          <article-title>Comparing a computer agent with a humanoid robot</article-title>
          .
          <source>Proceeding ACMIEEE Int. Conf. Hum</source>
          .-Robot
          <string-name>
            <surname>Interact</surname>
          </string-name>
          .
          <source>- HRI 07. 145</source>
          (
          <year>2007</year>
          ). https://doi.org/10.1145/1228716.1228736.
        </mixed-citation>
      </ref>
      <ref id="ref25">
        <mixed-citation>
          25.
          <string-name>
            <surname>Kiesler</surname>
            ,
            <given-names>S.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Powers</surname>
            ,
            <given-names>A.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Fussell</surname>
            ,
            <given-names>S.R.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Torrey</surname>
            ,
            <given-names>C.</given-names>
          </string-name>
          :
          <article-title>Anthropomorphic Interactions with a Robot and Robot-Like Agent</article-title>
          .
          <source>Soc. Cogn</source>
          .
          <volume>26</volume>
          ,
          <fpage>169</fpage>
          -
          <lpage>181</lpage>
          (
          <year>2008</year>
          ). http://dx.doi.org/10.1521/soco.
          <year>2008</year>
          .
          <volume>26</volume>
          .2.169.
        </mixed-citation>
      </ref>
      <ref id="ref26">
        <mixed-citation>
          26.
          <string-name>
            <surname>Heerink</surname>
            ,
            <given-names>M.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Kröse</surname>
            ,
            <given-names>B.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Evers</surname>
            ,
            <given-names>V.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Wielinga</surname>
            ,
            <given-names>B.</given-names>
          </string-name>
          :
          <article-title>Influence of Social Presence on Acceptance of an Assistive Social Robot and Screen Agent by Elderly Users</article-title>
          .
          <source>Adv. Robot. 23</source>
          ,
          <fpage>1909</fpage>
          -
          <lpage>1923</lpage>
          (
          <year>2009</year>
          ). https://doi.org/10.1163/016918609X12518783330289.
        </mixed-citation>
      </ref>
      <ref id="ref27">
        <mixed-citation>
          27.
          <string-name>
            <surname>Bainbridge</surname>
            ,
            <given-names>W.A.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Hart</surname>
            ,
            <given-names>J.W.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Kim</surname>
            ,
            <given-names>E.S.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Scassellati</surname>
            ,
            <given-names>B.</given-names>
          </string-name>
          :
          <article-title>The Benefits of Interactions with Physically Present Robots over Video-Displayed Agents</article-title>
          .
          <source>Int. J. Soc. Robot</source>
          .
          <volume>3</volume>
          ,
          <fpage>41</fpage>
          -
          <lpage>52</lpage>
          (
          <year>2011</year>
          ). https://doi.org/10.1007/s12369-010-0082-7.
        </mixed-citation>
      </ref>
      <ref id="ref28">
        <mixed-citation>
          28.
          <string-name>
            <surname>Jung</surname>
            ,
            <given-names>J.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Kanda</surname>
            ,
            <given-names>T.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Kim</surname>
          </string-name>
          , M.
          <article-title>-S.: Guidelines for Contextual Motion Design of a Humanoid Robot</article-title>
          .
          <source>Int. J. Soc. Robot</source>
          .
          <volume>5</volume>
          ,
          <fpage>153</fpage>
          -
          <lpage>169</lpage>
          (
          <year>2013</year>
          ). https://doi.org/10.1007/s12369-012-0175-6.
        </mixed-citation>
      </ref>
      <ref id="ref29">
        <mixed-citation>
          29.
          <string-name>
            <surname>Wang</surname>
            ,
            <given-names>I.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Ruiz</surname>
          </string-name>
          , J.:
          <article-title>Examining the Use of Nonverbal Communication in Virtual Agents</article-title>
          .
          <source>Int. J. Human-Computer Interact. 0</source>
          ,
          <fpage>1</fpage>
          -
          <lpage>26</lpage>
          (
          <year>2021</year>
          ). https://doi.org/10.1080/10447318.
          <year>2021</year>
          .
          <volume>1898851</volume>
          . 30.
          <string-name>
            <surname>Allwood</surname>
            ,
            <given-names>J.:</given-names>
          </string-name>
          <article-title>Feedback in Second Language Acquisition</article-title>
          .
          <source>Adult Lang. Acquis. Cross Linguist. Perspect. II Results</source>
          .
          <volume>196</volume>
          -
          <fpage>235</fpage>
          (
          <year>1993</year>
          ).
        </mixed-citation>
      </ref>
      <ref id="ref30">
        <mixed-citation>
          30.
          <string-name>
            <surname>Heylen</surname>
            ,
            <given-names>D.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Bevacqua</surname>
            ,
            <given-names>E.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Pelachaud</surname>
            ,
            <given-names>C.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Poggi</surname>
            ,
            <given-names>I.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Gratch</surname>
            ,
            <given-names>J.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Schröder</surname>
            ,
            <given-names>M.</given-names>
          </string-name>
          :
          <article-title>Generating Listening Behaviour</article-title>
          . In: Cowie,
          <string-name>
            <given-names>R.</given-names>
            ,
            <surname>Pelachaud</surname>
          </string-name>
          ,
          <string-name>
            <given-names>C.</given-names>
            , and
            <surname>Petta</surname>
          </string-name>
          , P. (eds.)
          <string-name>
            <surname>Emotion-Oriented Systems</surname>
          </string-name>
          . pp.
          <fpage>321</fpage>
          -
          <lpage>347</lpage>
          . Springer Berlin Heidelberg, Berlin, Heidelberg (
          <year>2011</year>
          ). https://doi.org/10.1007/978-3-
          <fpage>642</fpage>
          -15184-2_
          <fpage>17</fpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref31">
        <mixed-citation>
          31.
          <string-name>
            <surname>Buschmeier</surname>
            ,
            <given-names>H.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Kopp</surname>
            ,
            <given-names>S.</given-names>
          </string-name>
          :
          <article-title>Communicative Listener Feedback in Human-Agent Interaction: Artificial Speakers Need to Be Attentive</article-title>
          and Adaptive.
          <volume>9</volume>
          (
          <year>2018</year>
          ).
        </mixed-citation>
      </ref>
      <ref id="ref32">
        <mixed-citation>
          32.
          <string-name>
            <surname>Sakai</surname>
            ,
            <given-names>Y.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Nonaka</surname>
            ,
            <given-names>Y.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Yasuda</surname>
            ,
            <given-names>K.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Nakano</surname>
            ,
            <given-names>Y.I.</given-names>
          </string-name>
          :
          <article-title>Listener agent for elderly people with dementia</article-title>
          .
          <source>In: Proceedings of the seventh annual ACM/IEEE international conference on Human-</source>
        </mixed-citation>
      </ref>
      <ref id="ref33">
        <mixed-citation>
          33. Robot Interaction. pp.
          <fpage>199</fpage>
          -
          <lpage>200</lpage>
          . Association for Computing Machinery, New York, NY, USA (
          <year>2012</year>
          ). https://doi.org/10.1145/2157689.2157754.
        </mixed-citation>
      </ref>
      <ref id="ref34">
        <mixed-citation>
          34.
          <string-name>
            <surname>Oh</surname>
            ,
            <given-names>C.S.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Bailenson</surname>
            ,
            <given-names>J.N.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Welch</surname>
            ,
            <given-names>G.F.</given-names>
          </string-name>
          :
          <article-title>A Systematic Review of Social Presence: Definition, Antecedents, and Implications</article-title>
          . Front. Robot.
          <source>AI</source>
          .
          <volume>0</volume>
          , (
          <year>2018</year>
          ). https://doi.org/10.3389/frobt.
          <year>2018</year>
          .
          <volume>00114</volume>
          .
        </mixed-citation>
      </ref>
      <ref id="ref35">
        <mixed-citation>
          35.
          <string-name>
            <surname>Salem</surname>
            ,
            <given-names>M.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Eyssel</surname>
            ,
            <given-names>F.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Rohlfing</surname>
            ,
            <given-names>K.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Kopp</surname>
            ,
            <given-names>S.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Joublin</surname>
            ,
            <given-names>F.</given-names>
          </string-name>
          :
          <article-title>Effects of Gesture on the Perception of Psychological Anthropomorphism: A Case Study with a Humanoid Robot</article-title>
          . In: Mutlu,
          <string-name>
            <given-names>B.</given-names>
            , Bart- neck, C.,
            <surname>Ham</surname>
          </string-name>
          ,
          <string-name>
            <given-names>J.</given-names>
            ,
            <surname>Evers</surname>
          </string-name>
          ,
          <string-name>
            <given-names>V.</given-names>
            , and
            <surname>Kanda</surname>
          </string-name>
          , T. (eds.) Social Robotics. pp.
          <fpage>31</fpage>
          -
          <lpage>41</lpage>
          . Springer, Berlin, Heidelberg (
          <year>2011</year>
          ). https://doi.org/10.1007/978-3-
          <fpage>642</fpage>
          -25504-
          <issue>5</issue>
          _
          <fpage>4</fpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref36">
        <mixed-citation>
          36.
          <string-name>
            <surname>Allwood</surname>
            ,
            <given-names>J.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Cerrato</surname>
            ,
            <given-names>L.</given-names>
          </string-name>
          :
          <article-title>A study of gestural feedback expressions</article-title>
          .
          <volume>13</volume>
          .
        </mixed-citation>
      </ref>
      <ref id="ref37">
        <mixed-citation>
          37.
          <string-name>
            <surname>Nakano</surname>
            ,
            <given-names>Y.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Reinstein</surname>
            ,
            <given-names>G.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Stocky</surname>
            ,
            <given-names>T.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Cassell</surname>
          </string-name>
          , J.:
          <article-title>Towards a Model of Face-to-Face Grounding</article-title>
          .
          <source>Proc. 41st Annu. Meet. Assoc. Comput. Linguist</source>
          .
          <volume>553</volume>
          -
          <fpage>561</fpage>
          (
          <year>2003</year>
          ). https://doi.org/10.3115/1075096.1075166.
        </mixed-citation>
      </ref>
      <ref id="ref38">
        <mixed-citation>
          38.
          <string-name>
            <surname>Buschmeier</surname>
            ,
            <given-names>H.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Kopp</surname>
            ,
            <given-names>S.</given-names>
          </string-name>
          :
          <article-title>Towards Conversational Agents That Attend to and Adapt to Com- municative User Feedback</article-title>
          .
          <source>Intell. Virtual Agents</source>
          .
          <fpage>169</fpage>
          -
          <lpage>182</lpage>
          (
          <year>2011</year>
          ). https://doi.org/10.1007/978- 3-
          <fpage>642</fpage>
          -23974-8_
          <fpage>19</fpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref39">
        <mixed-citation>
          39.
          <string-name>
            <surname>Yamazaki</surname>
            ,
            <given-names>A.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Yamazaki</surname>
            ,
            <given-names>K.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Kuno</surname>
            ,
            <given-names>Y.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Burdelski</surname>
            ,
            <given-names>M.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Kawashima</surname>
            ,
            <given-names>M.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Kuzuoka</surname>
          </string-name>
          , H.:
          <article-title>Precision timing in human-robot interaction: coordination of head movement and utterance</article-title>
          .
          <source>Proc. SIGCHI Conf. Hum. Factors Comput. Syst</source>
          .
          <volume>131</volume>
          -
          <fpage>140</fpage>
          (
          <year>2008</year>
          ). https://doi.org/10.1145/1357054.1357077.
        </mixed-citation>
      </ref>
      <ref id="ref40">
        <mixed-citation>
          40.
          <string-name>
            <surname>Poppe</surname>
            ,
            <given-names>R.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Truong</surname>
            ,
            <given-names>K.P.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Reidsma</surname>
            ,
            <given-names>D.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Heylen</surname>
            ,
            <given-names>D.</given-names>
          </string-name>
          :
          <article-title>Backchannel Strategies for Artificial Listeners</article-title>
          . In: Allbeck,
          <string-name>
            <given-names>J.</given-names>
            ,
            <surname>Badler</surname>
          </string-name>
          ,
          <string-name>
            <given-names>N.</given-names>
            ,
            <surname>Bickmore</surname>
          </string-name>
          ,
          <string-name>
            <given-names>T.</given-names>
            ,
            <surname>Pelachaud</surname>
          </string-name>
          ,
          <string-name>
            <given-names>C.</given-names>
            , and
            <surname>Safonova</surname>
          </string-name>
          ,
          <string-name>
            <surname>A</surname>
          </string-name>
          . (eds.) Intelli- gent
          <source>Virtual Agents</source>
          . pp.
          <fpage>146</fpage>
          -
          <lpage>158</lpage>
          . Springer, Berlin, Heidelberg (
          <year>2010</year>
          ). https://doi.org/10.1007/978-3-
          <fpage>642</fpage>
          -15892-6_
          <fpage>16</fpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref41">
        <mixed-citation>
          41.
          <string-name>
            <surname>Bavelas</surname>
            ,
            <given-names>J.B.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Coates</surname>
            ,
            <given-names>L.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Johnson</surname>
          </string-name>
          , T.:
          <article-title>Listener Responses as a Collaborative Process: The Role of Gaze</article-title>
          .
          <source>J. Commun</source>
          .
          <volume>52</volume>
          ,
          <fpage>566</fpage>
          -
          <lpage>580</lpage>
          (
          <year>2002</year>
          ). https://doi.org/10.1111/j.1460-
          <fpage>2466</fpage>
          .
          <year>2002</year>
          .tb02562.x.
        </mixed-citation>
      </ref>
      <ref id="ref42">
        <mixed-citation>
          42.
          <string-name>
            <surname>Gratch</surname>
            ,
            <given-names>J.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Okhmatovskaia</surname>
            ,
            <given-names>A.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Lamothe</surname>
            ,
            <given-names>F.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Marsella</surname>
            ,
            <given-names>S.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Morales</surname>
          </string-name>
          , M.,
          <string-name>
            <surname>van der</surname>
            <given-names>Werf</given-names>
          </string-name>
          , R.J.,
          <string-name>
            <surname>Morency</surname>
            ,
            <given-names>L.-P.: Virtual</given-names>
          </string-name>
          <string-name>
            <surname>Rapport</surname>
          </string-name>
          . Intell. Virtual Agents.
          <fpage>14</fpage>
          -
          <lpage>27</lpage>
          (
          <year>2006</year>
          ). https://doi.org/10.1007/11821830_2.
        </mixed-citation>
      </ref>
    </ref-list>
  </back>
</article>