<!DOCTYPE article PUBLIC "-//NLM//DTD JATS (Z39.96) Journal Archiving and Interchange DTD v1.0 20120330//EN" "JATS-archivearticle1.dtd">
<article xmlns:xlink="http://www.w3.org/1999/xlink">
  <front>
    <journal-meta />
    <article-meta>
      <title-group>
        <article-title>Mixed-Initiative in Interactive Robotic Learning</article-title>
      </title-group>
      <contrib-group>
        <contrib contrib-type="author">
          <string-name>Julia Peltason</string-name>
          <xref ref-type="aff" rid="aff0">0</xref>
        </contrib>
        <contrib contrib-type="author">
          <string-name>Ingo Lu¨tkebohle</string-name>
          <xref ref-type="aff" rid="aff0">0</xref>
        </contrib>
        <contrib contrib-type="author">
          <string-name>Britta Wrede</string-name>
          <email>bwrede@techfak.uni-bielefeld.de</email>
          <xref ref-type="aff" rid="aff0">0</xref>
        </contrib>
        <contrib contrib-type="author">
          <string-name>Marc Hanheide</string-name>
          <xref ref-type="aff" rid="aff0">0</xref>
        </contrib>
        <aff id="aff0">
          <label>0</label>
          <institution>Applied Informatics, Bielefeld University</institution>
          ,
          <country country="DE">Germany</country>
        </aff>
      </contrib-group>
      <abstract>
        <p>In learning tasks, interaction is mostly about the exchange of knowledge. The interaction process shall be governed on the one hand by the knowledge the tutor wants to convey and on the other by the lacks of knowledge of the learner. In human-robot interaction (HRI), it is usually the human demonstrating or explicitly verbalizing her knowledge and the robot acquiring a respective representation. The ultimate goal in interactive robot learning is thus to enable inexperienced, untrained users to tutor robots in a most natural and intuitive manner. This goal is often impeded by a lack of knowledge of the human about the internal processing and expectations of the robot and by the inflexibility of the robot to understand open-ended, unconstrained tutoring or demonstration. Hence, we propose mixed-initiative strategies to allow both to mutually contribute to the interactive learning process as a bi-directional negotiation about knowledge. Along this line this paper discusses two initially different case studies on object manipulation and learning of spatial environments. We present different styles of mixedinitiative in these scenarios and discuss the merits in each case.</p>
      </abstract>
    </article-meta>
  </front>
  <body>
    <sec id="sec-1">
      <title>-</title>
      <p>
        Mixed-initiative human-robot communication is long being studied in rescue and
space mission tasks with the objective of finding the optimal balance between
teleoperation and full autonomy (see [
        <xref ref-type="bibr" rid="ref1">1</xref>
        ] for a survey). Under conditions where
the robot needs assistance, the human could take temporary control of the robot
putting it into a teleoperation mode until it is able to act autonomously again. It
has been shown that such mixed-initiative control solutions can decrease the task
completion time and the varying performance among users [
        <xref ref-type="bibr" rid="ref2">2</xref>
        ]. Mixed-initiative is
typically employed to resolve abnormal situations, and communication is mostly
realized via input devices such as keyboard or joystick, thus targeting at expert
users. In contrast, the presented work focuses on lay persons interacting with
robots using natural human modalities. Moreover, we regard mixed-initiative as
an integral interaction strategy within the robot’s regular operation process. We
present two case studies demonstrating that also within natural human robot
interaction, mixed-initiative has the potential to enhance both task performance
and usability. Both case studies illustrated in Fig. 1 follow the “learning by
interacting” paradigm [
        <xref ref-type="bibr" rid="ref3">3</xref>
        ] arguing that interaction needs to play a major role in
robotic learning. The first scenario is discussed under the aspect of enhancing
the learning process by mixed-initiative interaction. For the second scenario, we
focus on how to provide appropriate guidance for the user and thus facilitating
interaction by applying a mixed-initiative strategy.
2
      </p>
    </sec>
    <sec id="sec-2">
      <title>The Home Tour: Mixed-initiative facilitates learning</title>
      <p>
        Our first scenario [
        <xref ref-type="bibr" rid="ref4">4</xref>
        ] is embedded in the so-called Home Tour where a mobile
robot assistant has to become acquainted with human’s living environment by
interacting with a human during a guided tour. Here, a basic requirement for
the robot is being able to learn a spatial model of the environment and
integrate human and robotic representation as introduced in previous work [
        <xref ref-type="bibr" rid="ref5">5</xref>
        ].
Originally our system for this scenario employed a pure human initiative model,
where the human user had to explicitly teach the robot. As a consequence of
the user’s initiative modeling, the robot could only acquire a sparse model of
the environment as only those pieces of information have been included which
haven explicitly been presented by the user. To overcome this limitation, the
system was enhanced to comprise a more mixed-initiative interaction strategy.
This strategy enables active learning where the learner actively asks to reduce its
lack of knowledge. In the home tour the robot continuously observes the spatial
situation and asks if it changes significantly according to a novelty measure [
        <xref ref-type="bibr" rid="ref4">4</xref>
        ].
This obviously has the potential to speed up and optimize the learning process,
because the robot can ask for missing knowledge in a goal-directed manner. It
furthermore assures a most comprehensive and complete representation of space.
Hence, the robot can verify existing information, request new information and
resolve uncertainty as illustrated in table 1. Thus, this form of robotic learning
supports not only extension but also continuous revision of acquired knowledge.
This is especially important as the scenario targets at inexperienced users
without knowledge about the internal mode of operation, possibly representing sparse
or erroneous models.
      </p>
      <p>Fig. 1: The Mixed-initiative Home Tour and the Curious Robot scenario.</p>
      <p>Shaping the Interaction with Lay Users</p>
      <p>
        Interaction is modeled based on the principle of grounding [
        <xref ref-type="bibr" rid="ref6">6</xref>
        ] assuming that
both interaction partners aim to establish mutual understanding or common
ground during their interaction. As to the Home Tour, human and robot build
up a common understanding of the environment during a mixed-initiative
negotiation for which grounding provides the basis. Combining the two concepts
enables both interaction partners to engage in interaction in order to make the
robot’s spatial representation as correct, complete and consistent as possible.
This idea has successfully been demonstrated in a test-run ([
        <xref ref-type="bibr" rid="ref5">5</xref>
        ], [
        <xref ref-type="bibr" rid="ref4">4</xref>
        ]) where
despite limited classification quality a consistent environment representation could
be achieved due to the robot’s mixed-initiative capabilities.
3
      </p>
    </sec>
    <sec id="sec-3">
      <title>The Curious Robot: Mixed-initiative facilitates interaction</title>
      <p>
        Our second scenario is an interactive object learning and manipulation scenario
with a humanoid robot exploring objects that are interesting or salient for it,
assisted by the human tutor [
        <xref ref-type="bibr" rid="ref7">7</xref>
        ]. However, from experiences with the original
Home Tour, we observed that untrained users require a significant amount of
prior training to complete the task because the robot’s interaction model is not
immediately obvious. Therefore, their behaviors and interaction strategies are
varying enormously which is almost impossible for a system to cope with.
      </p>
      <p>
        Within this scenario, we consequently focus on how mixed-initiative allows
to structure interaction for the users and thus make their behavior more
predictable. Based on bottom-up visual salience [
        <xref ref-type="bibr" rid="ref8">8</xref>
        ] that we here consider as
perceptual context the robot takes initiative and asks the user about objects and how
to grasp them, as shown in table 2. By asking about objects on its own instead
of leaving it to the user to demonstrate them the robot provides guidance within
interaction which in particular unexperienced users could benefit from. Besides,
it communicates what is interesting for it which unexperienced users might also
not be aware of. Last but not least, with using robot initiative to determine the
objects to learn we avoid error-prone visual analysis of human demonstration
behavior which makes interaction even more robust. In a first evaluation
conducted as video study, we observed that in situations where the robot provided
guidance, participants answered quicker and required less clarification from
experimenter compared to situations with less guidance [
        <xref ref-type="bibr" rid="ref7">7</xref>
        ]. Further, in guided
situations, user behavior was more consistent and therefore more predictable.
      </p>
    </sec>
    <sec id="sec-4">
      <title>Discussion &amp; Outlook</title>
      <p>We briefly presented two case studies with initially different style of initiative
taking. Most of the interaction in the Home Tour was designed to be initiated by
the human interaction partner. But an extension towards more mixed-initiative
as presented in Sec. 2 already proved to be beneficial for the acquisition of
consistent representations. The latter scenario in contrast was following a robot
initiative paradigm from the very beginning. The robot strives to complete its
knowledge and itself decides what to learn. It controls the tasks and assures a
very well-shape interaction. For subjects in the studies the structure was easily
accessible. However, they had only limited abilities to decide and control what’s
going on. It is a consequent advancement to bring both mixed-initiative styles
closer together. For the Curious Robot scenario, this would mean to allow for
more human intervention, for instance to correct or modify grasps. The Home
Tour in contrast would benefit from more robot curiosity, for instance by adding
a bottom-up attention for interesting areas triggering robot initiative, not only
exploiting the spatial context as a data-driven cue.</p>
      <p>
        It shall be emphasized that despite the different initiative styles, both
scenarios use the same dialog framework and rely on the same principles considering
mixed-initiative and grounding as guide lines for system architecture [
        <xref ref-type="bibr" rid="ref4">4</xref>
        ]. For both
scenarios, the same negotiation protocol between components is used, reflecting
the grounding process between human and robot, making the presented
information consistent and thus establishing common ground. Moreover, the same
mechanisms for initiative taking are used. Components provide context
information, for instance about the spatial or perceptual context, triggering interaction
and the learning process.
This work was partially funded as part of the research project DESIRE by the German
Federal Ministry of Education and Research (BMBF) under grant no. 01IME01N and
the Cluster of Excellence “Cognitive Interaction Technology”.
      </p>
    </sec>
  </body>
  <back>
    <ref-list>
      <ref id="ref1">
        <mixed-citation>
          1.
          <string-name>
            <given-names>J.</given-names>
            <surname>Wang</surname>
          </string-name>
          and
          <string-name>
            <given-names>M.</given-names>
            <surname>Lewis</surname>
          </string-name>
          ,
          <article-title>Mixed-initiative multirobot control in USAR. I-Tech Education</article-title>
          and Publishing,
          <year>September 2007</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref2">
        <mixed-citation>
          2.
          <string-name>
            <given-names>C. W.</given-names>
            <surname>Nielsen</surname>
          </string-name>
          ,
          <string-name>
            <given-names>D. A.</given-names>
            <surname>Few</surname>
          </string-name>
          , and
          <string-name>
            <given-names>D. S.</given-names>
            <surname>Athey</surname>
          </string-name>
          , “
          <article-title>Using mixed-initiative human-robot interaction to bound performance in a search task</article-title>
          ,” in International Conference on Intelligent Sensors,
          <source>Sensor Networks and Information Processing</source>
          . IEEE,
          <year>2008</year>
          , pp.
          <fpage>195</fpage>
          -
          <lpage>200</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref3">
        <mixed-citation>
          3.
          <string-name>
            <given-names>M.</given-names>
            <surname>Hanheide</surname>
          </string-name>
          and G. Sagerer, “
          <article-title>Active memory-based interaction strategies for learning-enabling behaviors</article-title>
          ,”
          <source>International Symposium on Robot and Human Interactive Communication (ROMAN)</source>
          ,
          <volume>01</volume>
          /08/
          <year>2008</year>
          2008.
        </mixed-citation>
      </ref>
      <ref id="ref4">
        <mixed-citation>
          4.
          <string-name>
            <given-names>J.</given-names>
            <surname>Peltason</surname>
          </string-name>
          ,
          <string-name>
            <given-names>F. H. K.</given-names>
            <surname>Siepmann</surname>
          </string-name>
          ,
          <string-name>
            <given-names>T. P.</given-names>
            <surname>Spexard</surname>
          </string-name>
          ,
          <string-name>
            <given-names>B.</given-names>
            <surname>Wrede</surname>
          </string-name>
          ,
          <string-name>
            <given-names>M.</given-names>
            <surname>Hanheide</surname>
          </string-name>
          , and
          <string-name>
            <given-names>E. A.</given-names>
            <surname>Topp</surname>
          </string-name>
          , “
          <article-title>Mixedinitiative in human augmented mapping</article-title>
          ,” in International Conference on Robotics and Automation. Kobe, Japan: IEEE,
          <year>2009</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref5">
        <mixed-citation>
          5. E. Topp, “
          <article-title>Human-robot interaction and mapping with a service robot: Human augmented mapping</article-title>
          ,”
          <source>Doctoral Thesis</source>
          , KTH School of Computer Science and Communication (CSC), Stockholm, Sweden, Sept.
          <year>2008</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref6">
        <mixed-citation>
          6.
          <string-name>
            <given-names>H. H.</given-names>
            <surname>Clark</surname>
          </string-name>
          and
          <string-name>
            <given-names>S. A.</given-names>
            <surname>Brennan</surname>
          </string-name>
          , “
          <article-title>Grounding in communication,” in Perspectives on socially shared cognition</article-title>
          , L. B.
          <string-name>
            <surname>Resnick</surname>
            ,
            <given-names>J. M.</given-names>
          </string-name>
          <string-name>
            <surname>Levine</surname>
          </string-name>
          , and S. D. Teasley, Eds.,
          <year>1991</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref7">
        <mixed-citation>
          7. I. Lu¨tkebohle, J.
          <string-name>
            <surname>Peltason</surname>
            ,
            <given-names>L.</given-names>
          </string-name>
          <string-name>
            <surname>Schillingmann</surname>
            ,
            <given-names>C.</given-names>
          </string-name>
          <string-name>
            <surname>Elbrechter</surname>
            ,
            <given-names>B.</given-names>
          </string-name>
          <string-name>
            <surname>Wrede</surname>
            ,
            <given-names>S.</given-names>
          </string-name>
          <string-name>
            <surname>Wachsmuth</surname>
            , and
            <given-names>R.</given-names>
          </string-name>
          <string-name>
            <surname>Haschke</surname>
          </string-name>
          , “
          <article-title>The curious robot - structuring interactive robot learning</article-title>
          ,
          <source>” in International Conference on Robotics and Automation</source>
          . Kobe, Japan: IEEE,
          <year>2009</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref8">
        <mixed-citation>
          8.
          <string-name>
            <given-names>Y.</given-names>
            <surname>Nagai</surname>
          </string-name>
          ,
          <string-name>
            <given-names>K.</given-names>
            <surname>Hosada</surname>
          </string-name>
          ,
          <string-name>
            <given-names>A.</given-names>
            <surname>Morita</surname>
          </string-name>
          , and
          <string-name>
            <given-names>M.</given-names>
            <surname>Asada</surname>
          </string-name>
          , “
          <article-title>A constructive model for the development of joint attention</article-title>
          ,
          <source>” Connection Science</source>
          , vol.
          <volume>15</volume>
          , no.
          <issue>4</issue>
          , pp.
          <fpage>211</fpage>
          -
          <lpage>229</lpage>
          ,
          <year>December 2003</year>
          .
        </mixed-citation>
      </ref>
    </ref-list>
  </back>
</article>