<!DOCTYPE article PUBLIC "-//NLM//DTD JATS (Z39.96) Journal Archiving and Interchange DTD v1.0 20120330//EN" "JATS-archivearticle1.dtd">
<article xmlns:xlink="http://www.w3.org/1999/xlink">
  <front>
    <journal-meta />
    <article-meta>
      <title-group>
        <article-title>Engineering Design with Everyday Materials Multi- modal Dataset</article-title>
      </title-group>
      <contrib-group>
        <contrib contrib-type="author">
          <string-name>Marcelo Worsley</string-name>
          <email>marcelo.worsley@northwestern.edu</email>
          <xref ref-type="aff" rid="aff0">0</xref>
        </contrib>
        <aff id="aff0">
          <label>0</label>
          <institution>Northwestern University</institution>
          ,
          <addr-line>Evanston, IL 60208</addr-line>
          ,
          <country country="US">USA</country>
        </aff>
      </contrib-group>
      <fpage>2</fpage>
      <lpage>8</lpage>
      <abstract>
        <p>This paper describes a multi-modal dataset collected for studying collaborative, engineering design cognition among undergraduate students. Students worked in pairs to solve two engineering design challenges, and also participated in a variety of interventions aimed to improve the quality of the learning experience. While students completed these hands-on tasks, multimodal data was captured using Xbox Kinect, Leap motion, a high definition web camera, and Affectiva Q-sensor.</p>
      </abstract>
      <kwd-group>
        <kwd>Multimodal learning analytics</kwd>
        <kwd>engineering design cognition</kwd>
        <kwd>collaboration</kwd>
      </kwd-group>
    </article-meta>
  </front>
  <body>
    <sec id="sec-1">
      <title>Introduction</title>
      <p>Multimodal data capture capabilities and multimodal learning analytics [1]–[3] are
rapidly growing. Research and practitioners now have the opportunity to collect a wealth
of multimodal process data as students complete collaborative tasks in the physical
and/or digital world. What’s more, these tools may offer a different perspective into
students’ learning experiences. For example, sensors like the Xbox Kinect and high
definition web-cameras have the ability to store rich, high frequency data about how
learners are physically engaging with a given task. Moreover, different
bio-physiological sensors (e.g., Empatica E4 or the Affectiva Q-sensor) have the ability to capture
data that may be too fine-grained and minute for a human to detect.</p>
      <p>
        In this paper I describe a dataset that was recently collected to study engineering design
cognition as perceived through a host of multimodal sensors. However, before I delve
into describing the data, I will briefly provide the motivation for the collecting this data.
I will then move into describing the different pieces of data collected and some of the
challenges collecting the data. I will conclude with some of the on-going work being
done to clean the dataset and prepare it for dissemination.
Over the past few years there has been growing interest in improving K-16 engineering
education [4]–[7]. However, in the same way that Problem-Based Learning faced many
challenges, engineering design, and the Maker Movement more broadly, are also
susceptible to the phenomena of “doing without understanding” [
        <xref ref-type="bibr" rid="ref16">8</xref>
        ]–[
        <xref ref-type="bibr" rid="ref9">10</xref>
        ]. Hence, a primary
motivation for collecting this dataset is to study and compare different strategies for
promoting learning in the context of an open-ended, engineering design task. In
particular, we examine how different interventions can improve the quality of an engineering
design, or Making, experience, and the ways that this is evidenced through multimodal
data, in accordance with prior work [11], [12].
3
3.1
      </p>
    </sec>
    <sec id="sec-2">
      <title>Methodology</title>
      <p>Study Participants
54 students from a West Coast community college, who ranged in terms of age and
prior experience with engineering design, participated in this study. The students
included several majors and students with different career goals. The participants
received course credit for their involvement, and in this way were a sample of
convenience.
3.2</p>
      <sec id="sec-2-1">
        <title>Study Description</title>
        <p>The study uses a 2-by-2 design where students worked in pairs on two engineering
design tasks and participated in two interventions. Student pairing was based on
availability. Students also completed other activities to allow us to better understand and
analyze their learning experience. A diagrammatic representation of the events is shown
in Figure 1. The total experience lasted about one hour per group. This amount of total
time mirrors a number of engineering design oriented “maker” challenges that take
place at after-school programs in libraries and museums around the country. A detailed
description of the study is included in the following paragraphs.</p>
      </sec>
      <sec id="sec-2-2">
        <title>Intervention #1: Single Design vs. Multiple Designs. Pairs were randomly assigned</title>
        <p>to draw either one design (three minutes) or three very different designs (one minute
per design). We will refer to these conditions as Single and Multiple.</p>
        <p>Task #1: Paper and Textbook Task. Students were asked to use one sheet of printer
paper to construct a structure that could support one or more engineering textbooks at
least three inches off the table (see Figures 2-4 for examples). Students had six minutes
to complete this task. The task aimed to help students realize that the configuration of
the piece of paper was of significant import, and that the material was not the only
factor. The number of books that each pair’s design could support was recorded.
Question for an Expert. After the first task, students wrote down questions that they
would ask an expert in mechanical engineering or engineering design. Students had
approximately three minutes to write down their questions.</p>
        <p>Intervention #2: Video or Discussion. Pairs of students were randomly assigned to
watch a short informational video while others participated in a discussion about how
the first task had proceeded.</p>
        <p>Task #2: Paper Plate and Beans Task. Students had 10 minutes to use one paper
plate, two feet of tape, three wooden sticks and four straws. These materials were used
to build a structure that could support a mass of 0.5 lb. as high off the table as possible
(see Figures 8 &amp; 9 for examples). This task looked to build on similar principles as the
first task, but with greater variability in the materials, and with the added goal of height
optimization. Note: Both Task #1 and Task #2 sit at the fuzzy intersection of
engineering and making which is currently be advanced by a number of schools, museums,
government organizations, etc. [7], [13]. The height of each pair’s design was recorded.
Pre-, Mid- and Post-tests. Students were asked to identify principles or mechanisms
in three example structures (see Figures 5-7) as pre-, mid- and post-tests (note: they
weren’t framed as tests, but as ways to help the students do better on the tasks). Students
were given approximately 1 minute to write down their ideas for why each structure
was stable. Students had access to their prior responses and were allowed to copy or
update their previous ideas. Responses to these questions served as the basis for
documenting conceptual change.</p>
        <p>Following the post test, students also answered a number of questions about their
experience, and completed two transfer tasks that asked them to compare different
designs, and to describe how they would teach a younger student how to go about
completing a similar engineering design task.</p>
        <p>Pre-test
Intervention #1
Paper &amp; Textbook</p>
        <p>Questions for</p>
        <p>Expert
Intervention #2</p>
        <p>Mid-test
Paper Plate &amp; Beans</p>
        <p>Post-test</p>
        <p>EDA Stress Tests</p>
      </sec>
      <sec id="sec-2-3">
        <title>Multimodal Data</title>
        <p>In addition to the time-stamped task annotations and hand written artifacts collected for
this study, I was also able to collect data from a number of multimodal sensors. These
sensors include: 1 high definition web camera (audio/video data), Xbox Kinect
(multichannel audio, skeletal tracking and frontal images), Affectiva Q-sensor
(electro-dermal activation, and hand/wrist movement), three Leap motion controllers (3-axis
hand/wrist and prop movement). The high definition web camera was positioned
directly over top of the students and collected data at 30 frames per second (Figure 10).
The Xbox Kinect was positioned approximately 1.5 meters in front of the students and
collected skeletal tracking data at 10 frames per second, and frontal images (Figure 11)
at one frame per second. The audio from the Xbox Kinect included all four channels
and was captured at 16kHz. The Affectiva Q-sensor was worn on the wrist, and
collected data at a rate of 8 samples per second. Furthermore, two stress tests were
administered at the conclusion of the task in order to provide an individualized baseline for
each student under stress. Finally, a Leap motion sensor was positioned to the side of
each participant, and over the top (next to the web camera). Leap data was captured at
approximately 60 Hz.</p>
      </sec>
      <sec id="sec-2-4">
        <title>Multimodal Data Extraction</title>
        <p>The current dataset includes all of the raw data described above, in addition to several
pieces of data that were extracted from the different modalities. For example, transcripts
are available for all participants during both tasks. For Task #2 the transcripts are
timestamped, and have been aligned with the audio to simplify prosodic analysis, for
example. Additionally, the frontal images were used to provide second by second head pose
estimation and automatic facial expression analysis[14], [15]. In particular, I have an
estimate for the direction of each user’s gaze, evidence of facial action units, and
evidence of basic facial expressions. Finally, demographic information about each student,
their performance in school, high school grade point average, etc. is also available.
4</p>
      </sec>
    </sec>
    <sec id="sec-3">
      <title>Challenges</title>
      <p>A particular challenge in collecting this data set was synchronizing the data across the
different machines being used to collected the different modalities. Part of this process
was simplified by running synchronization tasks before several of experiments, but
considerable effort was still required to properly synchronize data from the different
modalities. Additionally, providing accurate real-time event annotation was a challenge
(though in the end this greatly eased the synchronization process). Another challenge
encountered was the sporadic nature of some of the data collection tools. For example,
overhead video, audio, and/or frontal images is missing from a number of pairs. This
lost data is largely the result of software failing to operate as expected.</p>
      <p>Similar challenges exist in analyzing and visualizing the current data in such a way
that is meaningful. Many of the analytic tools available cater towards working in a
particular modality, but there seem to be few tools that can effectively be used with
multiple modalities, outside of custom scripts in MATLAB and/or Python.</p>
    </sec>
  </body>
  <back>
    <ref-list>
      <ref id="ref1">
        <mixed-citation>
          <string-name>
            <surname>Blikstein</surname>
            ,
            <given-names>P.</given-names>
          </string-name>
          &amp;
          <string-name>
            <surname>Worsley</surname>
            ,
            <given-names>M.</given-names>
          </string-name>
          <article-title>Multimodal Learning Analytics: a methodological framework for research in constructivist learning</article-title>
          .
          <source>Journal of. Learning Analytics: 3</source>
          ,
          <fpage>220</fpage>
          -
          <lpage>238</lpage>
          (
          <year>2016</year>
          ).
        </mixed-citation>
      </ref>
      <ref id="ref2">
        <mixed-citation>
          <string-name>
            <surname>Ochoa</surname>
            ,
            <given-names>X.</given-names>
          </string-name>
          &amp;
          <string-name>
            <surname>Worsley</surname>
            ,
            <given-names>M.</given-names>
          </string-name>
          <article-title>Augmenting Learning Analytics with Multimodal Sensory Data</article-title>
          .
          <source>Journal of. Learning Analytics</source>
          .
          <volume>3</volume>
          ,
          <fpage>213</fpage>
          -
          <lpage>219</lpage>
          (
          <year>2016</year>
          ).
        </mixed-citation>
      </ref>
      <ref id="ref3">
        <mixed-citation>
          <string-name>
            <surname>Worsley</surname>
            ,
            <given-names>M.</given-names>
          </string-name>
          <article-title>Multimodal learning analytics: enabling the future of learning through multimodal data analysis and interfaces</article-title>
          .
          <source>in Proceedings of the 14th ACM international conference on Multimodal interaction 353-356</source>
          (
          <year>2012</year>
          ).
        </mixed-citation>
      </ref>
      <ref id="ref4">
        <mixed-citation>
          <string-name>
            <given-names>National</given-names>
            <surname>Research Council</surname>
          </string-name>
          .
          <article-title>A Framework for</article-title>
          K-12
          <source>Science Education: Practices</source>
          ,
          <string-name>
            <given-names>Crosscutting</given-names>
            <surname>Concepts</surname>
          </string-name>
          , and
          <string-name>
            <surname>Core Ideas.</surname>
          </string-name>
          (National Academies Press,
          <year>2012</year>
          ).
        </mixed-citation>
      </ref>
      <ref id="ref5">
        <mixed-citation>
          <string-name>
            <given-names>Next</given-names>
            <surname>Generation Science Standards: For States</surname>
          </string-name>
          , By States. (
          <year>2013</year>
          ).
        </mixed-citation>
      </ref>
      <ref id="ref6">
        <mixed-citation>
          <string-name>
            <surname>Martin</surname>
            ,
            <given-names>L.</given-names>
          </string-name>
          <article-title>The promise of the Maker Movement for education</article-title>
          .
          <source>J</source>
          .
          <string-name>
            <surname>Pre-College Eng</surname>
          </string-name>
          .
          <source>Educ. Res. 5</source>
          ,
          <issue>4</issue>
          (
          <year>2015</year>
          ).
        </mixed-citation>
      </ref>
      <ref id="ref7">
        <mixed-citation>
          <string-name>
            <surname>Honey</surname>
            ,
            <given-names>M.</given-names>
          </string-name>
          &amp;
          <string-name>
            <surname>Kanter</surname>
            ,
            <given-names>D. E.</given-names>
          </string-name>
          <string-name>
            <surname>Design</surname>
          </string-name>
          , Make,
          <source>Play: Growing the Next Generation of STEM Innovators. (Routledge</source>
          ,
          <year>2013</year>
          ).
        </mixed-citation>
      </ref>
      <ref id="ref8">
        <mixed-citation>
          <string-name>
            <surname>Kolodner</surname>
          </string-name>
          , J. L. et al.
          <article-title>Journal of the Learning Problem-Based Learning Meets Case-Based Reasoning in the Middle-School Science Classroom : Putting Learning by Design ( tm ) Into Practice</article-title>
          .
          <volume>37</volume>
          -
          <fpage>41</fpage>
          (
          <year>2009</year>
          ).
        </mixed-citation>
      </ref>
      <ref id="ref9">
        <mixed-citation>
          <source>doi:10</source>
          .1207/S15327809JLS1204 Schwartz,
          <string-name>
            <surname>D. L.</surname>
          </string-name>
          <article-title>Constructivism in an age of non-constructivist assessments</article-title>
          .
        </mixed-citation>
      </ref>
      <ref id="ref10">
        <mixed-citation>
          <string-name>
            <surname>Barron</surname>
            ,
            <given-names>B.</given-names>
          </string-name>
          et al.
          <article-title>Doing With Understanding: Lessons From Research on Problem and Project-Based Learning</article-title>
          .
          <source>J. Learn. Sci. 7</source>
          ,
          <fpage>271</fpage>
          -
          <lpage>311</lpage>
          (
          <year>1998</year>
          ).
        </mixed-citation>
      </ref>
      <ref id="ref11">
        <mixed-citation>
          <string-name>
            <surname>Worsley</surname>
            ,
            <given-names>M.</given-names>
          </string-name>
          &amp;
          <string-name>
            <surname>Blikstein</surname>
            ,
            <given-names>P.</given-names>
          </string-name>
          <article-title>Leveraging Multimodal Learning Analytics to Differentiate Student Learning Strategies</article-title>
          .
          <source>in Proceedings of the Fifth International Conference on Learning Analytics And Knowledge</source>
          <volume>360</volume>
          -
          <fpage>367</fpage>
          (
          <issue>ACM</issue>
          ,
          <year>2015</year>
          ). doi:
          <volume>10</volume>
          .1145/2723576.2723624 Worsley,
          <string-name>
            <given-names>M.</given-names>
            &amp;
            <surname>Blikstein</surname>
          </string-name>
          ,
          <string-name>
            <surname>P.</surname>
          </string-name>
          <article-title>Using learning analytics to study cognitive disequilibrium in a complex learning environment</article-title>
          .
          <source>Proc. Fifth Int. Conf. Learn.</source>
        </mixed-citation>
      </ref>
      <ref id="ref12">
        <mixed-citation>
          <string-name>
            <given-names>Anal.</given-names>
            <surname>Knowl</surname>
          </string-name>
          . - LAK '
          <volume>15</volume>
          <fpage>426</fpage>
          -
          <lpage>427</lpage>
          (
          <year>2015</year>
          ). doi:
          <volume>10</volume>
          .1145/2723576.2723659 Vossoughi,
          <string-name>
            <given-names>S.</given-names>
            &amp;
            <surname>Bevan</surname>
          </string-name>
          ,
          <string-name>
            <surname>B.</surname>
          </string-name>
          <article-title>Making and tinkering: A review of the literature</article-title>
          .
        </mixed-citation>
      </ref>
      <ref id="ref13">
        <mixed-citation>
          <source>National Research Council Committee on Out of School Time STEM</source>
          . (
          <year>2014</year>
          ).
        </mixed-citation>
      </ref>
      <ref id="ref14">
        <mixed-citation>
          <string-name>
            <surname>Baltrušaitis</surname>
            ,
            <given-names>T.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Robinson</surname>
            ,
            <given-names>P.</given-names>
          </string-name>
          &amp;
          <string-name>
            <surname>Morency</surname>
            ,
            <given-names>L.-P.</given-names>
          </string-name>
          <article-title>OpenFace: an open source facial behavior analysis toolkit</article-title>
          .
          <source>in IEEE Winter Conference on Applications of Computer Vision</source>
          (
          <year>2016</year>
          ).
        </mixed-citation>
      </ref>
      <ref id="ref15">
        <mixed-citation>
          <string-name>
            <surname>Bartlett</surname>
            ,
            <given-names>M.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Littlewort</surname>
            ,
            <given-names>G.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Wu</surname>
            ,
            <given-names>T.</given-names>
          </string-name>
          &amp;
          <string-name>
            <surname>Movellan</surname>
            ,
            <given-names>J.</given-names>
          </string-name>
          <article-title>Computer expression recognition toolbox</article-title>
          .
          <source>in Automatic Face &amp; Gesture Recognition</source>
          ,
          <year>2008</year>
          . FG'
          <volume>08</volume>
          .
        </mixed-citation>
      </ref>
      <ref id="ref16">
        <mixed-citation>
          <source>8th IEEE International Conference on 1-2</source>
          (
          <year>2008</year>
          ).
        </mixed-citation>
      </ref>
    </ref-list>
  </back>
</article>