<!DOCTYPE article PUBLIC "-//NLM//DTD JATS (Z39.96) Journal Archiving and Interchange DTD v1.0 20120330//EN" "JATS-archivearticle1.dtd">
<article xmlns:xlink="http://www.w3.org/1999/xlink">
  <front>
    <journal-meta />
    <article-meta>
      <title-group>
        <article-title>Teaching Conceptual Modeling in ER: Chen Worlds</article-title>
      </title-group>
      <contrib-group>
        <contrib contrib-type="author">
          <string-name>Natalya G. Keberle</string-name>
          <email>nkeberle@gmail.com</email>
          <xref ref-type="aff" rid="aff1">1</xref>
        </contrib>
        <contrib contrib-type="author">
          <string-name>Ivan V. Utkin</string-name>
          <xref ref-type="aff" rid="aff0">0</xref>
          <xref ref-type="aff" rid="aff1">1</xref>
        </contrib>
        <contrib contrib-type="editor">
          <string-name>Key Terms. TeachingProcess, TeachingMethodology, ConceptualModeling</string-name>
        </contrib>
        <aff id="aff0">
          <label>0</label>
          <institution>Prima Development Group</institution>
        </aff>
        <aff id="aff1">
          <label>1</label>
          <institution>Zaporozhye National University</institution>
          ,
          <addr-line>Zhukovskogo st. 66 69063 Zaporozhye</addr-line>
          <country country="UA">Ukraine</country>
        </aff>
      </contrib-group>
      <fpage>222</fpage>
      <lpage>227</lpage>
      <abstract>
        <p>Whilst algorithmic modeling is taught intensively both in school and higher education, conceptual modeling, or modeling of data to be used by algorithms, is less highlighted in the teaching curricula. However, understanding basic conceptual modeling principles plays a very important role in practice, as the cost of wrong solutions taken at the level of conceptual modeling is usually high. Tools, accelerating learning of conceptual modeling, are rare, at least freely available or mentioned in the literature. We present our work in progress - a system called “Chen Worlds” reflecting the focus on ER paradigm of conceptual modeling, describe its use cases, architecture and technical solutions undertaken.</p>
      </abstract>
      <kwd-group>
        <kwd />
        <kwd>Conceptual modeling</kwd>
        <kwd>entity-relationship diagram</kwd>
        <kwd>teaching ER</kwd>
      </kwd-group>
    </article-meta>
  </front>
  <body>
    <sec id="sec-1">
      <title>-</title>
      <p>
        A course in databases and information systems, even introductory, pays attention to
the architecture of a database system and to the process of a database system building,
part of which is conceptual modeling. The analysis of student works in conceptual
modeling for databases, particularly in ER notation [
        <xref ref-type="bibr" rid="ref1">1</xref>
        ], has envisaged a set of
common mistakes, coming from either total misunderstanding of ER notation (rather
seldom, nevertheless), or from misconception of specific use cases. Provision of
detailed feedback on resolving misconceptions for interested students raises the
syntactical validity of their ER diagrams produced essentially. Without clear
understanding of the syntax of ER notation it is hard to produce conceptual models
which correctly reflect a domain at the level of semantics. We acknowledge that
teaching semantically correct conceptual modeling requires much more time for
practice, and restrict ourselves with a modest aim – to create an easy-to-use and
friendly environment to learn syntax of ER diagrams.
      </p>
      <p>The paper is structured as follows: in the Section 2, we describe and classify
patterns of misconception of ER notation, detected during the analysis of students’
works. Section 3 presents the related work and fruitful ideas inspiring the presented
approach to teaching ER. The architecture of the system “Chen Worlds” and the use
cases are sketched in the Section 4. Section 5 presents concluding remarks.
2</p>
    </sec>
    <sec id="sec-2">
      <title>Common Mistakes in Understanding ER</title>
      <p>The notation of ER uses small and simple set of basic constructs: entities, attributes
(including primary, optional, multiple), relationships, weak entities and
generalizations. Well known complex entities – aggregations and associations are
derived from this set.</p>
      <p>All the mistakes in understanding ER are divided into syntactic and semantic.
Unsurprisingly, usage of incorrect syntax is conditioned by a poor understanding of a
domain modeled. However, we have assumed that domain modeling skill requires
some experience and time, and during the introductory course we can only teach some
well known heuristics for semantically correct conceptual modeling.</p>
      <p>Let’s illustrate the idea with typical examples, taken from works of students. The
absence of a primary key for an entity is a syntax mistake (see Fig. 1a), whereas usage
of single-valued attribute as a primary key for an entity having all other properties as
multiple-valued attributes shows a semantic misconception of a domain (see Fig. 1b).</p>
      <p>There are more ambiguous examples, like the one depicted on Fig. 1c.
Nondifferentiation between the values of a property (e.g. “rain”, “snow’, “hoar-frost” etc.
are some values of a property “name” of an entity “Precipitation”) and composite
properties (e.g. “address” is a composite property with atomic parts “city”, “street”,
“house’, “apartment”, “postal code”) could be a semantic misconception.</p>
      <p>However, under certain circumstances, namely, assuming that the values of
properties “rain”, “snow”, “hoar-frost” are ranging in Boolean data type, the diagram,
depicted on Fig. 1c, could be semantically correct.
2. Duplication of attributes across several entities is not allowed, especially if the
attributes are foreign keys.
3. Primary key should be introduced for every concept in an ER diagram.
4. Primary key attributes cannot be optional, they should be mandatory.
5. Primary keys of relationships are not allowed on an ER diagram.
6. Cardinalities of a relationship are from the set {1, M}.
7. Participation cardinalities (optionality) of a relationship are from the set {0,1}.
8. A relationship cannot be directly related to another relationship.
9. An entity cannot be directly related to another concept.
10. Composite properties should have as parts only properties, not values of
properties.
11. Weak entity cannot be related to strong entity with cardinalities, different from
(1,1):(1,M).
12. Participation cardinalities for aggregate entity and its parts are to be mandatory.
13.</p>
      <p>To resolve at least these misconceptions and to help tutor and students in
understanding conceptual modeling in ER at least at the level of syntax we exploit
several techniques, described in the next section.</p>
    </sec>
    <sec id="sec-3">
      <title>3 Accelerating</title>
    </sec>
    <sec id="sec-4">
      <title>Environment</title>
    </sec>
    <sec id="sec-5">
      <title>Learning,</title>
    </sec>
    <sec id="sec-6">
      <title>Root</title>
    </sec>
    <sec id="sec-7">
      <title>Questions and</title>
    </sec>
    <sec id="sec-8">
      <title>Gaming</title>
      <p>
        According to the review of motivations for achievements in mathematics [
        <xref ref-type="bibr" rid="ref2">2</xref>
        ]: “…If
students realize that their successes are meaningful and result both from their abilities
and from a high degree of effort, they are likely to believe that they can do
mathematics if they try…”. Understanding conceptual modeling from our point of
view is more to efforts put to understand the notation first and to apply it for different
domains.
      </p>
      <p>The idea of accelerating learning when teaching conceptual modeling as part of a
course in databases and information systems has roots in one of the works of
Professor Emeritus Jeffrey Ullman1 at Stanford University. Prof. Ullman with
colleagues had created Gradiance On-Line Accelerated Learning System2 – a service
for on-line training in solving simple-to-tricky exercises in various fields of Database
Systems, Compilers, Automata Theory and Operation Systems.</p>
      <p>
        Following the idea of a root question [
        <xref ref-type="bibr" rid="ref3">3</xref>
        ] they had introduced sets of exercises,
addressing the main problems of the particular field of study, sets of possible
mistakes, and sets of hints (and in some cases, solutions) to avoid those mistakes.
      </p>
      <p>
        The idea of immersion of learning to gaming environment is widely used in
practice for various fields of study. Of particular interest are the systems for teaching
programming, like ALICE [
        <xref ref-type="bibr" rid="ref4">4</xref>
        ] – for object-oriented programming, PictoMir [
        <xref ref-type="bibr" rid="ref5">5</xref>
        ] and
KuMir [
        <xref ref-type="bibr" rid="ref6">6</xref>
        ] – well known and respected environments aiming at teaching children and
secondary school pupils basics of algorithmic thinking. The main distinguishable
1 http://infolab.stanford.edu/~ullman/
2 http://www.gradiance.com/
feature of such systems is immediate visualization of the choices a scholar makes, and
consequent visualization of the solution by the system.
      </p>
      <p>For example, in PictoMir there are: Environment – a part of the Universe, usually,
a plain surface made of squares, Robot-Performer – a robot, able to move on that
surface one step back or forth, rotate 90 degrees to the left and to the right, and
Learner – usually a child, that carries out a task, e.g. “Let Robot-Performer fill in with
some color all the squares at the corners of the surface”, writing an algorithm of the
kind “Move 1 step forth, Fill the square, Turn 90 degrees left, Move 1 step forth…”.
Algorithm writing is also replaced with picking up the proper symbol of the robotic
language, like “left arrow”, “turn 90 degrees left arrow” and so forth. The solution
proposed by a scholar is executed “as is” step by step, and it’s easily seen at what step
of the algorithm Robot-Performer fails.</p>
      <p>For Conceptual Modeling in ER we initially have a set of graphic primitives and a
set of connection rules for primitives. Adopting the idea of a root question we propose
an environment for building, visualization and validation of conceptual models in ER
notation, described in the next section.
4</p>
    </sec>
    <sec id="sec-9">
      <title>Chen Worlds: Use Cases, Architecture and Technical Solutions</title>
      <p>Chen Worlds – is a cross-platform software system for learning, building,
visualization and validation of conceptual models in ER notation.</p>
      <p>We have identified two main roles of actors in Chen Worlds: a Scholar that is a
person, who learns conceptual modeling for databases, and a Tutor, who prepares
teaching materials.</p>
      <sec id="sec-9-1">
        <title>A Scholar use case</title>
        <p>A Scholar uses GUI to access the system. His/her aim is to obtain certain
knowledge on how to use Chen notation in conceptual modeling for databases. The
system proposes a Scholar a set of examples, each of which is oriented on answering
of one root question.</p>
        <p>
          Each example is presented as a text, shown to a Scholar, or as a picture, depending
on the type of question. There are one or several correct answers (there are situations
when several solutions are correct in Chen notation), which are usually not shown to a
Scholar until he or she submits at least one solution. The questions are created
according to the root question technique [
          <xref ref-type="bibr" rid="ref3">3</xref>
          ].
        </p>
        <p>A Scholar may also use the system as a guide to notation, choosing a theme from a
list of themes dedicated to conceptual modeling in ER.</p>
      </sec>
      <sec id="sec-9-2">
        <title>A Tutor use case</title>
        <p>A Tutor uses either GUI or usual text editor to create examples of ER diagrams.
His/her aim is to construct different examples, demonstrating the applicability of a
particular problem addressed with a root question. A Tutor edits a set of rules for
detection of core misconceptions (see Section 2).</p>
        <p>The architecture of the system is presented in Fig. 2. The system consists of
VISUALIZER, SOLVER, PARSER and GUI components.</p>
        <p>VISUALISER shows a task (a text, a picture, or both), shows a hint in the text of a
task, and shows a hint in a picture. VISUALIZER provides a palette of graphical
primitives of Chen notation.</p>
        <p>In order to correctly present an ER diagram after reading its XML encoding,
VISUALISER performs initial layout task, assigning each entity, relationship and
generalization particular absolute places on a working space. Attributes, cardinalities
and optionalities are placed relatively to entities/relationships they belong to.</p>
        <p>With respect to the user role, VISUALISER may consult SOLVER and restrict the
applicability of elements of the ER notation, avoiding syntactically incorrect diagrams
(suitable in Tutor mode to save time). For Scholar mode VISUALISER allows
arbitrary combinations of elements.</p>
        <p>A solution submitted by a user is parsed into XML presentation and evaluated by
SOLVER. SOLVER compares a solution submitted by a Scholar with respect to a
correct answer of the task, detects mistakes, using a set of predefined rules, already
created by a Tutor, and depending on which rules are violated, returns to
VISUALISER a hint code for a user. If several mistakes are appeared in one
submitted solution, SOLVER first returns hint code of highest priority (which means
the most serious mistake), and in case a Scholar correctly resolves the addressed
mistake, re-checks the rules again.</p>
        <p>PARSER reads a file of a task (in XML), saves a solution (both in XML and in
graphical presentation) created by a Scholar, saves a task (both in XML and in
graphical form) created by a Tutor, reads a file of rules for detecting mistakes.</p>
        <p>
          Examples of tasks are written as XML documents with an XML schema,
substantially extending developed in [
          <xref ref-type="bibr" rid="ref7">7</xref>
          ] XML Schema with cardinalities, optionalities
and generalizations. Additional XML Schema was developed for presentation of rules
for detecting mistakes.
5
        </p>
      </sec>
    </sec>
    <sec id="sec-10">
      <title>Concluding Remarks</title>
      <p>Understanding conceptual modeling plays an important role in building useful and
extensible database systems. There are many tools facilitating the process of
construction of conceptual models (e.g. Computer Associates ERWin Data Modeler,
IBM Rational, a lot of others), but most of them just use correspondent notation (e.g.
IDEF1X, Crow’s Foot, UML etc.) and do not teach it. Tools facilitating the process of
learning conceptual modeling are rare.</p>
      <p>Taking classical ER notation as an example it is shown that there exist both
syntactic and semantic misconceptions of the notation itself, blocking its proper usage
and leading to serious problems in future database schema construction.</p>
      <p>Chen Worlds system is currently oriented on teaching classical ER notation,
however the principles of the system could be applied to other complex graphical
notations, e.g. IDEF1X, UML Class diagrams, as the hardest part of their teaching is
in preparation of sets of core syntactic (and in general, semantic) misconceptions.</p>
    </sec>
  </body>
  <back>
    <ref-list>
      <ref id="ref1">
        <mixed-citation>
          1.
          <string-name>
            <surname>Garsia-Molino</surname>
            ,
            <given-names>H.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Ullman</surname>
            ,
            <given-names>J.D.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Widom</surname>
          </string-name>
          , J.:
          <source>Database Systems: the Complete Book. Pearson Prentice Hall</source>
          ,
          <volume>1203</volume>
          p. (
          <year>2009</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref2">
        <mixed-citation>
          2.
          <string-name>
            <surname>Middleton</surname>
            ,
            <given-names>J. A.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Spanias</surname>
            ,
            <given-names>P.A.</given-names>
          </string-name>
          :
          <article-title>Motivation for Achievement in Mathematics: Findings, Generalizations, and Criticisms of the Research</article-title>
          .
          <source>J. Research in Mathematics Education</source>
          ,
          <volume>30</volume>
          (
          <issue>1</issue>
          ), pp.
          <fpage>65</fpage>
          --
          <lpage>88</lpage>
          (
          <year>1999</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref3">
        <mixed-citation>
          3.
          <string-name>
            <surname>Ullman</surname>
            ,
            <given-names>J. D.</given-names>
          </string-name>
          <string-name>
            <surname>Gradiance</surname>
          </string-name>
          <article-title>On-Line Accelerated Learning Guide for Authors</article-title>
          . http://www.gradiance.com/downloads/auth-guide.pdf
        </mixed-citation>
      </ref>
      <ref id="ref4">
        <mixed-citation>4. ALICE, http://www.alice.org</mixed-citation>
      </ref>
      <ref id="ref5">
        <mixed-citation>5. PictoMir, http://www.pictomir.ru</mixed-citation>
      </ref>
      <ref id="ref6">
        <mixed-citation>6. KuMir, http://www.niisi.ru/kumir</mixed-citation>
      </ref>
      <ref id="ref7">
        <mixed-citation>
          7.
          <string-name>
            <surname>Jin</surname>
            ,
            <given-names>S.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Kang</surname>
            ,
            <given-names>W.</given-names>
          </string-name>
          :
          <article-title>Mapping Rules for ER to XML Using XML schema</article-title>
          .
          <source>In: Proc. 10th Southern Association for Information Systems Conference</source>
          . Jacksonville, Florida, USA,
          <fpage>9</fpage>
          -
          <issue>10</issue>
          <year>March</year>
          ,
          <year>2007</year>
          , pp.
          <fpage>211</fpage>
          -
          <lpage>216</lpage>
          (
          <year>2007</year>
          )
        </mixed-citation>
      </ref>
    </ref-list>
  </back>
</article>