<!DOCTYPE article PUBLIC "-//NLM//DTD JATS (Z39.96) Journal Archiving and Interchange DTD v1.0 20120330//EN" "JATS-archivearticle1.dtd">
<article xmlns:xlink="http://www.w3.org/1999/xlink">
  <front>
    <journal-meta />
    <article-meta>
      <title-group>
        <article-title>Collaborative Conceptual Exploration as a Tool for Crowdsourcing Domain Ontologies</article-title>
      </title-group>
      <contrib-group>
        <contrib contrib-type="author">
          <string-name>Sergei Obiedkov</string-name>
          <email>sergei.obj@gmail.com</email>
          <xref ref-type="aff" rid="aff0">0</xref>
        </contrib>
        <contrib contrib-type="author">
          <string-name>Nikita Romashkin</string-name>
          <email>romashkin.nikita@gmail.com</email>
          <xref ref-type="aff" rid="aff0">0</xref>
        </contrib>
        <aff id="aff0">
          <label>0</label>
          <institution>National Research University Higher School of Economics</institution>
          ,
          <addr-line>Moscow</addr-line>
          ,
          <country country="RU">Russia</country>
        </aff>
      </contrib-group>
      <fpage>58</fpage>
      <lpage>70</lpage>
      <abstract>
        <p>Domain ontologies are essential in disciplines as diverse as software engineering, medicine, or political science to name just a few. This paper describes an ongoing e ort to develop a methodology for collaborative ontology construction by geographically spread communities of experts and implement a web-based prototype supporting this methodology. A distinctive feature of the proposed approach is the use of conceptual exploration techniques, which make it possible to organize the process of ontology construction by automatically identifying and explicitly highlighting issues that remain to be addressed. Given a set of objects (facts, situations, etc.) of a subject domain, which is known to have considerably more such objects, and their uni ed descriptions in terms of presence or absence of certain attributes, a conceptual exploration system maintains a compact representation of implications behind the currently built ontology and o ers them for experts to accept or falsify by entering new objects or extending the description language with new attributes. Upon termination, exploration results in identi cation of a (relatively small) representative part of the domain from which a conceptual hierarchy of the entire domain can be automatically constructed. We consider theoretic, algorithmic, representational, and pragmatic issues of transforming the exploration methods into a toolset useful for domain experts.</p>
      </abstract>
    </article-meta>
  </front>
  <body>
    <sec id="sec-1">
      <title>-</title>
      <p>Domain ontologies provide common interfaces to knowledge accumulated in their
respective elds through achieving consensus on terminology used therein and
its meaning. Their construction is thus an important topic in knowledge
engineering. In this paper, we describe a project aimed at developing a platform for
online collaborative ontology construction by geographically distributed groups
of users that would not simply support sharing and common editing of formalized
knowledge, but would also include means for on-the- y validation of the ontology
being constructed and automatic generation of guidelines for its completion.</p>
      <p>
        The proposed approach is based on formal concept analysis, a mathematical
theory oriented at applications in knowledge representation, knowledge
acquisition, data analysis and visualization [
        <xref ref-type="bibr" rid="ref5">5</xref>
        ]. It provides tools for understanding
the structure of data given as a set of objects with certain descriptions, e.g., in
terms of their attributes, which is done by representing the data as a hierarchy
of concepts, or more exactly, a concept lattice (in the sense of lattice theory).
The objects, attributes, and relation between them constitute a formal context;
hence, the de nition of a concept is necessarily contextual. Every concept has
extent (the set of objects that fall under the concept) and intent (the set of
attributes or features that together are necessary and su cient for an object to
be an instance of the concept). Concepts are ordered in terms of being more
general or less general (i.e., covering more objects or fewer objects).
      </p>
      <p>
        The concept lattice, being a rather universal structure, provides a wealth
of information about the relations among objects and attributes, which made
possible applications in areas ranging from sociology [
        <xref ref-type="bibr" rid="ref8">8</xref>
        ] to ontology construction
[
        <xref ref-type="bibr" rid="ref14">14</xref>
        ]. Indeed, it can help in processing a wide class of data types providing a
framework in which various data analysis and knowledge acquisition techniques
can be formulated.
      </p>
      <p>
        One such technique is attribute exploration [
        <xref ref-type="bibr" rid="ref4">4</xref>
        ], which, in its basic version,
can be summarized as follows. Given a set of objects of a domain, which is known
to have many more such objects, and their descriptions in terms of presence or
absence of certain attributes, attribute exploration builds an implicational theory
of the entire domain and a representative set of its objects. The implicational
theory is the set of sentences of the form \if an object has all attributes from
set A, then it also has all attributes from set B" that are believed to hold for
all objects of the domain. It is always possible to nd a minimal representation
of such a set, the Duquenne{Guigues or canonical basis of implications, from
which all valid implications can be inferred [
        <xref ref-type="bibr" rid="ref6">6</xref>
        ]. The representative set of objects
must respect all implications from the generated canonical basis and provide
a counterexample for every implication that cannot be inferred from the basis
(in this case, the concept lattice of the domain is isomorphic, i.e., structurally
identical, to the concept lattice of this relatively small set of objects).
      </p>
      <p>
        The process of attribute exploration is interactive: it consists in computer
suggesting implications and the user accepting them or providing
counterexamples. Attribute exploration is designed to be e cient, i.e., to suggest as few
implications as possible without loss in completeness of the result. It can work even if
the initial set of objects is empty. Attribute exploration is domain-independent:
although its application is more straightforward in precise domains such as
mathematics [
        <xref ref-type="bibr" rid="ref11">11</xref>
        ], [
        <xref ref-type="bibr" rid="ref12">12</xref>
        ], it can certainly be used in other elds [
        <xref ref-type="bibr" rid="ref10">10</xref>
        ].
      </p>
      <p>
        Advanced versions of attribute exploration can take into account background
information, such as relations among attribute values (not necessarily given as
implications), thus, avoiding suggesting trivial implications [
        <xref ref-type="bibr" rid="ref4">4</xref>
        ]. There are also
methods of concept exploration, relational exploration, and rule exploration that,
each in its own way, generalize attribute exploration, making it possible to work
with a broader class of dependencies [
        <xref ref-type="bibr" rid="ref13 ref15 ref16">15, 13, 16</xref>
        ]. We refer to all such methods
collectively as conceptual exploration.
      </p>
      <p>Our aim is to extend these methods so as to make them suitable for
collaborative ontology construction over the web by a geographically spread community
of researchers working in the same domain. The idea is to provide tools that
would allow them to use attribute exploration and related techniques to
rene the language in which their domain is described and boost their knowledge
about it. Being suggested, an implication will be accepted only if no expert has
a counterexample for it. In the end, there will be a list of implications correctly
describing objects under study and a representative context. If, at a later stage,
a counterexample becomes known, it is added to the context and the implication
basis is modi ed accordingly. Thus, an up-to-date list of open problems of the
domain can be maintained.</p>
      <p>We start with a short introduction into the relevant aspects of formal concept
analysis and then de ne attribute exploration. After that, we discuss possible
ways to make this procedure collaborative. Finally, we describe the current state
of our web-based exploration system and the work still to be done.
2</p>
    </sec>
    <sec id="sec-2">
      <title>Concept Lattices</title>
      <p>
        We brie y introduce necessary mathematical de nitions [
        <xref ref-type="bibr" rid="ref5">5</xref>
        ] and then explain
them less formally. Given a (formal) context K = (G; M; I), where G is called a
set of objects, M is called a set of attributes, and the binary relation I G M
speci es which objects have which attributes, the derivation operators ( )I are
de ned for A G and B M as follows:
      </p>
      <p>AI = fm 2 M j 8g 2 A : gImg</p>
      <p>BI = fg 2 G j 8m 2 B : gImg
In words, AI is the set of attributes common to all objects of A and BI is the
set of objects sharing all attributes of B.</p>
      <p>If this does not result in ambiguity, ( )0 is used instead of ( )I . The double
application of ( )0 is a closure operator, i.e., ( )00 is extensive, idempotent, and
monotonous. Therefore, sets A00 and B00 are said to be closed.</p>
      <p>A (formal) concept of the context (G; M; I) is a pair (A; B), where A G,
B M , A = B0, and B = A0. In this case, we also have A = A00 and B = B00.
The set A is called the extent and B is called the intent of the concept (A; B).</p>
      <p>A concept (A; B) is a subconcept of (C; D) if A C (equivalently, D B).
The concept (C; D) is then called a superconcept of (A; B). We write (A; B)
(C; D). The set of all concepts ordered by forms a lattice, which is called the
concept lattice of the context K:</p>
      <p>The formal context makes precise the scope of the discussion by specifying
the domain to which it applies (listing all the objects of this domain) and de ning
the terms in which it is going to be discussed (listing the attributes to be used
in object descriptions).</p>
      <p>To esh this out a bit, we give a small example based on the data from
the O*NET Resource Center (http://www.onetcenter.org/), which essentially
provides an interface to a taxonomy of occupations, organizing occupations in
various groups and describing the knowledge, skills, and abilities required by
t
n
e
m
isc ega
on gn an
tr i</p>
      <sec id="sec-2-1">
        <title>Computer and Information</title>
      </sec>
      <sec id="sec-2-2">
        <title>Research Scientists</title>
      </sec>
      <sec id="sec-2-3">
        <title>Computer Programmers</title>
      </sec>
      <sec id="sec-2-4">
        <title>Mathematicians Fig. 1. A formal concept of some computer and mathematical occupations.</title>
        <p>each occupation. Here, we focus on the knowledge required by occupations from
the Computer and Mathematical job family.</p>
        <p>A formal context encompassing three occupations is shown in Fig. 1. Here,
rows correspond to objects (occupations) and columns correspond to attributes
(areas of knowledge). A cross indicates that the corresponding object has the
corresponding attribute; in our case, it means that an occupation requires
knowledge in a certain area.</p>
        <p>A line diagram of the concept lattice of this context is shown in Fig. 2.
Nodes correspond to formal concepts, with more general concepts placed above
less general ones. Two concepts are connected with a line if one is more general
than the other and there is no concept between the two. Every concept in the
diagram is described extensionally, by a group of objects, and intensionally, by
attributes shared by all the objects in the extent of this node. The names of the
objects in the extent of a node can be read o from the diagram by looking at
the labels immediately below this node and below all nodes that can be reached
from this node by downward arcs. Conversely, the set of attributes forming the
intent of a node consists of labels immediately above this node and those above
nodes that can be reached from this node by upward arcs. For example, the
bottom-right node corresponds to the concept whose extent consists of a single
occupation, Computer and Information Research Scientists, whereas its intent
includes all attributes but Physics. The top concept is labelled by Computers
and Electronics and by Mathematics, which means that all the three occupations
require knowledge in these two areas.</p>
        <p>It can be seen from this diagram that all occupations in our context requiring
knowledge in Education and Training also require knowledge in Administration
and Management. This is formally captured by the notion of an implication,</p>
        <sec id="sec-2-4-1">
          <title>Computers and Electronics</title>
        </sec>
        <sec id="sec-2-4-2">
          <title>Mathematics</title>
        </sec>
        <sec id="sec-2-4-3">
          <title>Physics</title>
        </sec>
      </sec>
      <sec id="sec-2-5">
        <title>Mathematicians</title>
        <sec id="sec-2-5-1">
          <title>Administration and</title>
        </sec>
        <sec id="sec-2-5-2">
          <title>Management</title>
        </sec>
      </sec>
      <sec id="sec-2-6">
        <title>Computer Programmers</title>
        <sec id="sec-2-6-1">
          <title>Education and Training</title>
        </sec>
      </sec>
      <sec id="sec-2-7">
        <title>Computer and Information</title>
      </sec>
      <sec id="sec-2-8">
        <title>Research Scientists</title>
        <p>which is, formally, an expression A ! B, where A; B M are attribute subsets.
It holds or is valid in the context if A0 B0, i.e., every object of the context
that has all attributes from A also has all attributes from B.</p>
        <p>An attribute subset X M respects or is a model of an implication A ! B
if A 6 X or B X. Obviously, an implication holds in a context (G; M; I) if
and only if fgg0 respects the implication for all g 2 G. If an object g 2 G is
such that fgg0 is not a model of A ! B, we will call g a counterexample to this
implication.</p>
        <p>If A0 = ? for A M , then the implication A ! M necessarily holds in the
context. We will sometimes write such an implication as A ! ?, with ? standing
for \contradiction", meaning that attributes of A never occur all together.</p>
        <p>All valid implications of the context can be summarized by means of the
Duquenne{Guigues basis :
fP ! P 00 n P j P</p>
        <p>
          M is pseudo-closedg;
where a set P M is recursively de ned to be pseudo-closed if P 6= P 00 and
Q00 P for every pseudo-closed Q P . The models of these implications are
precisely the models of all implications valid in the context, and the Duquenne{
Guigues basis has the smallest number of implications among all implication sets
with this property [
          <xref ref-type="bibr" rid="ref6">6</xref>
          ]. Any valid implication can be inferred from the basis using
Armstrong rules [
          <xref ref-type="bibr" rid="ref2">2</xref>
          ].
        </p>
        <p>The Duquenne{Guigues basis of the context in Fig. 1 consists of three
implications shown in Fig. 3. Although they look similar, it may be more intuitive
to read them di erently:
{ The rst implication says that all occupations require knowledge both in</p>
        <p>Computers and Electronics and in Mathematics.
{ The second implication says that, if an occupation requires knowledge in
Computers and Electronics, Mathematics, and Education and Training, then
it also requires knowledge in Administration and Management.
? ! fComputers and Electronics; Mathematicsg
fComputers and Electronics; Mathematics; Education and Trainingg
fAdministration and Managementg
fComputers and Electronics; Mathematics; Administration and Management; Physicsg
! fEducation and Trainingg 0
!
1
{ The third implication means that there are no occupations requiring at the
same time knowledge in Computers and Electronics, Mathematics,
Administration and Management, and Physics.</p>
        <p>These implications are valid in our context, but the context contains only three
occupations. How do we know if the three implications are valid for all computer
and mathematical occupations out there? This is where attribute exploration
becomes useful.
3</p>
      </sec>
    </sec>
    <sec id="sec-3">
      <title>Attribute Exploration</title>
      <p>The idea of attribute exploration is simple: consider every implication in the
Duquenne{Guigues basis of the context and add a counterexample if the
implication is not valid generally. To be more precise, we are dealing with two contexts
here: one corresponds to the entire subject domain (computer and mathematical
occupations, in our case) and may not be immediately observable in its entirety,
while the other contains only a selection of objects. This smaller context is the
one that we can put into a cross-table such as one in Fig. 1 and for which we can
compute the implication basis. The goal is to make the smaller context
representative of the larger context in a very precise sense: the two contexts must be
models of exactly the same implications. Implications valid in the larger context
are always valid in the smaller context; we can force the converse to become true
as well by adding counterexamples to implications valid in the smaller context,
but not in the larger one. When we are done, the two contexts will share the
same implication basis and, furthermore, have isomorphic concept lattices: in
other words, the concept intents will be the same in the two contexts.</p>
      <p>As an example, consider again the lattice in Fig. 2 and the corresponding
implication basis in Fig. 3. The implication</p>
      <p>? ! fComputers and Electronics; Mathematicsg
is valid in our current context in Fig. 1, but according to the O*NET data,
knowledge of Mathematics is not really essential for Clinical Data Managers,
whose role is to \apply knowledge of health care and database management to
analyze clinical data, and to identify and report trends". Out of the ve areas
we consider, they must have knowledge of Computers and Electronics and of
Administration and Management. Therefore, we add Clinical Data Managers to
our context as a counterexample to the above implication:</p>
      <p>Clinical Data Managers</p>
      <p>Note that Clinical Data Managers still must have knowledge of Computers
and Electronics; therefore, we still have to consider the implication
? ! fComputers and Electronicsg:
It seems that knowledge of Computers and Electronics is a requirement for all
computer and mathematical occupations; so, we accept the implication.</p>
      <p>Since we changed the context by adding a new object, the implication basis
has also changed. Now, it includes the implication</p>
      <p>fComputers and Electronics; Physicsg ! fMathematicsg
suggesting that, for an occupation within the job family under consideration,
knowledge of physics is useless without knowledge of mathematics. If we are to
believe the O*NET data, the knowledge of Physics is not required for anyone
within this job family but Mathematicians, for whom the knowledge of
mathematics is obviously a must. Thus, we accept the implication.</p>
      <p>For the same reason, we accept the third implication in Fig. 2: the only
occupation category requiring knowledge of Physics is Mathematicians, but it
does not require knowledge in Administration and Management; hence, there
is no occupation satisfying the premise of the implication and the implication
trivially holds.</p>
      <p>In our modi ed context, there remains only one implication to consider: Is it
true that, if an occupation requires knowledge both in Computers and Electronics
and in Education and Training, then it also requires knowledge in Mathematics
and in Administration and Management (as it is the case with Computer and
Information Research Scientists)? The answer is negative, because this does not
hold for Informatics Nurse Specialists, who are there to \apply knowledge of
nursing and informatics to assist in the design, development, and ongoing
modi cation of computerized health care systems". They \may educate sta [. . . ]
to promote the implementation of the health care system", and, therefore, need
knowledge in Education and Training, but knowledge of Mathematics is not a
requirement for them. We add a new object to our context:</p>
      <p>Informatics Nurse Specialists
Proceeding likewise, we add</p>
      <p>Statisticians
to give an example of those who, according to O*NET, typically need knowledge
of Education and Training, but not of Administration and Management; accept
the implication indicating that no occupation in this job family typically requires
the simultaneous knowledge of Physics and of Education and Training; and we
are done. The resulting concept lattice is shown in Fig. 4. Below, we present the
resulting implication basis:1
{ ? ! fComputers and Electronicsg
{ fPhysicsg ! fMathematicsg
{ fAdministration and Management; Physicsg ! ?
{ fEducation and Training; Physicsg ! ?</p>
      <sec id="sec-3-1">
        <title>Computers and Electronics</title>
        <p>Mathematics Education and Training aAnddmMinaisntaragteimonent</p>
        <sec id="sec-3-1-1">
          <title>Clinical Data</title>
        </sec>
        <sec id="sec-3-1-2">
          <title>Managers</title>
        </sec>
      </sec>
      <sec id="sec-3-2">
        <title>Physics</title>
        <sec id="sec-3-2-1">
          <title>Mathematicians Computer Programmers</title>
        </sec>
        <sec id="sec-3-2-2">
          <title>Statisticians Informatics Nurse</title>
        </sec>
        <sec id="sec-3-2-3">
          <title>Specialists</title>
        </sec>
        <sec id="sec-3-2-4">
          <title>Computer and Information</title>
        </sec>
        <sec id="sec-3-2-5">
          <title>Research Scientists Fig. 4. The concept lattice resulting from attribute exploration on the context in Fig. 1.</title>
          <p>Although we have considered only a small fraction of computer and
mathematical occupations, these four implications completely characterize the
implicational theory of all such occupations (at least, as long as we trust the O*NET
1 Here, implication premises are given in an abbreviated form using their so-called
minimal generating sets: instead of a pseudo-closed set P , we use its minimal subset
Q satisfying Q00 = P 00. The resulting implication set is equivalent to the Duquenne{
Guigues basis.
data) and the six occupations covered by the concept lattice in Fig. 4 are
representative of the job family in this sense. The concept lattice contains all relevant
combinations of knowledge areas that may be required for computer and
mathematical occupations, and each such occupation can be placed into one of the ten
concepts of this lattice. For example, Web Administrators must have knowledge
of Computers and Electronics and of Administration and Management, but not
of any of the other three areas; therefore, they are covered by the same concept
as Clinical Data Managers.</p>
          <p>But this is counterintuitive: surely, Web Administrators might need
knowledge in areas foreign to Clinical Data Managers. The problem is that the six
occupations we identi ed during attribute exploration and the concept lattice
generated from them are representative only with respect to the ve knowledge
areas we have chosen before. However, the choice of knowledge areas might not
be representative itself. To correct this, we may use object exploration, a process
dual to attribute exploration, which involves working with object implications,
such as
fInformatics Nurse Specialistsg ! fComputer and Information Research
Scientistsg.</p>
          <p>In our case, this implication is interpreted as follows: \Every knowledge area
essential for Informatics Nurse Specialists is also important for Computer and
Information Research Scientists." This is not so, since Informatics Nurse
Specialists are typically required to have knowledge also in Medicine and Dentistry,
which forces us to introduce a new attribute into the context. Thus, object
exploration ensures that the language we use to describe the domain is su ciently
rich to di erentiate between what must be di erentiated.</p>
          <p>
            Alternating between object and attribute exploration in this way, we can
arrive at a complete conceptual hierarchy (in the form of a concept lattice) of
the domain under consideration. Such a hierarchy modeling the subconcept{
superconcept relationship is an essential part of any domain ontology. Advanced
versions of attribute exploration, such as rule exploration [
            <xref ref-type="bibr" rid="ref16">16</xref>
            ], may be used to
identify other types of relations between objects or classes of objects.
4
          </p>
        </sec>
      </sec>
    </sec>
    <sec id="sec-4">
      <title>Making Exploration Collaborative</title>
      <p>The construction of the concept lattice is automatic, but the data needed for that
is gathered through exploration in an interactive fashion: the exploration system
asks the user questions and the user replies positively by accepting an implication
or negatively by providing a counterexample. Extending this approach to the case
of many users will make it possible to apply it in construction of ontologies for
large domains of which experts have only partial and mutually complementary
knowledge.</p>
      <p>It is important to note that users of such a system must not be knowledge
engineers: they never have to explicitly describe the ontology of the domain
under consideration using a specialized knowledge representation formalism.
Instead, they only supply data su cient for building such ontology automatically,
while the exploration system directs them through questions/implications. If
exploration is used within a scienti c project, the users will typically be domain
experts. In other cases, it is possible that they even come from the \general
public". For example, if we were to collect data about occupations, we would
welcome input not only from job market research analysts, but also from people
working in HR departments, who know their companies' requirements, as well as
from job seekers, who know about such requirements from their own experience.
Thus, conceptual exploration provides a foundation for a toolset for
crowdsourcing data based on which a domain ontology (or its part) can be automatically
constructed.</p>
      <p>
        One of the rst experiments of collaborative exploration was conducted within
lattice theory already twenty years ago. The aim was to study relations between
various properties of lattices: algebraic, atomistic, nite, etc., and, indeed,
interesting non-trivial dependencies have been discovered [
        <xref ref-type="bibr" rid="ref11">11</xref>
        ]. In that setting, valid
implications corresponded to theorems, which had to be formally proven. The
resulting concept lattice (of lattice properties) was found valuable by lattice
theorists.
      </p>
      <p>Obviously, the Internet together with appropriate software can make the
process much better organized and more realizable in practice. We experimented
with building a web system supporting attribute exploration within a joint
project with political scientists at Higher School of Economics in Moscow. The
project had as one of its goals developing a hierarchy of concepts of democracy
de ned in literature and implemented in practice. The focus was on so-called
\defective democracies", such as \delegative democracy" or \exclusive democracy".
To make the construction of this concept hierarchy easier and more transparent,
a website was set up in order to allow political scientists to collectively build a
corresponding formal context using attribute exploration and related tools.
Objects of this context were countries and/or regimes described in political science
literature, and attributes were their various properties. During exploration, users
consider implications of the form:
\If a regime is characterized by all attributes from set A, it is also
characterized by all attributes from set B,"
which they can accept or reject. In the latter case, the user must describe a
regime characterized by all attributes from A, but not by all attributes from B.
Similarly, object implications of the form
\Every attribute characterizing all regimes in A is shared by all regimes
in B"
propose the user to di erentiate between regimes from sets A and B by adding
a new attribute shared by all regimes in A but missing from some regimes in B.</p>
      <p>It turned out however that there are a number of issues that must be
resolved before exploration techniques become a truly useful tool for people not
experienced in knowledge representation and formal concept analysis. Therefore,
we settled down to developing a general paradigmatic model of collaborative
conceptual exploration and its prototype implementation. A website is under
development that currently supports only the basic version of attribute
exploration, at the same time, providing some standard facilities of a collaborative
environment such as user pro les, separate projects, etc.</p>
      <p>Upon signing up, the user can start a new exploration project or join an
existing project by contacting its owner. Of course, one user may participate
in several projects. A typical work ow is as follows: project members start
attribute exploration by de ning an initial list of attributes (which can afterwards
be altered in any way) and continue by reviewing the dynamically updated
implication basis as described above, i.e., by accepting implications they deem valid
and providing counterexamples to other implications. The formal context
maintained in the process can be modi ed directly, as long as the modi cations do
not con ict with the accepted implications.</p>
      <p>
        There are many issues that need to be addressed to allow for exploration of
complex domains:
{ Object exploration was discussed earlier; it is essential for identifying features
di erentiating objects from each other.
{ Incomplete speci cation of examples. Users must still be able to add a
counterexample for an implication to the system even if they are unsure about
its status with respect to attributes not occurring in this implication. Its
description will be completed at later stages, either manually or
automatically (if accepting an implication forces certain values for some attributes
that have been left unspeci ed). This requires a modi ed de nition of the
implication basis allowing for incompletely speci ed objects.
{ Background knowledge. There must be a way to explicitly add known facts
about the domain to speed up exploration and concentrate on new
knowledge. The implication basis must be built relative to these facts summarizing
the part of knowledge behind the context not covered by them.
{ Algorithmic issues. Since, for large ontologies, the canonical basis may also
be large, algorithms allowing for incremental update [
        <xref ref-type="bibr" rid="ref9">9</xref>
        ] of and navigation
through the basis must be developed.
{ The policy of collaboration. It is essential to develop methods for resolution
of con icts arising when di erent users have di erent views on the status
of an implication or when new information becomes available providing a
counterexample for an already accepted implication.
{ Supporting tools. There must be a way for the user initiating any modi
cation to annotate it: provide a proof for an accepted implication, describe
the meaning of the new attribute, upload a document with evidence for the
new object, etc. In some domains, it may be possible to automatically prove
implications or search the Internet for potential counterexamples.
{ More complex languages for object description. Attribute exploration can
easily be extended to the case of many-valued contexts (non-binary object{
attribute tables) by means of conceptual scaling [
        <xref ref-type="bibr" rid="ref5">5</xref>
        ]. However, which scales
to use for a many-valued attribute should be a matter of collective decision;
support for organizing the process of decision making must be envisioned.
Another possible extension is to the case when objects are described by
arbitrary formulas rather than only by conjunctions of attributes.
{ Integration with state-of-the-art tools for ontology development, especially
those based on description logics. Attribute exploration has been used for
completing description logic knowledge bases, see, e.g., [
        <xref ref-type="bibr" rid="ref3">3</xref>
        ], although in a
slightly di erent context. Another relevant line of research to be taken into
account is learning ontologies in the framework of exact learning with queries
[
        <xref ref-type="bibr" rid="ref1">1</xref>
        ], which has a lot in common with attribute exploration; see, e.g., [
        <xref ref-type="bibr" rid="ref7">7</xref>
        ].
{ Non-implicational knowledge. In some cases, it may be desirable to use a
richer subset of propositional logic to summarize the theory of the domain.
{ Relational exploration. Attribute/object exploration is, for the most part,
concerned with the subsumption hierarchy of domain concepts. It may be
desirable to extract knowledge about other relations between concepts;
relational or rule [
        <xref ref-type="bibr" rid="ref16">16</xref>
        ] exploration may be adapted for this purpose.
{ Merging results of several explorations. It may be more appropriate for users
to work with subcontexts corresponding to their eld of expertise rather
than with a larger context. Methods for combining the results of several
explorations must be devised.
5
      </p>
    </sec>
    <sec id="sec-5">
      <title>Conclusion</title>
      <p>We believe that conceptual exploration techniques provide a powerful
framework for collaborative development of domain ontologies. Recent developments
in Web applications open new prospects for tools based on exploration. A proper
implementation may result in a platform e ectively supporting online social
networks of experts exchanging knowledge in a structured way and working
together towards constructing an ontology of their eld. Nevertheless, a
considerable amount of research, development, and testing is still needed to unleash
the full potential of conceptual exploration in a distributive environment.</p>
    </sec>
    <sec id="sec-6">
      <title>Acknowledgment</title>
      <p>The rst author was supported by the Russian Foundation for Basic Research
grant no. 14-01-93960.</p>
    </sec>
  </body>
  <back>
    <ref-list>
      <ref id="ref1">
        <mixed-citation>
          1.
          <string-name>
            <surname>Angluin</surname>
            ,
            <given-names>D.</given-names>
          </string-name>
          :
          <article-title>Queries and concept learning</article-title>
          .
          <source>Machine Learning</source>
          <volume>2</volume>
          ,
          <volume>319</volume>
          {
          <fpage>342</fpage>
          (
          <year>1988</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref2">
        <mixed-citation>
          2.
          <string-name>
            <surname>Armstrong</surname>
          </string-name>
          , W.:
          <article-title>Dependency structure of data base relationships</article-title>
          .
          <source>Proc. IFIP</source>
          Congress pp.
          <volume>580</volume>
          {
          <issue>583</issue>
          (
          <year>1974</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref3">
        <mixed-citation>
          3.
          <string-name>
            <surname>Baader</surname>
            ,
            <given-names>F.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Ganter</surname>
            ,
            <given-names>B.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Sertkaya</surname>
            ,
            <given-names>B.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Sattler</surname>
            ,
            <given-names>U.</given-names>
          </string-name>
          :
          <article-title>Description logic knowledge bases using formal concept analysis</article-title>
          . In: Veloso, M.M. (ed.)
          <source>Proceedings of the Twentieth International Joint Conference on Arti cial Intelligence (IJCAI'07)</source>
          . pp.
          <volume>230</volume>
          {
          <issue>235</issue>
          (
          <year>2007</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref4">
        <mixed-citation>
          4.
          <string-name>
            <surname>Ganter</surname>
            ,
            <given-names>B.</given-names>
          </string-name>
          :
          <article-title>Attribute exploration with background knowledge</article-title>
          .
          <source>Theoretical Computer Science</source>
          <volume>217</volume>
          (
          <issue>2</issue>
          ),
          <volume>215</volume>
          {
          <fpage>233</fpage>
          (
          <year>1999</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref5">
        <mixed-citation>
          5.
          <string-name>
            <surname>Ganter</surname>
            ,
            <given-names>B.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Wille</surname>
          </string-name>
          , R.:
          <source>Formal Concept Analysis: Mathematical Foundations</source>
          . Springer, Berlin/Heidelberg (
          <year>1999</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref6">
        <mixed-citation>
          6.
          <string-name>
            <surname>Guigues</surname>
            ,
            <given-names>J.L.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Duquenne</surname>
            ,
            <given-names>V.</given-names>
          </string-name>
          :
          <article-title>Famille minimale d'implications informatives resultant d'un tableau de donnees binaires</article-title>
          .
          <source>Mathematiques et Sciences Humaines</source>
          <volume>24</volume>
          (
          <issue>95</issue>
          ),
          <volume>5</volume>
          {
          <fpage>18</fpage>
          (
          <year>1986</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref7">
        <mixed-citation>
          7.
          <string-name>
            <surname>Konev</surname>
            ,
            <given-names>B.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Ozaki</surname>
            ,
            <given-names>A.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Wolter</surname>
            ,
            <given-names>F.</given-names>
          </string-name>
          :
          <article-title>A model for learning description logic ontologies based on exact learning</article-title>
          .
          <source>In: Proceedings of AAAI-16</source>
          (
          <year>2016</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref8">
        <mixed-citation>
          8.
          <string-name>
            <surname>Mohr</surname>
            ,
            <given-names>J.W.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Duquenne</surname>
            ,
            <given-names>V.</given-names>
          </string-name>
          :
          <article-title>The duality of culture and practice: Poverty relief</article-title>
          in New York City,
          <year>1888</year>
          {
          <year>1917</year>
          .
          <source>Theory and Society</source>
          <volume>26</volume>
          , 305{
          <fpage>356</fpage>
          (
          <year>1997</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref9">
        <mixed-citation>
          9.
          <string-name>
            <surname>Obiedkov</surname>
            ,
            <given-names>S.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Duquenne</surname>
            ,
            <given-names>V.</given-names>
          </string-name>
          :
          <article-title>Attribute-incremental construction of the canonical implication basis</article-title>
          .
          <source>Annals of Mathematics and Arti cial Intelligence</source>
          <volume>49</volume>
          (
          <issue>1-4</issue>
          ),
          <volume>77</volume>
          {99 (April
          <year>2007</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref10">
        <mixed-citation>
          10.
          <string-name>
            <surname>Obiedkov</surname>
            ,
            <given-names>S.A.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Kourie</surname>
            ,
            <given-names>D.G.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Elo</surname>
            ,
            <given-names>J.H.P.</given-names>
          </string-name>
          :
          <article-title>Building access control models with attribute exploration</article-title>
          .
          <source>Computers &amp; Security</source>
          <volume>28</volume>
          (
          <issue>1-2</issue>
          ), 2{
          <issue>7</issue>
          (
          <year>2009</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref11">
        <mixed-citation>
          11.
          <string-name>
            <surname>Reeg</surname>
            ,
            <given-names>S.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Wei</surname>
            ,
            <given-names>W.</given-names>
          </string-name>
          :
          <article-title>Properties of nite lattices</article-title>
          .
          <source>Diplomarbeit, TH Darmstadt</source>
          (
          <year>1990</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref12">
        <mixed-citation>
          12.
          <string-name>
            <surname>Revenko</surname>
            ,
            <given-names>A.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Kuznetsov</surname>
            ,
            <given-names>S.O.</given-names>
          </string-name>
          :
          <article-title>Attribute exploration of properties of functions on sets</article-title>
          .
          <source>Fundam. Inform</source>
          .
          <volume>115</volume>
          (
          <issue>4</issue>
          ),
          <volume>377</volume>
          {
          <fpage>394</fpage>
          (
          <year>2012</year>
          ), http://dx.doi.org/10.3233/ FI-2012-660
        </mixed-citation>
      </ref>
      <ref id="ref13">
        <mixed-citation>
          13.
          <string-name>
            <surname>Rudolph</surname>
            ,
            <given-names>S.</given-names>
          </string-name>
          : Relational Exploration|
          <article-title>Combining Description Logics and Formal Concept Analysis for Knowledge Speci cation</article-title>
          .
          <source>Universitatsverlag Karlsruhe (Dec</source>
          <year>2006</year>
          ), dissertation
        </mixed-citation>
      </ref>
      <ref id="ref14">
        <mixed-citation>
          14.
          <string-name>
            <surname>Stumme</surname>
            ,
            <given-names>G.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Maedche</surname>
            ,
            <given-names>A.</given-names>
          </string-name>
          :
          <article-title>FCA-MERGE: Bottom-Up Merging of Ontologies</article-title>
          .
          <source>In: Proc. of the 17th International Joint Conference on Arti cial Intelligence</source>
          . pp.
          <volume>225</volume>
          {
          <issue>234</issue>
          (
          <year>2001</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref15">
        <mixed-citation>
          15.
          <string-name>
            <surname>Stumme</surname>
          </string-name>
          , G.:
          <article-title>Concept exploration: knowledge acquisition in conceptual knowledge systems</article-title>
          .
          <source>Ph.D. thesis, TH Darmstadt</source>
          (
          <year>1997</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref16">
        <mixed-citation>
          16.
          <string-name>
            <surname>Zickwol</surname>
            ,
            <given-names>M.</given-names>
          </string-name>
          :
          <article-title>Rule exploration: rst order logic in formal concept analysis</article-title>
          .
          <source>Ph.D. thesis, TH Darmstadt</source>
          (
          <year>1991</year>
          )
        </mixed-citation>
      </ref>
    </ref-list>
  </back>
</article>