<!DOCTYPE article PUBLIC "-//NLM//DTD JATS (Z39.96) Journal Archiving and Interchange DTD v1.0 20120330//EN" "JATS-archivearticle1.dtd">
<article xmlns:xlink="http://www.w3.org/1999/xlink">
  <front>
    <journal-meta />
    <article-meta>
      <title-group>
        <article-title>A Multi-Ontology Approach for Personal Information Management ?</article-title>
      </title-group>
      <contrib-group>
        <contrib contrib-type="author">
          <string-name>Huiyong Xiao</string-name>
          <xref ref-type="aff" rid="aff0">0</xref>
        </contrib>
        <contrib contrib-type="author">
          <string-name>Isabel F. Cruz</string-name>
          <xref ref-type="aff" rid="aff0">0</xref>
        </contrib>
        <aff id="aff0">
          <label>0</label>
          <institution>Department of Computer Science University of Illinois at Chicago</institution>
          ,
          <country country="US">USA</country>
        </aff>
      </contrib-group>
      <abstract>
        <p>The increasingly huge volume of personal information stored in a desktop computer is characterized by disparate models, unstructured contents, and implicit knowledge. Aiming at a semantic rich environment, a number of Semantic Desktop frameworks have been proposed, concentrating on different aspects, including organization, manipulation, and visualization of the data. In this paper, we propose a layered and semantic ontology-based framework for personal information management, and we discuss its annotations, associations, and navigation. We also discuss query processing in two cases: query rewriting in a single personal information application, PIA, and that between two PIAs.</p>
      </abstract>
    </article-meta>
  </front>
  <body>
    <sec id="sec-1">
      <title>1 Introduction</title>
      <p>
        In 1945, Vannevar Bush put forward the first vision of personal information
management (PIM) system, Memex, by pointing out that the human mind “operates by
associations”, and we should “learn from it” in building Memex [
        <xref ref-type="bibr" rid="ref4">4</xref>
        ]. The Hypertext systems
(see the survey of Conklin [
        <xref ref-type="bibr" rid="ref6">6</xref>
        ]), which flourished in the 80’s, reinforced this vision and
yielded the current World Wide Web, in a broader scope. Recently, with the
Semantic Web vision [
        <xref ref-type="bibr" rid="ref2">2</xref>
        ], a number of PIM systems associated with that vision, hence called
Semantic Desktop, have been proposed. By summarizing these proposals and taking
into account the characteristics of personal information (PI), we propose the following
principles that a PIM system should follow:
Semantic data organization. Almost all existing approaches are trying to go beyond
the hierarchical directory model. The critical factors of semantic data organization
include adequate annotations, explicit semantics, meaningful associations, and a uniform
representation. A semantic-rich data organization has several advantages. First, the
annotations and associations (as the superimposed information over the coarse data [
        <xref ref-type="bibr" rid="ref19">19</xref>
        ])
form the context of the PI, thus making the data more easily understandable. Second, the
superimposed information also allows for a finer and more flexible manipulation (e.g.,
browsing and querying) of the data. Third, an explicit formal semantics for the data
can facilitate reasoning on the data and deriving new knowledge. Finally, the uniform
representation can support the integration of data that may be heterogeneous.
Flexible data manipulation. A PIM system can provide integration, exchange,
navigation, and query processing of the stored personal information. The framework of PIM,
? This work was partially supported by NSF Awards ITR IIS-0326284 and IIS-0513553.
      </p>
      <p>PI Space (C:\)
papers</p>
      <p>WISE03-1.pdf
WISE03-submission.pdf
WISE03-camera.pdf
JoDS05.pdf
photos</p>
      <p>WISE
myself.jpeg
talk.jpeg
with sam.jpeg
talks</p>
      <p>WISE03.ppt
IDEAS04.ppt
AP2PC04.ppt
Super-invited.ppt
emails</p>
      <p>Final submission of WISE.eml
Meeting on Monday.eml
WISE photos.eml</p>
      <p>
        Register for WISE.eml
including the data model, query language, and user interface, should provide multiple
ways to manipulate data in a powerful and flexible manner. Furthermore, a PIM system
should possess the capability for seamless communication (or interoperability) with
external sources (possibly in another PIM system), e.g., in a peer-to-peer (P2P) way [
        <xref ref-type="bibr" rid="ref26">26</xref>
        ].
Rich visualization. Multiple visualizations can help the user in understanding data.
Instead of providing separate views of the data as most traditional applications do, a
PIM system should support data visualization from different perspectives, to offer a
comprehensive view. Examples include association-centric visualization [
        <xref ref-type="bibr" rid="ref24">24</xref>
        ] and
timecentric visualization [
        <xref ref-type="bibr" rid="ref12 ref13">13, 12</xref>
        ].
      </p>
      <p>Example 1. Figure 1 presents a fragment of PI space, which consists of four directories
of files in the hard drive C:\. The papers directory contains four papers of the format
pdf, photos\WISE contains three pictures taken at the WISE ’03 conference, talks
contains four Powerpoint files that are respectively the slides of four talks, and emails
contains four saved email messages. Even if the concrete contents of all these files are
unknown, we can tell from their names (or the names of their respective directories)
that several of them appear to be related to one another. Unfortunately, their storage in
different and possibly unrelated directories does not show such inter-relationships, thus
resulting in possible difficulties in locating the wanted information. Some
keywordbased searching techniques, e.g., offered by the Google Desktop Search,1 can retrieve
all files that are relevant to WISE. However, without further inspection of the contents
of each file, the user may not be able to discover certain associations between them,
e.g., that file JoDS05.pdf is an extended journal paper of WISE03-camera.pdf.</p>
      <p>From this example, we can see that the lack of semantic associations among the
stored data could be a handicap for data and knowledge discovery. In this paper, we
focus on issues of semantic data organization and management in PIM, by taking the
following approach:</p>
      <p>
        1) We propose a layered framework for PIM, in which multiple ontologies playing a
variety of roles are employed. Specifically, the resource layer stores all the PI resources
(using URIs), metadata of the PI, and all kinds of associations using RDF. The domain
layer contains the ontologies specific to various domains that are used to structure the
data and categorize the resources. The application layer, built on top of the domain
layer, is where the user constructs different application ontologies for different purposes
of data usage. This layered architecture enables: i) a semantics-rich environment for
1 http://desktop.google.com
personal information management; ii) a flexible and reusable system, by decoupling the
domain and application ontologies, so that the construction of application ontologies
for different applications can reuse the underlying domain ontologies. We argue that
this provides certain advantages over the use of a single domain model for all the PI
(e.g., [
        <xref ref-type="bibr" rid="ref9">9</xref>
        ]).
      </p>
      <p>
        2) We discuss in detail how to utilize superimposed information for semantic
organization, focusing on the construction of resource-file and resource-resource
associations. We also present the idea of 3D navigation, which is a combination of the vertical,
horizontal and temporal navigation in the PI space. The idea is inspired by some
existing PIM systems including MyLifeBits [
        <xref ref-type="bibr" rid="ref13">13</xref>
        ] and Placeless Documents [
        <xref ref-type="bibr" rid="ref10">10</xref>
        ], and is
demonstrated in a browser.
      </p>
      <p>3) In our framework, the basic unit for the user to manage the Semantic Desktop
is the personal information application (PIA). Each PIA aims to accomplish or assist
a specific task (e.g., bibliography management, paper composition, and trip planning).
The PIAs can be standalone, with their own application ontology, user interface, and
workflows. Meanwhile, they can communicate with each other as if in a P2P network,
by means of the connections (mappings) established between their application
ontologies. In this sense, different PIAs interoperate at a semantic level. We describe query
processing in our framework in two cases: within a single PIA or between two PIAs, in
a P2P query processing mode.</p>
      <p>
        Among the existing approaches to PIM, the Gnowsis project from DFKI2 aims at
a Semantic Desktop environment, which supports P2P (or distributed) data
management based on Desktop services [
        <xref ref-type="bibr" rid="ref26">26</xref>
        ]. Like in our framework, Gnowsis uses ontologies
for expressing semantic associations and RDF for data modeling. However, the
emphasis of Gnowsis is more on the flexible integration of a large number of applications
than on semantic data organization and manipulation. SEMEX [
        <xref ref-type="bibr" rid="ref9">9</xref>
        ] is another personal
data integration framework that uses data annotation (i.e., schemas), similarly to our
ontology-based framework. A single domain model is provided as the unified interface
for data access.
      </p>
      <p>
        MyLifeBits [
        <xref ref-type="bibr" rid="ref13">13</xref>
        ], Haystack [
        <xref ref-type="bibr" rid="ref24">24</xref>
        ], and Placeless Documents [
        <xref ref-type="bibr" rid="ref10">10</xref>
        ] are three PIM
systems that support annotations and collections. Here, the concept of collection is
essentially the same as the conceptualization (using ontologies) of resources in our
framework. MyLifeBits supports easy annotation and multiple visualizations (e.g., detail,
thumbnail, timeline, and cluster-time views on the data). For this purpose, the resources
are enriched by a number of properties, including the standard ones (e.g., size and
creation date) and more specific ones (e.g., time interval) [
        <xref ref-type="bibr" rid="ref13">13</xref>
        ]. Haystack aims to create,
manipulate, and visualize arbitrary RDF data, in a comprehensive platform. For
visualization, it uses an ontological/agent approach, where user interfaces and views are
constructed by agents using predefined ontologies [
        <xref ref-type="bibr" rid="ref24">24</xref>
        ]. Placeless Documents
introduces “active properties”, where documents can have executable codes that provide
document-based services [
        <xref ref-type="bibr" rid="ref10">10</xref>
        ].
      </p>
      <p>
        Other approaches to PIM include Chandler,3 Lifestreams [
        <xref ref-type="bibr" rid="ref12">12</xref>
        ], Stuff I’ve Seen (SIS)
[
        <xref ref-type="bibr" rid="ref11">11</xref>
        ], and Xanadu [
        <xref ref-type="bibr" rid="ref22">22</xref>
        ], which focus on different aspects.
      </p>
      <sec id="sec-1-1">
        <title>2 http://www.dfki.de/web/</title>
        <p>3 http://www.osafoundation.org/
PIM 1</p>
        <p>Application Layer</p>
        <p>Application</p>
        <p>Ontology 1
Domain Layer</p>
        <p>Domain</p>
        <p>Ontology 1
Resource Layer</p>
        <p>Resource
repository
(RDF)</p>
        <p>Association</p>
        <p>Application
Ontology 2
Domain
Ontology 2
Resource
-file index</p>
        <p>PIM 2</p>
        <p>Application Layer</p>
        <p>AOpnptloicloagtiyoni . . .</p>
        <p>Application
Ontology n
Domain</p>
        <p>Ontology m
File Metadata</p>
        <p>PIM 3</p>
        <p>Application Layer</p>
        <p>AOpnptloicloagtiyonj . . .</p>
        <p>PI Space</p>
        <p>Contacts
Simple-content Bibtex</p>
        <p>...</p>
        <p>Textual</p>
        <p>Papers</p>
        <p>Reports
Complex-content Emails</p>
        <p>Slides
...</p>
        <p>Nontextual: (Video, Audio, Pictures, ...)</p>
        <p>Relational database, XML</p>
        <p>The rest of the paper is structured as follows. In Section 2, we describe the layered
framework and its main components. The semantic organization of the PI (including the
concepts of annotation, association, and representation) is discussed in Section 3.
Section 4 and Section 5 focus on two main ways of data manipulation, namely, navigation
and query processing. Finally, we conclude in Section 6.
2</p>
      </sec>
    </sec>
    <sec id="sec-2">
      <title>Framework</title>
      <p>
        Our framework follows the principle of superimposed information, i.e., data or metadata
“placed over” existing information sources [
        <xref ref-type="bibr" rid="ref19">19</xref>
        ]. This concept seems particularly useful
for the organization, access, interconnection, and reuse of the information elements. We
propose for PIM a layered ontology-based framework, as shown in Figure 2, with the
following data components:
Personal information space. The personal information space may contain structured
data (e.g., relational), semi-structured data (e.g., XML), or unstructured data.
Unstructured data can be textual or non-textual (as in video, audio, or picture files).
Furthermore, textual files can be classified as simple-content or complex-content. More
specifically, simple-content files have no references to other files. Typical examples include
people contacts and Bibtex entries. In contrast, complex-content files have a flexible
scheme of presentation, and may contain references to other files, e.g., by means of
citations or hypertext links [
        <xref ref-type="bibr" rid="ref6">6</xref>
        ]. For example, a paper in the PI space may cite another
paper (existing in the PI space or an external space), which, in turn, could cite other
papers.
File description. We annotate each file using a file description (or metadata) consisting
of a set of properties of the file. Each item in the file description is a property-value
pair. The file description is the first-level (direct) annotation for the individual files, and
has the same scheme (structure) for the same type of files. For example, the following
fragment contains a typical description of a JPEG file.
      </p>
      <p>Dimensions: 3072 £ 2048 pixels
Device make: Canon
Color space: RGB
Focal Length: 75
......</p>
      <p>Domain ontologies. A number of ontologies are published on the Web. Examples of
such ontology libraries include DAML Ontology Library,4 the Semantic Web
Ontologies,5 and the Prote´ge´ OWL ontologies.6 The ontologies in these libraries are typically
designed and organized for different domains such as Conference, Person, Photo, and
Email. In our framework, the domain ontology layer is designed to be loosely-coupled
with the other layers, to enable the insertion and removal of ontologies as “plug-ins”.
Resource-file index and RDF repository. One of the roles of domain ontologies is to
provide the basis for data classification. In order to establish the connections between
the files and the concepts in the domain ontologies, we treat each file as a resource,
which is then classified as an instance of one or more concepts. The resource-file index
is a local database storing these connections between resources and files. Furthermore,
the various types of associations among resources (as instances of association of
concepts in the domain ontologies) are stored in an RDF repository. The resource-file index
and the RDF repository are both in the resource layer, providing resource instances for
the domain ontologies in the domain layer above.</p>
      <p>Application ontology. Above the domain layer is the application layer, which contains
the ontologies for different applications. The domain ontologies, as an intermediate
layer between the applications and the data, are meant to enhance the reusability and
flexibility of the framework. More specifically, the application ontologies are defined
as views of the domain ontologies, which can be reused for the construction of
different application ontologies. In our framework, each personal information application
(PIA), is associated with an application ontology, has access to relevant data, and is
functionally independent of other applications. It may be infeasible to have a single
ontology to cover various applications, e.g., for trip planning and paper writing. Instead, as
many PIAs as needed can be designed in one or more PIM systems, where the PIAs can
interoperate (e.g., through P2P query processing) for the purpose of integrating relevant
information. This issue is elaborated on in Section 5.</p>
      <p>Besides the data components described above, a PIM system also needs some
functional components to perform all kinds of data and metadata processing, to make the
framework work as a whole. Such components include an indexer (for establishing and
managing the indexes of the files), a wrapper (for identifying and extracting resources
4 http://www.daml.org/ontologies/
5 http://www.schemaweb.info
6 http://protege.stanford.edu/plugins/owl/owl-library/
from the files), and an ontology designer (for importing and editing an ontology).
Because of space limitations, we do not elaborate further on these components.
3</p>
    </sec>
    <sec id="sec-3">
      <title>Semantic Data Organization</title>
      <p>The layered architecture of our PIM framework described previously enables the
reusability and the organization of semantically rich data for PIM. In this section, we discuss in
detail the mechanisms that our framework uses to support the semantic organization of
the PI space, including those for semantic annotation, association, and representation.
3.1</p>
      <sec id="sec-3-1">
        <title>Annotation</title>
        <p>Given that the data in the PI space is the base information, all the other data components
in our framework are actually superimposed information over this base. The most
fundamental function of the superimposed information is to provide semantic annotations
of the base information to enable powerful and accurate data access. We discuss the
following two aspects:
File description. It is especially important to provide the searcher with a detailed
description of the nontextual files. When performing a keyword-based searching, the
searcher matches the submitted keywords (e.g., “Canon”) or key-value pairs (e.g.,
“Maker:Canon”) with the property-value pairs of the file description, to find the right
files requested by the user. Even for textual files, taking into account such metadata will
improve the effectiveness of full-text searching.</p>
        <p>Domain ontologies. Given that a file is identified as a resource, we are able to annotate
the file using a domain ontology, by associating the resource with a concept of an
ontology. The domain ontology provides not only a context for understanding the data, but
also semantic clues for the precise data retrieval. For example, the user can query the PI
using a query language for RDF instead of using keywords. We note that a file can be
an instance of more than one concept, according to different classification criteria.
3.2</p>
      </sec>
      <sec id="sec-3-2">
        <title>Association</title>
        <p>In our framework, semantic associations are used to relate all the data (base
information) and metadata (superimposed information). There are two classes of
associations: the resource-file associations that are actually the resource-file indexes and
the resource-resource associations that are instances of the domain ontologies and are
stored in the RDF repository.</p>
        <p>Resource-file associations. In addition to the ontological resources that are used to
identify (through data classification) the files, a (textual) file may contain and refer
to a number of resources. Therefore, the resource-file associations can be one of the
following: identification, containment, and reference.</p>
        <p>Example 2. Suppose that the user has saved an email message, which is an
announcement of a seminar, as shown in Figure 3. First, the email message can be classified as
an instance of the concept Email, provided that the concept exists in some domain
ontology. Then, the system can generate for the concept SeminarAnnouncement and its
properties a new instance (i.e., resource), which is associated with the saved email by the
relationship containment. Finally, a reference association can be established between
the resource http://www.tliap.nus.edu.sg/ (e.g., of the concept WebsiteAddress) and
the email message.</p>
        <p>
          The process of setting up the resource-file associations is the one of recognizing
resources from the file description and/or the file content and then mapping them to the
ontological concepts. The user may determine the degree to which the resources should
be extracted from a file and its description. For instance, in the previous example, the
user can further create resources for the title and abstract of the seminar, and for the
biography of the presenter. It is expected that this process (as well as the process of
discovering resource-resource associations, as discussed later) can be maximally
automated, to reduce the user’s burden. For this purpose, we may utilize the following
methods:
– Keyword extraction. From the text of a file, keywords can be extracted based on a
thesaurus or be highlighted manually by the user. Each keyword can be considered
a resource contained by the file. The matching of the resources with the concepts
in the domain ontologies can be guided by a thesaurus such as WordNet.7
– Hyperlink analysis. For the textual files that include hyperlinks to classified
resources (e.g., a citation of a paper or a link to a webpage), we create for each
hyperlink a reference-type resource-file association, as well as a resource-resource
association between the referring resource and the referred one.
– Natural language processing. We can utilize known techniques (e.g., [
          <xref ref-type="bibr" rid="ref1">1</xref>
          ]) to parse
each sentence of a text or its summary obtained by means of text summarization
[
          <xref ref-type="bibr" rid="ref20">20</xref>
          ]. For each resulting triple hsubject; predicate; objecti, we try to match it with
the patterns hs; p; oi in the domain ontologies, where p is a property of the concept
s and has a value typed of o. If such pattern exists, a resource-resource association
of type property and of the form hsubject; predicate; objecti is generated.
– History. As the framework proceeds with such classification and cognition, more
and more knowledge about this process can be accumulated and reused by a new
process.
        </p>
        <sec id="sec-3-2-1">
          <title>7 http://wordnet.princeton.edu</title>
          <p>Resource-resource associations. We borrow from the Object Oriented Design (OOD)
techniques the following four types of relationships between objects: instantiation (i.e.,
membership), property, aggregation (i.e., whole/part), and generalization (i.e.,
inheritance). These four relationships, which are used in object models, are adopted to
describe the associations among concepts as well as resources in our framework. Note that
“property” refers to a pattern as identified, for example, using natural language
processing techniques, which corresponds to a user-defined property. For example, writes can
be a property of the concept Author, connecting Author to the concept Book. Table 1
summarizes the resource-resource associations in our framework.</p>
          <p>By using the previously described techniques, we can discover the resources and
their associations implied in the PI, and classify them into the domain ontologies, thus
populating the ontologies. In the example of Figure 3, it is possible to extract a pattern
hSingapore; implements; ITSi, which could then be classified as an instance of an
ontological pattern such as hOrganization; implements; Systemi, where Organization
and System are two concepts, and implements is a property. Note that the user is
allowed to choose the granularity of this knowledge (resource and associations) discovery
process, ranging from only taking the whole file as a single resource to analyzing the
detailed contents of the file.</p>
          <p>
            In addition, ontology mappings may be established between correspondences that
connect concepts in different domain and application ontologies. Currently, we consider
equivalence as the only semantics for the mapping between two concepts, although
richer semantics of the mappings could be considered [
            <xref ref-type="bibr" rid="ref17">17</xref>
            ].
In our framework, all information, including file descriptions, the resources in the
repository, and the resource-file indexes, are represented in the Resource Description
Framework (RDF),8 a W3C proposed standard. For the schema of these data (i.e., the
application and domain ontologies), we use the vocabulary language for RDF, RDF Schema
(RDFS).9 The RDF model is a semantic network, where the nodes denote the resources
and the edges are properties that represent the relations between resources. The
network can also be seen as a set of statements (triples) in the form of (subject, predicate,
object). RDFS is used to define the vocabulary (in terms of classes and properties) of
          </p>
        </sec>
        <sec id="sec-3-2-2">
          <title>8 http://www.w3.org/RDF/</title>
          <p>9 http://www.w3.org/TR/rdf-schema
the RDF data, such as rdfs:Class, rdf:Property, and rdf:type. Table 2 summarizes the
RDFS vocabularies that are used to represent different types of associations.</p>
          <p>
            The use of RDF as the data model and RDFS as the ontology language in our
framework is motivated by the nature of the RDF as a Web resources description mechanism
and the fact that the PI is represented as a set of interrelated resources. In contrast, XML
is not chosen because it cannot represent semantic associations [
            <xref ref-type="bibr" rid="ref8">8</xref>
            ]. Certainly, OWL
(Web Ontology Language), as built on top of RDFS, is more expressive for ontology
representation. However, the use of a slightly extended version of RDFS is adequate for
representing resource-file and resource-resource associations.
          </p>
          <p>The extension to RDFS is as follows: we define in a namespace (abbreviated using
the prefix rdfx) a new RDF property, contains, which is used to represent the
aggregation relationship. For the representation of the instantiation and generalization
relationships, we use rdf:type and rdfs:subClassOf, respectively. The property relationship
is represented naturally by an RDF property defined in the user-defined namespace.
Figure 4 gives a concrete example of an RDF representation.
4</p>
        </sec>
      </sec>
    </sec>
    <sec id="sec-4">
      <title>Semantic Navigation</title>
      <p>It is critical for a Semantic Desktop to provide the user with the capability to access the
stored data in a variety of ways. The user may want to browse the information by means
of the flexible and intelligent navigation in the information space, including the base and
superimposed information. The user may also desire that certain query facilities (e.g.,
keyword-based searching or certain query languages) be provided by the framework. In
this section, we discuss the navigation in the data space of a Semantic Desktop. Query
processing is discussed in the next section.</p>
      <p>The semantic data organization in our framework enables the navigation in the PI
space, making use of useful hints (e.g., the context of a concept being browsed) so as
to facilitate the user’s understanding of data. More specifically, by taking into account
the layered architecture, the semantic navigation in our framework can be performed in
three directions: (1) In vertical navigation, the user follows a path across layers. Two
cases are possible for this way of navigation: top-down from the application ontologies
to the stored files and bottom-up from the stored files to the application ontologies. (2)
Ontology for attending a conference</p>
      <p>Person writtenBy Paper extendedVersion Journal
attends publishedAt
Conference presentedAt Talk</p>
      <p>presentedBy
r
e
y
a
L Email
n
o
it
a
c
li
p
p
A
receivedBy
sentBy</p>
      <p>Person</p>
      <p>In horizontal navigation, the user follows links of concepts (or resources) within one
layer. Typically, there are three cases of horizontal navigation, corresponding to each
layer: application-to-application navigation, domain-to-domain navigation, and
file-tofile navigation. (3) In temporal navigation, the user can navigate by following
references in chronological order, each being a resource for the same real world object with
a time stamp associated with it. For example, the user may want to look at different
versions of a research paper.</p>
      <p>All the base and superimposed information in the framework forms a directed graph,
where the vertices are the resources in the ontologies and the files stored in the PI space,
and the edges are the associations between the resources and files. We say that the three
directions of navigation together provide the capacity of a 3-dimension (3D) navigation
mechanism, which can facilitate the construction of a browser. For instance, suppose
the user is browsing a specific application ontology in a visualized browser. When the
user clicks on the node of a concept in the ontology, the browser can then choose to
Application Ontology
Email
display the instances of the concept thus selected (by vertical navigation), the context
of the concept in the domain (also by vertical navigation), and the associated concepts
in other application ontologies (by horizontal navigation). Compared to the traditional
navigation approach that is based on hierarchical directories, 3D navigation is based on
semantic associations, similarly to those that humans establish between concepts.
Example 3. Consider the scenario shown in Figure 4. The spirit of 3D navigation is
demonstrated in the browser of Figure 5. The current resource (file) that the user is
browsing is an email message (i.e., wisephotomsg), which has some photos attached,
which were taken at WISE ’03. The concepts that this resource belong to are highlighted
(in white) so as to show the contexts to which they belong. All associated resources are
categorized and shown on the right tabbed pane, which provides a guidance for the
user in navigating the PI space. The bottom-right pane shows the timeline of different
versions of the current resource (if they exist) or all the resources belonging to the same
concept as the current resource.
5</p>
    </sec>
    <sec id="sec-5">
      <title>Semantic Query Processing</title>
      <p>Unlike navigation, which is an interactive process, query processing is performed
without further intervention from the user. To retrieve relevant data from the PI space, the
user’s request may be posed as a sequence of keywords or as a query formulated in a
certain query language.</p>
      <p>
        The keyword-based search matches the input keywords and the vector of words in
the candidate documents, calculates the similarity for each of the matches, and returns
to the user the results after ranking them [
        <xref ref-type="bibr" rid="ref25">25</xref>
        ]. The results of a search are usually
evaluated using the statistical criteria such as precision, recall, or a combination of them.
The shortcoming of keyword-based search is that the semantic associations between
relevant data are not considered. In contrast, query languages can provide a semantically
richer access interface, thus facilitating the data retrieval and improving the accuracy of
the answers. However, a query is usually performed based on an exact match between
the query and the data, so that the recall of the answers is influenced, in the sense that
some relevant but not matched data is not retrieved.
      </p>
      <p>
        Since the two approaches complement each other, it is desirable to provide both of
them. In this section, however, we mainly focus on query processing in our framework.
We choose to express the queries in RDQL [
        <xref ref-type="bibr" rid="ref16">16</xref>
        ]; they can query both the resources
and their associations. We discuss how to process a query submitted by the user in two
cases: within a PIA and across different PIAs.
5.1
      </p>
      <sec id="sec-5-1">
        <title>Query processing in a PIA</title>
        <p>
          In our framework, the user query is formulated in RDQL (RDF Data Query Language),
which uses an SQL-like syntax [
          <xref ref-type="bibr" rid="ref16">16</xref>
          ]. To reduce the user’s burden, a graphic means can
be used to facilitate the user’s query formulation. For simplicity, we use a subset of
RDQL that we call conjunctive RDQL (c-RDQL), which can be expressed as a
conjunctive formula: ans(X) :- p1(X1); :::; pn(Xn), where Xi = (xi; x0i) and pi is an
RDF property of xi having the value x0 .
        </p>
        <p>i</p>
        <p>
          In our framework, an application ontology is constructed over one or more domain
ontologies, and the files in the PI space are formalized as instances of the concepts in
the domain ontologies. If we consider the application ontology as the global ontology
(since the user query is posed on it), the whole system can be seen as a GaV data
integration system [
          <xref ref-type="bibr" rid="ref18">18</xref>
          ]. Therefore query processing in a single PIA is performed as in a
GaV system. In particular, when the user poses a query (in RDQL) over the application
ontology, the RDQL query is then rewritten into a new RDQL query in terms of the
domain ontologies, based on the mappings between the global ontology and domain
ontologies. By executing the rewritten query on the corresponding domain ontologies,
resources (files) that match the query are then returned as answers to the query.
        </p>
        <p>
          There are a number of algorithms for query rewriting in relational or XML data
integration systems [
          <xref ref-type="bibr" rid="ref14">14</xref>
          ]. In a GaV based integration system, query processing is
performed using a “unfolding” strategy [
          <xref ref-type="bibr" rid="ref18">18</xref>
          ]. More specifically, for rewriting a query (e.g.,
a conjunctive query) that is posed on the global schema or ontology, we simply
substitute the predicates in the body of the query with the corresponding view definitions. In
our framework, where the mappings between the application ontology and the domain
ontologies are expressed as RDF class or property correspondences, the algorithm for
query rewriting is similar to this strategy.
        </p>
        <p>By assuming that there are no integrity constraints over the application ontologies
and the user queries are formulated in c-RDQL, we give the formal description of our
query rewriting algorithm in a single PIA, which we call ADREWRITING (for
rewriting from Application ontologies to Domain ontologies), as follows. We note that we do
not consider the namespaces of ontologies for simplicity of the description.
Algorithm ADREWRITING
Input:1. q1 over the application ontology G: ans(X) :- p1(X1); :::; pm(Xm);
2. M: the mapping table between G and domain ontologies S1; :::; Sn.</p>
        <p>Output: q2: A c-RDQL query over S1; :::; Sn.
1. headq2 = ans(X); bodyq2 = null;
2. For i = 1 to m do
3. (c1; c2) = name of the classes referred to by (x1; x2), for Xj = (x1; x2);
4. Search M to find (d1; d2) such that f(c1; d1); (c2; d2)g are two class
correspondences in M;
5. Traverse S1; :::; and Sn by following all kinds of associations, to find the vertices,
v1; :::; vk, connecting from d1 to d2;
6. If k = 0 then add p(x1; x2) (or p(x2; x1)) to bodyq2 , if there exists p
connecting d1 to d2 (or d2 to d1);
7. Else for j = 1 to k ¡ 1 do
8. Add p(x^j ; x^j+1) (or p(x^j + 1; x^j )) to bodyq2 , if p is not a mapping and
connects vj to vj+1 (or vj+1 to vj );
9. Add p(x1; x^1) (or p(x^1; x1)) to bodyq2 , if p is not a mapping and connects d1
to v1 (or v1 to d1);
10. Add p(x^k; x2) (or p(x2; x^k)) to bodyq2 , if p is not a mapping and connects vk
to d2 (or d2 to vk);
11. q2 = headq2 :- bodyq2 ;
Example 4. Suppose the user wants to list all conference papers with their authors and
journal version, using the query q1 : ans(x; y; z) :- writtenBy(x; y), extendedV ersion(x; z),
which is posed on the application ontology of publication management. For the
variables (x; y; z), we get the classes that they refer to as (Paper, Person, Journal), as
indicated by Line 3. By looking into M, we find the corresponding class sequence
as (Publication:InProceedings, Publication:Person, Publication:Article), where the
names before the colons are domain ontology names. From Lines 5 to 10, we compute
the predicates in the body of q2 as follows.</p>
        <p>q2: ans(x; y; z) :- editor(x; y), extends(z; x)</p>
        <p>By executing q2 over the RDF repository as shown in Figure 4, we get the answer
f(#wise03papercamera, #xiao, #jods05), (#wise03papercamera, #cruz, #jods05)g.
5.2</p>
      </sec>
      <sec id="sec-5-2">
        <title>A2A query processing</title>
        <p>
          Application to application (A2A) query processing occurs when an application is
attempting to retrieve relevant data from another semantically related application, to
answer a query. If the PIAs are considered as connected peers (i.e., service providers for
certain data access), the A2A query processing is similar to that in peer-to-peer (P2P)
systems [
          <xref ref-type="bibr" rid="ref15 ref7">7, 15</xref>
          ]. Whether the PIAs exist in a single desktop or are physically distributed
makes no differences to the A2A query processing.
        </p>
        <p>A2A query processing consists of two steps of query rewriting. First, we rewrite the
original query q, which is posed on the application ontology G1, to a query q0 on the
other application ontology G2, according to the mappings between G1 and G2. Then, q0
is rewritten to a query q00 on the domain ontologies, to which G2 is mapped. Answers
are obtained by executing q00 on the RDF repository. The second query rewriting is
exactly the one described by the algorithm ADREWRITING, whereas the first rewriting
is slightly different from ADREWRITING. In particular, unlike the total mapping from
an application ontology to the domain ontologies, some of the concepts in G1 may not
be mapped to those in G2. Therefore, the answers returned by q00 may contain null values
or Skolem functions for the unmapped concepts or properties.</p>
        <p>
          The A2A mappings can be derived by composing the mappings between G1 and the
domain ontologies, inter-domain mappings, and those between G2 and the domain
ontologies. To evaluate both query rewriting processes, we need to check the equivalence
(or containment) between a query and its rewriting. A correct query rewriting is the one
that is equivalent to (or maximally contained in) the query. These two issues (reasoning
on mappings [
          <xref ref-type="bibr" rid="ref23 ref3">3, 23</xref>
          ] and reasoning on queries [
          <xref ref-type="bibr" rid="ref21 ref5">5, 21</xref>
          ]) have been extensively studied
and are beyond the scope of this paper.
6
        </p>
      </sec>
    </sec>
    <sec id="sec-6">
      <title>Conclusions and Future Work</title>
      <p>In this paper, we present our design of a PIM system. We propose a layered
ontologybased framework, which aims to provide a semantics-rich environment for personal
information organization and manipulation. The multiple ontologies existing in
different layers of the architecture explicitly support the data semantics. Furthermore, the
decoupling of the domain layer and the application layer enhances the flexibility and
reusability of the framework. Specifically, we discuss in detail the semantic-enriched
data organization, including the use of file descriptions and domain ontologies as
annotations, and the construction of resource-file and resource-resource associations. We
also introduce the idea of 3D navigation, which is used in a desktop browser. We discuss
query processing in our framework in two cases: within a single personal information
application, PIA, and between two PIAs, using application to application (A2A)
communication. A formal query rewriting algorithm is presented for the single PIA case.</p>
      <p>In the future, we will continue the study and implementation of our framework. It
is clear that a lot of the success of PIM systems lies on the successful automation of
the different mechanisms that are needed. In particular, we will look further into the
automation of the conceptualization of full-text files and that of matching resources
to ontological concepts. Also, we will elaborate on the idea of 3D navigation both by
studying a model for temporal navigation and by carrying out user studies. The study
of A2A communication, including data exchange, collaboration, and query processing
will also be continued. While RDQL queries are expressive, they may not be suitable
for most users. We are therefore exploring visual queries that can express a class of
RDQL queries “appropriate” for the semantic desktop.</p>
    </sec>
  </body>
  <back>
    <ref-list>
      <ref id="ref1">
        <mixed-citation>
          1.
          <string-name>
            <given-names>M.</given-names>
            <surname>Berland</surname>
          </string-name>
          and
          <string-name>
            <given-names>E.</given-names>
            <surname>Charniak</surname>
          </string-name>
          .
          <article-title>Finding Parts in Very Large Corpora</article-title>
          .
          <source>In ACL</source>
          ,
          <year>1999</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref2">
        <mixed-citation>
          2.
          <string-name>
            <given-names>T.</given-names>
            <surname>Berners-Lee</surname>
          </string-name>
          ,
          <string-name>
            <given-names>J.</given-names>
            <surname>Hendler</surname>
          </string-name>
          , and
          <string-name>
            <surname>O. Lassila.</surname>
          </string-name>
          <article-title>The Semantic Web</article-title>
          . Scientific American, May
          <year>2001</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref3">
        <mixed-citation>
          3.
          <string-name>
            <given-names>P. A.</given-names>
            <surname>Bernstein</surname>
          </string-name>
          .
          <article-title>Applying Model Management to Classical Meta Data Problems</article-title>
          . In CIDR,
          <year>2003</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref4">
        <mixed-citation>
          4.
          <string-name>
            <given-names>V.</given-names>
            <surname>Bush</surname>
          </string-name>
          . As We May Think.
          <source>The Atlantic Monthly</source>
          ,
          <volume>176</volume>
          (
          <issue>1</issue>
          ):
          <fpage>101</fpage>
          -
          <lpage>108</lpage>
          ,
          <year>1945</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref5">
        <mixed-citation>
          5.
          <string-name>
            <given-names>D.</given-names>
            <surname>Calvanese</surname>
          </string-name>
          ,
          <string-name>
            <given-names>G. D.</given-names>
            <surname>Giacomo</surname>
          </string-name>
          ,
          <string-name>
            <given-names>M.</given-names>
            <surname>Lenzerini</surname>
          </string-name>
          , and
          <string-name>
            <given-names>M. Y.</given-names>
            <surname>Vardi</surname>
          </string-name>
          .
          <article-title>View-based Query Containment</article-title>
          .
          <source>In PODS</source>
          , pages
          <fpage>56</fpage>
          -
          <lpage>67</lpage>
          ,
          <year>2003</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref6">
        <mixed-citation>
          6.
          <string-name>
            <given-names>J.</given-names>
            <surname>Conklin</surname>
          </string-name>
          .
          <article-title>Hypertext: An Introduction and Survey</article-title>
          . IEEE Computer,
          <volume>20</volume>
          (
          <issue>9</issue>
          ):
          <fpage>17</fpage>
          -
          <lpage>41</lpage>
          ,
          <year>1987</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref7">
        <mixed-citation>
          7.
          <string-name>
            <given-names>I. F.</given-names>
            <surname>Cruz</surname>
          </string-name>
          ,
          <string-name>
            <given-names>H.</given-names>
            <surname>Xiao</surname>
          </string-name>
          , and
          <string-name>
            <given-names>F.</given-names>
            <surname>Hsu</surname>
          </string-name>
          .
          <article-title>Peer-to-Peer Semantic Integration of XML and RDF Data Sources</article-title>
          . In AP2PC,
          <year>July 2004</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref8">
        <mixed-citation>
          8.
          <string-name>
            <given-names>S.</given-names>
            <surname>Decker</surname>
          </string-name>
          ,
          <string-name>
            <given-names>S.</given-names>
            <surname>Melnik</surname>
          </string-name>
          , F. van
          <string-name>
            <surname>Harmelen</surname>
            ,
            <given-names>D.</given-names>
          </string-name>
          <string-name>
            <surname>Fensel</surname>
            ,
            <given-names>M. C. A.</given-names>
          </string-name>
          <string-name>
            <surname>Klein</surname>
            ,
            <given-names>J.</given-names>
          </string-name>
          <string-name>
            <surname>Broekstra</surname>
            ,
            <given-names>M.</given-names>
          </string-name>
          <string-name>
            <surname>Erdmann</surname>
            ,
            <given-names>and I. Horrocks.</given-names>
          </string-name>
          <article-title>The Semantic Web: The Roles of XML and RDF</article-title>
          .
          <source>IEEE Internet Computing</source>
          ,
          <volume>4</volume>
          (
          <issue>5</issue>
          ):
          <fpage>63</fpage>
          -
          <lpage>74</lpage>
          ,
          <year>2000</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref9">
        <mixed-citation>
          9.
          <string-name>
            <given-names>X.</given-names>
            <surname>Dong</surname>
          </string-name>
          and
          <string-name>
            <given-names>A. Y.</given-names>
            <surname>Halevy</surname>
          </string-name>
          .
          <article-title>A Platform for Personal Information Management and Integration</article-title>
          .
          <source>In CIDR</source>
          , pages
          <fpage>119</fpage>
          -
          <lpage>130</lpage>
          ,
          <year>2005</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref10">
        <mixed-citation>
          10.
          <string-name>
            <given-names>P.</given-names>
            <surname>Dourish</surname>
          </string-name>
          ,
          <string-name>
            <given-names>W. K.</given-names>
            <surname>Edwards</surname>
          </string-name>
          , A. LaMarca, J. Lamping,
          <string-name>
            <given-names>K.</given-names>
            <surname>Petersen</surname>
          </string-name>
          ,
          <string-name>
            <given-names>M.</given-names>
            <surname>Salisbury</surname>
          </string-name>
          ,
          <string-name>
            <given-names>D. B.</given-names>
            <surname>Terry</surname>
          </string-name>
          , and
          <string-name>
            <given-names>J.</given-names>
            <surname>Thornton</surname>
          </string-name>
          .
          <article-title>Extending Document Management Systems with User-specific Active Properties</article-title>
          .
          <source>ACM Transaction of Information System</source>
          ,
          <volume>18</volume>
          (
          <issue>2</issue>
          ):
          <fpage>140</fpage>
          -
          <lpage>170</lpage>
          ,
          <year>2000</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref11">
        <mixed-citation>
          11.
          <string-name>
            <given-names>S. T.</given-names>
            <surname>Dumais</surname>
          </string-name>
          ,
          <string-name>
            <given-names>E.</given-names>
            <surname>Cutrell</surname>
          </string-name>
          ,
          <string-name>
            <given-names>J. J.</given-names>
            <surname>Cadiz</surname>
          </string-name>
          , G. Jancke,
          <string-name>
            <given-names>R.</given-names>
            <surname>Sarin</surname>
          </string-name>
          , and
          <string-name>
            <given-names>D. C.</given-names>
            <surname>Robbins</surname>
          </string-name>
          .
          <article-title>Stuff I've Seen: A System for Personal Information Retrieval and Re-use</article-title>
          .
          <source>In SIGIR</source>
          , pages
          <fpage>72</fpage>
          -
          <lpage>79</lpage>
          ,
          <year>2003</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref12">
        <mixed-citation>
          12.
          <string-name>
            <given-names>E.</given-names>
            <surname>Freeman</surname>
          </string-name>
          and
          <string-name>
            <given-names>D.</given-names>
            <surname>Gelernter</surname>
          </string-name>
          .
          <article-title>Lifestreams: A Storage Model for Personal Data</article-title>
          .
          <source>SIGMOD Record</source>
          ,
          <volume>25</volume>
          (
          <issue>1</issue>
          ):
          <fpage>80</fpage>
          -
          <lpage>86</lpage>
          ,
          <year>1996</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref13">
        <mixed-citation>
          13.
          <string-name>
            <surname>J. Gemmell</surname>
            , G. Bell,
            <given-names>R.</given-names>
          </string-name>
          <string-name>
            <surname>Lueder</surname>
            ,
            <given-names>S. M.</given-names>
          </string-name>
          <string-name>
            <surname>Drucker</surname>
            , and
            <given-names>C.</given-names>
          </string-name>
          <string-name>
            <surname>Wong</surname>
          </string-name>
          .
          <article-title>MyLifeBits: Fulfilling the Memex Vision</article-title>
          . In ACM Multimedia, pages
          <fpage>235</fpage>
          -
          <lpage>238</lpage>
          ,
          <year>2002</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref14">
        <mixed-citation>
          14.
          <string-name>
            <given-names>A. Y.</given-names>
            <surname>Halevy. Answering Queries Using Views: A Survey. VLDB</surname>
          </string-name>
          J.,
          <volume>10</volume>
          (
          <issue>4</issue>
          ):
          <fpage>270</fpage>
          -
          <lpage>294</lpage>
          ,
          <year>2001</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref15">
        <mixed-citation>
          15.
          <string-name>
            <surname>A. Y. Halevy</surname>
            ,
            <given-names>Z. G.</given-names>
          </string-name>
          <string-name>
            <surname>Ives</surname>
            ,
            <given-names>P.</given-names>
          </string-name>
          <string-name>
            <surname>Mork</surname>
            ,
            <given-names>and I. Tatarinov.</given-names>
          </string-name>
          <article-title>Piazza: Data Management Infrastructure for Semantic Web Applications</article-title>
          . In WWW, pages
          <fpage>556</fpage>
          -
          <lpage>567</lpage>
          ,
          <year>2003</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref16">
        <mixed-citation>
          16.
          <string-name>
            <given-names>HP</given-names>
            <surname>Labs. RDQL - RDF Data Query</surname>
          </string-name>
          <article-title>Language</article-title>
          . http://www.hpl.hp.com/semweb/rdql.htm,
          <year>2005</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref17">
        <mixed-citation>
          17.
          <string-name>
            <given-names>Y.</given-names>
            <surname>Kalfoglou</surname>
          </string-name>
          and
          <string-name>
            <given-names>M.</given-names>
            <surname>Schorlemmer</surname>
          </string-name>
          .
          <article-title>Ontology Mapping: the State of the Art</article-title>
          .
          <source>The Knowledge Engineering Review</source>
          ,
          <volume>18</volume>
          (
          <issue>1</issue>
          ):
          <fpage>1</fpage>
          -
          <lpage>31</lpage>
          ,
          <year>2003</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref18">
        <mixed-citation>
          18.
          <string-name>
            <given-names>M.</given-names>
            <surname>Lenzerini</surname>
          </string-name>
          .
          <article-title>Data Integration: A Theoretical Perspective</article-title>
          . In PODS, pages
          <fpage>233</fpage>
          -
          <lpage>246</lpage>
          , Madison, Wisconsin,
          <year>June 2002</year>
          . ACM.
        </mixed-citation>
      </ref>
      <ref id="ref19">
        <mixed-citation>
          19.
          <string-name>
            <given-names>D.</given-names>
            <surname>Maier</surname>
          </string-name>
          and
          <string-name>
            <given-names>L. M. L.</given-names>
            <surname>Delcambre</surname>
          </string-name>
          .
          <article-title>Superimposed Information for the Internet</article-title>
          . In WebDB, pages
          <fpage>1</fpage>
          -
          <lpage>9</lpage>
          ,
          <year>1999</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref20">
        <mixed-citation>
          20. I. Mani.
          <article-title>Recent Developments in Text Summarization</article-title>
          .
          <source>In CIKM</source>
          , pages
          <fpage>529</fpage>
          -
          <lpage>531</lpage>
          ,
          <year>2001</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref21">
        <mixed-citation>
          21.
          <string-name>
            <surname>T. D. Millstein</surname>
            ,
            <given-names>A. Y.</given-names>
          </string-name>
          <string-name>
            <surname>Halevy</surname>
            , and
            <given-names>M.</given-names>
          </string-name>
          <string-name>
            <surname>Friedman</surname>
          </string-name>
          .
          <article-title>Query Containment for Data Integration Systems</article-title>
          .
          <source>Journal of Computer and System Sciences</source>
          ,
          <volume>66</volume>
          (
          <issue>1</issue>
          ):
          <fpage>20</fpage>
          -
          <lpage>39</lpage>
          ,
          <year>2003</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref22">
        <mixed-citation>
          22.
          <string-name>
            <given-names>T. H.</given-names>
            <surname>Nelson</surname>
          </string-name>
          .
          <article-title>Xanalogical Structure, Needed Now More than Ever: Parallel Documents</article-title>
          , Deep Links to Content, Deep Versioning, and
          <string-name>
            <surname>Deep</surname>
          </string-name>
          Re-use.
          <source>ACM Computer Surveys</source>
          ,
          <volume>31</volume>
          (4es):
          <fpage>33</fpage>
          ,
          <year>1999</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref23">
        <mixed-citation>
          23.
          <string-name>
            <given-names>N. F.</given-names>
            <surname>Noy</surname>
          </string-name>
          .
          <article-title>Semantic Integration: A Survey Of Ontology-Based Approaches</article-title>
          .
          <source>SIGMOD Record</source>
          ,
          <volume>33</volume>
          (
          <issue>4</issue>
          ):
          <fpage>65</fpage>
          -
          <lpage>70</lpage>
          ,
          <year>2004</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref24">
        <mixed-citation>
          24.
          <string-name>
            <given-names>D.</given-names>
            <surname>Quan</surname>
          </string-name>
          ,
          <string-name>
            <given-names>D.</given-names>
            <surname>Huynh</surname>
          </string-name>
          , and
          <string-name>
            <given-names>D. R.</given-names>
            <surname>Karger</surname>
          </string-name>
          .
          <article-title>Haystack: A Platform for Authoring End User Semantic Web Applications</article-title>
          . In ISWC, pages
          <fpage>738</fpage>
          -
          <lpage>753</lpage>
          ,
          <year>2003</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref25">
        <mixed-citation>
          25.
          <string-name>
            <given-names>G.</given-names>
            <surname>Salton. Automatic Text</surname>
          </string-name>
          <article-title>Processing: The Transformation, Analysis, and Retrieval of Information by Computer</article-title>
          . Addison-Wesley,
          <year>1989</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref26">
        <mixed-citation>
          26.
          <string-name>
            <given-names>L.</given-names>
            <surname>Sauermann</surname>
          </string-name>
          .
          <article-title>The Gnowsis Semantic Desktop for Information Integration</article-title>
          .
          <source>In The 3rd Conference on Professional Knowledge Management</source>
          , pages
          <fpage>39</fpage>
          -
          <lpage>42</lpage>
          ,
          <year>2005</year>
          .
        </mixed-citation>
      </ref>
    </ref-list>
  </back>
</article>