<!DOCTYPE article PUBLIC "-//NLM//DTD JATS (Z39.96) Journal Archiving and Interchange DTD v1.0 20120330//EN" "JATS-archivearticle1.dtd">
<article xmlns:xlink="http://www.w3.org/1999/xlink">
  <front>
    <journal-meta />
    <article-meta>
      <title-group>
        <article-title>How to build a Snippet Manager</article-title>
      </title-group>
      <contrib-group>
        <contrib contrib-type="author">
          <string-name>Steve Cayzer</string-name>
          <email>steve.cayzer@hp.com</email>
          <xref ref-type="aff" rid="aff0">0</xref>
        </contrib>
        <contrib contrib-type="author">
          <string-name>Paolo Castagna</string-name>
          <email>paolo.castagna@hp.com</email>
          <xref ref-type="aff" rid="aff0">0</xref>
        </contrib>
        <aff id="aff0">
          <label>0</label>
          <institution>Hewlett-Packard Laboratories</institution>
          ,
          <addr-line>Filton Road, Stoke Gifford, Bristol BS34 8QZ</addr-line>
        </aff>
      </contrib-group>
      <abstract>
        <p>In our research group, there is a need to capture, organize and share resources associated with a domain of exploration. We are building a tool for this task, based on previous experience in the knowledge management domain. In this position paper, we present our thoughts on what works (and what doesn't work), together with details of our initial implementation.</p>
      </abstract>
    </article-meta>
  </front>
  <body>
    <sec id="sec-1">
      <title>1 Introduction</title>
      <p>
        The snippet manager idea is not a new one [
        <xref ref-type="bibr" rid="ref1">1</xref>
        ]. It refers to the idea of a small peer
group capturing ‘snippets’ of information in a lightweight manner, categorizing and
sharing them. There have been a number of approaches to this problem over the
years; in this paper we present our opinions on what works, and what doesn’t work.
We also define the scope – even the ideal snippet manager would not be a panacea
for knowledge management generally. Rather, it is a useful tool (or at least a useful
concept) for a specific task. We have started to implement a tool using the principles
outlined here, and we present some design details. We also introduce the success
factors that we intend to adopt, and that we hope will be more generally useful.
      </p>
    </sec>
    <sec id="sec-2">
      <title>2 The snippet manager problem domain</title>
      <p>
        Knowledge management is defined widely, for example achieving a global sharing of
knowledge within a company [
        <xref ref-type="bibr" rid="ref2">2</xref>
        ]. However, our domain of interest is rather more
tightly defined:
“In [my group] we frequently circulate items of interest (such as news articles,
software tools, links to Web sites, and competitor information). We call them snippets,
or information nuggets, we would like to store, annotate, and share. Email is not the
ideal medium for these tasks; its transient nature means the snippets are effectively
lost over time. Yet the risk from using a more formal process, like a centralized
database, is that it is both cumbersome to use (a barrier to entry) and overly rigid in
its data model (not amenable to storing different types of information). Our need
illustrates what I call decentralized, informal knowledge management…” [
        <xref ref-type="bibr" rid="ref3">3</xref>
        ]
The point here is that the domain allows us to make some simplifying assumptions.
Firstly, the group is small, often co-located. Members can ‘pop round’ to discuss
ideas or gather for an informal discussion. In the particular group we are designing
for, there are regular, weekly meetings. So we assume that information flow is
unhindered, and that conflicts or inconsistencies can be quickly ironed out. We need
not rely on snippet manager as the sole conduit for communication. A small group
also makes the job of converging on a domain model much easier. We do not assume
that we will get the model(s) right first time, but the first pass should be good enough
to get general buy-in, and improvements can occur by means of incremental
evolution.
      </p>
      <p>Secondly, the users are technically literate. This means that they are likely to
quickly get to grips with a new tool, and may be motivated to make some small
changes to behaviour for sufficient added value (a good example of this would be the
use of ‘graffiti’ writing on PDAs). However, getting the balance right is not always
easy; we note that in our (internal) semantic wiki, people hardly ever use the
supplied wiki syntax to add RDF metadata. This may be due to the lack of instant, direct
reward for adding metadata, a point which we are trying to address in our work.</p>
      <p>
        Finally, the domain of interest is tightly focused, even more so than that envisaged
by Cayzer [
        <xref ref-type="bibr" rid="ref3">3</xref>
        ]. Our laboratory is interested in a myriad of semantic web related
topics, but the user group for the current incarnation of the snippet manager is
specifically and actively looking at a particular topic, that of enterprise information
management.
      </p>
      <sec id="sec-2-1">
        <title>2.1 Use cases</title>
        <p>What is it, then, that the snippet manager is expected to achieve? We are in the
early stages of this project, but we have engaged the user community and gathered
some initial use cases in order to inform our guiding principles, a few of which are
described here. We reiterate that these principles are relevant for our domain of
interest – small group, tightly focused, technically literate researchers. We don’t expect
the principles to necessarily generalize to the whole of knowledge management, or to
the web at large. However, we believe that our problem domain is sufficiently
common for these principles to be of value for the semantic web community.</p>
      </sec>
      <sec id="sec-2-2">
        <title>Easy Capture</title>
        <p>“I need a way to collect evidence (web pages, PDFs, emails, forum posts), and to
categorise parts of these so they can be linked together for post-hoc search.”
“I need a way of annotating resources with evidence - 'why have you written this'”
It should be ludicrously simple to collect snippets, using familiar methods such as
bookmarklets or email. Snippet Manager should also pull in snippets from other
sources (eg intranet databases) and handle provenance.</p>
      </sec>
      <sec id="sec-2-3">
        <title>Editable Ontologies</title>
        <p>I need to create a category or classification for a new area of interest. Now, I need
to add new companies, products, documents or links and tag them with this
classification. I want to create relationships between these instances - e.g. competitor links.
Our users will certainly want to change the ontologies on the fly. Although this
sounds like a tall order, in our case both the structure of the ontologies (effectively
taxonomies) and the nature of the changes (adding/removing/renaming a node) can
simplify the implementation enormously. Of course the UI for such changes may not
be trivial.</p>
      </sec>
      <sec id="sec-2-4">
        <title>Export</title>
        <p>I want a regular alert showing the results of a web search for a topic. [OR I want to
produce a report that shows all relevant products or technologies for a given topic]
From a technical point of view, the ability to export (meta)data in a standard,
machine readable way is a future-proofing mechanism, intended to prevent the portal
becoming yet another information silo. From a user point of view, export in a
human-readable form is equally important.</p>
      </sec>
      <sec id="sec-2-5">
        <title>Web Application</title>
        <p>
          Our experience with the early snippet manager prototype [
          <xref ref-type="bibr" rid="ref1">1</xref>
          ] taught us that there is
considerable reluctance to download software, let alone to standardize on it across a
group. In addition, the snippet manager should be integrated into a users’ normal
work pattern. For our group, this suggests a web application such as a portal.
        </p>
      </sec>
      <sec id="sec-2-6">
        <title>Immediate Feedback</title>
        <p>There should be an instant reward for the user who adds metadata. The community
aspect should be (from the user’s point of view) a beneficial side effect.</p>
      </sec>
    </sec>
    <sec id="sec-3">
      <title>4 Implementation Details</title>
      <p>
        We have used the semantic portal [
        <xref ref-type="bibr" rid="ref4">4</xref>
        ] idea and codebase to provide a browsing
interface over the group’s snippets. Essentially, this portal uses the metadata to drive a
facet browser, so that users can find what they are looking for using a variety of
search paths. We are building simple capture modalities such as bookmarklets, mail
processors and web forms; and importers for other systems such as blogs, technical
reports, people databases and the group’s official wiki. For export, we plan RSS
feeds, email alerts, customizable reports and a SPARQL[
        <xref ref-type="bibr" rid="ref5">5</xref>
        ] interface for
programmatic access.
      </p>
      <sec id="sec-3-1">
        <title>4.1 Success Factors</title>
        <p>We have previously built several semantic web applications whose primary function
was to demonstrate a particular aspect of the technology. Our focus here is to build a
tool. Therefore the simplest way of assessing its success is to measure its usage:
1.
2.
3.
4.
5.
6.</p>
        <p>At what rate is new content added to the snippet manager?
What proportion of the user group use the snippet manager as a day to day
tool
How well does the snippet manager integrate with other tools in use?
How often is the snippet manager consulted for information or report
generation?
What is the satisfaction level of the users?</p>
        <p>How quickly can new user requirements be integrated into the tool
These measures are largely qualitative in nature. Yet they get to the heart of what of
means to build a semantic web tool for personal, and group, productivity. We intend
to assess our work using these criteria.
5</p>
      </sec>
    </sec>
    <sec id="sec-4">
      <title>Related Work</title>
      <p>
        Simile’s Semantic Bank [
        <xref ref-type="bibr" rid="ref6">6</xref>
        ] is a snippet repository that lets you persist, share and
publish data collected by individuals, groups or communities. Data capture is
accomplished using a Firefox extension called Piggy Bank [
        <xref ref-type="bibr" rid="ref7">7</xref>
        ], and the information is
accessed using a faceted browser. It is probably the closest system in philosophy to
ours, but there are some important differences. Firstly, Semantic Bank is intended to
be a general purpose, potentially global scale snippet repository. This means that
there are significant research challenges in making the ontologies both sufficiently
compact and understandable. In snippet manager we chose to have a small number
of tightly focused facets. Secondly, our aim is to allow both the gathering of snippets
and the linking of these snippets with data from other sources.
      </p>
      <p>
        The broader idea of a semantically enabled website is explored in a number of
public portals, notably the Semantic Web Community Portal [
        <xref ref-type="bibr" rid="ref8">8</xref>
        ] and MindSwap [
        <xref ref-type="bibr" rid="ref9">9</xref>
        ],
both of which use metadata for filtering and querying. As the number of items
increases, the value of our faceted browsing approach becomes more apparent. There
are other public portals such as Ontaria [
        <xref ref-type="bibr" rid="ref10">10</xref>
        ] and SchemaWeb [
        <xref ref-type="bibr" rid="ref11">11</xref>
        ], which are
primarily intended for browsing ontology data.
      </p>
      <p>
        Many people use their weblog as a knowledge management tool and we think that
structuring the content of a post by adding some metadata could be useful for a group
of people. But the chronological view that weblog gives to the content not always it
the best solution to let users move through information. Our solution to this, which
we call semantic blogging [
        <xref ref-type="bibr" rid="ref3">3</xref>
        ], uses metadata guided views, such as record cards or
tables. A similar approach has been taken by the structured blogging community
[
        <xref ref-type="bibr" rid="ref12">12</xref>
        ]. A more subtle point is the information model, in which the blog entry is no
longer the primary object. Rather, the information item (such as web page, report or
person) which is being blogged about takes centre stage. The blog entry is an
annotation attached to this item. Armed with this perspective, the semantic blog becomes a
useful personal knowledge management tool, and a source of data for the snippet
manager.
      </p>
      <p>
        Wikis are also interesting tools for collaboratively building knowledge, and there
are examples [
        <xref ref-type="bibr" rid="ref13 ref14">13, 14</xref>
        ] that use metadata to enhance navigation and to provide
multiple views. In some ways the snippet manager idea is similar (although our data entry
mechanism is different); however we integrate information from a number of
sources. Just like blogs, wikis are a valuable source of data for the snippet manager.
6
      </p>
    </sec>
    <sec id="sec-5">
      <title>Conclusion</title>
      <p>In this position paper, we have outlined our thoughts on what it would take to build a
snippet manager for small group domain-focused knowledge sharing. We have
shared some design principles which we hope will prove generally useful. We have
also explained how we are going about building a system using these ideas. We have
high hopes that our user-centred approach will function less as an interesting demo
and more as a genuinely useful tool.</p>
    </sec>
  </body>
  <back>
    <ref-list>
      <ref id="ref1">
        <mixed-citation>
          1.
          <string-name>
            <surname>Banks</surname>
            ,
            <given-names>D.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Cayzer</surname>
            ,
            <given-names>S.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Dickinson</surname>
            ,
            <given-names>I.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Reynolds</surname>
            ,
            <given-names>D.</given-names>
          </string-name>
          <article-title>The ePerson Snippet Manager: a Semantic Web Application</article-title>
          .
          <source>Hewlett-Packard Laboratories Technical Report HPL-2002-328</source>
          (
          <year>2002</year>
          ): http://www.hpl.hp.com/techreports/2002/HPL-2002-328.html
        </mixed-citation>
      </ref>
      <ref id="ref2">
        <mixed-citation>
          2.
          <string-name>
            <surname>Davenport</surname>
            ,
            <given-names>T. H.</given-names>
          </string-name>
          <article-title>and</article-title>
          <string-name>
            <surname>Prusak</surname>
          </string-name>
          , L. Working Knowledge:
          <article-title>How Organizations Manage What They Know</article-title>
          . (
          <year>1997</year>
          ) Harvard Business School Press.
        </mixed-citation>
      </ref>
      <ref id="ref3">
        <mixed-citation>
          3.
          <string-name>
            <surname>Cayzer</surname>
            ,
            <given-names>S.</given-names>
          </string-name>
          <article-title>Semantic blogging and decentralized knowledge management</article-title>
          .
          <source>Communications of the ACM</source>
          <volume>47</volume>
          ,
          <issue>12</issue>
          (Dec.
          <year>2004</year>
          ),
          <fpage>47</fpage>
          -
          <lpage>52</lpage>
          . DOI= http://doi.acm.
          <source>org/10</source>
          .1145/1035134.1035164
        </mixed-citation>
      </ref>
      <ref id="ref4">
        <mixed-citation>
          4.
          <string-name>
            <surname>Reynolds</surname>
            ,
            <given-names>D.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Shabajee</surname>
            ,
            <given-names>P.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Cayzer</surname>
            ,
            <given-names>S.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Steer</surname>
            ,
            <given-names>D.</given-names>
          </string-name>
          <string-name>
            <surname>Semantic Portals Demonstrator - Lessons Learnt</surname>
          </string-name>
          .
          <source>Sept</source>
          <year>2004</year>
          . http://www.w3.org/2001/sw/Europe/reports/demo_2_report/
        </mixed-citation>
      </ref>
      <ref id="ref5">
        <mixed-citation>
          5.
          <string-name>
            <surname>Prud'hommeaux</surname>
          </string-name>
          , E.,
          <string-name>
            <surname>Seaborne</surname>
            ,
            <given-names>A. SPARQL</given-names>
          </string-name>
          <article-title>Query Language for RDF</article-title>
          .
          <source>W3C Working Draft 21 July</source>
          <year>2005</year>
          http://www.w3.org/TR/rdf-sparql-query/
        </mixed-citation>
      </ref>
      <ref id="ref6">
        <mixed-citation>6. The Semantic Bank project http://simile.mit.edu/semantic-bank/</mixed-citation>
      </ref>
      <ref id="ref7">
        <mixed-citation>
          7.
          <string-name>
            <surname>Huynh</surname>
            ,
            <given-names>D.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Mazzocchi</surname>
            ,
            <given-names>S.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Karger</surname>
          </string-name>
          . D. Piggy Bank:
          <article-title>Experience the Semantic Web Inside Your Web Browser</article-title>
          . Submitted to 4th
          <source>International Semantic Web Conference (ISWC</source>
          <year>2005</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref8">
        <mixed-citation>8. The Semantic Web Community Portal http://beta.semanticweb.org/</mixed-citation>
      </ref>
      <ref id="ref9">
        <mixed-citation>9. The MindSwap Group http://www.mindswap.org/</mixed-citation>
      </ref>
      <ref id="ref10">
        <mixed-citation>10. The Ontaria project http://www.w3.org/2004/ontaria/</mixed-citation>
      </ref>
      <ref id="ref11">
        <mixed-citation>11. SchemaWeb: RDF Schemas directory http://www.schemaweb.info/</mixed-citation>
      </ref>
      <ref id="ref12">
        <mixed-citation>12. Structured blogging http://structuredblogging.org/</mixed-citation>
      </ref>
      <ref id="ref13">
        <mixed-citation>
          13.
          <string-name>
            <surname>Aumueller</surname>
            ,
            <given-names>D.</given-names>
          </string-name>
          <article-title>Semantic authoring and retrieval within a Wiki</article-title>
          .
          <source>Demonstration track, 2nd European Semantic Web Conference (ESWC</source>
          <year>2005</year>
          ). http://wiki.navigable.info
        </mixed-citation>
      </ref>
      <ref id="ref14">
        <mixed-citation>
          14.
          <string-name>
            <surname>Tazzoli</surname>
            ,
            <given-names>R.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Castagna</surname>
            ,
            <given-names>P.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Campanini</surname>
            ,
            <given-names>S. E.</given-names>
          </string-name>
          <string-name>
            <surname>Towards</surname>
          </string-name>
          <article-title>a Semantic Wiki Wiki Web</article-title>
          . 3rd
          <source>International Semantic Web Conference, Poster Track (ISWC</source>
          <year>2004</year>
          ). http://platypuswiki.sourceforge.net/
        </mixed-citation>
      </ref>
    </ref-list>
  </back>
</article>