<?xml version="1.0" encoding="UTF-8"?>
<TEI xml:space="preserve" xmlns="http://www.tei-c.org/ns/1.0" 
xmlns:xsi="http://www.w3.org/2001/XMLSchema-instance" 
xsi:schemaLocation="http://www.tei-c.org/ns/1.0 https://raw.githubusercontent.com/kermitt2/grobid/master/grobid-home/schemas/xsd/Grobid.xsd"
 xmlns:xlink="http://www.w3.org/1999/xlink">
	<teiHeader xml:lang="en">
		<fileDesc>
			<titleStmt>
				<title level="a" type="main">OntoTag: A Semantic Web Page Linguistic Annotation Model</title>
			</titleStmt>
			<publicationStmt>
				<publisher/>
				<availability status="unknown"><licence/></availability>
			</publicationStmt>
			<sourceDesc>
				<biblStruct>
					<analytic>
						<author>
							<persName><forename type="first">Guadalupe</forename><surname>Aguado De Cea</surname></persName>
							<affiliation key="aff0">
								<orgName type="department">Department of Applied Linguistics to Science and Technology (DLACT)</orgName>
							</affiliation>
							<affiliation key="aff1">
								<orgName type="department">Computer Science Faculty</orgName>
								<orgName type="institution">UPM</orgName>
								<address>
									<settlement>Madrid</settlement>
									<country key="ES">Spain</country>
								</address>
							</affiliation>
						</author>
						<author>
							<persName><forename type="first">Inmaculada</forename><surname>Álvarez De Mon</surname></persName>
							<affiliation key="aff2">
								<orgName type="department" key="dep1">DLACT</orgName>
								<orgName type="department" key="dep2">Telecommunications Engineering College</orgName>
								<orgName type="institution">UPM</orgName>
								<address>
									<settlement>Madrid</settlement>
									<country key="ES">Spain</country>
								</address>
							</affiliation>
						</author>
						<author>
							<persName><forename type="first">Antonio</forename><surname>Pareja-Lora</surname></persName>
							<affiliation key="aff3">
								<orgName type="department" key="dep1">Department of Computer Systems and Programming (DSIP)</orgName>
								<orgName type="department" key="dep2">Computer Science Faculty</orgName>
								<orgName type="institution">UCM</orgName>
								<address>
									<settlement>Madrid</settlement>
									<country key="ES">Spain</country>
								</address>
							</affiliation>
						</author>
						<author>
							<persName><forename type="first">Rosario</forename><surname>Plaza-Arteche</surname></persName>
							<affiliation key="aff4">
								<orgName type="department">Department of Applied Linguistics to Science and Technology (DLACT)</orgName>
							</affiliation>
							<affiliation key="aff5">
								<orgName type="department">Computer Science Faculty</orgName>
								<orgName type="institution">UPM</orgName>
								<address>
									<settlement>Madrid</settlement>
									<country key="ES">Spain</country>
								</address>
							</affiliation>
						</author>
						<title level="a" type="main">OntoTag: A Semantic Web Page Linguistic Annotation Model</title>
					</analytic>
					<monogr>
						<imprint>
							<date/>
						</imprint>
					</monogr>
					<idno type="MD5">CD24FE03EBA5B1D5BA6DDE01DEBD437B</idno>
				</biblStruct>
			</sourceDesc>
		</fileDesc>
		<encodingDesc>
			<appInfo>
				<application version="0.7.2" ident="GROBID" when="2023-03-25T03:45+0000">
					<desc>GROBID - A machine learning software for extracting information from scholarly documents</desc>
					<ref target="https://github.com/kermitt2/grobid"/>
				</application>
			</appInfo>
		</encodingDesc>
		<profileDesc>
			<abstract>
<div xmlns="http://www.tei-c.org/ns/1.0"><p>Although with the Semantic Web initiative much research on web page semantic annotation has already been done by AI researchers, linguistic text annotation, including the semantic one, was originally developed in Corpus Linguistics and its results have been somehow neglected by AI. The purpose of the research presented in this proposal is to prove that integration of results in both fields is not only possible, but also highly useful in order to make Semantic Web pages more machine-readable. A multi-level (possibly multi-purpose and multi-language) annotation model based on EAGLES standards and Ontological Semantics, implemented with last generation Semantic Web languages is being developed to fit the needs of both communities. 1 2 3 4</p></div>
<div xmlns="http://www.tei-c.org/ns/1.0"><head n="1.">INTRODUCTION.</head><p>All of us are by now used to making extensive use of the so-called World Wide Web (WWW) which we might consider a great source of information, accessible through computers but, hitherto, only understandable to human beings. In its beginning, web pages were hand made, intended and oriented to the exchange of information among human beings. All of these documents contained a huge amount of text, images and even sounds, meaningless to a computer. In this way, they put the burden of extracting and interpreting the relevant information on the reader. Due to the astonishing growth of Internet use, new technologies emerged and, with them, machine-aided web page generation appeared.</p><p>Currently, web page presentation in the WWW is being handled independently from its content, mainly through the use of XML <ref type="bibr" target="#b0">[1]</ref> or other resource-oriented languages as XOL [2], SHOE [3], OML [4], RDF [5], RDF Schema [6], OIL <ref type="bibr" target="#b6">[7]</ref> or DAML+OIL <ref type="bibr" target="#b7">[8]</ref>. But even though the automatic process of information is being eased, still the above-mentioned tasks -relevant information access, extraction and interpretation-cannot be wholly performed by computers. Hence, the goal of enabling computers to understand the meaning (the semantics) of written texts and web pages is the main pillar sustaining the development of the Semantic Web <ref type="bibr" target="#b8">[9]</ref>.</p></div>
			</abstract>
		</profileDesc>
	</teiHeader>
	<text xml:lang="en">
		<body>
<div xmlns="http://www.tei-c.org/ns/1.0"><p>this context, the semantic annotation of texts, since it makes meaning explicit, has become a relevant topic and, therefore, advanced design and application of models and formalisms for the semantic annotation of web pages are needed.</p><p>Lately, much research has already been carried out by ontologists on the semantic annotation of web pages <ref type="bibr" target="#b2">[3]</ref>, <ref type="bibr" target="#b9">[10]</ref>, <ref type="bibr" target="#b10">[11]</ref>, <ref type="bibr" target="#b11">[12]</ref>. However, such works have somehow neglected the results obtained on corpus annotation in the field of Corpus Linguistics, not only in the semantic level, but also in other linguistic levels. These other linguistic levels, whilst not being intrinsically semantic, can add extra semantic information to help a computer understand a text or, in our case, web pages.</p><p>The goal of this paper is to present the results of our research in which special efforts are being devoted to finding a way of bringing together and identifying complementarities between the semantic annotation models from AI and the annotations proposed by Corpus Linguistics.</p><p>This paper is organised as follows: firstly, an introduction to the state of the art in semantic annotation in corpus linguistics is presented (section 2). In section 3, some brief notes on the use of ontologies in semantic annotation are sketched. In section 4, an example of the integration of both paradigms (AI's and Corpus Linguistics') is presented in the scope of our project goals. The main advantages of this integration are then analysed -section 5and, finally, further work to be done is included -section 6-.</p></div>
<div xmlns="http://www.tei-c.org/ns/1.0"><head n="2.">SEMANTIC ANNOTATION IN CORPUS LINGUISTICS.</head><p>The idea of text annotation was originally developed in Corpus Linguistics. Traditionally, linguists have defined corpus as "a body of naturally occurring (authentic) language data which can be used as a basis for linguistic research" <ref type="bibr" target="#b12">[13]</ref>. From this point of view, Corpus Linguistics <ref type="bibr" target="#b13">[14]</ref> may not be considered a branch of Linguistics in itself, like syntax or semantics. The latter are focused on describing or explaining an aspect of language use; the former is rather a methodology or an approach which can be taken by these branches to explain or describe their particular aspect of language use. Following the same authors, Corpus Linguistics was first applied to research on language acquisition, to the teaching of a second language, to the elaboration of descriptive grammars, etc.. With the arrival of computers, the number of potential studies to which corpora could be applied increased exponentially.</p><p>So, nowadays, the term corpus is being applied to "a body of language material which exists in electronic form, and which may be processed by computer for various purposes such as linguistic research and language engineering" <ref type="bibr" target="#b12">[13]</ref>. An annotated corpus "may be considered to be a repository of linguistic information [...] made explicit through concrete annotation" <ref type="bibr" target="#b13">[14]</ref>. The benefit of such an annotation is clear: it makes retrieving and analysing information about what is contained in the corpus quicker and easier. Let us now see the recommendations stated in Corpus Linguistics for text semantic annotation.</p><p>As asserted in <ref type="bibr" target="#b13">[14]</ref>, two broad types of semantic annotation may be identified, related to: 1. Semantic relationships between items in the text (i.e., the agents or patients of particular actions). This type of annotation has scarcely begun to be applied. 2. The semantic features of words in a text, essentially the annotation of word senses in one form or another. There is no universal agreement in semantics about which features of words should be annotated 5 .</p><p>Although some preliminary recommendations on lexical semantic encoding have already been posited <ref type="bibr" target="#b14">[15]</ref>, no EAGLES semantic corpus annotation standard has yet been published; nevertheless, for choosing or devising a corpus semantic field 6  annotation system (second type of semantic annotation above mentioned) a set of reference criteria has been proposed by Schmidt and is presented in <ref type="bibr" target="#b15">[16]</ref>. These criteria are: 1. It should make sense in linguistic or psycholinguistic terms. It is known from psycholinguistic experiments that certain basic categories exist in the mind. At present, in general, there is a good agreement between many basic categories we already know about from neuropsychology (for example colours, body parts, topography and so on); but still an exhaustive set of categories is to be determined. Overabstraction must be avoided, in any case. 2. It should be able to account exhaustively for the vocabulary in the corpus, not just for a part of it. If a term cannot readily be classified in the existing annotation system, then the system clearly needs to be amended.</p></div>
<div xmlns="http://www.tei-c.org/ns/1.0"><head n="3.">It should be sufficiently flexible to allow for those</head><p>emendations that are necessary for treating a different period, language, register or textbase. The treatment of specialised texts (such as computer-related, commerce, etc.) may require considerably more detailed subclassification of the domain in question than other texts.</p></div>
<div xmlns="http://www.tei-c.org/ns/1.0"><head n="4.">It should operate at an appropriate level of granularity (or delicacy of detail) -related to criteria (3)</head><p>. What level of granularity is correct for an annotation system is an open question and depends partly on the aims of the end user. For this reason, the next criterion is posited. 5. It should, where appropriate, possess a hierarchical structure.</p><p>If a semantic category system has a hierarchical structure, based on increasingly general levels of relatedness between terms, the end user can look at all the different levels and</p><formula xml:id="formula_0">------------------------------------------------------<label>5</label></formula><p>See, for example, the controversies within the SENSEVAL initiative meetings - <ref type="bibr" target="#b28">[30]</ref>, <ref type="bibr" target="#b29">[31]</ref>. 6 A semantic field (sometimes also called a conceptual field, a semantic domain or a lexical domain) is a theoretical construct which groups together words that are related by virtue of their being connected -at some level of generality-with the same mental concept <ref type="bibr" target="#b15">[16]</ref>.</p><p>decide which one must employ, simply by moving up or down to the next level in the hierarchy. 6. It should conform to a standard, if one exists. A hard-and-fast system of categories, even being the result of a consensual work, may be rejected by many researchers. However, a standard in this level could lay, like EAGLES standards have done in other levels, a broad framework of principles and major categories. Such a standard would facilitate comparability and, at the same time, could be modified as necessary for individual needs 7 .</p></div>
<div xmlns="http://www.tei-c.org/ns/1.0"><head n="3.">ONTOLOGIES AND SEMANTIC WEB ANNOTATIONS.</head><p>AI researchers have found in ontologies <ref type="bibr" target="#b16">[17]</ref>, <ref type="bibr" target="#b17">[18]</ref> the ideal knowledge model to formally describe web resources and its vocabulary and, hence, to make explicit in some way the underlying meaning of the terms included in web pages. With Ontological Semantics <ref type="bibr" target="#b18">[19]</ref> as a support theory 8 , the annotation of these web resources with ontological information should allow intelligent access to them, should ease searching and browsing within them and should exploit new web inference approaches from them. Many systems and projects have been developed: SHOE <ref type="bibr" target="#b2">[3]</ref>; the (KA) 2 initiative <ref type="bibr" target="#b9">[10]</ref>; PlanetOnto <ref type="bibr" target="#b10">[11]</ref> and the Semantic Community Web Portals project <ref type="bibr" target="#b11">[12]</ref>. Semantic annotation tools have also been developed so far: COHSE <ref type="bibr" target="#b19">[20]</ref>, MnM <ref type="bibr" target="#b20">[21]</ref>, OntoMat-Annotizer <ref type="bibr" target="#b21">[22]</ref>, SHOE Knowledge Annotator [23] and AeroDAML <ref type="bibr" target="#b22">[24]</ref>.</p></div>
<div xmlns="http://www.tei-c.org/ns/1.0"><head n="4.">INTEGRATION OF PARADIGMS: AN EXAMPLE.</head><p>As we have already mentioned, the goal of this paper is to present the complementarities of linguistic and ontological annotation for the Semantic Web. The purpose of the project we are presenting, ContentWeb, is the creation of an ontology-based platform to enable users to query e-commerce applications by using natural language, performing the automatic retrieval of information from web documents annotated with ontological and linguistic information. ContentWeb objectives can be enunciated as follows:</p><p>1. Semi-automatic building of ontologies in the domains of ecommerce and of entertainment, reusing existing ontologies and international e-commerce standards and joint initiatives. 2. Elaboration of OntoTag, a model and environment for the hybrid -linguistic and ontological-annotation of web documents.</p></div>
<div xmlns="http://www.tei-c.org/ns/1.0"><head n="3.">Development of OntoConsult, a natural language interface</head><p>based on ontologies.</p><p>------------------------------------------------------ 7 Once again the SENSEVAL initiatives <ref type="bibr" target="#b28">[30]</ref>, <ref type="bibr" target="#b29">[31]</ref> must be mentioned: they reveal the demand for semantic standardization in the field of word sense disambiguation. 8 Ontological Semantics <ref type="bibr" target="#b18">[19]</ref> uses a constructed world model -the ontology-as the central resource for extracting and representing meaning of natural language texts, reasoning about knowledge derived from texts as well as generating natural language texts based on representations of their meaning. </p></div>
<div xmlns="http://www.tei-c.org/ns/1.0"><head n="4.">Creation of OntoAdvice, an ontology-based system</head><p>for querying and retrieving information from annotated web documents in the entertainment domain. One of the tasks performed to reach goal 2 is the manual annotation of a Spanish sentence "Tras cinco años de espera y después de muchas habladurías, llega a nuestras pantallas la película más esperada de los últimos tiempos." ("After five years of expectation and gossiping, here comes the most expected film for the time being.") on the languages XML and RDF(S). The RDF(S) annotation of this sentence in the first three levels is shown in Figure <ref type="figure" target="#fig_0">1</ref>, Figure <ref type="figure">2</ref> and Figure <ref type="figure">3</ref>.</p><p>In the morphosyntactic level (Figure <ref type="figure" target="#fig_0">1</ref>) every word or lexical token is given a different Uniform Resource Identifier (URI). The morphosyntactic annotation of the article "la", according to three different tagsets and systems is presented. Each tagset has been assigned a different class in the morphAnnot namespace: TradAnnot (CRATER tagset), MBTAnnot (MBT tagset <ref type="bibr" target="#b23">[25]</ref>) and ConstrAnnot (Constraint Grammar -CONEXOR tagset <ref type="bibr" target="#b24">[26]</ref>). For the sake of space, just the annotation of the article "la" has been included in the figure <ref type="figure">.</ref> In the syntactic level (Figure <ref type="figure">2</ref>) every syntactic relationship between morpho-syntactic items is given a new URI, so that it can be referenced in higher-level relationships or by other levels of the annotation model (i.e. &lt;synAnnot:Chunk rdf:ID="1_510"&gt;). The annotation of the phrase "la película más esperada de los últimos tiempos" has been included in the figure <ref type="figure">.</ref> In the semantic level (see Figure <ref type="figure">3</ref>) some components of lower level annotations are tagged with semantic references to the concepts, attributes and relationships determined by our (domain) ontology, implemented in the language DAML+OIL.</p></div>
<div xmlns="http://www.tei-c.org/ns/1.0"><head n="5.">ADVANTAGES OF THE INTEGRATED MODEL.</head><p>As shown in the example from section 4, it seems that AI and Corpus Linguistics, far from being irreconcilable, can join together to give birth to an integrated annotation model. This conjunct annotation scheme would be very useful and valuable in the development of the Semantic Web and would benefit from the results of both disciplines in many ways. Let us now see the benefits at the semantic level of a hybrid annotation model, first from a linguistic point of view and, then, from an ontological point of view. </p></div>
<div xmlns="http://www.tei-c.org/ns/1.0"><head n="5.1.">Regarding ontology-based annotations from a linguistic point of view.</head><p>The first result of our work is that the use of ontologies as a basis for a semantic annotation scheme fits perfectly and accomplishes the criteria posited by Schmidt. Clearly, its mostly hierarchical structure fulfils by itself criterion (5) and, as a side effect, criteria (2) and ( <ref type="formula">4</ref>), since an ontology can grow horizontally (in breadth) and vertically (in depth). Criterion (3) is also satisfied by an ontology-based semantic annotation scheme, since we can always specialise the concepts in the ontology according to specific periods, languages, registers and textbases. Ontologies are, by definition, consensual and, thus, are closer to becoming a standard than many other knowledge models, as criteria (6) requires. Concerning criterion <ref type="bibr" target="#b0">(1)</ref>, quite a lot of groups developing ontologies are characterized by a strong interdisciplinary approach that combines Computer Science, Linguistics and (sometimes) Philosophy; then, an ontology-based approach should also make sense in linguistic terms.</p></div>
<div xmlns="http://www.tei-c.org/ns/1.0"><head>5.2.</head><p>Regarding linguistic annotations from an ontological point of view.</p><p>The main drawback for AI researchers to adopt a linguistically motivated annotation model would lie on the fact that (section 2) "there is no universal agreement in semantics about which features of words should be annotated" or on Schmidt's criterion (1): "still an exhaustive set of categories is to be determined". But ontology researchers are trying to fill this gap with initiatives such as the UNSPSC <ref type="bibr" target="#b25">[27]</ref> or RosettaNet <ref type="bibr" target="#b26">[28]</ref> in specific domains (i.e. ecommerce). In any case, linguistic annotations at the semantic level are more ambitious and potentially wider than the strictly ontologybased ones. Establishing a link between semantic annotation and discourse annotation and text construction following the RST approach, which has already been applied in text generation <ref type="bibr" target="#b27">[29]</ref>, seems a fairly promising linguistic enhancement.</p></div>
<div xmlns="http://www.tei-c.org/ns/1.0"><head n="6.">CONCLUSIONS AND FURTHER WORK.</head><p>This paper has shown the results of the research carried out on how linguistic annotation can help computers understand the text contained in a document -a Semantic Web page-bringing together semantic annotation models from AI and the annotations proposed for every linguistic level from Corpus Linguistics.</p><p>Further elements susceptible of semantic annotation are presently being sought and research is being done towards their determination by the team of linguists in our project. The pragmatic counterpart of OntoTag has not yet been tackled at this phase of the project.</p><p>Still, much work must be done in order to fully specify, implement and assess the whole model. Besides, many efforts are being devoted to developing OntoAdvice, the ontology-based information retrieval system, in order to validate this model. </p></div>
<div xmlns="http://www.tei-c.org/ns/1.0"><head>ACKNOWLEDGEMENTS.</head></div><figure xmlns="http://www.tei-c.org/ns/1.0" xml:id="fig_0"><head>Figure 1 :</head><label>1</label><figDesc>Figure 1: Morphosyntactic annotation of the article "la".</figDesc></figure>
<figure xmlns="http://www.tei-c.org/ns/1.0" xml:id="fig_1"><head>Figure 2 :Figure 3 :</head><label>23</label><figDesc>Figure 2: Syntactic annotation of the chunk "la película más esperada de los últimos tiempos" in RDF(S).</figDesc></figure>
<figure xmlns="http://www.tei-c.org/ns/1.0" xml:id="fig_2"><head></head><label></label><figDesc>The research described in this paper is supported by MCyT (Spanish Ministry of Science and Technology) under the project name: ContentWeb: "PLATAFORMA TECNOLÓGICA PARA LA WEB SEMÁNTICA: ONTOLOGÍAS, ANÁLISIS DE LENGUAJE NATURAL Y COMERCIO ELECTRÓNICO" -TIC2001-2745 ("ContentWeb: Semantic Web Technologic &lt;!--Semantic annotation excerpt --&gt; &lt;onto:PremiereEvent rdf:ID="_anon27"&gt; &lt;semSynAnnot:includes rdf:about="#1_13"&gt;llega&lt;/semSynAnnot:includes&gt; &lt;semSynAnnot:includes rdf:about="#1_509"&gt;a nuestras pantallas&lt;/semSynAnnot:includes&gt; &lt;onto:hasFilm rdf:about="#_anon30"/&gt; &lt;/onto:PremiereEvent&gt; &lt;onto:Film rdf:ID="_anon30"&gt; &lt;semAnnot:includes rdf:about="#1_18"&gt;película&lt;/semAnnot:includes&gt; &lt;onto:comment rdf:about="#_anon40"&gt; &lt;onto:comment rdf:about="#_anon41"&gt; &lt;/onto:Film&gt; &lt;onto:ControversialFilm rdf:ID="_anon40"&gt; &lt;semSynAnnot:includes rdf:about="#1_506"&gt;después de muchas habladurías&lt;/semSynAnnot:includes&gt; &lt;/onto:ControversialFilm&gt; &lt;onto:AwaitedFilm rdf:ID="_anon41"&gt; &lt;semSynAnnot:includes rdf:about="#1_503"&gt;Tras cinco años de espera&lt;/semSynAnnot:includes&gt; &lt;semSynAnnot:includes rdf:about="#1_512"&gt;más esperada de los últimos tiempos&lt;/semSynAnnot:includes&gt; &lt;/onto:ControversialFilm&gt; &lt;onto:Film rdf:about="#_anon30"&gt; &lt;semSynAnnot:includes rdf:about="#3_507"&gt;El Señor de los Anillos&lt;/semSynAnnot:includes&gt; &lt;onto:filmTitle&gt;El Señor de los Anillos&lt;/onto:filmTitle&gt; &lt;/onto:Film&gt;</figDesc></figure>
		</body>
		<back>

			<div type="acknowledgement">
<div xmlns="http://www.tei-c.org/ns/1.0"><p>Platform: Ontologies, Natural Language Analysis and E-Business"). We would also like to thank Socorro Bernardos, Óscar Corcho and Mariano Fernández for their help with the ontological aspects of this paper.</p></div>
			</div>

			<div type="references">

				<listBibl>

<biblStruct xml:id="b0">
	<monogr>
		<author>
			<persName><forename type="first">T</forename><surname>Bray</surname></persName>
		</author>
		<author>
			<persName><forename type="first">J</forename><surname>Paoli</surname></persName>
		</author>
		<author>
			<persName><forename type="first">C</forename><surname>Sperberg</surname></persName>
		</author>
		<ptr target="http://www.w3.org/TR/REC-xml" />
		<title level="m">Extensible Markup Language (XML) 1.0. W3C Recommendation</title>
				<imprint>
			<date type="published" when="1998">1998</date>
		</imprint>
	</monogr>
</biblStruct>

<biblStruct xml:id="b1">
	<monogr>
		<title level="m" type="main">XOL: An XML-Based Ontology Exchange Language</title>
		<author>
			<persName><forename type="first">R</forename><surname>Karp</surname></persName>
		</author>
		<author>
			<persName><forename type="first">V</forename><surname>Chaudhri</surname></persName>
		</author>
		<author>
			<persName><forename type="first">J</forename><surname>Thomere</surname></persName>
		</author>
		<ptr target="http://www.ai.sri.com/~pkarp/xol/xol.html" />
		<imprint>
			<date type="published" when="1999">1999</date>
		</imprint>
	</monogr>
	<note type="report_type">Technical Report</note>
</biblStruct>

<biblStruct xml:id="b2">
	<monogr>
		<author>
			<persName><forename type="first">S</forename><surname>Luke</surname></persName>
		</author>
		<author>
			<persName><forename type="first">J</forename><surname>Heflin</surname></persName>
		</author>
		<ptr target="http://www.cs.umd.edu/projects/plus/SHOE/spec1.01.htm" />
		<title level="m">Proposed Specification. SHOE Project</title>
				<imprint>
			<date type="published" when="2000">2000</date>
			<biblScope unit="page">1</biblScope>
		</imprint>
	</monogr>
	<note>SHOE 1</note>
</biblStruct>

<biblStruct xml:id="b3">
	<monogr>
		<title level="m" type="main">Conceptual Knowledge Markup Language (version 0</title>
		<author>
			<persName><forename type="first">R</forename><surname>Kent</surname></persName>
		</author>
		<ptr target="http://sern.ucalgary.ca/KSI/KAW/KAW99/papers/Kent1/CKML.pdf" />
		<imprint>
			<date type="published" when="1998">1998</date>
		</imprint>
	</monogr>
</biblStruct>

<biblStruct xml:id="b4">
	<monogr>
		<title level="m" type="main">Resource Description Framework (RDF) Model and Syntax Specification</title>
		<author>
			<persName><forename type="first">O</forename><surname>Lassila</surname></persName>
		</author>
		<author>
			<persName><forename type="first">R</forename><surname>Swick</surname></persName>
		</author>
		<ptr target="http://www.w3.org/TR/PR-rdf-syntax" />
		<imprint>
			<date type="published" when="1999">1999</date>
		</imprint>
	</monogr>
	<note>W3C Recommendation</note>
</biblStruct>

<biblStruct xml:id="b5">
	<analytic>
		<title level="a" type="main">Resource Description Framework (RDF) Schema Specification</title>
		<author>
			<persName><forename type="first">D</forename><surname>Brickley</surname></persName>
		</author>
		<author>
			<persName><forename type="first">R</forename><forename type="middle">V</forename><surname>Guha</surname></persName>
		</author>
		<ptr target="http://www.w3.org/TR/PR-rdf-schema" />
	</analytic>
	<monogr>
		<title level="m">W3C Candidate Recommendation</title>
				<imprint>
			<date type="published" when="2000">2000</date>
		</imprint>
	</monogr>
</biblStruct>

<biblStruct xml:id="b6">
	<analytic>
		<title level="a" type="main">OIL in a Nutshell</title>
		<author>
			<persName><forename type="first">I</forename><surname>Horrocks</surname></persName>
		</author>
		<author>
			<persName><forename type="first">D</forename><surname>Fensel</surname></persName>
		</author>
		<author>
			<persName><forename type="first">F</forename><surname>Harmelen</surname></persName>
		</author>
		<author>
			<persName><forename type="first">S</forename><surname>Decker</surname></persName>
		</author>
		<author>
			<persName><forename type="first">M</forename><surname>Erdmann</surname></persName>
		</author>
		<author>
			<persName><forename type="first">M</forename><surname>Klein</surname></persName>
		</author>
		<ptr target="http://www.cs.vu.nl/~ontoknow/oil/downl/oilnutshell.pdf" />
	</analytic>
	<monogr>
		<title level="m">12 th International Conference in Knowledge Engineering and Knowledge Management</title>
		<title level="s">Lecture Notes in Artificial Intelligence</title>
		<meeting><address><addrLine>Berlin, Germany</addrLine></address></meeting>
		<imprint>
			<publisher>Springer-Verlag</publisher>
			<date type="published" when="2000">2000</date>
			<biblScope unit="page" from="1" to="16" />
		</imprint>
	</monogr>
</biblStruct>

<biblStruct xml:id="b7">
	<monogr>
		<title level="m" type="main">Reference description of the DAML+OIL ontology markup language</title>
		<author>
			<persName><forename type="first">I</forename><surname>Horrocks</surname></persName>
		</author>
		<author>
			<persName><forename type="first">F</forename><surname>Van Harmelen</surname></persName>
		</author>
		<ptr target="http://www.daml.org/2000/12/reference.html" />
		<imprint>
			<date type="published" when="2001">2001. 2001</date>
		</imprint>
	</monogr>
	<note type="report_type">Draft report</note>
</biblStruct>

<biblStruct xml:id="b8">
	<monogr>
		<title level="m" type="main">Weaving the Web: The Original Design and Ultimate Destiny of the World Wide Web by its Inventor</title>
		<author>
			<persName><forename type="first">T</forename><surname>Berners-Lee</surname></persName>
		</author>
		<author>
			<persName><forename type="first">M</forename><surname>Fischetti</surname></persName>
		</author>
		<imprint>
			<date type="published" when="1999">1999</date>
			<publisher>Harper</publisher>
			<pubPlace>San Francisco</pubPlace>
		</imprint>
	</monogr>
</biblStruct>

<biblStruct xml:id="b9">
	<analytic>
		<title level="a" type="main">KA) 2 : Building Ontologies for the Internet: a Mid Term Report</title>
		<author>
			<persName><forename type="first">V</forename><forename type="middle">R</forename><surname>Benjamins</surname></persName>
		</author>
		<author>
			<persName><forename type="first">D</forename><surname>Fensel</surname></persName>
		</author>
		<author>
			<persName><forename type="first">S</forename><surname>Decker</surname></persName>
		</author>
		<author>
			<persName><forename type="first">A</forename><surname>Gómez-Pérez</surname></persName>
		</author>
	</analytic>
	<monogr>
		<title level="j">IJHCS, International Journal of Human Computer Studies</title>
		<imprint>
			<biblScope unit="volume">51</biblScope>
			<biblScope unit="page" from="687" to="712" />
			<date type="published" when="1999">1999</date>
		</imprint>
	</monogr>
</biblStruct>

<biblStruct xml:id="b10">
	<analytic>
		<title level="a" type="main">Case Studies in Ontology-Driven Document Enrichment</title>
		<author>
			<persName><forename type="first">E</forename><surname>Motta</surname></persName>
		</author>
		<author>
			<persName><forename type="first">S</forename><surname>Buckingham Shum</surname></persName>
		</author>
		<author>
			<persName><forename type="first">J</forename><surname>Domingue</surname></persName>
		</author>
	</analytic>
	<monogr>
		<title level="m">Proceedings of the 12th Banff Knowledge Acquisition Workshop</title>
				<meeting>the 12th Banff Knowledge Acquisition Workshop<address><addrLine>Banff, Alberta, Canada</addrLine></address></meeting>
		<imprint>
			<date type="published" when="1999">1999</date>
		</imprint>
	</monogr>
</biblStruct>

<biblStruct xml:id="b11">
	<monogr>
		<title level="m" type="main">Semantic Community Web Portals</title>
		<author>
			<persName><forename type="first">S</forename><surname>Staab</surname></persName>
		</author>
		<author>
			<persName><forename type="first">J</forename><surname>Angele</surname></persName>
		</author>
		<author>
			<persName><forename type="first">S</forename><surname>Decker</surname></persName>
		</author>
		<author>
			<persName><forename type="first">M</forename><surname>Erdmann</surname></persName>
		</author>
		<author>
			<persName><forename type="first">A</forename><surname>Hotho</surname></persName>
		</author>
		<author>
			<persName><forename type="first">A</forename><surname>Mädche</surname></persName>
		</author>
		<author>
			<persName><forename type="first">H.-P</forename><surname>Schnurr</surname></persName>
		</author>
		<author>
			<persName><forename type="first">R</forename><surname>Studer</surname></persName>
		</author>
		<imprint>
			<date type="published" when="2000">2000</date>
			<publisher>WWW´9</publisher>
			<pubPlace>Amsterdam</pubPlace>
		</imprint>
	</monogr>
</biblStruct>

<biblStruct xml:id="b12">
	<analytic>
		<title level="a" type="main">Introducing corpus annotation</title>
		<author>
			<persName><forename type="first">G</forename><surname>Leech</surname></persName>
		</author>
	</analytic>
	<monogr>
		<title level="m">Corpus Annotation: Linguistic Information from Computer Text Corpora</title>
				<editor>
			<persName><forename type="first">R</forename><surname>Garside</surname></persName>
		</editor>
		<editor>
			<persName><forename type="first">G</forename><surname>Leech</surname></persName>
		</editor>
		<editor>
			<persName><forename type="first">A</forename><forename type="middle">M</forename><surname>Mcenery</surname></persName>
		</editor>
		<meeting><address><addrLine>London</addrLine></address></meeting>
		<imprint>
			<publisher>Longman</publisher>
			<date type="published" when="1997">1997a</date>
		</imprint>
	</monogr>
</biblStruct>

<biblStruct xml:id="b13">
	<monogr>
		<author>
			<persName><forename type="first">A</forename><forename type="middle">M</forename><surname>Mcenery</surname></persName>
		</author>
		<author>
			<persName><forename type="first">A</forename><surname>Wilson</surname></persName>
		</author>
		<title level="m">Corpus Linguistics: An Introduction</title>
				<meeting><address><addrLine>Edinburgh</addrLine></address></meeting>
		<imprint>
			<publisher>Edinburgh University Press</publisher>
			<date type="published" when="2001">2001</date>
		</imprint>
	</monogr>
</biblStruct>

<biblStruct xml:id="b14">
	<monogr>
		<author>
			<persName><surname>Eagles</surname></persName>
		</author>
		<ptr target="http://www.ilc.pi.cnr.it/EAGLES/EAGLESLE.PDF" />
		<title level="m">EAGLES LE3-4244: Preliminary Recommendations on Semantic Encoding</title>
				<imprint>
			<date type="published" when="1999">1999</date>
		</imprint>
	</monogr>
	<note>Final Report</note>
</biblStruct>

<biblStruct xml:id="b15">
	<analytic>
		<title level="a" type="main">Semantic Annotation</title>
		<author>
			<persName><forename type="first">A</forename><surname>Wilson</surname></persName>
		</author>
		<author>
			<persName><forename type="first">J</forename><surname>Thomas</surname></persName>
		</author>
	</analytic>
	<monogr>
		<title level="m">Corpus Annotation: Linguistic Information from Computer Text Corpora</title>
				<editor>
			<persName><forename type="first">R</forename><surname>Garside</surname></persName>
		</editor>
		<editor>
			<persName><forename type="first">G</forename><surname>Leech</surname></persName>
		</editor>
		<editor>
			<persName><forename type="first">A</forename><forename type="middle">M</forename><surname>Mcenery</surname></persName>
		</editor>
		<meeting><address><addrLine>London</addrLine></address></meeting>
		<imprint>
			<publisher>Longman</publisher>
			<date type="published" when="1997">1997</date>
		</imprint>
	</monogr>
</biblStruct>

<biblStruct xml:id="b16">
	<analytic>
		<title level="a" type="main">A translation approach to portable ontology specification</title>
		<author>
			<persName><forename type="first">R</forename><surname>Gruber</surname></persName>
		</author>
	</analytic>
	<monogr>
		<title level="j">Knowledge Acquisition</title>
		<imprint>
			<biblScope unit="volume">5</biblScope>
			<biblScope unit="page" from="199" to="220" />
			<date type="published" when="1993">1993</date>
		</imprint>
	</monogr>
</biblStruct>

<biblStruct xml:id="b17">
	<analytic>
		<title level="a" type="main">Knowledge Engineering: Principles and Methods</title>
		<author>
			<persName><forename type="first">R</forename><surname>Studer</surname></persName>
		</author>
		<author>
			<persName><forename type="first">R</forename><surname>Benjamins</surname></persName>
		</author>
		<author>
			<persName><forename type="first">D</forename><surname>Fensel</surname></persName>
		</author>
	</analytic>
	<monogr>
		<title level="j">DKE</title>
		<imprint>
			<biblScope unit="volume">25</biblScope>
			<biblScope unit="issue">1-2</biblScope>
			<biblScope unit="page" from="161" to="197" />
			<date type="published" when="1998">1998</date>
		</imprint>
	</monogr>
</biblStruct>

<biblStruct xml:id="b18">
	<monogr>
		<title level="m" type="main">Ontological Semantics</title>
		<author>
			<persName><forename type="first">S</forename><surname>Nirenburg</surname></persName>
		</author>
		<author>
			<persName><forename type="first">V</forename><surname>Raskin</surname></persName>
		</author>
		<ptr target="http://crl.nmsu.edu/Staff.pages/Technical/sergei/book/index-book.html" />
		<imprint>
			<date type="published" when="2001">2001</date>
		</imprint>
	</monogr>
</biblStruct>

<biblStruct xml:id="b19">
	<monogr>
		<title/>
		<author>
			<persName><surname>Cohse</surname></persName>
		</author>
		<ptr target="http://cohse.semanticweb.org/" />
		<imprint>
			<date type="published" when="2002">2002</date>
		</imprint>
	</monogr>
</biblStruct>

<biblStruct xml:id="b20">
	<analytic>
		<title level="a" type="main">Knowledge Extraction by Using an Ontology-based Annotation Tool</title>
		<author>
			<persName><forename type="first">M</forename><surname>Vargas-Vera</surname></persName>
		</author>
		<author>
			<persName><forename type="first">E</forename><surname>Motta</surname></persName>
		</author>
		<author>
			<persName><forename type="first">J</forename><surname>Domingue</surname></persName>
		</author>
		<author>
			<persName><forename type="first">S</forename><forename type="middle">B</forename><surname>Shum</surname></persName>
		</author>
		<author>
			<persName><forename type="first">M</forename><surname>Lanzoni</surname></persName>
		</author>
	</analytic>
	<monogr>
		<title level="m">Proceedings of the K-CAP&apos;01 Workshop on Knowledge Markup and Semantic Annotation</title>
				<meeting>the K-CAP&apos;01 Workshop on Knowledge Markup and Semantic Annotation<address><addrLine>Victoria B.C., Canada</addrLine></address></meeting>
		<imprint>
			<date type="published" when="2001">2001</date>
		</imprint>
	</monogr>
</biblStruct>

<biblStruct xml:id="b21">
	<monogr>
		<title/>
		<author>
			<persName><surname>Ontomat</surname></persName>
		</author>
		<ptr target="http://www.cs.umd.edu/projects/plus/SHOE/KnowledgeAnnotator.html" />
		<imprint>
			<date type="published" when="2002">2002. 2002</date>
		</imprint>
	</monogr>
</biblStruct>

<biblStruct xml:id="b22">
	<monogr>
		<title/>
		<author>
			<persName><surname>Aerodaml</surname></persName>
		</author>
		<ptr target="http://ubot.lockheedmartin.com/ubot/hotdaml/aerodaml.html" />
		<imprint>
			<date type="published" when="2002">2002</date>
		</imprint>
	</monogr>
</biblStruct>

<biblStruct xml:id="b23">
	<monogr>
		<title/>
		<author>
			<persName><surname>Mbt</surname></persName>
		</author>
		<ptr target="http://ilk.kub.nl/~zavrel/tagtest.html" />
		<imprint>
			<date type="published" when="2002">2002</date>
		</imprint>
	</monogr>
</biblStruct>

<biblStruct xml:id="b24">
	<monogr>
		<title/>
		<author>
			<persName><forename type="first">O</forename><forename type="middle">Y</forename><surname>Conexor</surname></persName>
		</author>
		<ptr target="http://www.conexoroy.com/products.htm" />
		<imprint>
			<date type="published" when="2002">2002</date>
		</imprint>
	</monogr>
</biblStruct>

<biblStruct xml:id="b25">
	<monogr>
		<ptr target="http://www.unspsc.org/" />
		<title level="m">Universal Standard Products and Services Classification (UNSPSC</title>
				<imprint>
			<date type="published" when="2002">2002</date>
		</imprint>
		<respStmt>
			<orgName>UNSPSC</orgName>
		</respStmt>
	</monogr>
</biblStruct>

<biblStruct xml:id="b26">
	<monogr>
		<author>
			<persName><surname>Rosettanet</surname></persName>
		</author>
		<ptr target="http://www.rosettanet.org/" />
		<title level="m">RosettaNet: Lingua Franca for eBusiness</title>
				<imprint>
			<date type="published" when="2002">2002</date>
		</imprint>
	</monogr>
</biblStruct>

<biblStruct xml:id="b27">
	<analytic>
		<title level="a" type="main">Rhetorical Structure Theory: Toward a functional theory of text organization</title>
		<author>
			<persName><forename type="first">W</forename><surname>Mann</surname></persName>
		</author>
		<author>
			<persName><forename type="first">S</forename><surname>Thomson</surname></persName>
		</author>
	</analytic>
	<monogr>
		<title level="j">Text</title>
		<imprint>
			<biblScope unit="volume">18</biblScope>
			<biblScope unit="issue">3</biblScope>
			<biblScope unit="page" from="243" to="281" />
			<date type="published" when="1988">1988</date>
		</imprint>
	</monogr>
</biblStruct>

<biblStruct xml:id="b28">
	<analytic>
		<title level="a" type="main">SENSEVAL: An Exercise in Evaluating Word Sense Disambiguation Programs</title>
		<author>
			<persName><forename type="first">A</forename><surname>Kilgarriff</surname></persName>
		</author>
	</analytic>
	<monogr>
		<title level="m">Proceedings of LREC</title>
				<meeting>LREC<address><addrLine>Granada, Spain</addrLine></address></meeting>
		<imprint>
			<date type="published" when="1998">1998</date>
			<biblScope unit="page" from="581" to="588" />
		</imprint>
	</monogr>
</biblStruct>

<biblStruct xml:id="b29">
	<analytic>
		<title level="a" type="main">English SENSEVAL: Report and Results</title>
		<author>
			<persName><forename type="first">A</forename><surname>Kilgarriff</surname></persName>
		</author>
		<author>
			<persName><forename type="first">J</forename><surname>Rosenzweig</surname></persName>
		</author>
	</analytic>
	<monogr>
		<title level="m">Proceedings of LREC</title>
				<meeting>LREC<address><addrLine>Athens, Greece</addrLine></address></meeting>
		<imprint>
			<date type="published" when="2000">2000</date>
		</imprint>
	</monogr>
</biblStruct>

				</listBibl>
			</div>
		</back>
	</text>
</TEI>
