<!DOCTYPE article PUBLIC "-//NLM//DTD JATS (Z39.96) Journal Archiving and Interchange DTD v1.0 20120330//EN" "JATS-archivearticle1.dtd">
<article xmlns:xlink="http://www.w3.org/1999/xlink">
  <front>
    <journal-meta />
    <article-meta>
      <title-group>
        <article-title>University of Hagen at GeoCLEF 2007: Exploring Location Indicators for Geographic Information Retrieval</article-title>
      </title-group>
      <contrib-group>
        <contrib contrib-type="author">
          <string-name>Johannes Leveling</string-name>
          <xref ref-type="aff" rid="aff0">0</xref>
          <xref ref-type="aff" rid="aff1">1</xref>
        </contrib>
        <contrib contrib-type="author">
          <string-name>Sven Hartrumpf</string-name>
          <xref ref-type="aff" rid="aff0">0</xref>
          <xref ref-type="aff" rid="aff1">1</xref>
        </contrib>
        <contrib contrib-type="author">
          <string-name>General Terms</string-name>
          <xref ref-type="aff" rid="aff0">0</xref>
          <xref ref-type="aff" rid="aff1">1</xref>
        </contrib>
        <aff id="aff0">
          <label>0</label>
          <institution>58084 Hagen</institution>
          ,
          <country country="DE">Germany</country>
        </aff>
        <aff id="aff1">
          <label>1</label>
          <institution>University of Hagen, FernUniversita ̈t in Hagen</institution>
        </aff>
      </contrib-group>
      <abstract>
        <p>Location indicators are text segments from which a geographic scope can be inferred, e.g. adjectives, demonyms (names for inhabitants of a place), geographic codes, orthographic variants, and abbreviations can be mapped to location names in one or more inferential steps. In this paper, the normalization of location indicators and treating morphology of location indicators for geographic information retrieval (GIR) within the system GIRSA (Geographic Information Retrieval by Semantic Annotation) are explored. Several retrieval experiments are performed on the German GeoCLEF 2007 data, including a baseline IR experiment on stemmed text (0.119 mean average precision, MAP). Results for this experiment are compared to results for experiments with normalized location indicators. Additionally, the latter approach was combined with an approach using semantic networks for retrieval (an extension of an experiment performed for GeoCLEF 2005). When using the topic title and description, the best performance was achieved by the combination of approaches (0.196 MAP); adding location names from the narrative part increased MAP to 0.258. Results indicate that 1) employing normalized location indicators improves MAP and increases the number of relevant documents found; 2) additional location names from the narrative increase MAP and recall, and 3) the semantic network approach has a high initial precision and even adds some relevant documents which were previously not found. For bilingual (English-German) experiments, queries were first translated into German before utilizing the translation as input to GIRSA. Performance for these experiments is generally lower, but reflect results for monolingual German. The baseline experiment (0.114 MAP) is clearly outperformed by all other experiments, achieving the best performance for a setup using title, description, and narrative (0.209 MAP).</p>
      </abstract>
      <kwd-group>
        <kwd>H</kwd>
        <kwd>3</kwd>
        <kwd>1 [Information Storage and Retrieval]</kwd>
        <kwd>Content Analysis and Indexing</kwd>
        <kwd>Indexing methods</kwd>
        <kwd>Linguistic processing</kwd>
        <kwd>H</kwd>
        <kwd>3</kwd>
        <kwd>3 [Information Storage and Retrieval]</kwd>
        <kwd>Information Search and Retrieval</kwd>
        <kwd>Query formulation</kwd>
        <kwd>Search process</kwd>
        <kwd>H</kwd>
        <kwd>3</kwd>
        <kwd>4 [Information Storage and Retrieval]</kwd>
        <kwd>Systems and Software</kwd>
        <kwd>Performance evaluation (efficiency and effectiveness)</kwd>
      </kwd-group>
    </article-meta>
  </front>
  <body>
    <sec id="sec-1">
      <title>1 Introduction</title>
      <p>Traditional information retrieval applies stemming to all words in a text. In the context of geographical
information retrieval (GIR) on textual information, named entity recognition and classification play an
important role to identify location names and to avoid stemming them. GIR is concerned with
facilitating geographically-aware retrieval of information. This awareness often results from identifying proper
nouns in the text, disambiguating them further into person names, organization names, and location names
(geographic entities). Thus, identification of location names is typically restricted to proper nouns only.</p>
      <p>The main goal of this paper is to investigate if one should aim at a broader GIR approach which is not
solely based on proper nouns corresponding to location names. To test end, the notion of location indicators
is introduced and retrieval experiments are performed by the system GIRSA (Geographic Information
Retrieval by Semantic Annotation).1 The experiments are based on documents and topics for GeoCLEF
2007, the geographic information retrieval task at CLEF 2007 (Cross Language Evaluation Forum).
2
2.1</p>
    </sec>
    <sec id="sec-2">
      <title>Location Indicators</title>
      <sec id="sec-2-1">
        <title>Definition</title>
        <p>In this paper, location indicators are investigated. Location indicators are text segments from which the
geographic scope of a document can be inferred. They include, but are not limited to:
• Adjectives corresponding to a location.</p>
        <sec id="sec-2-1-1">
          <title>Examples: “tunesisch”/“Tunisian” for “Tunesien”/“Tunisia”; “irisch”/“Irish” for “Irland”/“Ireland”; “bayrisch, bayerisch”/“Bavarian” for “Bayern”/“Bavaria”.</title>
          <p>• Demonyms, e.g. the name for inhabitants originating from a location.</p>
        </sec>
        <sec id="sec-2-1-2">
          <title>Examples: “Franzose, Franzo¨sin”/“Frenchman, Frenchwoman” for “Frankreich”/“France”; “Mongole, Mongolin”/“Mongolian” for “Mongolei”/“Mongolia”; “Du¨sseldorfer, Du¨sseldorferin”/“inhabitant of Du¨sseldorf ” for “Du¨sseldorf ”.</title>
          <p>• Codes for a location name, including ISO region codes, postal and zip codes.</p>
        </sec>
        <sec id="sec-2-1-3">
          <title>Examples: “HU21” for “Tolna County, Hungary” (FIPS region code); “GUY” for “Guyana” (ISO</title>
          <p>3166-1 alpha-3); “GY” for “Guyana” (ISO 3166-1 alpha-2); “EGLL” for “Heathrow Airport,
London, UK” or “LPBJ” for “Beja Air Base, Beja, Portugal” (International Civil Aviation Organization
codes).
• Abbreviations and acronyms for a location name, including abbreviations of adjectives.</p>
        </sec>
        <sec id="sec-2-1-4">
          <title>Examples: “franz.” for “franzo¨sisch”/“French” (mapped to “Frankreich”/“France”); “ital.” for</title>
          <p>“italienisch”/“Italian” (“Italien”/“Italy”); “Whv.” for “Wilhelmshaven”; “NRW” for
“Nordrhein</p>
        </sec>
        <sec id="sec-2-1-5">
          <title>Westfalen”/“North Rhine-Westphalia” .</title>
          <p>• Orthographic variants, including exonyms and historic names.</p>
        </sec>
        <sec id="sec-2-1-6">
          <title>Examples: “Cologne” for “Ko¨ln”; “Lower Saxony” for “Niedersachsen”.</title>
          <p>• Language names in the text.</p>
        </sec>
        <sec id="sec-2-1-7">
          <title>Example: “Portuguese” for “Portuguese speaking countries” (mapped to “Portugal, Angola, Cape</title>
        </sec>
        <sec id="sec-2-1-8">
          <title>Verde, East Timor, Mozambique, and Brazil”).</title>
          <p>1The research described is part of the IRSAW project (Intelligent Information Retrieval on the Basis of a Semantically Annotated
Web; LIS 4 – 554975(2) Hagen, BIB 48 HGfu 02-01), which is funded by the DFG (Deutsche Forschungsgemeinschaft).
• Meta-information for a document, i.e. the language a document is written in.</p>
        </sec>
        <sec id="sec-2-1-9">
          <title>Example: “Die Katze jagt die Maus” for “German language” (mapped to “Germany, Austria, and</title>
        </sec>
        <sec id="sec-2-1-10">
          <title>Switzerland”).</title>
          <p>• Unique entities associated with a geographic location, i.e. headquarters of an organization, persons,
and buildings.</p>
        </sec>
        <sec id="sec-2-1-11">
          <title>Examples: “Boeing” for “Seattle, Washington”; “Molie´re” for “France”; “Galileo Galilei” for “Italy”; “Eiffel Tower” for “Paris”; “Pentagon” for “Washington, D.C.”.</title>
          <p>• The location names itself, including full names and short forms.</p>
        </sec>
        <sec id="sec-2-1-12">
          <title>Example: “Republik Korea”/“Republic of Korea” for “Su¨dkorea”/“South Korea”.</title>
          <p>Typically, location indicators are not included in gazetteers, e.g. the morphology and lexical
knowledge for adjectives is missing completely. Distinct location indicators contribute differently to the task of
assigning a geographic scope to a document. Their importance depends on their usage and frequency in
the corpus (e.g. adjectives are generally frequent) and the correctness of identifying them, because new
ambiguities arise (e.g. the ISO 3166-1 code for Tuvalu (TV) is also the abbreviation for television).
2.2</p>
        </sec>
      </sec>
      <sec id="sec-2-2">
        <title>Location Indicator Normalization</title>
        <p>The normalization of location indicators to location names takes place on different levels of linguistic
analysis in GIRSA.</p>
        <p>• Character level: In all entries of the name lexicons, diacritical marks are replaced with non-accented
characters to create orthographic variants of names. These resulting orthographic variants are used as
elements of a synonym set and normalized by selecting a representative for the synonym set (synset).</p>
        <p>Example: “Que´bec” →“Quebec”.
• Morphologic level: Inflectional endings for adjective and noun forms are identified and separated
using a set of manually created rules and large lists of exceptions. Typical German inflectional
endings of a word form (e.g. “-s” , “-es” , “-er” , “-en” , “e”) are removed before the lookup in name
lexicons. (Note that location names usually do not have a plural form.)
More complex cases are multi-word expressions which may contain inflectional morphology.
Morphologic variations of location names are reduced to its base form.</p>
        <p>Examples: “Berlins” →“Berlin”; “das Rote Meer” →“Rote Meer”; “des Roten Meer(e)s” →“Rote</p>
        <sec id="sec-2-2-1">
          <title>Meer”.</title>
          <p>Derivational morphology is part of connecting adjectives to location names.</p>
          <p>Example: “bayrisch” →“Bayern”; “da¨nisch” →“Da¨nemark”.
• Semantic level: Prefixes indicating compass directions are separated from the name. A database
management system may view the hyphenated result as either one or two terms, depending on the
search options. Thus, a search for “Norddeutschland” will also return documents containing the
phrase “im Norden Deutschlands”. Also on the semantic level, a mapping between location
indicators and location names takes place.</p>
          <p>Examples: “Norddeutschland” →“Nord-Deutschland” ; “Su¨d-Frankreich”
exception: “Su¨dafrika” →“Su¨dafrika”.
→“Su¨d-Frankreich” ;
• Lexical level: Name variations are normalized using synset representatives. The synsets contain
elements referencing the same geographic location.</p>
          <p>Example: “Burma”, “Birma” →“Myanmar”.</p>
          <p>Of course, there is an implicit ordering of normalization steps: morphological variations are
identified first, removing inflectional endings before lookup. Then, complex named entities are recognized and
represented as a single term. Next, adjectives and acronyms are mapped to the expanded location name.
Normalization by mapping to a synset representative is the last operation.
2.3</p>
        </sec>
      </sec>
      <sec id="sec-2-3">
        <title>Semantic Analysis for GIR</title>
        <p>
          This year, the approach of semantic representation matching (GIR-InSicht, derived from the deep QA
system InSicht, [
          <xref ref-type="bibr" rid="ref5">5</xref>
          ]) was tried again for GeoCLEF. See [
          <xref ref-type="bibr" rid="ref8">8</xref>
          ] for details on the first experiment in this direction
at GeoCLEF 2005. GIR-InSicht matches reduced semantic representations of the topic description (or topic
title) to the semantic representations of sentences from the document collection. This process is quite strict
and proceeds sentence by sentence.2 Before matching starts, the query semantic network was allowed to
be split in parts at specific semantic relations, e.g. at a LOC relation (location of a situation or object) of
the MultiNet formalism (multilayered extended semantic networks; [
          <xref ref-type="bibr" rid="ref6">6</xref>
          ]), to increase recall while not losing
too much precision.
        </p>
        <p>For GeoCLEF 2007, query decomposition was implemented, i.e. a query can be decomposed into two
dependent queries, the subquery and the main query. The subquery was answered by the QA system
InSicht; the answers were integrated into the main query on the semantic network level (thereby avoiding the
complicated or problematic integration on the surface level). For example, the title of topic 10.2452/57-GC
“Whiskyherstellung auf den schottischen Inseln” (‘Whiskey production on the Scottish Islands’) and
similarly the description of this topic lead to the subquery “Nenne schottische Inseln” (‘Name Scottish islands’).
Decomposition is also applied to the alternative query semantic networks derived by inferential query
expansion. In the above example, this leads to the subquery “Nenne Inseln in Schottland” (‘Name islands in
Scotland’). InSicht answers the subqueries on the semantic representations of the GeoCLEF document
collection and the German Wikipedia. For the above subqueries, it correctly delivered islands like “Iona” and
“Islay”, which in turn lead to main query semantic networks which could be paraphrased as
“Whiskyherstellung auf Iona” (‘Whiskey production on Iona’) and “Whiskyherstellung auf Islay” (‘Whiskey production
on Islay’). Note that the decomposed queries are processed only as alternatives to the original query.</p>
        <p>Another decomposition strategy produces questions aiming at meronymy knowledge based on the
geographical type of a location, e.g. for a country C in the original query a subquery like “Name cities in
C.” is generated, whose results are integrated into the main query semantic network. This strategy led to
interesting questions like “Welcher Staat/Welche Region/Welche Stadt liegt im Himalaya?” (‘Which
country/region/city is located in the Himalaya?’). In total, both decomposition strategies led to 80 different
subqueries for the 25 topics. After the title and description of a topic have been processed independently,
GIR-InSicht combines the results. If a document occurs in the title results and the description results, the
highest score was taken for the combination.</p>
        <p>The semantic matching approach is completely independent of the main approach in GIRSA. Some
of the functionality of the main approach is also realized in the matching approach, e.g. some of the
location indicators described above are also exploited in GIR-InSicht (adjectives; demonyms for regions
and countries). They are not normalized, but the query semantic network is extended by many alternative
semantic networks that are in part derived by symbolic inference rules using the semantic knowledge about
location indicators. In contrast, the main approach exploits this information on the level of terms.
2.4</p>
      </sec>
      <sec id="sec-2-4">
        <title>Related Work</title>
        <p>
          Nagel [
          <xref ref-type="bibr" rid="ref11">11</xref>
          ] describes the manual construction of a place name ontology containing 17,000 geographic
entities as a prerequisite for analyzing German sentences. He states that in German, toponyms have a
simple inflectional morphology, but a complex (idiosyncratic) derivational morphology.
        </p>
        <p>
          Buscaldi, Rosso et al. [
          <xref ref-type="bibr" rid="ref1">1</xref>
          ] investigate semi-automatic creation of a geographical ontology, using gazetteer
data and resources like Wikipedia and WordNet.
        </p>
        <p>
          Wang et al. [
          <xref ref-type="bibr" rid="ref12">12</xref>
          ] introduce the concept of dominant locations (later called implicit locations, [
          <xref ref-type="bibr" rid="ref10">10</xref>
          ]).
Implicit locations are locations not explicitly mentioned in a text. The only case explored are locations that
are closely related to other locations.
        </p>
        <p>
          Previous work on GIR by members of the IICS includes experiments with documents and queries
represented as semantic networks [
          <xref ref-type="bibr" rid="ref8">8</xref>
          ], and experiments dealing with linguistic phenomena, such as cases of
regular metonymy of location names [
          <xref ref-type="bibr" rid="ref7">7</xref>
          ], which was utilized as a means to increase precision in GIR. Due
2But documents can also be found if the information is distributed across several sentences because a coreference resolver
processed all document representations.
to time constraints, metonymy recognition was not included in GIRSA. For the experiments for GeoCLEF
2007, we focused on investigating means to increase recall.
3
        </p>
      </sec>
    </sec>
    <sec id="sec-3">
      <title>Experimental Setup</title>
      <p>
        The GeoCLEF 2007 documents constitute a corpus of more than 275,000 German newspaper articles from
‘Frankfurter Rundschau’, ‘Schweizerische Depeschenagentur’, and ‘Der Spiegel’ from the years 1994 and
1995 (see [
        <xref ref-type="bibr" rid="ref3">3</xref>
        ]). The performance of GIRSA is evaluated on the test set from GeoCLEF 2007, containing
25 topics with a title, a short description, and a narrative part. As in a setup for previous GIR experiments
on GeoCLEF data [
        <xref ref-type="bibr" rid="ref8 ref9">8, 9</xref>
        ], the documents were indexed with the Zebra database management system [
        <xref ref-type="bibr" rid="ref4">4</xref>
        ],
which supports a standard relevance ranking (tf-idf IR model).
      </p>
      <p>
        Documents are preprocessed as follows to produce different indexes:
1. S: As in traditional IR, all words in the document text (including location names) are stemmed, using
an implementation of the German snowball stemmer.
2. SL: Location indicators are identified and normalized to a base form of a location name.
3. SLD: In addition, decompounding is applied to the words in the text. German decompounding
follows the frequency-based approach described in [
        <xref ref-type="bibr" rid="ref2">2</xref>
        ].
4. O: Documents and queries are represented as semantic networks and GIR is seen as a form of
question answering (see Sect. 2.3).
      </p>
      <p>The following location indicators were normalized in documents and queries for the GIR experiments:
adjectives corresponding to locations, demonyms, abbreviations, orthographic variants, language names,
and location names. Normalization consists of applying a set of transformation rules (covering regular
variations) and looking up locations in specialized exception lists for each type of location indicator.</p>
      <p>Basically, queries and documents are processed in the same way. The title and short description were
used for creating a query. GeoCLEF topics contain a narrative part describing documents which are to
be assessed as relevant. Instead of employing a large gazetteer containing location names as a knowledge
base for query expansion, additional location names were extracted from the narrative part of the topic. No
meronymy information is utilized for direct query expansion (because these may be just the terms the blind
feedback may find and there would be a combined effect).</p>
      <p>For the bilingual (English-German) experiments, the queries were translated using the Promt web
service for machine translation.3 Query processing then follows the setup for monolingual German
experiments.</p>
      <p>The following parameter settings were used in different retrieval experiments:
1. query language: German (DE) or English (EN);
2. index type: stemming only (S), identification of locations, not stemmed (SL), decomposition of</p>
      <p>German compounds (SLD), hybrid (SLD/O), and based on semantic networks (O);
3. query fields: combinations of title (T), description (D), and locations from narrative (N).</p>
      <p>Parameters and results for monolingual German and bilingual English-German experiments are shown
in Table 1. The table shows relevant and retrieved documents (rel ret), MAP and precision at five, ten,
and twenty documents. In total, 904 documents were assessed as relevant for the 25 topics. For the
run FUHtd6de, results from GIR-InSicht were merged with results from the experiment FUHtd3de in a
straightforward way, using the maximum score.
index
S
SL
SLD
SL
SLD
SLD/O
O
S
SL
SLD
SL
SLD
fields
TD
TD
TD
TDN
TDN
TD
TD
TD
TD
TD
TDN
TDN
rel ret
Identifying and indexing normalized location indicators, decompounding, and adding location names from
the narrative part improves performance considerably, i.e. 120 additional relevant documents are found and
MAP is increased from 0.119 (FUHtd1de) to 0.258 (FUHtdn5de) in comparison to the baseline experiment.</p>
      <p>Decompounding German nouns seems to have different effects on precision and recall (FUHtd2de
vs. FUHtd3de and FUHtdn4de vs. FUHtdn5de): while more relevant documents are retrieved without
decompounding, initial precision is higher when utilizing decompounding.</p>
      <p>Topic 10.2452/55-GC contains a negation in the topic title and description ( “but not in the Alps”).
However, adding the location names from the narrative part of the topic (“Scotland, Norway, Iceland”) did
not notably improve precision for this topic (0.005 MAP in FUGtd3de vs. 0.013 MAP in FUHtdn5de).</p>
      <p>A small analysis of results found by GIR-InSicht in comparison with the main GIR system reveals that
for ten topics, GIR-InSicht retrieved documents, and for seven topics, it returned relevant documents (see
Fig. 1 and Table 2). This approach, originating from question answering and based on a strict matching
of semantic representations, returns three additional relevant documents for the combination (FUHtd6de).
However, the MAP for some topics in the combined run indicates that merging by taking the maximum
of two scores might be too simple. For a single topic (10.2452/52-GC), zero relevant documents were
retrieved in all experiments.</p>
      <p>Results for the bilingual (English-German) experiments are generally lower. As for German, all other
experiments outperform the baseline (0.114 MAP). The best performance is achieved by an experiment
using topic title, description, and location names from the narrative (0.209 MAP). In comparison with results
for the monolingual German experiments, the performance drop lies between 4.2% (first experiment) and
27.1% (fifth experiment).
5</p>
    </sec>
    <sec id="sec-4">
      <title>Conclusion and Outlook</title>
      <p>In this paper location indicators were introduced as text segments from which location names can be
inferred. For the GeoCLEF 2007 experiments, different indexes containing stemmed words and location
indicators normalized to location names were created. Results of the GIR experiments show that MAP
is higher when using location indicators instead of location names to represent the geographic scope of a
document. A broader approach to identify the geographic scope of a document is needed because proper
nouns or location names do not alone imply the geographic scope of a document.</p>
      <p>In addition, we investigated using location names extracted from the narrative part of a topic (instead
of looking up additional location names in large gazetteers). The narrative contains a detailed description
about which documents are to be assessed as relevant (and which not), including additional location names.
Adding these location names to the query notably improves performance. This result is seemingly in
contrast to some results from GeoCLEF 2006, were it was found that additional query terms (from gazetteers)
degrade performance. A possible explanation is that in this experiment, only a few location names were
added (3.16 location names on average for fifteen of the 25 topics with a maximum of thirteen additional
location names). When using a gazetteer, one has to decide which terms are the most useful in query
expansion. If this decision is based on the importance of a location, a semantic shift in the results may occur,
which degrades performance. In contrast, selecting terms from the narrative part increases the chance to
expand a query with relevant terms only.</p>
      <p>The hybrid approach for GIR proved interesting, and even a few additional relevant documents were
found in the combined run. As GIR-InSicht originates from a deep (read: semantic) QA approach, it returns
documents with a high initial precision, which may prove useful in combination with a geographic blind
feedback strategy. GIR-InSicht performs worse than the IR baseline, because only 102 documents were
retrieved for ten of the 25 topics. However, more than half (56 documents) turned out to be relevant.</p>
      <p>Several improvements are planned for GIRSA. These include using estimates for the importance (weight)
of different location indicators, possibly depending on the context (e.g. “Danish coast” →“Denmark”,
but “German shepherd” 6→ “Germany”), and using a part-of-speech tagger and named entity recognizer
to identify location names. Finally, we plan to investigate the combination of means to increase
precision (e.g. recognizing metonymic location names) with means to increase recall (e.g. recognizing and
normalizing location indicators).</p>
    </sec>
  </body>
  <back>
    <ref-list>
      <ref id="ref1">
        <mixed-citation>
          [1]
          <string-name>
            <given-names>Davide</given-names>
            <surname>Buscaldi</surname>
          </string-name>
          , Paolo Rosso, and Piedachu Peris Garcia.
          <article-title>Inferring geographical ontologies from multiple resources for geographical information retrieval</article-title>
          .
          <source>In Proceedings of the 3rd Workshop on Geographical Information Retrieval (GIR</source>
          <year>2006</year>
          ), pages
          <fpage>52</fpage>
          -
          <lpage>55</lpage>
          , Seattle, USA,
          <year>2006</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref2">
        <mixed-citation>
          [2]
          <string-name>
            <given-names>Aitao</given-names>
            <surname>Chen</surname>
          </string-name>
          .
          <article-title>Cross-language retrieval experiments at CLEF 2002</article-title>
          . In Carol Peters, Martin Braschler, Julio Gonzalo, and Michael Kluck, editors,
          <source>Advances in Cross-Language Information Retrieval, Third Workshop of the Cross-Language Evaluation Forum</source>
          ,
          <string-name>
            <surname>CLEF</surname>
          </string-name>
          <year>2002</year>
          , volume
          <volume>2785</volume>
          <source>of LNCS</source>
          , pages
          <fpage>28</fpage>
          -
          <lpage>48</lpage>
          , Berlin,
          <year>2002</year>
          . Springer.
        </mixed-citation>
      </ref>
      <ref id="ref3">
        <mixed-citation>
          [3]
          <string-name>
            <given-names>Fredric</given-names>
            <surname>Gey</surname>
          </string-name>
          , Ray Larson, Mark Sanderson, Hideo Joho, Paul Clough, and Vivien Petras.
          <article-title>GeoCLEF: the CLEF 2005 cross-language geographic information retrieval track overview</article-title>
          . In Carol Peters, Fredric C. Gey, Julio Gonzalo, Henning Mu¨ller,
          <string-name>
            <surname>Gareth</surname>
            <given-names>J. F.</given-names>
          </string-name>
          <string-name>
            <surname>Jones</surname>
          </string-name>
          , Michael Kluck, Bernardo Magnini, and Maarten de Rijke, editors,
          <source>Accessing Multilingual Information Repositories, 6th Workshop of the Cross-Language Evaluation Forum</source>
          ,
          <string-name>
            <surname>CLEF</surname>
          </string-name>
          <year>2005</year>
          , volume
          <volume>4022</volume>
          <source>of LNCS</source>
          , pages
          <fpage>908</fpage>
          -
          <lpage>919</lpage>
          . Springer, Berlin,
          <year>2006</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref4">
        <mixed-citation>
          [4]
          <string-name>
            <given-names>Sebastian</given-names>
            <surname>Hammer</surname>
          </string-name>
          , Adam Dickmeiss, Heikki Levanto, and
          <string-name>
            <given-names>Mike</given-names>
            <surname>Taylor</surname>
          </string-name>
          .
          <article-title>Zebra - User's Guide and Reference</article-title>
          . Copenhagen, Denmark,
          <year>2005</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref5">
        <mixed-citation>
          [5]
          <string-name>
            <given-names>Sven</given-names>
            <surname>Hartrumpf</surname>
          </string-name>
          and
          <string-name>
            <given-names>Johannes</given-names>
            <surname>Leveling</surname>
          </string-name>
          . University of Hagen at QA@
          <article-title>CLEF 2006: Interpretation and normalization of temporal expressions</article-title>
          . In Alessandro Nardi, Carol Peters, and Jose´ Luis Vicedo, editors,
          <source>Results of the CLEF 2006 Cross-Language System Evaluation Campaign, Working Notes for the CLEF 2006 Workshop</source>
          , Alicante, Spain,
          <year>2006</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref6">
        <mixed-citation>
          [6]
          <string-name>
            <given-names>Hermann</given-names>
            <surname>Helbig</surname>
          </string-name>
          .
          <source>Knowledge Representation and the Semantics of Natural Language</source>
          . Springer, Berlin,
          <year>2006</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref7">
        <mixed-citation>
          [7]
          <string-name>
            <given-names>Johannes</given-names>
            <surname>Leveling</surname>
          </string-name>
          and
          <string-name>
            <given-names>Sven</given-names>
            <surname>Hartrumpf</surname>
          </string-name>
          .
          <article-title>On metonymy recognition for GIR</article-title>
          .
          <source>In Proceedings of the 3rd Workshop on Geographical Information Retrieval (GIR</source>
          <year>2006</year>
          ), pages
          <fpage>9</fpage>
          -
          <lpage>13</lpage>
          , Seattle, USA,
          <year>2006</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref8">
        <mixed-citation>
          [8]
          <string-name>
            <given-names>Johannes</given-names>
            <surname>Leveling</surname>
          </string-name>
          , Sven Hartrumpf, and
          <string-name>
            <given-names>Dirk</given-names>
            <surname>Veiel</surname>
          </string-name>
          .
          <article-title>Using semantic networks for geographic information retrieval</article-title>
          . In Carol Peters, Fredric C. Gey, Julio Gonzalo,
          <string-name>
            <given-names>Gareth J. F.</given-names>
            <surname>Jones</surname>
          </string-name>
          , Michael Kluck, Bernardo Magnini, Henning Mu¨ller, and Maarten de Rijke, editors,
          <source>Accessing Multilingual Information Repositories, 6th Workshop of the Cross-Language Evaluation Forum</source>
          ,
          <string-name>
            <surname>CLEF</surname>
          </string-name>
          <year>2005</year>
          , volume
          <volume>4022</volume>
          <source>of LNCS</source>
          , pages
          <fpage>977</fpage>
          -
          <lpage>986</lpage>
          . Springer, Berlin,
          <year>2006</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref9">
        <mixed-citation>
          [9]
          <string-name>
            <given-names>Johannes</given-names>
            <surname>Leveling</surname>
          </string-name>
          and
          <string-name>
            <given-names>Dirk</given-names>
            <surname>Veiel</surname>
          </string-name>
          . University of Hagen at GeoCLEF 2006:
          <article-title>Experiments with metonymy recognition in documents</article-title>
          . In Alessandro Nardi, Carol Peters, and Jose´ Luis Vicedo, editors,
          <source>Results of the CLEF 2006 Cross-Language System Evaluation Campaign, Working Notes for the CLEF 2006 Workshop</source>
          , Alicante, Spain,
          <year>2006</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref10">
        <mixed-citation>
          [10]
          <string-name>
            <given-names>Zhisheng</given-names>
            <surname>Li</surname>
          </string-name>
          ,
          <string-name>
            <given-names>Chong</given-names>
            <surname>Wang</surname>
          </string-name>
          , Xing Xie,
          <string-name>
            <given-names>Xufa</given-names>
            <surname>Wang</surname>
          </string-name>
          , and
          <string-name>
            <surname>Wei-Ying Ma</surname>
          </string-name>
          .
          <article-title>Indexing implicit locations for geographical information retrieval</article-title>
          .
          <source>In Proceedings of the 3rd Workshop on Geographical Information Retrieval (GIR</source>
          <year>2006</year>
          ), pages
          <fpage>68</fpage>
          -
          <lpage>70</lpage>
          , Seattle, USA,
          <year>2006</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref11">
        <mixed-citation>
          [11]
          <string-name>
            <given-names>Sebastian</given-names>
            <surname>Nagel</surname>
          </string-name>
          .
          <article-title>An ontology of German place names</article-title>
          . Corela - Cognition, Repre´sentation, Langage,
          <article-title>Le traitement lexicographique des noms propres (Nume´ros spe´ciaux</article-title>
          ),
          <year>2005</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref12">
        <mixed-citation>
          [12]
          <string-name>
            <surname>Lee</surname>
            <given-names>Wang</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Chuang</surname>
            <given-names>Wang</given-names>
          </string-name>
          , Xing Xie, Josh Forman, Yansheng Lu,
          <string-name>
            <surname>Wei-Ying Ma</surname>
            , and
            <given-names>Ying</given-names>
          </string-name>
          <string-name>
            <surname>Li</surname>
          </string-name>
          .
          <article-title>Detecting dominant locations from search queries</article-title>
          .
          <source>In Proceedings of the 28th annual international ACM SIGIR conference on research and development in information retrieval (SIGIR '05)</source>
          , pages
          <fpage>424</fpage>
          -
          <lpage>431</lpage>
          , New York, USA,
          <year>2005</year>
          . ACM Press.
        </mixed-citation>
      </ref>
    </ref-list>
  </back>
</article>