<!DOCTYPE article PUBLIC "-//NLM//DTD JATS (Z39.96) Journal Archiving and Interchange DTD v1.0 20120330//EN" "JATS-archivearticle1.dtd">
<article xmlns:xlink="http://www.w3.org/1999/xlink">
  <front>
    <journal-meta />
    <article-meta>
      <title-group>
        <article-title>LCC's PowerAnswer at QA@CLEF 2006</article-title>
      </title-group>
      <contrib-group>
        <contrib contrib-type="author">
          <string-name>Mitchell Bowden</string-name>
          <xref ref-type="aff" rid="aff0">0</xref>
        </contrib>
        <contrib contrib-type="author">
          <string-name>Marian Olteanu</string-name>
          <xref ref-type="aff" rid="aff0">0</xref>
        </contrib>
        <contrib contrib-type="author">
          <string-name>Pasin Suriyentrakorn</string-name>
          <xref ref-type="aff" rid="aff0">0</xref>
        </contrib>
        <contrib contrib-type="author">
          <string-name>Jonathan Clark</string-name>
          <xref ref-type="aff" rid="aff0">0</xref>
        </contrib>
        <contrib contrib-type="author">
          <string-name>Texas</string-name>
          <xref ref-type="aff" rid="aff0">0</xref>
        </contrib>
        <contrib contrib-type="author">
          <string-name>United States of America</string-name>
          <xref ref-type="aff" rid="aff0">0</xref>
        </contrib>
        <contrib contrib-type="author">
          <string-name>General Terms</string-name>
          <xref ref-type="aff" rid="aff0">0</xref>
        </contrib>
        <aff id="aff0">
          <label>0</label>
          <institution>Measurement</institution>
          ,
          <addr-line>Performance, Experimentation</addr-line>
        </aff>
      </contrib-group>
      <pub-date>
        <year>2006</year>
      </pub-date>
      <abstract>
        <p>This paper reports on Language Computer Corporation's rst QA@CLEF participation. For this exercise, we integrated our open-domain PowerAnswer question answering system with our statistical machine translation engine. For 2006, we participated in the English-to-Spanish, French and Portuguese cross-language tasks. We took the approach of intermediate translation, only processing English within the QA system regardless of the input or source languages. The output snippets were then mapped back into the source language documents for the nal output of the system and submission. What follows is a description of our system and methodology.</p>
      </abstract>
      <kwd-group>
        <kwd>Open-domain Question Answering</kwd>
        <kwd>Questions beyond factoids</kwd>
        <kwd>Statistical machine translation</kwd>
      </kwd-group>
    </article-meta>
  </front>
  <body>
    <sec id="sec-1">
      <title>Introduction</title>
      <p>For 2006, Language Computer's open-domain question answering system PowerAnswer [5]
participated in QA@CLEF for the rst time. PowerAnswer has previously participated in many
other evaluations, notably TREC [1], however, this is the rst Multilingual QA evaluation the
system has entered. We have developed our own statistical machine translation system, which
we integrated with PowerAnswer for this evaluation. Since PowerAnswer is a very modular and
extensible system, we were able to make a minimum of modi cations for this integration for our
initial approach.</p>
      <p>
        Our goals for this year's participation were (
        <xref ref-type="bibr" rid="ref1">1</xref>
        ) to examine how well the current QA system
performs when given noisy data, such as that from automatic translation and (
        <xref ref-type="bibr" rid="ref2">2</xref>
        ) to examine the
performance of the machine translation system in a question answering environment. To that end,
we adopted an approach of intermediate translation instead of adapting the QA system to process
target languages natively.
      </p>
      <p>The paper presents a summary of the PowerAnswer system, our machine translation engine,
the integration of the two for QA@CLEF 2006, and then follows with a discussion of our results
and challenges in this year's CLEF question topics. Table 1 lists the cross-lingual tasks in which
we participated.</p>
      <sec id="sec-1-1">
        <title>Source</title>
        <sec id="sec-1-1-1">
          <title>English</title>
        </sec>
        <sec id="sec-1-1-2">
          <title>English</title>
        </sec>
        <sec id="sec-1-1-3">
          <title>English</title>
        </sec>
      </sec>
      <sec id="sec-1-2">
        <title>Target</title>
        <sec id="sec-1-2-1">
          <title>French</title>
        </sec>
        <sec id="sec-1-2-2">
          <title>Spanish</title>
        </sec>
        <sec id="sec-1-2-3">
          <title>Portuguese</title>
          <p>2</p>
        </sec>
      </sec>
    </sec>
    <sec id="sec-2">
      <title>Overview of LCC's PowerAnswer</title>
      <p>Automatic question answering requires a system that has a wide range of tools available. There is
no one monolithic solution for all question types or even data sources. In realization of this, LCC
developed PowerAnswer 2 as a fully-modular and distributed multi-strategy question answering
system that integrates semantic relations, advanced inferencing abilities, syntactically constrained
lexical chains, and temporal contexts. This section presents an outline of the system and how it
was modi ed to meet the challenges of QA@CLEF 2006.</p>
      <p>PowerAnswer comprises a set of strategies that are selected based on advanced question
processing, and each strategy is developed to solve a speci c class of questions either independently
or together. A Strategy Selection module automatically analyzes the question and chooses a set
of strategies with the algorithms and tools that are tailored to the class of the given question.
PowerAnswer can distribute the strategies across workers in the case of multiple strategies being
selected, alleviating the increase in the complexity of the question answering process by splitting
the workload across machines and processors.</p>
      <p>Syntactic
Parsing</p>
      <p>Named Entity
Recognition</p>
      <p>Reference</p>
      <p>Resolution
Question</p>
      <p>Internet</p>
      <p>Question
Processing</p>
      <p>Module
(QP)</p>
      <p>Passage
Retrieval
Module</p>
      <p>(PR)
Web−Boosting</p>
      <p>Strategy</p>
      <p>Answer
Processing
Module
(AP)</p>
      <p>Question Logic
Transformation
Answer Logic
Transformation</p>
      <p>QLF
COGEX</p>
      <p>ALF
World Kowledge</p>
      <p>Axioms
Documents</p>
      <p>Answer
Selection</p>
      <p>(AS)
Answer</p>
      <p>
        Each strategy is a collection of components, (
        <xref ref-type="bibr" rid="ref1">1</xref>
        ) Question Processing, (
        <xref ref-type="bibr" rid="ref2">2</xref>
        ) Passage Retrieval,
and (
        <xref ref-type="bibr" rid="ref3">3</xref>
        ) Answer Processing. Each of these components constitute one or more modules, which
interface to a library of generic NLP tools. These NLP tools are the building blocks of the
PowerAnswer 2 system that, through a well-de ned set of interfaces, allow for rapid integration
and testing of new tools and third-party software such as IR systems, syntactic parsers, named
entity recognizers, logic provers, semantic parsers, ontologies, word sense disambiguation modules,
and more. Furthermore, the components that make up each strategy can be interchanged to quickly
create new strategies, if needed, they can also be distributed [10].
      </p>
      <p>
        As illustrated in Figure 1, the role of the QP module is to determine (
        <xref ref-type="bibr" rid="ref1">1</xref>
        ) the expected answer
type, (
        <xref ref-type="bibr" rid="ref2">2</xref>
        ) to select the keywords used in retrieving relevant passages, and (
        <xref ref-type="bibr" rid="ref3">3</xref>
        ) perform any
preliminary questions as necessary for resolving question ambiguity. The PR module ranks passages that
are retrieved by the IR system, while the AP module extracts and scores the candidate answers.
All modules have access to a syntactic parser, a named entity recognizer and a reference
resolution system through LCC's generic NLP tool libraries. To improve the answer selection, we take
advantage of redundancy in large corpora, speci cally in this case, the Internet. As the size of
a document collection grows, a question answering system is more likely to pinpoint a candidate
answer that closely resembles the surface structure of the question. These features have the role
of correcting the errors in answer processing that are produced by the selection of keywords, by
syntactic and semantic processing and by the absence of pragmatic information. Usually, the nal
decision for selecting answers is based on logical proofs from our inference engine COGEX. For this
year's QA@CLEF, however, we disabled the logic prover in order to better evaluate the individual
components of this QA architecture. COGEX's evaluation on multilingual data was performed in
the CLEF Answer Validation Exercise [13].
3
      </p>
    </sec>
    <sec id="sec-3">
      <title>Overview of Translation Engine</title>
      <p>The translation system used at LCC implements phrase-based statistical machine translation [2],
the core translation engine is the open-source Phramer [12] system, developed by one of LCC's
engineers. Phramer in turn implements and extends the phrase-based machine translation
algorithms implemented by Pharaoh [4]. A more detailed description of the MT solution that we
adopted for Multilingual QA@CLEF can be found in [11]. We trained the translation system
using the European Parliament Proceedings Parallel Corpus 1996-2003 (EUROPARL) [3], which
provides between 600k and 800k pairs of sentences (sentences in English paired with the
translation in another European language). We followed the training procedure described in the Pharaoh
training manual1 to generate the phrase table required for translation.</p>
      <p>
        In order to translate entire documents, we augmented the core translation engine with (
        <xref ref-type="bibr" rid="ref1">1</xref>
        )
tokenization, (
        <xref ref-type="bibr" rid="ref2">2</xref>
        ) capitalization, and (
        <xref ref-type="bibr" rid="ref3">3</xref>
        ) de-tokenization.
      </p>
      <p>The tokenization process was performed on the original documents (in French, Portuguese or
Spanish), in order to convert the sentences to space-separated entities, in which the punctuation
and the words are isolated. The step was required because the statistical machine translation core
engine accepts only lowercased tokenized input.</p>
      <p>The capitalization process follows the translation process and it restores the casing of the
words. The capitalization tool uses three-gram statistics extracted from 150 million words from
the English GigaWord Second Edition2 corpus, augmented with two heuristics:</p>
      <sec id="sec-3-1">
        <title>1. rst word will always be uppercased</title>
      </sec>
      <sec id="sec-3-2">
        <title>2. if the words appear also in the foreign documents, the casing is preserved (this rule is very e ective for proper nouns and named entities)</title>
        <p>4</p>
      </sec>
    </sec>
    <sec id="sec-4">
      <title>PowerAnswer-Phramer Integration</title>
      <p>Our cross-language solution for Question Answering was based on automatic translation of the
documents in the source language (English). QA was performed on a collection consisting only
of English documents. The answers were converted back into the target language (the original
language of the documents) by aligning the translation with the original document (checking to
see what was the original phrase in the original document that generated the answer in English);
when this method failed, the system falls back to machine translation (source ! target).
1http://www.iccs.inf.ed.ac.uk/ pkoehn/training.tgz
2http://www.ldc.upenn.edu/Catalog/CatalogEntry.jsp?catalogId=LDC2005T12
Passage Retrieval
Making use of PowerAnswer's modular design, we developed three di erent retrieval methods,
settling on the rst for our nal experiment.</p>
      <sec id="sec-4-1">
        <title>1. use an index of English words, created from the translated documents</title>
      </sec>
      <sec id="sec-4-2">
        <title>2. use an index of foreign words (French, Spanish or Portuguese), created from the original documents</title>
      </sec>
      <sec id="sec-4-3">
        <title>3. use an index of English words, created from the original documents in correlation with the translation table</title>
        <p>The rst solution is the default solution. The entire target language document collection is
translated into English, processed through the set of NLP tools and indexed for querying. Its major
disadvantage is the computational e ort required to translate the entire collection. It also requires
updating the English version of the collection when one improves the quality of the translation. Its
major advantage is that there are no additional costs during question answering (the documents
are already translated). This passage retrieval method is illustrated in Figure 2.</p>
        <p>EN Questions</p>
        <p>EN query
EN passages</p>
        <p>PowerAnswer
EN</p>
        <p>TL</p>
        <p>EN Answers</p>
        <p>Answer Aligner
Target Language</p>
        <p>Passages</p>
        <p>Target Language Answers</p>
        <p>The second solution, as seen in Figure 3, requires minimum e ort during indexing (the
document collection is indexed in its native language). In order to retrieve the relevant documents,
we translate the keywords of the IR query (the query submitted by PowerAnswer to the
Lucenebased 3 IR system) with alternations as the new IR query (step 1). The translation of keywords is
performed using Phramer, by generating n-best translations. This translated query is submitted
to the target language index (step 2). The documents retrieved by this query are then dynamically
translated into English using Phramer (step 3). We use a cache to store translated documents so
that IR query reformulations and other questions that might retrieve the same documents will
not need to be translated again. The set of translated documents is indexed into a mini-collection
(step 4) and the mini-collection is re-queried using the original English-based IR query (step 5).
(s0)
EN Questions
(s1)
Keyword translation
(s6)
EN Answers
(s7)
(s8)</p>
        <p>Target Language Answers
Answer Aligner</p>
        <p>EN query (s5)</p>
        <p>EN passages
PowerAnswer
Phramer</p>
        <p>(s2)
Target language query
(s3)
EN mini−index creation
(s4)</p>
        <p>TL</p>
        <p>EN</p>
        <p>For example, the boolean IR query in English (\poem" AND \love" AND \1922") is translated
into French as (\poeme" AND (\aiment" OR \aimer" OR \aimez" OR \amour") AND \1922")
with the alternations. This new query will return 85 French documents. Some of them do not
contain \love" in their automatic translation (but the original document contains \aiment", \aimer",
\aimez" or \amour"). Thus, by re-querying the translated sub-collection (that contains only the
translation of those 85 documents) we retrieve only 72 English documents that will be passed to
PowerAnswer.</p>
        <p>The advantage of the second method is that minimum e ort is required during collection
preparation. Also, the collection preparation might not be under the control of the QA system
(i.e. it can be web-based). Also, improvements in the MT engine can be re ected immediately in
the output of the integrated system. The disadvantage is that more computation is required at
run-time for translating the IR query and the documents dynamically.</p>
        <p>The third alternative extracts during indexing the English words that might be part of the
translation and indexes the collection accordingly. The process doesn't involve lexical choice - all
choices are considered possible. The set of keywords is determined using the translation table,
and collects all words that are part of the translation lattice ([4]). Determining only the words
according to the translation table (semi-translation) is approximately 10 times faster than the
full translation. The index is queried using the original IR query generated by PowerAnswer
(with English keywords). After the initial retrieval, the algorithm is similar to the second method:
translate the retrieved documents, re-query the mini-collection. The advantage is the much smaller
indexing time when compared with the rst method, besides all the advantages of the second
method. Also, it has all the disadvantages of the second method, except that it doesn't require
IR query translation.</p>
        <p>Because preliminary testing proved that there aren't signi cant di erences in recall between
the three methods and because the rst method is fastest after the document collection is prepared,
we used only the rst method for the nal evaluation.</p>
        <p>Answer Processing
For each of the above methods, PowerAnswer returns the exact answer and the supporting sentence
answer the exact was extracted from (all in English). Then, these answers are then aligned to
the corresponding text in the target language documents. The nal output of the system is
the converted responses in the target language with the appropriate supporting snippet. If the
alignment method fails, the English answer is converted into the target language as the nal
response.
5</p>
      </sec>
    </sec>
    <sec id="sec-5">
      <title>Results</title>
      <p>Our integrated multilingual PowerAnswer system was tested on 190 English ! Spanish, 190
English ! French and 188 English ! Portuguese factoid and de nition questions and 10 English
! Spanish, 10 English ! French and 12 English ! Portuguese list questions. For QA@CLEF,
the main score is the overall accuracy, the average of SCORE(q), where SCORE(q) is de ned for
factoids and de nition questions as 1 if the top answer for q is assessed as correct, 0 otherwise.
Also included are the Mean Reciprocal Rank (MRR) and the Con dence Weighted Score (CWS)
that judges how well a system returns correct answers higher in the ranked list of answers.</p>
      <p>Table 2 illustrates the nal results of Language Computer's e orts in our rst participation at
QA@CLEF for 2006.</p>
      <sec id="sec-5-1">
        <title>Source</title>
        <sec id="sec-5-1-1">
          <title>Spanish</title>
        </sec>
        <sec id="sec-5-1-2">
          <title>French</title>
        </sec>
        <sec id="sec-5-1-3">
          <title>Portuguese</title>
        </sec>
      </sec>
      <sec id="sec-5-2">
        <title>Accuracy</title>
      </sec>
    </sec>
    <sec id="sec-6">
      <title>Error Analysis and Challenges in 2006</title>
      <p>There were several sources of errors in LCC's submission, including one major source that accounts
for a 50% loss in accuracy. The sources of errors include: translation misalignments, tokenization
errors, and data processing errors - questions and passages.</p>
      <p>Translation misalignments
Because the version of PowerAnswer used this year is monolingual, the design we used for multilingual
question answering involved translating documents dynamically for processing through the QA
system and mapping the responses back into the source language documents. This resulted in many
places where errors could occur. While the translation of the documents into English did
introduce noise into the data such as mistranslations, words that were not translated and should have
been or words that should not have been translated and were, it did not a ect the QA system as
much as we suspected. By far the greatest source of errors was the alignment between the English
answers and the source documents which produced the nal response list. This source accounts for
roughly a 50% loss of accuracy for the tasks we participated in, as Table 3 shows. For Portuguese,
there was also an error in the submission that accounts for the great di erence in accuracy when
compared to the Spanish and French results.</p>
      <sec id="sec-6-1">
        <title>Source</title>
        <sec id="sec-6-1-1">
          <title>Spanish</title>
        </sec>
        <sec id="sec-6-1-2">
          <title>French</title>
        </sec>
        <sec id="sec-6-1-3">
          <title>Portuguese</title>
        </sec>
      </sec>
      <sec id="sec-6-2">
        <title>Position 1 Acc. Top 5 Acc.</title>
        <p>Submission Acc.
Data errors
Questions that contained common nouns in the source language question were one example of data
processing errors. Question 83 in the EN-PT task is Who is the director of the lm \Caro diario"?.
Here, the noun \Caro" was translated \expensive" when being indexed in English. Hence, the
query keyword \Caro" as it is in the question could not be found in the translated collection by
the IR system, causing the passage to receive a lowered score when retrieved on less keywords.</p>
        <p>Sometimes the wording of the questions also a ected the system in a negative way. There
were some spelling changes that were unrecoverable by our system, such \Huan Karlos" for Juan
Carlos (EN-PT #5). There were also a few instances of keywords for which PowerAnswer was
unable to generate the correct alternation, for example \celebrated" in EN-ES #75 In which year
was the Football World Cup celebrated in the United States?, where the correct alternation would
have been a synonym for \to host".</p>
        <p>The scoring of de nition questions tends to be subjective, so there were cases where we believe
an answer returned warranted at worst an inexact but was judged wrong, such as EN-PT #84 What
is the Unhcr?. PowerAnswer responded with \O Acnur (Alto Comissariado das Naco~es Unidas
para Refugiados) consultou o Brasil e mais 30 pa ses sobre a possibilidade de acolher um grupo
de 5.000 refugiados da ex-Iugoslavia", which was judged as wrong. Additionally, the de nition
question strategy often returned answer snippets that were a full sentence and were judged as
inexact because of their length. One example is EN-PT question 34 Who was Alexander Graham
Bell?. PowerAnswer returns the full sentence containing the important nugget \A empresa foi
fundada em 1885 e entre os socios estava Alexander Graham Bell, o inventor do telefone", where
the nal part \inventor of the telephone" was all that was necessary.
7</p>
      </sec>
    </sec>
    <sec id="sec-7">
      <title>Conclusions</title>
      <p>While this year's performance was not what we hoped due to some major errors in the nal stages
of processing answers, we look forward to better results in the following years. One large step
we will be taking is to make the core of PowerAnswer more language-independent. The English
dependency comes from the NLP tools more than the modules of PowerAnswer, where the language
dependence occurs primarily in Question Processing. By developing a set of multilingual NLP
tools, including syntactic parsers and semantic relations extractors, and abstracting the
Englishdependent components of PowerAnswer out and building language-independent as well as some
language-speci c question processing, we hope to see great improvement in our multilingual QA
capabilities. This work will be eased by the exible design of PowerAnswer and our libraries
of generic NLP tool interfaces. For next year's CLEF, we plan on submitting results from an
improved and corrected system using the same method we have discussed in this paper, as well
as from a more language-independent PowerAnswer that can process the source language text
natively without intermediate translation.</p>
    </sec>
  </body>
  <back>
    <ref-list>
      <ref id="ref1">
        <mixed-citation>
          [1]
          <string-name>
            <given-names>Sanda</given-names>
            <surname>Harabagiu</surname>
          </string-name>
          , Dan Moldovan, Christine Clark, Mitchell Bowden, Andy Hickl,
          <string-name>
            <given-names>Patrick</given-names>
            <surname>Wang</surname>
          </string-name>
          .
          <source>Employing Two Question Answering Systems in TREC-2005. In Text REtrieval Conference</source>
          ,
          <year>2005</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref2">
        <mixed-citation>
          [2]
          <string-name>
            <given-names>Philipp</given-names>
            <surname>Koehn</surname>
          </string-name>
          , Franz Josef Och and
          <string-name>
            <given-names>Daniel</given-names>
            <surname>Marcu</surname>
          </string-name>
          .
          <article-title>Statistical phrase-based translation</article-title>
          .
          <source>Proceedings of HLT/NAACL</source>
          2003 Edmonton, Canada,
          <year>2003</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref3">
        <mixed-citation>
          [3]
          <string-name>
            <given-names>Philipp</given-names>
            <surname>Koehn</surname>
          </string-name>
          .
          <article-title>Europarl: A Parallel Corpus for Statistical Machine Translation</article-title>
          .
          <source>MT Summit</source>
          <year>2005</year>
          ,
          <year>2005</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref4">
        <mixed-citation>
          [4]
          <string-name>
            <given-names>Philipp</given-names>
            <surname>Koehn</surname>
          </string-name>
          .
          <article-title>Pharaoh: A beam search decoder for phrase-based statistical machine translation models</article-title>
          .
          <source>In Proceedings of AMTA</source>
          <year>2004</year>
          ,
          <year>2004</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref5">
        <mixed-citation>
          [5]
          <string-name>
            <given-names>Dan</given-names>
            <surname>Moldovan</surname>
          </string-name>
          , Sanda Harabagiu, Christine Clark and Mitchell Bowden.
          <article-title>PowerAnswer 2: Experiments and Analysis over TREC 2004</article-title>
          . In Text REtrieval Conference,
          <year>2004</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref6">
        <mixed-citation>
          [6]
          <string-name>
            <given-names>Dan</given-names>
            <surname>Moldovan</surname>
          </string-name>
          , Christine Clark, and
          <string-name>
            <given-names>Sanda</given-names>
            <surname>Harabagiu</surname>
          </string-name>
          .
          <article-title>Temporal Context Representation and Reasoning</article-title>
          .
          <source>In Proceedings of IJCAI, Edinburgh</source>
          , Scotland,
          <year>2005</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref7">
        <mixed-citation>
          [7]
          <string-name>
            <given-names>Dan</given-names>
            <surname>Moldovan</surname>
          </string-name>
          , Christine Clark, Sanda Harabagiu, and
          <string-name>
            <given-names>Steve</given-names>
            <surname>Maiorano</surname>
          </string-name>
          .
          <article-title>COGEX A Logic Prover for Question Answering</article-title>
          .
          <source>In Proceedings of the HLT/NAACL</source>
          ,
          <year>2003</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref8">
        <mixed-citation>
          [8]
          <string-name>
            <given-names>Dan</given-names>
            <surname>Moldovan</surname>
          </string-name>
          and
          <string-name>
            <given-names>Adrian</given-names>
            <surname>Novischi</surname>
          </string-name>
          .
          <article-title>Lexical chains for Question Answering</article-title>
          .
          <source>In Proceedings of COLING</source>
          , Taipei, Taiwan,
          <year>August 2002</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref9">
        <mixed-citation>
          [9]
          <string-name>
            <given-names>Dan</given-names>
            <surname>Moldovan</surname>
          </string-name>
          and
          <string-name>
            <given-names>Vasile</given-names>
            <surname>Rus</surname>
          </string-name>
          .
          <article-title>Logic Form Transformation of WordNet and its Applicability to Question Answering</article-title>
          .
          <source>In Proceedings of ACL</source>
          , France,
          <year>2001</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref10">
        <mixed-citation>
          [10]
          <string-name>
            <surname>Dan</surname>
            <given-names>Moldovan</given-names>
          </string-name>
          , Munirathnam Srikanth, Abraham Fowler, Altaf Mohammed,
          <string-name>
            <given-names>Eric</given-names>
            <surname>Jean</surname>
          </string-name>
          . Synergist:
          <article-title>Tools for Intelligence Analysis</article-title>
          . NIMD Conference, Arlington,
          <string-name>
            <surname>VA</surname>
          </string-name>
          ,
          <year>2006</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref11">
        <mixed-citation>
          [11]
          <string-name>
            <surname>Marian</surname>
            <given-names>Olteanu</given-names>
          </string-name>
          , Pasin Suriyentrakorn and
          <string-name>
            <given-names>Dan</given-names>
            <surname>Moldovan</surname>
          </string-name>
          .
          <article-title>Language Models and Reranking for Machine Translation</article-title>
          .
          <source>In NAACL 2006 Workshop On Statistical Machine Translation</source>
          ,
          <year>2006</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref12">
        <mixed-citation>
          [12]
          <string-name>
            <surname>Marian</surname>
            <given-names>Olteanu</given-names>
          </string-name>
          , Chris Davis, Ionut Volosen and
          <string-name>
            <given-names>Dan</given-names>
            <surname>Moldovan</surname>
          </string-name>
          .
          <article-title>Phramer - An Open Source Statistical Phrase-Based Translator</article-title>
          .
          <source>In NAACL 2006 Workshop On Statistical Machine Translation</source>
          ,
          <year>2006</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref13">
        <mixed-citation>
          [13]
          <string-name>
            <surname>Marta</surname>
            <given-names>Tatu</given-names>
          </string-name>
          , Brandon Iles,
          <string-name>
            <given-names>Dan</given-names>
            <surname>Moldovan</surname>
          </string-name>
          .
          <article-title>Automatic Answer Validation using COGEX. Cross-Language Evaluation Forum (CLEF</article-title>
          ),
          <year>2006</year>
          .
        </mixed-citation>
      </ref>
    </ref-list>
  </back>
</article>