<!DOCTYPE article PUBLIC "-//NLM//DTD JATS (Z39.96) Journal Archiving and Interchange DTD v1.0 20120330//EN" "JATS-archivearticle1.dtd">
<article xmlns:xlink="http://www.w3.org/1999/xlink">
  <front>
    <journal-meta />
    <article-meta>
      <title-group>
        <article-title>TeamHCMUS: A Concept-Based Information Retrieval Approach for Web Medical Documents</article-title>
      </title-group>
      <contrib-group>
        <contrib contrib-type="author">
          <string-name>Nghia Huynh</string-name>
          <email>huynhnghiavn@gmail.com</email>
          <xref ref-type="aff" rid="aff1">1</xref>
        </contrib>
        <contrib contrib-type="author">
          <string-name>Thanh Tuan Nguyen</string-name>
          <email>tuannt@fit.hcmute.edu.vn</email>
          <xref ref-type="aff" rid="aff0">0</xref>
        </contrib>
        <contrib contrib-type="author">
          <string-name>Quoc Ho</string-name>
          <email>hbquoc@fit.hcmus.edu.vn</email>
          <xref ref-type="aff" rid="aff1">1</xref>
        </contrib>
        <aff id="aff0">
          <label>0</label>
          <institution>Faculty of Information Technology, HCMC University of Technology and Education</institution>
          ,
          <country country="VN">Vietnam</country>
        </aff>
        <aff id="aff1">
          <label>1</label>
          <institution>Faculty of Information Technology, University of Science</institution>
          ,
          <addr-line>Ho Chi Minh City</addr-line>
          ,
          <country country="VN">Vietnam</country>
        </aff>
      </contrib-group>
      <abstract>
        <p>It's difficult for laypeople, even clinicians, to understand eHealth contents found in the web medical documents. With the objective to build a health search engine, task 2 of 2015 CLEF eHealth aims to detect levels of accuracy of information retrieval systems when searching for web medical documents. In this task, our approach is to integrate a retrieval of medical concepts into the preprocessing of corpora. This means that all terms in the documents that are not related to medicine are removed before indexing. We also expand queries for searching more effectively. In general, our results are not better than other participants' in doing task 2 except some queries. When using integration of extracting medical concepts and query expansion based on laypeople's queries, searching retrieval is also lower. It can be explained partly that laypeople's queries are not commonly included medical terms or only contain features painting their health situations. In addition, we also give a brief statement of the main points of an estimation of readability which is a significant assessment referred by CLEF eHealth near future.</p>
      </abstract>
      <kwd-group>
        <kwd>Concept-based</kwd>
        <kwd>Medical Information Retrieval</kwd>
        <kwd>Medical Documents</kwd>
        <kwd>Language Model</kwd>
      </kwd-group>
    </article-meta>
  </front>
  <body>
    <sec id="sec-1">
      <title>-</title>
      <p>Laypeople as well as clinicians find it hard to comprehend the eHealth
documents which are retrieved from searching for necessary information on
the Internet. Their problems are how to understand professional terms more
exactly. It’s the third year when CLEF eHealth continues the purpose of
promotion in doing research and finding out advanced methods to build a search
engine system for meeting users’ requirements of searching for medical
information.</p>
      <p>
        CLEF eHealth pointed out two tasks this year1 . Task 1 is a mission
statement of information extraction from clinical text. It is split into two sub-tasks:
task 1a and task 1b. Specifically, task 1a is clinical speech recognition related
to converting verbal nursing handover to written free-text records and task 1b
is named entity recognition in clinical reports. Task 2 is user-centered health
information retrieval [
        <xref ref-type="bibr" rid="ref15 ref16">15-16</xref>
        ]. In this paper, we proposed methods to meet
some requirements of task 2.
      </p>
      <p>There are some changes of type of queries and measure methods of
relevance assessments this year. One of them is queries that should be made by
laypeople do not come from experts in the heath domain because the
laypeople want to find out information to help them to be clearer their related
medical conditions. Another change is to apply two measures to assess results from
participants’ submissions.</p>
      <p>
        As a further matter, CLEF eHealth has also considered a readability-biased
assessment [
        <xref ref-type="bibr" rid="ref14">14</xref>
        ], the factor of understandability of information (or readability)
within the evaluation of submissions. However, because the measure has been
still at levels of experiment consideration, observations that were carried out
might not be conclusive.
      </p>
      <p>
        In this paper, our approach is to integrate retrieval of medical concepts
from the web medical documents into the preprocessing of corpora which is
based on a list of medical concepts built by [
        <xref ref-type="bibr" rid="ref2">2</xref>
        ]. Process of building a search
engine system can be summarized as following descriptions: At first, we used
some tools to remove tags of HTML files in the corpus provided by 2015
CLEF eHealth2 for task 2. We collected a set of raw data from the corpus. We
then removed stopwords and got stemming of terms in each document in the
set. All data extracted from this process was indexed and called Index A.
Another indexed corpus called Index B was also created from the data that only
includes terms related to medical domain. Building the Index B is described in
detail in Section 2.3.
      </p>
      <p>
        We obtained a baseline run and other runs after doing experiments in
searching for queries in Index A and B corpus with Dirichlet smooth
coefficients [
        <xref ref-type="bibr" rid="ref13">13</xref>
        ].
      </p>
      <p>We also expanded queries to get more information for searching.
Techniques to do this are described in Section 3.</p>
    </sec>
    <sec id="sec-2">
      <title>1 https://sites.google.com/site/clefehealth2015</title>
      <p>2 https://sites.google.com/site/clefehealth2015/task-2</p>
      <p>
        In general, most of our results are not better than other participants’ in
doing task 2 except some queries (see Figure 2, 3). In addition, when integrating
the method of extracting medical concepts [
        <xref ref-type="bibr" rid="ref2">2</xref>
        ] in building Index B as well as
expending laypeople’s queries, searching results are also lower (see Figure 4,
5). This can be explained that their queries are not commonly included
medical terms or only contain terms painting their health situations.
      </p>
      <p>The rest of the paper is organized as follows: Section 2 outlines the CLEF
eHealth dataset and methods of preprocessing. Section 3 describes the
structure of a query and some techniques for query expansion. Section 4 presents
relevance assessments. Section 5 demonstrates description of our runs.
Section 6 explains experiments done by task participants. Finally, Section 7
concludes the paper.
2</p>
      <sec id="sec-2-1">
        <title>Dataset and Preprocessing</title>
        <p>The dataset for Task 2 is provided by Khresmoi project3. It has about one
million documents, a set of documents in the HTML (Hyper Text Markup
Language) format. All documents are collected from well-known health and
medical sites and databases in 2012. The size of the dataset is about 6.3G in
compressed status and approximately 43.6 GB after extracting.</p>
        <p>Each file in the dataset is in the format of .dat files and contains a set of
web pages and metadata where shows the original information of each web
page as described below:
 a unique identifier (#UID) for a web page in this document collection,
 the date of crawl in the form YYYYMM (#DATE),
 the URL (#URL) to the original path of a web page, and
 the raw HTML content (#CONTENT) of the web page
2.1</p>
      </sec>
      <sec id="sec-2-2">
        <title>Parse HTML to Text</title>
        <p>Majestic-12, Distributed Search Engine (DSearch) projects4, built an
opensource tool called HTML parser v3.1.4 for parsing tags in the HTML files.
Basing on this tool, we extracted text contents from tags of HTML documents
in the dataset. There are some tags in the documents that contain unnecessary</p>
      </sec>
    </sec>
    <sec id="sec-3">
      <title>3 http://khresmoi.eu/</title>
      <p>4 http://www.majestic12.co.uk/projects/
texts for building a search engine system in the medical field. Thus we tried to
ignore all those.
2.2</p>
      <sec id="sec-3-1">
        <title>Content Cleaning</title>
        <p>The text contents extracted from HTML documents are not always good
for building the system. Example, tags contain text of sitemap, update
information and some items on the main and popup menu in the HTML
documents. Therefore, all of them should be removed.</p>
        <p>
          Next, we had a process of removing stopwords because they were common
words in English language [
          <xref ref-type="bibr" rid="ref4">4</xref>
          ] and did not play an important role in
determining the meaning of sentences. To get more efficient in preprocessing, we used
the Porter algorithm [
          <xref ref-type="bibr" rid="ref9">9</xref>
          ] for stemming words. Finishing all above work, we
had a corpus ready for indexing. We used Lucene-5.0.0 tool5 to index this
corpus and name Index A.
2.3
        </p>
      </sec>
      <sec id="sec-3-2">
        <title>Extracting Concepts</title>
        <p>
          We used a list of medical concepts built by [
          <xref ref-type="bibr" rid="ref2">2</xref>
          ] to extract medical concepts
from the documents in the dataset by removing all their terms that are not in
the list and also not in UMLS6 (Unified Medical Language System).
        </p>
        <p>After this processing, we had a dataset that contains terms related to
medical field. We also used Lucene-5.0.0 to index this dataset and named Index B.
3</p>
      </sec>
      <sec id="sec-3-3">
        <title>Queries and Query Expansion</title>
        <p>Because of consideration in building a search engine system for English
documents, we only concentrated on English queries. The number of queries
for task 2 of 2015 CLEF eHealth includes 66 queries along with their
narrative fields. The narrative fields are used to provide information to the
assessors when performing relevance evaluations. Here is the typical structure of a
query7:
5 http://lucene.apache.org/
6 http://www.nlm.nih.gov/research/umls/
7 https://sites.google.com/site/clefehealth2015/task-2
&lt;top&gt;
&lt;num&gt;clef2015.training.1&lt;/num&gt;
&lt;query&gt;loss of hair on scalp in an inch width round&lt;/query&gt;
&lt;narr&gt;Documents should contain information allowing the user to
understand they have alopecia&lt;/narr&gt;
&lt;/top&gt;</p>
        <p>
          With purpose of getting more information for searching, with some tasks
of CLEF eHealth before, participants or teams found out synonym of query’s
terms in the UMLS or MeSH8 (Medical Subject Headings) for expanding their
queries [
          <xref ref-type="bibr" rid="ref10">10</xref>
          ], [
          <xref ref-type="bibr" rid="ref12">12</xref>
          ]. Other participants used pseudo-relevance feedback (PRF)
as a method for expansion [
          <xref ref-type="bibr" rid="ref6">6</xref>
          ], [
          <xref ref-type="bibr" rid="ref11">11</xref>
          ], or took Wikipedia9 for making a
semantic query expansion [
          <xref ref-type="bibr" rid="ref1">1</xref>
          ]. Because the set of queries of task 2 of 2015 CLEF
eHealth is user-centered queries (i.e. they are made by laypeople rather than
done by experts in the medical field), Terms of the queries are usually short
and not much relative to medical concepts. So we lacked evidences to expand
the queries. Thus, to get more information for every query expansion, we
searched for queries in each Index (A and B) to get the relevant document at
the top of each searching result. To get an expanded query, we connected that
top document to the query.
4
        </p>
      </sec>
      <sec id="sec-3-4">
        <title>Relevance Assessments</title>
        <p>
          Methods of relevance assessments are provided by the Share/CLEF
eHealth 2015 TASK 2 and described as follows:
 Result of runs is the top 1000 of relevant documents returned by
searching for 66 queries that based on LM (language model) [
          <xref ref-type="bibr" rid="ref8">8</xref>
          ] with
specification of Dirichlet smooth coefficients.
 Relevance is assessed as following descriptions:
─ Evaluation with standard trec_eval metrics10:
        </p>
        <p>
          o 2 point scale: non relevant (label 0); relevant (label 1)
─ Evaluation with nDCG: [
          <xref ref-type="bibr" rid="ref3">3</xref>
          ]
        </p>
        <p>
          o 3 point scale: gain 0 (label 0), gain 1 (label 1), gain 2 (label 2)
─ Readability-biased evaluation: [
          <xref ref-type="bibr" rid="ref14">14</xref>
          ]
        </p>
      </sec>
    </sec>
    <sec id="sec-4">
      <title>8 http://www.ncbi.nlm.nih.gov/mesh</title>
      <p>9 https://en.wikipedia.org
10 http://trec.nist.gov/trec_eval/
o 4 point scale: very technical and difficult (label 0), somewhat
technical and difficult (1), somewhat easy (label 2), very easy (label 3)
5</p>
      <sec id="sec-4-1">
        <title>Description of Runs</title>
        <p>
          With given 66 queries, we retrieved the top 1000 of relevant documents for
each query when using LM in Lucene 5.0 for matching the queries with each
document in the Index A or B corpus along with specific values of Dirichlet
smooth coefficients [
          <xref ref-type="bibr" rid="ref13">13</xref>
          ]. We submitted 8 runs in the task that are described in
summary as follows: (see Table 1)
Run 1 (baseline run): We applied the default value (2000) of smooth
coefficient to LM and search in Index A corpus.
        </p>
        <p>Run 2: It is a variant of run 1 in which value of smooth coefficient is 10000.
Run 3: We used the corpus of medical concepts (i.e. Index B) for searching
with the same smooth coefficient as run 1.</p>
        <p>Run 4: It is a variant of run 3 in which value of smooth coefficient is 10000.
Run 5: Each query, we took the top 01 of relevant documents at Run 3 for
expanding the query. Executing the same run 1, but for this expanded query.
Run 6: It is a variant of run 5 in which the top 01 of relevant documents at
run 4 was used to expand the query. Executing the same Run 2, but for this
expanded query.</p>
        <p>Run 7: Each query, we took the top 01 of relevant documents at run 3 for
expanding the query. Executing the same run 3, but for this expanded query.
Run 8: It is a variant of Run 7 in which the top 01 of relevant documents at
run 4 was used for expanding the query. Executing the same run 4, but for this
expanded query.</p>
      </sec>
      <sec id="sec-4-2">
        <title>Experiments</title>
      </sec>
      <sec id="sec-4-3">
        <title>Evaluation with standard trec_eval metrics and nDCG</title>
        <p>Two primary evaluation parameters for task 2 of 2015 CLEF eHealth are
the precision at 10 (P@10) and Normalized Discounted Cumulative Gain
(NDCG) at rank 10. Figure 1 shows the results of the submitted eight runs. It
can be seen clearly from Figure 1 that run 1 (baseline run) is the best run
yielding the highest values followed by run 2 to 5, 7 respectively. Whereas
run 8 is the least performing run and following it is run 6. It is observed that
the values of nDCG are higher than P@10 in processing with integration of
query expansion into the system (run 5 to 8).</p>
        <p>1
2
3
4
5
6
7</p>
        <p>8</p>
        <p>
          Runs
Demonstration of Figure 1 shows that our approach of integrating a
retrieval of medical concepts [
          <xref ref-type="bibr" rid="ref2">2</xref>
          ] into the system (i.e. run 3, 4) does not perform
better than the baseline run. This is because documents in Index B only
contain medical concepts while queries have few terms related to the medical
field.
        </p>
        <p>The Figure 1 also indicates that run 5 to 8 with performance of query
expansion techniques is worse than other runs. Thus it can be explained that
laypeople’s queries are not commonly included medical terms or only contain
terms painting their health situations. In case of that the relevant document for
query expansion evaluated by CLEF eHealth is not related to the query,
retrieval documents from searching for that query are not also in higher
relevance assessments. Another reason for explanation of bad results is that query
expansion is only a connection between a query and the relevant document at
the top of the searching result of the query. As the result, new queries are too
long to apply to LM.</p>
        <p>Figure 2 shows the graph structure of the participants’ performance of
baseline run (run 1) in task 2 of 2015 CLEF eHealth. For the baseline run, it is
observed that our experiment performed most of queries in under the best &amp;
median cases of all participated groups. But some queries, we reached
outperformance more than other systems as queries: 13, 26, 29, and 65. A few
queries are in lags such as 3, 32, 33, and 66 respectively.</p>
        <p>
          Figure 3 shows the performance graph of the participants’ run 2 in the task
2 of 2015 CLEF eHealth. For the run 2, it is observed that our experiment
only out-performs more than other systems in queries 15, 16, and 25
respectively. On the other hand, our run 2 system lags in queries 5, 20, 30, 32, 34,
46, 51, 57, 61, and 65 respectively. Figure 3 points out that increasing value
of smooth coefficient is not efficient.
When applying some techniques as retrieval of medical concepts [
          <xref ref-type="bibr" rid="ref2">2</xref>
          ] (i.e.
run 3, 4) and query expansion to do more experiments (i.e. run 5 to 8), we
reached retrieval results that are not better than other participated groups’
systems although there are still a few out-performed queries as indicated in
Figure 4 (i.e. run 3) and Figure 5 (i.e. run 7). The reason of those can be
explained that query expansion made new queries with large length and Index B
only contains medical concepts. So treatments should be proposed and
experimented on in future work.
        </p>
        <p>
          For this year task, it’s the first time CLEF eHealth has considered the
factor of readability (understandability) within the evaluation of the submissions
with assumption that readability is assessed independently of relevance
assessments. To account for readability in the evaluation, we have computed an
understandability biased measure, uRBP [
          <xref ref-type="bibr" rid="ref14">14</xref>
          ].
        </p>
        <p>
          We used the ubire-v0.1.0 tool11 to point out the values of RBP [
          <xref ref-type="bibr" rid="ref5">5</xref>
          ], and two
versions of uRBP. The user persistence parameter p of RBP (and uRBP) was
set to 0.8 [
          <xref ref-type="bibr" rid="ref7">7</xref>
          ]. Values of uRBP were computed by using user model 1 of [
          <xref ref-type="bibr" rid="ref14">14</xref>
          ]
with threshold=2, i.e. documents with a understandability score of 0 or 1
where deemed unreadable and had P(U|k)=0, while documents with a
understandability score of 2 or 3 where deemed readable and had P(U|k)=1. Values
of uRBPgr were computed by mapping graded understandability scores to
different probability values, in particular: readability of 0 was assigned
P(U|k)=0, readability of 1 was assigned P(U|k)=0.4, readability of 2 was
assigned P(U|k)=0.8, readability of 3 was assigned P(U|k)=1.
        </p>
        <p>CLEF eHealth also notes that these readability-biased measures are still
being experimented and observations that were made with the provided
measures may not be conclusive.</p>
        <p>1
2
3
4
5
6
7</p>
        <p>8</p>
        <p>Runs
11 https://github.com/ielab/ubire
throughout runs. This is a reductive trend except a fluctuation at run 7. That
the line of uRBP is under the uRBPgr line shows that mapping graded
understandability scores to different probability values reaches a bit
outperformance.
7</p>
      </sec>
      <sec id="sec-4-4">
        <title>Conclusion</title>
        <p>It’s necessary to build an efficient retrieval system for searching
laypeople’s queries. However, this is not easy and still a challenge for participating
groups in CLEF eHealth. This is partly because the features in laypeople’s
queries are not utterly medical terms or only terms of description of their
health situations. Thus, searching of retrieval systems for laypeople’s queries
should be returned.</p>
        <p>
          We carried out some experiments in applying our approach of retrieval of
medical concepts [
          <xref ref-type="bibr" rid="ref2">2</xref>
          ] and query expansion techniques to establish a retrieval
system. However, efficiency of the system is not still out-performance in
general except few queries.
        </p>
        <p>In future work, we will keep finding out advanced methods to refine
corpora so that they only contain suitable features. Simultaneously, we will
consider query expansion techniques, and experiment on various models of
matching documents. In addition, we also do research on the results of other
participating groups’ systems to make them better.</p>
      </sec>
    </sec>
  </body>
  <back>
    <ref-list>
      <ref id="ref1">
        <mixed-citation>
          1.
          <string-name>
            <surname>M. Almasri J. Chevallet</surname>
            ,
            <given-names>C.</given-names>
          </string-name>
          <article-title>Berrut: Exploiting Wikipedia Structure for Short Query Expansion in Cultural Heritage</article-title>
          ,
          <string-name>
            <surname>CORIA</surname>
          </string-name>
          (
          <year>2014</year>
          ).
        </mixed-citation>
      </ref>
      <ref id="ref2">
        <mixed-citation>
          2.
          <string-name>
            <given-names>Nghia</given-names>
            <surname>Huynh</surname>
          </string-name>
          , Quoc Ho:
          <article-title>TeamHCMUS: Analysis of Clinical Text</article-title>
          ,
          <source>Proceedings of the 9th International Workshop on Semantic Evaluation (SemEval</source>
          <year>2015</year>
          ), pages pp.
          <fpage>370</fpage>
          -
          <lpage>374</lpage>
          (
          <year>2015</year>
          ).
        </mixed-citation>
      </ref>
      <ref id="ref3">
        <mixed-citation>
          3.
          <string-name>
            <given-names>K.</given-names>
            <surname>Järvelin</surname>
          </string-name>
          ,
          <string-name>
            <surname>J.</surname>
          </string-name>
          <article-title>Kekäläinen: Cumulated Gain-Based Evaluation of IR Techniques</article-title>
          .
          <source>ACM Transactions on Information Systems (TOIS) 20(4)</source>
          , pp.
          <fpage>422</fpage>
          -
          <lpage>446</lpage>
          (
          <year>2002</year>
          ).
        </mixed-citation>
      </ref>
      <ref id="ref4">
        <mixed-citation>
          4.
          <string-name>
            <given-names>J.</given-names>
            <surname>Leskovec</surname>
          </string-name>
          ,
          <string-name>
            <given-names>A.</given-names>
            <surname>Rajaraman</surname>
          </string-name>
          ,
          <string-name>
            <given-names>J. D.</given-names>
            <surname>Ullman</surname>
          </string-name>
          : Mining of Massive Datasets. Cambridge University Press, chapter
          <volume>1</volume>
          (
          <year>2011</year>
          ).
        </mixed-citation>
      </ref>
      <ref id="ref5">
        <mixed-citation>
          5.
          <string-name>
            <given-names>A.</given-names>
            <surname>Moffat</surname>
          </string-name>
          , J. Zobel:
          <article-title>Rank-biased precision for measurement of retrieval effectiveness</article-title>
          ,
          <source>ACM Transactions on Information Systems (TOIS)</source>
          , vol.
          <volume>27</volume>
          no.
          <issue>1</issue>
          , pp.
          <fpage>1</fpage>
          -
          <lpage>27</lpage>
          (
          <year>2008</year>
          ).
        </mixed-citation>
      </ref>
      <ref id="ref6">
        <mixed-citation>
          6.
          <string-name>
            <given-names>E</given-names>
            <surname>Noguera</surname>
          </string-name>
          , F Llopis:
          <article-title>Applying Query Expansion Techniques to Ad Hoc Monolingual tasks with the IR-n system</article-title>
          ,
          <source>CEUR Workshop Proceedings</source>
          , Vol-
          <volume>1173</volume>
          (
          <year>2007</year>
          ).
        </mixed-citation>
      </ref>
      <ref id="ref7">
        <mixed-citation>
          7.
          <string-name>
            <given-names>L. A.</given-names>
            <surname>Park</surname>
          </string-name>
          ,
          <string-name>
            <surname>Y.</surname>
          </string-name>
          <article-title>Zhang: On the distribution of user persistence for rank-biased precision</article-title>
          ,
          <source>Proceedings of the 12th Australasian Document Computing Symposium</source>
          (
          <year>2007</year>
          ).
        </mixed-citation>
      </ref>
      <ref id="ref8">
        <mixed-citation>
          8.
          <string-name>
            <surname>J. M. Ponte</surname>
          </string-name>
          , W. B.
          <string-name>
            <surname>Croft</surname>
          </string-name>
          :
          <article-title>A language modeling approach to information retrieval</article-title>
          ,
          <source>Proceedings of the 21st annual international ACM SIGIR conference on Research and development in information retrieval, ACM</source>
          , pp.
          <fpage>275</fpage>
          -
          <lpage>281</lpage>
          (
          <year>1998</year>
          ).
        </mixed-citation>
      </ref>
      <ref id="ref9">
        <mixed-citation>
          9.
          <string-name>
            <given-names>M. F.</given-names>
            <surname>Porter</surname>
          </string-name>
          :
          <article-title>An algorithm for suffix stripping</article-title>
          ,
          <source>Program</source>
          , Vol.
          <volume>14</volume>
          , no.
          <issue>3</issue>
          , pp.
          <fpage>130</fpage>
          -
          <lpage>137</lpage>
          (
          <year>1980</year>
          ).
        </mixed-citation>
      </ref>
      <ref id="ref10">
        <mixed-citation>
          10. H.
          <string-name>
            <surname>Thakkar</surname>
            , G. Iyer,
            <given-names>K.</given-names>
          </string-name>
          <string-name>
            <surname>Shah</surname>
          </string-name>
          , P. Majumder:
          <article-title>Team IRLabDAIICT at ShARe/CLEF eHealth 2014 Task 3: User-centered Information Retrieval system for Clinical Documents</article-title>
          ,
          <source>CEUR Workshop Proceedings</source>
          , Vol-
          <volume>1180</volume>
          , pp.
          <fpage>248</fpage>
          -
          <lpage>259</lpage>
          (
          <year>2014</year>
          ).
        </mixed-citation>
      </ref>
      <ref id="ref11">
        <mixed-citation>
          11.
          <string-name>
            <given-names>O.</given-names>
            <surname>Thesprasith</surname>
          </string-name>
          ,
          <string-name>
            <surname>C.</surname>
          </string-name>
          <article-title>Jaruskulchai: CSKU GPRF-QE for Medical Topic Web Retrieval</article-title>
          ,
          <source>CEUR Workshop Proceedings</source>
          , Vol-
          <volume>1180</volume>
          , pp.
          <fpage>260</fpage>
          -
          <lpage>268</lpage>
          (
          <year>2007</year>
          ).
        </mixed-citation>
      </ref>
      <ref id="ref12">
        <mixed-citation>
          12. S. Verberne:
          <article-title>A language-modelling approach to User-Centred Health Information Retrieval</article-title>
          ,
          <source>CEUR Workshop Proceedings</source>
          , Vol-
          <volume>1180</volume>
          , pp.
          <fpage>269</fpage>
          -
          <lpage>275</lpage>
          (
          <year>2014</year>
          ).
        </mixed-citation>
      </ref>
      <ref id="ref13">
        <mixed-citation>
          13.
          <string-name>
            <given-names>C.</given-names>
            <surname>Zhai</surname>
          </string-name>
          and
          <string-name>
            <given-names>J.</given-names>
            <surname>Lafferty</surname>
          </string-name>
          :
          <article-title>A study of smoothing methods for language models applied to ad hoc information retrieval</article-title>
          ,
          <source>Proceedings of the ACM-SIGIR</source>
          <year>2001</year>
          , pp.
          <fpage>334</fpage>
          -
          <lpage>342</lpage>
          (
          <year>2001</year>
          ).
        </mixed-citation>
      </ref>
      <ref id="ref14">
        <mixed-citation>
          14. G .Zuccon,
          <string-name>
            <surname>B.</surname>
          </string-name>
          <article-title>Koopman: Integrating Understandability in the Evaluation of Consumer Health Search Engines</article-title>
          ,
          <source>Proceedings of the SIGIR workshop on Medical Information Retrieval</source>
          (
          <year>2014</year>
          ).
        </mixed-citation>
      </ref>
      <ref id="ref15">
        <mixed-citation>
          15.
          <string-name>
            <surname>Lorraine</surname>
            <given-names>Goeuriot</given-names>
          </string-name>
          , Liadh Kelly, Hanna Suominen, Leif Hanlen, Aurélie Névéol, Cyril Grouin, João Palotti,
          <string-name>
            <given-names>Guido</given-names>
            <surname>Zuccon</surname>
          </string-name>
          .
          <article-title>Overview of the CLEF eHealth Evaluation Lab 2015</article-title>
          .
          <source>CLEF 2015 - 6th Conference and Labs of the Evaluation Forum, Lecture Notes in Computer Science (LNCS)</source>
          , Springer,
          <year>September 2015</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref16">
        <mixed-citation>
          16.
          <string-name>
            <surname>Palotti</surname>
          </string-name>
          , João and Zuccon, Guido and Goeuriot, Lorraine and Kelly, Liadh and Hanbury, Allan and Jones,
          <string-name>
            <surname>Gareth</surname>
            <given-names>JF</given-names>
          </string-name>
          , and
          <string-name>
            <surname>Lupu</surname>
          </string-name>
          , Mihai and Pecina, Pavel.
          <source>CLEF eHealth Evaluation Lab</source>
          <year>2015</year>
          ,
          <article-title>task 2: Retrieving Information about Medical Symptoms</article-title>
          .
          <source>CLEF 2015 Online Working Notes, CEUR-WS</source>
          ,
          <year>2015</year>
          .
        </mixed-citation>
      </ref>
    </ref-list>
  </back>
</article>