<!DOCTYPE article PUBLIC "-//NLM//DTD JATS (Z39.96) Journal Archiving and Interchange DTD v1.0 20120330//EN" "JATS-archivearticle1.dtd">
<article xmlns:xlink="http://www.w3.org/1999/xlink">
  <front>
    <journal-meta />
    <article-meta>
      <title-group>
        <article-title>Demonstration of Improved Search Result Relevancy Using Real-Time Implicit Relevance Feedback</article-title>
      </title-group>
      <contrib-group>
        <contrib contrib-type="author">
          <string-name>David Hardtke</string-name>
          <email>hardtke@surfcanyon.com</email>
          <xref ref-type="aff" rid="aff0">0</xref>
        </contrib>
        <contrib contrib-type="author">
          <string-name>Mike Wertheim</string-name>
          <email>mikew@surfcanyon.com</email>
          <xref ref-type="aff" rid="aff0">0</xref>
        </contrib>
        <contrib contrib-type="author">
          <string-name>Mark Cramer</string-name>
          <email>mcramer@surfcanyon.com</email>
          <xref ref-type="aff" rid="aff0">0</xref>
        </contrib>
        <aff id="aff0">
          <label>0</label>
          <institution>Surf Canyon</institution>
          ,
          <addr-line>Incorporated, 274 14th St., Oakland, CA 94612</addr-line>
          ,
          <country country="US">USA</country>
        </aff>
      </contrib-group>
      <pub-date>
        <year>2009</year>
      </pub-date>
      <fpage>19</fpage>
      <lpage>23</lpage>
      <abstract>
        <p>Surf Canyon has developed real-time implicit personalization technology for web search and implemented the technology in a browser extension that can dynamically modify search engine results pages (Google, Yahoo!, and Live Search). A combination of explicit (queries, reformulations) and implicit (clickthroughs, skips, page reads, etc.) user signals are used to construct a model of instantaneous user intent. This user intent model is combined with the initial search result rankings in order to present recommended search results to the user as well as to reorder subsequent search engine results pages after the initial page. This paper will use data from the first three months of Surf Canyon usage to show that a user intent model built from implicit user signals can dramatically improve the relevancy of search results.</p>
      </abstract>
    </article-meta>
  </front>
  <body>
    <sec id="sec-1">
      <title>INTRODUCTION</title>
      <p>
        It has long since been demonstrated that explicit relevance
feedback can improve both precision and recall in
information retrieval[
        <xref ref-type="bibr" rid="ref1">1</xref>
        ]. An initial query is used to retrieve a set of
documents. The user is then asked to manually rate a
subset of the documents as relevant or not relevant. The terms
appearing in the relevant document are then added to the
initial query to produce a new query. Additionally,
nonrelevant documents can be used to remove or de-emphasize
terms for the reformulated query. This process can be
repeated iteratively, but it was found that after a few iterations
very few new relevant documents are found [
        <xref ref-type="bibr" rid="ref2">2</xref>
        ].
      </p>
      <p>Explicit relevance feedback as described above requires
active user participation. An alternative method that does not
require specific user participation is pseudo relevance
feedback. In this scheme, the top N documents from the initial
query are assumed to be relevant. The important terms in
these documents are then used to expand the original query.</p>
      <p>Implicit Relevance Feedback aims to improve the precision
and recall of information retrieval by utilizing user actions
to infer the relevance or non-relevance of documents. Many
different user behavior signals can contribute to a
probabilistic evaluation of document relevance. Explicit
document relevance determinations are more accurate, but
implicit relevance determinations are more easily obtained as
they require no additional user effort.
2.</p>
    </sec>
    <sec id="sec-2">
      <title>IMPLICIT SIGNALS AND USER INFOR</title>
    </sec>
    <sec id="sec-3">
      <title>MATION NEED</title>
      <p>
        With the large, open nature of the World Wide Web it is
very difficult to evaluate the quality of search engine
algorithms using explicit human evaluators. Hence, there have
been numerous investigations into using implicit user
signals for evaluation and optimization of search engine quality.
Several studies have investigated the extent to which a
clickthrough on a specific search engine result can be interpreted
as a user indication of document relevancy (for a review see
[
        <xref ref-type="bibr" rid="ref3">3</xref>
        ]). The primary issue involving clickthrough data is that
users are most likely to click on higher ranked documents
because they tend to read the SERP (search engine results
page) from top to bottom. Additionally, users trust that
a search engine places the most relevant documents at the
highest positions on the SERP.
      </p>
      <p>
        Joachims et al used eye tracking studies combined with
manual relevance judgements to investigate the accuracy of
clickthrough data for implicit relevance feedback [
        <xref ref-type="bibr" rid="ref4">4</xref>
        ]. They
conclude that clickthrough data can be used to accurately
determine relative document relevancies. If, for instance,
a user clicks on a search result after skipping other search
results, subsequent evaluation by human judges show that
in ∼80% of cases the clicked document is more relevant to
the query than the documents that were skipped.
      </p>
      <p>
        In addition to clickthroughs, other user behaviors can be
related to document relevancy. Fox et al. used a browser
add-in to track user behavior for a volunteer sample of
office workers[
        <xref ref-type="bibr" rid="ref5">5</xref>
        ]. In addition to tracking their search and web
usage, the browser add-in would prompt the user for
specific relevance evaluations for pages they had visited. Using
the observed user behavior and subsequent relevance
evaluations, they were able to correlate implicit user signals with
explicit user evaluations and determine what user signals
are most likely to indicate document relevance. For pages
clicked by the user, the user indicated that they were either
satisfied or partially satisfied with the document nearly 70%
of the time. In the study, two other variables were found
to be most important for predicting user satisfaction with
a result page visit. The first was the duration of time that
the user spent away from the SERP before returning – if
the user was away from the SERP for a short period of time
they tended to be dissatisfied with the document. The other
important variable for predicting user satisfaction was the
“Exit type” – users that closed the browser on a result page
tended to be satisfied with that result page. The
important outcome of this and other studies is that implicit user
behavior can be used instead of explicit user feedback to
determine the user’s information need.
      </p>
    </sec>
    <sec id="sec-4">
      <title>IMPLICIT REAL-TIME PERSONALIZA</title>
    </sec>
    <sec id="sec-5">
      <title>TION</title>
      <p>As discussed in the previous section, it has been shown
that implicit user behavior can often infer satisfaction with
visited results pages. The goal of the Surf Canyon
technology is to use implicit user behavior to predict which unseen
documents in a collection are most relevant to the user and
to recommend these documents to the user.</p>
      <p>
        Shen, Tan, and Zhai1 have investigated context-sensitive
adaptive information retrieval systems [
        <xref ref-type="bibr" rid="ref6">6</xref>
        ]. They use both
clickthrough information and query history information to
update the retrieval and ranking algorithm. A TREC
collection was used since manual relevancy judgements are
available. They built an adaptive search interface to this
collection, and had 3 volunteers conduct searches on 30 relatively
difficult TREC topics. The users could query, re-query,
examine document summaries, and examine documents. To
quantify the retrieval algorithms, they used Mean Average
Precision (MAP) or Precision at 20 documents. As these
were difficult TREC topics, users submitted multiple queries
for each topic. They found that including query history
produced a marginal improvement in MAP, while use of
clickthrough information produced dramatic increases (up
to nearly 100%) in MAP.
      </p>
      <p>
        Shen et al. also built an experimental adaptive search
interface called UCAIR (User-Centered Adaptive Information
Retrieval) [
        <xref ref-type="bibr" rid="ref7">7</xref>
        ]. Their client-side search agent has the
capability of automatic query reformulation and active reranking of
unseen search results based on a context driven user model.
They evaluated their system by asking 6 graduate students
to work on TREC topic distillation tasks. At the end of
each topic, the volunteers were asked to manually evaluate
the relevance of 30 top ranked search results displayed by the
system. The top results shown are mixed between Google
rankings and UCAIR rankings (some results overlap), and
the evaluators could not distinguish the two. UCAIR
rankings show a 20% increase in precision for the top 20 results.
      </p>
      <p>The Surf Canyon browser extension represents the first
attempt to integrate implicit relevance feedback directly into
the major commercial search engines. Hence, we are able to
evaluate this technology outside of controlled studies. From
a research perspective, this is the first study to investigate
this technology in the context of normal searches by normal
users. The drawback is that we have no chance to collect
a posteori relevancy judgements from the searchers or to
conduct surveys to evaluate the user experience. We can,
however, quickly collect large amounts of user data in order
to evaluate the technology.
1Shen, Tan, and Zhai are co-authors on one Surf Canyon
patent application but were not actively involved in the work
presented here
4.</p>
    </sec>
    <sec id="sec-6">
      <title>TECHNOLOGICAL DETAILS</title>
      <p>Surf Canyon’s technology can be used as both a
traditional web search engine and as a browser extension that
dynamically modifies the search results page from commercial
search engines (currently Google, Yahoo!, and Live Search).
The underlying algorithms in the two cases are mostly
identical. As the data presented was gathered using the browser
extension, we will describe that here.</p>
      <p>Surf Canyon’s browser extension was publicly launched
on February 19, 2008. From that point forward visitors to
the Surf Canyon website2 were invited to download a small
piece of free software that is installed in their browser. The
software works with both Internet Explorer and Firefox.
Although the implementation differs for the two browsers, the
functionality is identical.</p>
      <p>Internet Explorer leads in all current studies of web browser
market share with March 2008 market share estimated
between 60% and 90%. Among users of the Surf Canyon
browser extension, however, about 75% use Firefox. Among
users who merely visit the extension download page, the
breakdown by browser type is nearly 50/50. Part of the
skew towards Firefox in both website visitors and users of the
product can be attributed to the fact that marketing of the
product has been mainly via technology blogs. Readers of
technology blogs are more likely to use operating systems for
which Internet Explorer is not available (e.g. Mac, Linux).
Additionally, we speculate that Firefox may be more
prevalent among readers of technology blogs. The difference
between the fraction of visitors to the site using Firefox (∼50%)
and the fraction of people who install and use the product
using Firefox (∼75%) is likely due to the more widespread
acceptance towards browser extensions in the Firefox
community. The Firefox browser was specifically designed to
have minimal core functionality augmented by browser
addons submitted by the developer community. The
technologies used to implement Internet Explorer browser extensions
are also often used to distribute malware so there may be a
higher level of distrust among IE users.</p>
      <p>Once the browser extension is installed, the user never
needs to visit the company web site again to use the
product. The user enters a Google, Yahoo!, or Live Search web
search query just as they would for any search (using either
the search bar built into the browser or by navigating to
the URL of the search engine). After the initial query, the
search engine results page is returned exactly as it would be
were Surf Canyon not installed (for most users who have not
specified otherwise, the default number of search results is
10). Two minor modifications are made to the SERP. Small
bull’s eyes are placed next to the title hyperlink for each
search result (see Figure 1). Also, the numbered links to
subsequent search engine results pages at the bottom of the
SERP are replaced by a single “More Results” link.</p>
      <p>The client side browser extension is used to communicate
with the central Surf Canyon servers and to dynamically
update the search engine results page. The personalization
algorithms currently reside on the Surf Canyon servers. This
client-server architecture is used primarily to facilitate
optimization of the algorithm and to support active research
studies. Since web search patterns vary widely by user, the
best way to evaluate personalized search algorithms is to
vary the algorithms on the same set of users while
main</p>
      <sec id="sec-6-1">
        <title>2http://www.surfcanyon.com</title>
        <p>implicit relevance feedback - Google Search</p>
        <sec id="sec-6-1-1">
          <title>Web Images Maps News Shopping Gmail more ▼</title>
        </sec>
        <sec id="sec-6-1-2">
          <title>Sign in</title>
          <p>Google
Web
implicit relevance feedback
Search</p>
          <p>Advanced Search
Preferences
Reset recommendations</p>
          <p>Results 1 - 10 of about 1,180,000 for implicit relevance feedback. (0.04 seconds)
Relevance feedback - Wikipedia, the free encyclopedia
The idea behind relevance feedback is to take the results that are initially ... Implicit
feedback is inferred from user behavior, such as noting which ...
en.wikipedia.org/wiki/Relevance_feedback - 19k - Cached - Similar pages
Implicit Relevance Feedback from Eye Movements (ResearchIndex)
We explore the use of eye movements as a source of implicit relevance feedback
information. We construct a controlled information retrieval experiment where ...
citeseer.ist.psu.edu/730378.html - 20k - Cached - Similar pages
Click data as implicit relevance feedback in web search
In this article, we address three issues related to using click data as implicit relevance
feedback: (1) How click data beyond the search results page might ...
portal.acm.org/citation.cfm?id=1224561.1224720 - Similar pages</p>
          <p>Surf Canyon recommends 3 search results:
Using Implicit Relevance Feedback in a Web (ResearchIndex)
The explosive growth of information on the World Wide Web demands effective intelligent
search and filtering methods. Consequently, techniques have been ...
citeseer.ist.psu.edu/572595.html - 20k - Cached - Similar pages</p>
        </sec>
        <sec id="sec-6-1-3">
          <title>More results from citeseer.ist.psu.edu »</title>
          <p>Implicit relevance feedback in interactive music (from page 2)</p>
        </sec>
        <sec id="sec-6-1-4">
          <title>This paper presents methods for correlating a human performer and a synthetic accompaniment based on Implicit Relevance Feedback (IRF) using Graugaard's ... portal.acm.org/citation.cfm?id=1164845 - Similar pages</title>
        </sec>
        <sec id="sec-6-1-5">
          <title>More results from portal.acm.org »</title>
          <p>Scalable Relevance Feedback Using Click-Through Data for Web Image ... (from page 2)</p>
        </sec>
        <sec id="sec-6-1-6">
          <title>File Format: PDF/Adobe Acrobat - View as HTML</title>
          <p>In this paper, we have presented a scalable relevance feedback. mechanism for web
image retrieval. Click-through data is used as. implicit relevance ...
research.microsoft.com/users/leizhang/Paper/ACMMM06-Cheng.pdf - Similar pages</p>
        </sec>
        <sec id="sec-6-1-7">
          <title>More results from research.microsoft.com »</title>
          <p>[PPT] LBSC 796/INFM 718R: Week 8 Relevance Feedback
1 of 3
03/20/2008 10:14 AM
taining an identical user interface. With the client-server
architecture, the implicit relevance feedback algorithms can
be modified without alerting the user to any changes.
Nothing fundamental prevents the technology from becoming
exclusively client side.</p>
          <p>In addition to the ten results displayed by the search
engine to the user, a larger set of results (typically 200) for
the same query is gathered by the server. With few
exceptions, the top 10 links in the larger result set are identical
to the results displayed by the search engine. While the
user reads the search result page, the back-end servers parse
the larger result set and prepare to respond to user actions.
Each user action on the search result page is sent to the
back-end server (note that we are only using the user’s
actions on the SERP for personalization and do not follow the
user after they leave the SERP). For certain actions (select
a link, select a Surf Canyon bull’s eye, ask for more results)
the back end server sends recommended search results to
the browser. The Surf Canyon real-time implicit
personalization algorithm incorporates both the initial rank of the
result and personalized instantaneous relevancies. The
implicit feedback signals used to calculate the real-time search
result ranks are cumulative across all recent related queries
by that user. The algorithm does not, however, utilize any
long-term user profiling or collaborative filtering. The
precise details of the Surf Canyon algorithm are proprietary
and are not important for the evaluation of the technology
presented below. If an undisplayed result from the larger set
of results is deemed by Surf Canyon’s algorithm to be more
relevant than other results displayed below the last selected
link, it is shown as an indented recommendation below the
last selected link.</p>
          <p>The resulting page is shown in Figure 1. Here, the user
entered a query for “implicit relevance feedback” on Google3.
Google returned 10 organic search results (only three of
which are displayed in Figure 1) of the 1,180,000 documents
in their web index that satisfy the query. The user then
selected the third organic search result, a paper from an
ACM conference entitled “Click data as implicit relevance
feedback in web search”. Based on the implicit user signals
(which include interactions with this SERP, recent similar
queries, and interactions with those results pages) the Surf
Canyon algorithm recommends three search results. These
links were initially given a higher initial rank (&gt; 10) by
the Google algorithm in response to the query “implicit
relevance feedback”. The real-time personalization algorithm
has determined, however, that the three recommended links
are more pertinent to this user’s information need at this
particular time than the results displayed by Google with
initial ranks 4-10.</p>
          <p>Recommendations are also generated when a user clicks
on the small bull’s eyes next to the link title. We assume
that a selection of a bull’s eye indicates that the linked
document is similar to but not precisely what the user is looking
for. For the analysis below, up to three recommendations
are generated for each link selection or bull’s eye selection.
Unless the user specifically removes recommended search
results by clicking on the bull’s eye or by clicking the close box,
they remain displayed on the page. Recommendations can
nest up to three levels deep – if the user clicks on the first
recommended result then up to three recommendations are</p>
        </sec>
      </sec>
      <sec id="sec-6-2">
        <title>3http://www.google.com</title>
        <p>generated immediately below this search result.</p>
        <p>At the bottom of the 10 organic search results, there is a
link to get “More Results”. If the user requests the next page
of results, all results shown on the second and subsequent
pages are determined using Surf Canyon’s instantaneous
relevancy algorithm. Unlike the default search engine behavior,
subsequent pages of results are added to the existing page.
After selecting “More Results” links 1-20 are displayed in the
browser, with link 11 focused at the top of the window (the
user needs to scroll up to see links 1-10).
5.</p>
      </sec>
    </sec>
    <sec id="sec-7">
      <title>ANALYSIS OF USER BEHAVIOR</title>
      <p>Most previous studies of Interactive Information Retrieval
systems have used post-search user surveys to evaluate the
efficacy of the systems. These studies also tended to
recruit test subjects and use closed collections and/or
specific research topics. The data presented here was collected
from an anonymous (but not necessarily representative) set
of web surfers during the course of their interactions with
the three leading search engines (Google, Yahoo, and Live
Search). The majority of searches were conducted using
Google. Where possible, we have analyzed the user data
independently for each of the search engines and have not
found any cases where the conclusions drawn from this study
would differ depending on the user’s choice of search
engine. The total number of unique search queries analyzed
was ∼700,000.</p>
      <p>Since the users in this study were acquired primarily from
technology web blogs, their search behavior can be expected
to be significantly different than the average web surfer.
Thus, we cannot evaluate the real-time personalization
technology by comparing to previous studies of web user
behavior. Also, since we have changed the appearance of the
SERP and also dynamically modify the SERP, any metrics
calculated from our data cannot be directly compared to
historical data due to the different user interface.</p>
      <p>
        Surf Canyon only shows recommendations after a bull’s
eye or search result is selected. It is therefore interesting
to investigate how many actions a user makes for a given
query as this tells us how frequently implicit personalization
within the same query can be of benefit. Jansen and Spink
[
        <xref ref-type="bibr" rid="ref8">8</xref>
        ] found from a meta-analysis of search engine log studies
that user interaction with the search engine results pages is
decreasing. In 1997, 71% of searchers viewed beyond the first
page of search results. In 2002 only 27% of searchers looked
past the first page of search results. There is a paucity of
data on the number of web pages visited per search. Jansen
and Spink [
        <xref ref-type="bibr" rid="ref9">9</xref>
        ] reported the mean number of web pages
visited per query to be 2.5 for AllTheWeb searches in 2001,
but they exclude queries where no pages were visited in this
estimate. Analysis of the AOL query logs from 2006 [
        <xref ref-type="bibr" rid="ref10">10</xref>
        ]
gives a mean number of web pages viewed per unique query
of 0.97. For the current data sample, the mean number of
search results visited is 0.56. The comparatively low
number of search results that were selected in the current study
has multiple partial explanations. The search results page
now contains multiple additional links (news, videos) that
are not counted in this study. Additionally, the information
that the user is looking for is often on the SERP (e.g. a
search for a restaurant often produces the map, phone
number, and address). Search engines have replaced bookmarks
and direct URL typing for re-visiting web sites. For such
navigational searches the user will have either one or zero
      </p>
      <p>Number Of Search Results Selected
clicks depending on whether the specific web page is listed
on the SERP. Additionally, it may be that the current
sample of users is biased towards searchers who are less likely to
click on links.</p>
      <p>
        Figure 2 shows the distribution of the total number of
selections per query. 62% of all queries lead to the selection
of zero search results. Since Surf Canyon does nothing until
after the first selection, this number is intrinsic to the current
users interacting with these particular search engines. A
recent study by Downey, Dumais and Horvitz also showed
that after a query the user’s next action is to re-query or end
the search session about half the time [
        <xref ref-type="bibr" rid="ref11">11</xref>
        ]. In our study, only
12% of queries lead to more than one user selection. A goal
of implicit real-time personalization would be to decrease
direct query reformulation and to increase the number of
informational queries that lead to multiple selections. The
current data sample is insufficient to study whether this goal
has been achieved.
      </p>
      <p>In order to evaluate the implicit personalization
technology developed by Surf Canyon we chose to compare the
actions of the same set of users with and without the implicit
personalization technology enabled. Our baseline control
sample was created by randomly replacing recommended
search results with random search results selected from among
the results with initial ranks 11-200. These “Random
Recommendations” were only shown for 5% of the cases where
recommendations were generated. The position (1, 2, or 3)
in the recommendation list was also random. These
random recommendations were not necessarily poor, as they do
come from the list of results generated by the search engine
in response to the query.</p>
      <p>Figure 3 shows the click frequency for Surf Canyon
recommendations as a function of the position of the
recommendation relative to the last selected search result.
Position 1 is immediately below the last selected search result.
Also shown are the click frequencies for “Random
Recommendations” placed at the same positions. In both cases,
the frequency is relative to the total number of
recommendations shown at that position. The increase in click rate
(∼60%) is constant within statistical uncertainties for all
recommended link positions. Note that the
recommendations are generated each time a user selects a link and are
considered to be shown even if the user does not return to the
SERP. The low absolute click rates (3% or less) are due to
the fact that users do not often click on more than one search
result as discussed above. The important point, however, is
that the Surf Canyon implicit relevance feedback
technology increases the click frequency by ∼80% compared to the
links presented without any real-time user-intent modelling.
The relative increase in clickthrough rate is constant (within
statistical errors) for all display positions even though the
absolute clickthrough rates rapidly drop as funciton of
display position.
w/ Implicit Feedback
w/o Implicit Feedback
1
2</p>
      <p>3</p>
      <p>Recommended Link Position</p>
      <p>Figure 4 shows the per query distribution of initial search
result ranks for all selected search links in the current data
sample. The top 10 links are selected most frequently. Search
results beyond 10 are all displayed using Surf Canyon’s
algorithm (either through a bull’s eye selection, a link
selection, or when the user selects more results). For the
results displayed by Surf Canyon (initial ranks &gt; 10), the
selection frequency follows a power-law distribution with
P (IR) = 38% ∗ IR−1.8, where IR is the initial rank.</p>
      <p>As Surf Canyon’s algorithm favors links with higher initial
rank, the click frequency distribution does not fully reflect
the relevancy of the links as a function of initial rank.
Figure 5 shows the probability that a shown recommendation
is clicked as a function of the initial rank. This is only
for recommendations shown in the first position below the
last selected link. After using Surf Canyon’s instantaneous
relevancy algorithm, this probability shows at most a weak
dependence on the initial rank of the search result. The
dotted link shows the result of a linear regression to the data,
P (IR) = 3.2 − (0.0025 ± 0.00101) ∗ IR. When sufficient data
is available we will repeat the same analysis for “Random
Recommendations” as that will give us a user-interface
independent estimate of the relative relevance for deep links
in the search result set before the application of the implicit
feedback algorithms.</p>
      <p>For the second and subsequent results pages, the browser
extension has complete control over all displayed search
results. For a short period of time we produced search
results pages that mixed Surf Canyon’s top ranked results
with results having the top initial ranks from the search
Google w/ Surf Canyon
50
100</p>
      <p>
        150 200
Initial SERP Rank
engine. This procedure was proposed by Joachims as a way
to use clickthrough data to determine relative user
preference between two search engine retrieval algorithms [
        <xref ref-type="bibr" rid="ref12">12</xref>
        ].
Each time a user requests “More Results”, two lists are
generated. The first list (SC) contains the remaining search
results as ranked by the Surf Canyon’s instantaneous
relevancy algorithm. The second list (IR) contains the same set
of results ranked by their initial display rank from the search
engine. The list of results shown to the user is such that the
top kSC and kIR results are displayed from each list, with
|kSC − kIR| &lt; 1. Whenever kSC = kIR the next search
result is taken from one of the lists chosen at random. Thus,
the topmost search result on the second page will reflect
Surf Canyon’s ranking half the time and the initial search
result order half the time. By mixing the search results
this way, the user will see, on average, an equal number of
search results from each ranking algorithm in each position
on the page. The users have no way of determining which
algorithm produced each search result. If the users select
more search results from one ranking algorithm compared
to the other ranking algorithm it demonstrates an absolute
user preference for the retrieval function that led to more
selections.
      </p>
      <p>Figure 6 shows the ratio of link clicks for the two retrieval
functions. IR is the retrieval function based on the result
rank returned from the search engine. SC is the retrieval
function incorporating Surf Canyon’s implicit relevance
feedback technology. The ratio is plotted as a function of the
number of links selected previously for that query.
Previously selected links are generally considered to be positive
content feedback. If, on the other had, no links were selected
then the algorithm bases its decision exclusively on negative
feedback indications (skipped links) and on the user intent
model that may have been developed for similar recent
related queries.</p>
      <p>]
R
I
/C1.4
S
[o1.3
i
t
aR1.2</p>
      <p>We observe that, independent of the number of previous
user link selections in the same query, the number of clicks on
links from the relevance feedback algorithm is higher than
links displayed because of their higher initial rank. This
demonstrates an absolute user preference for the ranking
algorithm that utilizes implicit relevance feedback.
Remarkably, the significant user preference for search results
retrieved using the implicit feedback algorithm is also
apparent when the user had zero positive clickthrough actions on
the first 10 results. After skipping the first 10 results and
asking for a subsequent set of search links, the users are
∼35% more likely to click on the top ranked Surf Canyon
result compared to result # 11 from Google. Clearly, the
searcher is not so interested in search results produced by
the identical algorithm that produced the 10 skipped links
and an update of the user intent model for this query is
appropriate.</p>
    </sec>
    <sec id="sec-8">
      <title>CONCLUSIONS AND FUTURE DIREC</title>
    </sec>
    <sec id="sec-9">
      <title>TIONS</title>
      <p>Surf Canyon is an interactive information retrieval system
that dynamically modifies the SERP from major search
engines based on implicit relevance feedback. This was built
with the goal of relieving the growing user frustration with
the search experience and to help searchers “find what they
need right now”. The system presents recommended search
results based on an instantaneous user-intent model. By
comparing clickthrough rates, it was shown that real-time
implicit personalization can dramatically increase the
relevancy of presented search results.</p>
      <p>Users of web search engines learn to think like the search
engines they are using. As an example, searchers tend to
select words with high IDF (inverse document frequency)
when formulating queries – they naturally select the rarest
terms that they can think of that would be in all documents
they desire. Excellent searchers can often formulate
sufficiently specific queries after multiple iterations such that
they eventually find what they need. Properly implemented
implicit relevance feedback would reduce the need for query
reformulations, but it should be noted that in the current
study most users had not yet adjusted their browsing habits
to the modified behavior of the search engine. By tracking
the current users in the future we hope to see changes in
user behavior that can further improve the utility of this
technology. As the user-intent model is cumulative, more
interaction will produce better recommendations once the
users learn to trust the system.</p>
    </sec>
  </body>
  <back>
    <ref-list>
      <ref id="ref1">
        <mixed-citation>
          [1]
          <string-name>
            <given-names>J.J.</given-names>
            <surname>Rocchio</surname>
          </string-name>
          .
          <article-title>The Smart Retrieval System Experiments in Automatic Document Processing</article-title>
          . Prentice Hall,
          <year>1971</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref2">
        <mixed-citation>
          [2]
          <string-name>
            <given-names>D.</given-names>
            <surname>Harman</surname>
          </string-name>
          .
          <article-title>Relevance feedback revisited</article-title>
          .
          <source>In Proceedings of the Fifteenth International ACM SIGIR Conference</source>
          , pages
          <fpage>1</fpage>
          -
          <lpage>10</lpage>
          ,
          <year>1992</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref3">
        <mixed-citation>
          [3]
          <string-name>
            <given-names>D.</given-names>
            <surname>Kelly</surname>
          </string-name>
          and
          <string-name>
            <given-names>J.</given-names>
            <surname>Teevan</surname>
          </string-name>
          .
          <article-title>Implicit feedback for inferring user preference: A bibliography</article-title>
          .
          <source>In SIGIR Forum</source>
          <volume>37</volume>
          (
          <issue>2</issue>
          ), pages
          <fpage>18</fpage>
          -
          <lpage>28</lpage>
          ,
          <year>2003</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref4">
        <mixed-citation>
          [4]
          <string-name>
            <given-names>Thorsten</given-names>
            <surname>Joachims</surname>
          </string-name>
          , Laura Granka, Bing Pan, Helene Hembrooke, and
          <string-name>
            <given-names>Geri</given-names>
            <surname>Gay</surname>
          </string-name>
          .
          <article-title>Accurately interpreting clickthrough data as implicit feedback</article-title>
          .
          <source>In SIGIR '05</source>
          ,
          <year>2005</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref5">
        <mixed-citation>
          [5]
          <string-name>
            <given-names>Steve</given-names>
            <surname>Fox</surname>
          </string-name>
          , Kuldeep Karnawat, Mark Mydland, Susan Dumais, and
          <string-name>
            <given-names>Thomas</given-names>
            <surname>White</surname>
          </string-name>
          .
          <article-title>Evaluating implicit measures to improve web search</article-title>
          .
          <source>ACM Transactions on Information Systems</source>
          ,
          <volume>23</volume>
          (
          <issue>2</issue>
          ):
          <fpage>147</fpage>
          -
          <lpage>168</lpage>
          ,
          <year>April 2005</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref6">
        <mixed-citation>
          [6]
          <string-name>
            <given-names>Xuehua</given-names>
            <surname>Shen</surname>
          </string-name>
          ,
          <string-name>
            <given-names>Bin</given-names>
            <surname>Tan</surname>
          </string-name>
          , and ChengXiang Zhai.
          <article-title>Context-sensitive information retrieval using implicit feedback</article-title>
          .
          <source>In SIGIR '05</source>
          ,
          <year>2005</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref7">
        <mixed-citation>
          [7]
          <string-name>
            <given-names>Xuehua</given-names>
            <surname>Shen</surname>
          </string-name>
          ,
          <string-name>
            <given-names>Bin</given-names>
            <surname>Tan</surname>
          </string-name>
          , and ChengXiang Zhai.
          <article-title>Implicit user modelling for personalized search</article-title>
          .
          <source>In CIKM '05</source>
          ,
          <year>2005</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref8">
        <mixed-citation>
          [8]
          <string-name>
            <given-names>B.</given-names>
            <surname>Jansen</surname>
          </string-name>
          and
          <string-name>
            <given-names>A.</given-names>
            <surname>Spink</surname>
          </string-name>
          .
          <article-title>How are we searching the world wide web?: a comparison of nine search engine transaction logs</article-title>
          .
          <source>Information Processing and Management</source>
          ,
          <volume>42</volume>
          (
          <issue>1</issue>
          ):
          <fpage>248</fpage>
          -
          <lpage>263</lpage>
          ,
          <year>2006</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref9">
        <mixed-citation>
          [9]
          <string-name>
            <given-names>B.</given-names>
            <surname>Jansen</surname>
          </string-name>
          and
          <string-name>
            <given-names>A.</given-names>
            <surname>Spink</surname>
          </string-name>
          .
          <article-title>An analysis of web documents retrieved and viewed</article-title>
          .
          <source>In The 4th International Conference on Internet Computing</source>
          , pages
          <fpage>65</fpage>
          -
          <lpage>69</lpage>
          ,
          <year>2003</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref10">
        <mixed-citation>
          [10]
          <string-name>
            <given-names>G.</given-names>
            <surname>Pass</surname>
          </string-name>
          ,
          <string-name>
            <given-names>A.</given-names>
            <surname>Chowdhury</surname>
          </string-name>
          , and
          <string-name>
            <given-names>C.</given-names>
            <surname>Torgeson</surname>
          </string-name>
          .
          <article-title>A picture of search</article-title>
          .
          <source>In The First International Conference on Scalable Information Systems</source>
          ,
          <year>2006</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref11">
        <mixed-citation>
          [11]
          <string-name>
            <given-names>D.</given-names>
            <surname>Downey</surname>
          </string-name>
          ,
          <string-name>
            <given-names>S.</given-names>
            <surname>Dumais</surname>
          </string-name>
          , and
          <string-name>
            <given-names>E.</given-names>
            <surname>Horvitz</surname>
          </string-name>
          .
          <article-title>Studies of web search with common and rare queries</article-title>
          .
          <source>In SIGIR '07</source>
          ,
          <year>2007</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref12">
        <mixed-citation>
          [12]
          <string-name>
            <given-names>T.</given-names>
            <surname>Joachims</surname>
          </string-name>
          .
          <article-title>Unbiased evaluation of retrieval quality using clickthrough data</article-title>
          .
          <source>In SIGIR Workshop on Mathematical/Formal Methods in Information Retrieval</source>
          ,
          <year>2002</year>
          .
        </mixed-citation>
      </ref>
    </ref-list>
  </back>
</article>