Evaluating diversified search results using per-intent graded relevance

Tetsuya Sakai, Ruihua Song

研究成果: Conference contribution

77 被引用数 (Scopus)

抄録

Search queries are often ambiguous and/or underspecified. To accomodate different user needs, search result diversification has received attention in the past few years. Accordingly, several new metrics for evaluating diversification have been proposed, but their properties are little understood. We compare the properties of existing metrics given the premises that (1) queries may have multiple intents; (2) the likelihood of each intent given a query is available; and (3) graded relevance assessments are available for each intent. We compare a wide range of traditional and diversified IR metrics after adding graded relevance assessments to the TREC 2009 Web track diversity task test collection which originally had binary relevance assessments. Our primary criterion is discriminative power, which represents the reliability of a metric in an experiment. Our results show that diversified IR experiments with a given number of topics can be as reliable as traditional IR experiments with the same number of topics, provided that the right metrics are used. Moreover, we compare the intuitiveness of diversified IR metrics by closely examining the actual ranked lists from TREC. We show that a family of metrics called D#-measures have several advantages over other metrics such as α-nDCG and Intent-Aware metrics.

元の言語English
ホスト出版物のタイトルSIGIR'11 - Proceedings of the 34th International ACM SIGIR Conference on Research and Development in Information Retrieval
出版者Association for Computing Machinery
ページ1043-1052
ページ数10
ISBN(印刷物)9781450309349
DOI
出版物ステータスPublished - 2011 1 1
外部発表Yes
イベント34th International ACM SIGIR Conference on Research and Development in Information Retrieval, SIGIR 2011 - Beijing, China
継続期間: 2011 7 242011 7 28

出版物シリーズ

名前SIGIR'11 - Proceedings of the 34th International ACM SIGIR Conference on Research and Development in Information Retrieval

Conference

Conference34th International ACM SIGIR Conference on Research and Development in Information Retrieval, SIGIR 2011
China
Beijing
期間11/7/2411/7/28

ASJC Scopus subject areas

  • Information Systems

フィンガープリント 「Evaluating diversified search results using per-intent graded relevance」の研究トピックを掘り下げます。これらがまとまってユニークなフィンガープリントを構成します。

引用スタイル