Pooling-based continuous evaluation of information retrieval systems

Tonon, Alberto; Demartini, Gianluca; Cudré-Mauroux, Philippe

Information

Fulltext

Pooling-based continuous evaluation of information retrieval systems

Tonon, Alberto ; Demartini, Gianluca ; Cudré-Mauroux, Philippe

In: Information Retrieval Journal, 2015, vol. 18, no. 5, p. 445-472

Add to personal list

Title

Pooling-based continuous evaluation of information retrieval systems

Author

Tonon, Alberto. University of Fribourg, Fribourg, Switzerland
Demartini, Gianluca. University of Sheffield, South Yorkshire, UK
Cudré-Mauroux, Philippe. University of Fribourg, Fribourg, Switzerland

Document Type

Postprint

Language

English

Published in

Information Retrieval Journal, 2015, vol. 18, no. 5, p. 445-472. Springer Netherlands

Other Version

Publisher's version : https://doi.org/10.1007/s10791-015-9266-y

Classification

Computer science

Keywords

Information retrieval evaluation ; Crowdsourcing ; Continuous evaluation ; Poolingtechniques

OAI-PMH Identifier

oai:doc.rero.ch:331676

Summary

The dominant approach to evaluate the effectiveness of information retrieval (IR) systems is by means of reusable test collections built following the Cranfield paradigm. In this paper, we propose a new IR evaluation methodology based on pooled test-collections and on the continuous use of either crowdsourcing or professional editors to obtain relevance judgements. Instead of building a static collection for a finite set of systems known a priori, we propose an IR evaluation paradigm where retrieval approaches are evaluated iteratively on the same collection. Each new retrieval technique takes care of obtaining its missing relevance judgements and hence contributes to augmenting the overall set of relevance judgements of the collection. We also propose two metrics: Fairness Score, and opportunistic number of relevant documents, which we then use to define new pooling strategies. The goal of this work is to study the behavior of standard IR metrics, IR system ranking, and of several pooling techniques in a continuous evaluation context by comparing continuous and non-continuous evaluation results on classic test collections. We both use standard and crowdsourced relevance judgements, and we actually run a continuous evaluation campaign over several existing IR systems.

Pooling-based continuous evaluation of information retrieval systems

Tonon, Alberto ; Demartini, Gianluca ; Cudré-Mauroux, Philippe

In: Information Retrieval Journal, 2015, vol. 18, no. 5, p. 445-472

See also

Export as

Pooling-based continuous evaluation of information retrieval systems

Tonon, Alberto ; Demartini, Gianluca ; Cudré-Mauroux, Philippe

In: Information Retrieval Journal, 2015, vol. 18, no. 5, p. 445-472

See also

Links

Share

Export as