Note: This bibliographic page is archived and will no longer be updated. For an up-to-date list of publications from the Music Technology Group see the Publications list .

Toward Estimating the Rank Correlation between the Test Collection Results and the True System Performance

Title Toward Estimating the Rank Correlation between the Test Collection Results and the True System Performance
Publication Type Conference Paper
Year of Publication 2016
Conference Name International ACM SIGIR Conference on Research and Development in Information Retrieval
Authors Urbano, J. , & Marrero M.
Pagination 1033-1036
Abstract The Kendall tau and AP rank correlation coefficients have become mainstream in Information Retrieval research for comparing the rankings of systems produced by two different evaluation conditions, such as different effectiveness measures or pool depths. However, in this paper we focus on the expected rank correlation between the mean scores observed with a test collection and the true, unobservable means under the same conditions. In particular, we propose statistical estimators of tau and AP correlations following both parametric and non-parametric approaches, and with special emphasis on small topic sets. Through large scale simulation with TREC data, we study the error and bias of the estimators. In general, such estimates of expected correlation with the true ranking may accompany the results reported from an evaluation experiment, as an easy to understand figure of reliability. All the results in this paper are fully reproducible with data and code available online.