conference-paper Open access

Monte Carlo Estimates of Evaluation Metric Error and Bias

Research footprint

At a glance

Citations
1
References
17
Comments
0
Paper overview

Abstract

Traditional offline evaluations of recommender systems apply metrics from machine learning and information retrieval in settings where their underlying assumptions no longer hold. This results in significant error and bias in measures of top-N recommendation performance, such as precision, recall, and nDCG. Several of the specific causes of these errors, including popularity bias and misclassified decoy items, are well-explored in the existing literature. In this paper we survey a range of work on identifying and addressing these problems, and report on our work in progress to simulate the recommender data generation and evaluation processes to quantify the extent of evaluation metric errors and assess their sensitivity to various assumptions.

Record transparency

Publication details

DOI
10.18122/cs_facpubs/148/boisestate
OpenAlex
W2889128918
Document type
conference-paper
Language
EN
Last metadata update
Community

Comments

Log in to join the discussion.

  1. No comments yet. Start the discussion.