arXiv · 1508.05503
Empirical AUC for evaluating probabilistic forecasts
Abstract
Scoring functions are used to evaluate and compare partially probabilistic forecasts. We investigate the use of rank-sum functions such as empirical Area Under the Curve (AUC), a widely-used measure of classification performance, as a scoring function for the prediction of probabilities of a set of binary outcomes. It is shown that the AUC is not generally a proper scoring function, that is, under certain circumstances it is possible to improve on the expected AUC by modifying the quoted probabilities from their true values. However with some restrictions, or with certain modifications, it can be made proper.
Explore related subjects
Keep this discovery
Simon Byrne. 2015-08-22. Empirical AUC for evaluating probabilistic forecasts. https://doi.org/10.1214/16-ejs1109
Cite the original work for its findings. Save a collection to share your selection of sources.