arXiv · 2602.02229
Prediction-Powered Risk Monitoring of Deployed Models for Detecting Harmful Distribution Shifts
Abstract
We study the problem of monitoring model performance in dynamic environments where labeled data are limited. To this end, we propose prediction-powered risk monitoring (PPRM), a semi-supervised risk-monitoring approach based on prediction-powered inference (PPI). PPRM constructs anytime-valid lower bounds on the running risk by combining synthetic labels with a small set of true labels. Harmful shifts are detected via a threshold-based comparison with an upper bound on the nominal risk, satisfying assumption-free finite-sample guarantees on the type-I error. We demonstrate the effectiveness of PPRM through extensive experiments on image classification, large language model (LLM), and telecommunications monitoring tasks.
Explore related subjects
Keep this discovery
Guangyi Zhang, Yunlong Cai, Guanding Yu, Osvaldo Simeone. 2026-02-02. Prediction-Powered Risk Monitoring of Deployed Models for Detecting Harmful Distribution Shifts. https://arxiv.org/abs/2602.02229
Cite the original work for its findings. Save a collection to share your selection of sources.