arXiv · 2610.03142
CalCErt: Bin-wise Certification of Confidence Calibration in Medical Image Classification
Abstract
Deep neural networks remain vulnerable to adversarial perturbations, which can distort not only predictions but also confidence scores, undermining uncertainty calibration. While existing certification methods focus on preserving the predicted category, providing guarantees on how calibration behaves under adversarial attacks remains overlooked. In this work, we introduce CalCErt, a simple and efficient post-hoc strategy that certifies bin-wise confidence calibration for any pretrained differentiable classifier. Our approach combines empirical calibration estimates, statistical concentration bounds, and local Lipschitz estimates of the confidence function to derive data-dependent upper bounds on worst-case miscalibration within an $ell_2$-ball of radius R. We evaluate CalCErt across 11 medical image classification tasks and multiple adversarial perturbations, demonstrating substantially higher certified coverage than baseline strategies while maintaining competitive tightness. Our code is available at https://github.com/leofillioux/calcert.
Explore related subjects
Keep this discovery
Explore connections, maps & timelines
Leo Fillioux, Stergios Christodoulidis, Maria Vakalopoulou, Jose Dolz. 2026-10-02. CalCErt: Bin-wise Certification of Confidence Calibration in Medical Image Classification. https://arxiv.org/abs/2610.03142
Cite the original work for its findings. Save a collection to share your selection of sources.