arXiv · 2501.11795
Provably effective detection of effective data poisoning attacks
Abstract
This paper establishes a mathematically precise definition of dataset poisoning attack and proves that the very act of effectively poisoning a dataset ensures that the attack can be effectively detected. On top of a mathematical guarantee that dataset poisoning is identifiable by a new statistical test that we call the Conformal Separability Test, we provide experimental evidence that we can adequately detect poisoning attempts in the real world.
Explore related subjects
Keep this discovery
Jonathan Gallagher, Yasaman Esfandiari, Callen MacPhee, Michael Warren. 2025-01-21. Provably effective detection of effective data poisoning attacks. https://arxiv.org/abs/2501.11795
Cite the original work for its findings. Save a collection to share your selection of sources.