arXiv · 2403.19050
Detecting Generative Parroting through Overfitting Masked Autoencoders
Abstract
The advent of generative AI models has revolutionized digital content creation, yet it introduces challenges in maintaining copyright integrity due to generative parroting, where models mimic their training data too closely. Our research presents a novel approach to tackle this issue by employing an overfitted Masked Autoencoder (MAE) to detect such parroted samples effectively. We establish a detection threshold based on the mean loss across the training dataset, allowing for the precise identification of parroted content in modified datasets. Preliminary evaluations demonstrate promising results, suggesting our method's potential to ensure ethical use and enhance the legal compliance of generative models.
Explore related subjects
Keep this discovery
Explore connections, maps & timelines
Saeid Asgari Taghanaki, Joseph Lambourne. 2024-03-27. Detecting Generative Parroting through Overfitting Masked Autoencoders. https://arxiv.org/abs/2403.19050
Cite the original work for its findings. Save a collection to share your selection of sources.