arXiv · 2005.00190
Evaluating Neural Machine Comprehension Model Robustness to Noisy Inputs and Adversarial Attacks
Abstract
We evaluate machine comprehension models' robustness to noise and adversarial attacks by performing novel perturbations at the character, word, and sentence level. We experiment with different amounts of perturbations to examine model confidence and misclassification rate, and contrast model performance in adversarial training with different embedding types on two benchmark datasets. We demonstrate improving model performance with ensembling. Finally, we analyze factors that effect model behavior under adversarial training and develop a model to predict model errors during adversarial attacks.
Explore related subjects
Keep this discovery
Winston Wu, Dustin Arendt, Svitlana Volkova. 2020-05-01. Evaluating Neural Machine Comprehension Model Robustness to Noisy Inputs and Adversarial Attacks. https://arxiv.org/abs/2005.00190
Cite the original work for its findings. Save a collection to share your selection of sources.