ImageNet-Patch: A Dataset for Benchmarking Machine Learning Robustness against Adversarial Patches
arXiv:2203.04412 · doi:10.1016/j.patcog.2022.109064
Abstract
Adversarial patches are optimized contiguous pixel blocks in an input image that cause a machine-learning model to misclassify it. However, their optimization is computationally demanding, and requires careful hyperparameter tuning, potentially leading to suboptimal robustness evaluations. To overcome these issues, we propose ImageNet-Patch, a dataset to benchmark machine-learning models against adversarial patches. It consists of a set of patches, optimized to generalize across different models, and readily applicable to ImageNet data after preprocessing them with affine transformations. This process enables an approximate yet faster robustness evaluation, leveraging the transferability of adversarial perturbations. We showcase the usefulness of this dataset by testing the effectiveness of the computed patches against 127 models. We conclude by discussing how our dataset could be used as a benchmark for robustness, and how our methodology can be generalized to other domains. We open source our dataset and evaluation code at https://github.com/pralab/ImageNet-Patch.
Published in Pattern Recognition. DOI: https://doi.org/10.1016/j.patcog.2022.109064
References in corpus (5)
Cited by in corpus (5)
- CARLA-GeAR: a Dataset Generator for a Systematic Evaluation of Adversarial Robustness of Vision Models
- Topological safeguard for evasion attack interpreting the neural networks' behavior
- Benchmarking the Spatial Robustness of DNNs via Natural and Adversarial Localized Corruptions
- SpaNN: Detecting Multiple Adversarial Patches on CNNs by Spanning Saliency Thresholds
- Unveiling Hidden Vulnerabilities in Digital Human Generation via Adversarial Attacks