Multitask Learning Strengthens Adversarial Robustness
arXiv:2007.07236
Abstract
Although deep networks achieve strong accuracy on a range of computer vision benchmarks, they remain vulnerable to adversarial attacks, where imperceptible input perturbations fool the network. We present both theoretical and empirical analyses that connect the adversarial robustness of a model to the number of tasks that it is trained on. Experiments on two datasets show that attack difficulty increases as the number of target tasks increase. Moreover, our results suggest that when models are trained on multiple tasks at once, they become more robust to adversarial attacks on individual tasks. While adversarial defense remains an open challenge, our results suggest that deep networks are vulnerable partly because they are trained on too few tasks.
References in corpus (6)
- Theoretically Principled Trade-off between Robustness and Accuracy
- One Model To Learn Them All
- Houdini: Fooling Deep Structured Prediction Models
- AdvSPADE: Realistic Unrestricted Attacks for Semantic Segmentation
- Generalization in multitask deep neural classifiers: a statistical physics approach
- Live Trojan Attacks on Deep Neural Networks