Enhancing generalization in high energy physics using white-box adversarial attacks
arXiv:2411.09296 · doi:10.1103/PhysRevD.112.016004
Abstract
Machine learning is becoming increasingly popular in the context of particle physics. Supervised learning, which uses labeled Monte Carlo (MC) simulations, remains one of the most widely used methods for discriminating signals beyond the Standard Model. However, this paper suggests that supervised models may depend excessively on artifacts and approximations from Monte Carlo simulations, potentially limiting their ability to generalize well to real data. This study aims to enhance the generalization properties of supervised models by reducing the sharpness of local minima. It reviews the application of four distinct white-box adversarial attacks in the context of classifying Higgs boson decay signals. The attacks are divided into weight-space attacks and feature-space attacks. To study and quantify the sharpness of different local minima, this paper presents two analysis methods: gradient ascent and reduced Hessian eigenvalue analysis. The results show that white-box adversarial attacks significantly improve generalization performance, albeit with increased computational complexity.
14 pages, 7 figures, 10 tables, 3 algorithms, published in Physical Review D (PRD), presented at the ML4Jets 2024 conference
References in corpus (17)
- An Introduction to PYTHIA 8.2
- Herwig++ Physics and Manual
- Herwig 7.0 / Herwig++ 3.0 Release Note
- Energy Flow Networks: Deep Sets for Particle Jets
- The Machine Learning Landscape of Top Taggers
- An Efficient Lorentz Equivariant Graph Neural Network for Jet Tagging
- Classifying Anomalies THrough Outer Density Estimation (CATHODE)
- DisCo Fever: Robust Networks Through Distance Correlation
- PC-JeDi: Diffusion for Particle Cloud Generation in High Energy Physics
- Explainable Equivariant Neural Networks for Particle Physics: PELICAN
- The Landscape of Unfolding with Machine Learning
- -Flows: Fast and improved neutrino reconstruction in multi-neutrino final states with conditional normalizing flows
- Re-Simulation-based Self-Supervised Learning for Pre-Training Foundation Models
- Enhancing searches for resonances with machine learning and moment decomposition
- Accuracy versus precision in boosted top tagging with the ATLAS detector
- Improving Robustness of Jet Tagging Algorithms with Adversarial Training
- A search for R-parity-violating supersymmetry in final states containing many jets in collisions at with the ATLAS detector