Showing hep-phShow all
2 papers · 1 filter
hep-ph2025
Why Is Attention Sparse In Particle Transformer?
Timothy Legge, Aaron Wang, Jacob Ortiz +7
Transformer-based models have achieved state-of-the-art performance in jet tagging at the CERN Large Hadron Collider (LHC), with the Particle Transformer (ParT) representing a lead…
hep-ph2024
Interpreting Transformers for Jet Tagging
Aaron Wang, Abhijith Gandrakota, Jennifer Ngadiuba +4
Machine learning (ML) algorithms, particularly attention-based transformer models, have become indispensable for analyzing the vast data generated by particle physics experiments l…