Interpreting Transformers for Jet Tagging

December 04, 2024 · Declared Dead · 🏛 arXiv.org

"No code URL or promise found in abstract"

Evidence collected by the PWNC Scanner

Authors Aaron Wang, Abhijith Gandrakota, Jennifer Ngadiuba, Vivekanand Sahu, Priyansh Bhatnagar, Elham E Khoda, Javier Duarte arXiv ID 2412.03673 Category hep-ph Cross-listed cs.LG, hep-ex, physics.data-an Citations 6 Venue arXiv.org Last Checked 3 months ago

Abstract

Machine learning (ML) algorithms, particularly attention-based transformer models, have become indispensable for analyzing the vast data generated by particle physics experiments like ATLAS and CMS at the CERN LHC. Particle Transformer (ParT), a state-of-the-art model, leverages particle-level attention to improve jet-tagging tasks, which are critical for identifying particles resulting from proton collisions. This study focuses on interpreting ParT by analyzing attention heat maps and particle-pair correlations on the $η$-$φ$ plane, revealing a binary attention pattern where each particle attends to at most one other particle. At the same time, we observe that ParT shows varying focus on important particles and subjets depending on decay, indicating that the model learns traditional jet substructure observables. These insights enhance our understanding of the model's internal workings and learning process, offering potential avenues for improving the efficiency of transformer architectures in future high-energy physics applications.

📄 View on arXiv 🌐 View on ar5iv 📑 PDF 🎉 Report Code Found

Community Contributions

Found the code? Know the venue? Think something is wrong? Let us know!

📜 Similar Papers

In the same crypt — hep-ph

R.I.P. 👻 Ghosted

How to GAN away Detector Effects

Marco Bellagente, Anja Butter, ... (+3 more)

hep-ph 🏛 SciPost Physics 📚 100 cites 6 years ago

R.I.P. 👻 Ghosted

CaloMan: Fast generation of calorimeter showers with density estimation on learned manifolds

Jesse C. Cresswell, Brendan Leigh Ross, ... (+4 more)

hep-ph 🏛 arXiv 📚 53 cites 3 years ago

R.I.P. 👻 Ghosted

An unfolding method based on conditional Invertible Neural Networks (cINN) using iterative training

Mathias Backes, Anja Butter, ... (+2 more)

hep-ph 🏛 SciPost Physics Core 📚 52 cites 3 years ago

R.I.P. 👻 Ghosted

PELICAN: Permutation Equivariant and Lorentz Invariant or Covariant Aggregator Network for Particle Physics

Alexander Bogatskiy, Timothy Hoffman, ... (+2 more)

hep-ph 🏛 arXiv 📚 39 cites 3 years ago

R.I.P. 👻 Ghosted

Stacking machine learning classifiers to identify Higgs bosons at the LHC

Alexandre Alves

hep-ph 🏛 arXiv 📚 32 cites 9 years ago

R.I.P. 👻 Ghosted

The Power of Genetic Algorithms: what remains of the pMSSM?

Steven Abel, David G. Cerdeno, Sandra Robles

hep-ph 🏛 arXiv 📚 20 cites 8 years ago

Died the same way — 👻 Ghosted

R.I.P. 👻 Ghosted

Federated Learning: Strategies for Improving Communication Efficiency

Jakub Konečný, H. Brendan McMahan, ... (+4 more)

cs.LG 🏛 arXiv 📚 5.2K cites 9 years ago

R.I.P. 👻 Ghosted

In-Datacenter Performance Analysis of a Tensor Processing Unit

Norman P. Jouppi, Cliff Young, ... (+73 more)

cs.AR 🏛 ISCA 📚 5.1K cites 9 years ago

R.I.P. 👻 Ghosted

Deep Convolutional Neural Networks for Computer-Aided Detection: CNN Architectures, Dataset Characteristics and Transfer Learning

Hoo-Chang Shin, Holger R. Roth, ... (+7 more)

cs.CV 🏛 IEEE TMI 📚 4.9K cites 10 years ago

R.I.P. 👻 Ghosted

Explanation in Artificial Intelligence: Insights from the Social Sciences

Tim Miller

cs.AI 🏛 AI 📚 4.9K cites 8 years ago