AG-2024.12-1290·hep-ph·cross-listed: cs.LGhep-exphysics.data-an
Interpreting Transformers for Jet Tagging
Authors
- Aaron Wang
- Abhijith Gandrakota
- Jennifer Ngadiuba
- Vivekanand Sahu
- Priyansh Bhatnagar
- Elham E Khoda
- Javier Duarte
Abstract
Machine learning (ML) algorithms, particularly attention-based transformer models, have become indispensable for analyzing the vast data generated by particle physics experiments like ATLAS and CMS at the CERN LHC. Particle Transformer (ParT), a state-of-the-art model, leverages particle-level attention to improve jet-tagging tasks, which are critical for identifying particles resulting from proton collisions. This study focuses on interpreting ParT by analyzing attention heat maps and particle-pair correlations on the $η$-$φ$ plane, revealing a binary attention pattern where each particle attends to at most one other particle. At the same time, we observe that ParT shows varying focus on important particles and subjets depending on decay, indicating that the model learns traditional jet substructure observables. These insights enhance our understanding of the model's internal workings and learning process, offering potential avenues for improving the efficiency of transformer architectures in future high-energy physics applications.
Submitted
4 December 20241 year ago
Version
v1
License
CC-BY-4.0
DOI
10.48550/arXiv.2412.03673
Summary
Researchers decoded how Particle Transformer, an AI model used to identify particles from collisions at CERN, actually works by analyzing its internal attention patterns, finding it learns to focus on physically meaningful particle groupings.
- The transformer's 'attention' mechanism—how it weighs relationships between particles—follows a simple binary pattern where each particle essentially pairs with just one other, making the model's decision-making more interpretable than a black box.
- The model implicitly rediscovers classical jet substructure observables (mathematical features physicists have long used to categorize jets), suggesting transformers naturally converge on physically meaningful representations without explicit instruction.
- Understanding what transformers learn internally could help physicists redesign these architectures to be more efficient for particle physics, where computing power is precious and interpretability matters for discovery.
curious · generated by claude-haiku-4-5
Chat with this PDF
Ask questions, probe assumptions, request a plain-English summary. Answers cite sections from the preprint itself.
Community
Questions and answers about this paper from other readers. No formal peer review — just a place to think out loud.