Particle Transformer for Jet Tagging
Huilin Qu, Congqiao Li, Sitian Qian
Abstract
Jet tagging is a critical yet challenging classification task in particle physics. While deep learning has transformed jet tagging and significantly improved performance, the lack of a large-scale public dataset impedes further enhancement. In this work, we present JETCLASS, a new comprehensive dataset for jet tagging. The JETCLASS dataset consists of 100 M jets, about two orders of magnitude larger than existing public datasets. A total of 10 types of jets are simulated, including several types unexplored for tagging so far. Based on the large dataset, we propose a new Transformer-based architecture for jet tagging, called Particle Transformer (ParT). By incorporating pairwise particle interactions in the attention mechanism, ParT achieves higher tagging performance than a plain Transformer and surpasses the previous state-of-the-art, ParticleNet, by a large margin. The pre-trained ParT models, once fine-tuned, also substantially enhance the performance on two widely adopted jet tagging benchmarks. The dataset, code and models are publicly available at https: //github.com/jet-universe/ particle_transformer.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 21ca5959-945e-49f0-975a-45fa1de9bf23Cited by top-tier papers6
- Lorentz-Equivariant Geometric Algebra Transformers for High-Energy PhysicsJonas Spinner, Victor Bresó, Pim de Haan, Tilman Plehn et al.NeurIPS 2024 · 65 citations
- A General Framework for Equivariant Neural Networks on Reductive Lie GroupsIlyes Batatia, Mario Geiger, Jose M. Munoz, Tess E. Smidt et al.NeurIPS 2023 · 28 citations
- Lorentz Local Canonicalization: How to make any Network Lorentz-EquivariantJonas Spinner, Luigi Favaro, Peter Lippmann, Sebastian Pitz et al.NeurIPS 2025 · 19 citations
- Locality-Sensitive Hashing-Based Efficient Point Transformer with Applications in High-Energy PhysicsSiqi Miao, Zhiyuan Lu, Mia Liu, Javier M. Duarte et al.ICML 2024 · 13 citations
- AutoSciDACT: Automated Scientific Discovery through Contrastive Embedding and Hypothesis TestingSamuel Bright-Thonney, Christina Reissel, Gaia Grosso, Nathaniel Woodward et al.NeurIPS 2025 · 4 citations
Builds on8
- Swin Transformer: Hierarchical Vision Transformer using Shifted WindowsZe Liu, Yutong Lin, Yue Cao, Han Hu et al.ICCV 2021 · 31,683 citations
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn et al.ICLR 2021 · 21,477 citations
- Informer: Beyond Efficient Transformer for Long Sequence Time-Series ForecastingHaoyi Zhou, Shanghang Zhang, Jieqi Peng, Shuai Zhang et al.AAAI 2021 · 7,289 citations
- Reformer: The Efficient TransformerNikita Kitaev, Lukasz Kaiser, Anselm LevskayaICLR 2020 · 2,878 citations
- On the Variance of the Adaptive Learning Rate and BeyondLiyuan Liu, Haoming Jiang, Pengcheng He, Weizhu Chen et al.ICLR 2020 · 2,210 citations
Related papers
- Particle Cloud Generation with Message Passing Generative Adversarial NetworksRaghav Kansal, Javier M. Duarte, Hao Su, Breno Orzari et al.NeurIPS 2021 · 89 citations
- Lorentz Group Equivariant Neural Network for Particle PhysicsAlexander Bogatskiy, Brandon M. Anderson, Jan T. Offermann, Marwah Roussi et al.ICML 2020 · 164 citations
- Jetfire: Efficient and Accurate Transformer Pretraining with INT8 Data Flow and Per-Block QuantizationHaocheng Xi, Yuxiang Chen, Kang Zhao, Kai Jun Teh et al.ICML 2024 · 35 citations
- FM4NPP: A Scaling Foundation Model for Nuclear and Particle PhysicsDavid Keetae Park, Shuhang Li, Yi Huang, Xihaier Luo et al.ICLR 2026 · 6 citations
- Scaling Rich Style-Prompted Text-to-Speech DatasetsAnuj Diwan, Zhisheng Zheng, David Harwath, Eunsol ChoiEMNLP 2025 · 2 citations
