REVIEW 5 cited by
Explainable Equivariant Neural Networks for Particle Physics: PELICAN
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
Signed reviews
abstract
PELICAN is a novel permutation equivariant and Lorentz invariant or covariant aggregator network designed to overcome common limitations found in architectures applied to particle physics problems. Compared to many approaches that use non-specialized architectures that neglect underlying physics principles and require very large numbers of parameters, PELICAN employs a fundamentally symmetry group-based architecture that demonstrates benefits in terms of reduced complexity, increased interpretability, and raw performance. We present a comprehensive study of the PELICAN algorithm architecture in the context of both tagging (classification) and reconstructing (regression) Lorentz-boosted top quarks, including the difficult task of specifically identifying and measuring the $W$-boson inside the dense environment of the Lorentz-boosted top-quark hadronic final state. We also extend the application of PELICAN to the tasks of identifying quark-initiated vs.~gluon-initiated jets, and a multi-class identification across five separate target categories of jets. When tested on the standard task of Lorentz-boosted top-quark tagging, PELICAN outperforms existing competitors with much lower model complexity and high sample efficiency. On the less common and more complex task of 4-momentum regression, PELICAN also outperforms hand-crafted, non-machine learning algorithms. We discuss the implications of symmetry-restricted architectures for the wider field of machine learning for physics.
Forward citations
Cited by 5 Pith papers
-
Predict before you train: Scaling Laws for particle physics foundation models
A Chinchilla-style law fit on ParticleViT runs below 10^19 FLOPs predicts held-out pretraining loss within ~1% at >100× compute and tracks downstream jet-tagging rejection.
-
Explicit or Implicit? Encoding Physics at the Precision Frontier
On three precision classification tasks — reweighting-based unfolding, likelihood-ratio estimation, and weakly supervised anomaly detection — a Lorentz-equivariant transformer and a pretrained foundation model perform...
-
Graph theory inspired anomaly detection at the LHC
Sparse globally rigid graph representations of jets, combined with roughly 30 reclustered subjets, improve graph autoencoder anomaly detection on the LHC Olympics benchmark.
-
Tagging fully hadronic exotic decays of the vectorlike $\mathbf{B}$ quark using a graph neural network
A GNN-based search could give the HL-LHC an exclusion reach near 2.4 TeV for vectorlike B quarks decaying fully hadronically through b plus a singlet scalar, with performance comparable to semileptonic searches.
-
Transformer networks for Heavy flavor jet tagging
A review of transformer-based jet tagging that highlights the authors' CA-Mixer network as a state-of-the-art, faster alternative to Particle Transformer.
Discussion (0). Continue with ORCID to comment.