Pith. sign in

REVIEW 12 cited by

Masked Particle Modeling on Sets: Towards Self-Supervised High Energy Physics Foundation Models

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2401.13537 v3 pith:IRPPFDYJ submitted 2024-01-24 hep-ph cs.LGhep-exphysics.data-an

classification hep-phcs.LGhep-exphysics.data-an
keywords maskedsetsdataenergyhighmodelingphysicsself-supervised
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

We propose masked particle modeling (MPM) as a self-supervised method for learning generic, transferable, and reusable representations on unordered sets of inputs for use in high energy physics (HEP) scientific data. This work provides a novel scheme to perform masked modeling based pre-training to learn permutation invariant functions on sets. More generally, this work provides a step towards building large foundation models for HEP that can be generically pre-trained with self-supervised learning and later fine-tuned for a variety of down-stream tasks. In MPM, particles in a set are masked and the training objective is to recover their identity, as defined by a discretized token representation of a pre-trained vector quantized variational autoencoder. We study the efficacy of the method in samples of high energy jets at collider physics experiments, including studies on the impact of discretization, permutation invariance, and ordering. We also study the fine-tuning capability of the model, showing that it can be adapted to tasks such as supervised and weakly supervised jet classification, and that the model can transfer efficiently with small fine-tuning data sets to new classes and new data domains.

Discussion (0). Sign in to comment.

Forward citations

Cited by 12 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score.

  1. Predict before you train: Scaling Laws for particle physics foundation models

    hep-ex 2026-07 conditional novelty 7.0 of 10

    A Chinchilla-style law fit on ParticleViT runs below 10^19 FLOPs predicts held-out pretraining loss within ~1% at >100× compute and tracks downstream jet-tagging rejection.

  2. Learning Standard Model structure from LHC data with Riemannian flow matching

    hep-ph 2026-07 conditional novelty 7.0 of 10

    ShellFlow, a Riemannian flow-matching transformer fed only on-shell and invariant-mass priors and ~8×10^8 recorded ATLAS events, reproduces the SM's dilepton resonances, Weinberg angle, and top/W mass peaks in a singl...

  3. Masked-Token Prediction for Anomaly Detection at the Large Hadron Collider

    hep-ph 2026-04 unverdicted novelty 7.0 of 10

    The work demonstrates masked-token prediction with transformers for model-independent anomaly detection in LHC data, achieving strong results on top-rich BSM signatures like four-top production using VQ-VAE tokenization.

  4. Learning transferable event representations for charmed baryon physics at BESIII

    physics.data-an 2026-07 conditional novelty 6.0 of 10

    A Particle Transformer pre-trained on simulated Lambda_c events transfers across 12 decay channels, improving classification and momentum-direction regression over training from scratch in low-statistics regimes.

  5. D$e^+e^-$ffusion: Capturing the Beam-Beam Physics of $e^+e^-$ Collisions with Diffusion Models

    hep-ph 2026-07 conditional novelty 6.0 of 10

    A diffusion model trained on GuineaPig++ reproduces FCC-ee beam-induced pair-production distributions at particle and detector level, about 10^4 times faster.

  6. Explicit or Implicit? Encoding Physics at the Precision Frontier

    hep-ph 2026-03 conditional novelty 6.0 of 10

    On three precision classification tasks — reweighting-based unfolding, likelihood-ratio estimation, and weakly supervised anomaly detection — a Lorentz-equivariant transformer and a pretrained foundation model perform...

  7. A universal vision transformer for fast calorimeter simulations

    hep-ph 2026-01 conditional novelty 6.0 of 10

    A vision-transformer flow-matching model generates calorimeter showers across regular and irregular detector geometries at millisecond speeds, and pretraining plus fine-tuning cuts training cost by about half.

  8. Enhancing next token prediction based pre-training for jet foundation models

    hep-ph 2025-12 conditional novelty 6.0 of 10

    Using continuous particle features as input and combining next-token with masked-token pre-training markedly improves classification accuracy of the OmniJet jet foundation model without visibly hurting its generative quality.

  9. Pretrained Event Classification Model for High Energy Physics Analysis

    hep-ph 2024-12 unverdicted novelty 6.0 of 10

    A GNN pretrained on 120M simulated HEP events generalizes to unseen processes and ATLAS data; fine-tuning boosts accuracy especially with small datasets, with CKA showing preserved encoders but altered intermediate layers.

  10. A Lightweight Foundation Model for Collider Physics with Multi-Domain Adaptation

    cs.LG 2026-07 conditional novelty 5.0 of 10

    A lightweight autoencoder pre-trained on LHC track data transfers to collider and out-of-domain scientific tasks, matching a transformer within ~2% at about 46x lower per-epoch training cost.

  11. Are We Ready for AI-Driven Discovery? AI Verification Before the Next Fundamental Physics Breakthrough

    physics.data-an 2026-07 accept novelty 4.0 of 10

    Verification of ML in fundamental physics is essential precisely when models enter statistical modeling, inference, or hypothesis testing, and is bounded by unavoidable inductive bias, sample complexity, and experimen...

  12. HEPTAPOD: Orchestrating High Energy Physics Workflows Towards Autonomous Agency

    hep-ph 2025-12 conditional novelty 4.0 of 10

    HEPTAPOD uses LLM agents to drive FeynRules, MadGraph, Pythia, and analysis tools through schema-validated tool calls and run-card templates, demonstrated on a leptoquark signal scan.

Pith tools