REVIEW 4 cited by
Incorporating Physical Priors into Weakly-Supervised Anomaly Detection
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
read the original abstract
We propose a new machine-learning-based anomaly detection strategy for comparing data with a background-only reference (a form of weak supervision). The sensitivity of previous strategies degrades significantly when the signal is too rare or there are many unhelpful features. Our Prior-Assisted Weak Supervision (PAWS) method incorporates information from a class of signal models to significantly enhance the search sensitivity of weakly supervised approaches. As long as the true signal is in the pre-specified class, PAWS matches the sensitivity of a dedicated, fully supervised method without specifying the exact parameters ahead of time. On the benchmark LHC Olympics anomaly detection dataset, our mix of semi-supervised and weakly supervised learning is able to extend the sensitivity over previous methods by a factor of 10 in cross section. Furthermore, if we add irrelevant (noise) dimensions to the inputs, classical methods degrade by another factor of 10 in cross section while PAWS remains insensitive to noise. This new approach could be applied in a number of scenarios and pushes the frontier of sensitivity between completely model-agnostic approaches and fully model-specific searches.
Forward citations
Cited by 4 Pith papers
-
Enhancing anomaly detection with topology-aware autoencoders
Autoencoders with latent spaces shaped like S^2, S^2×S^2, or RP^2, matched to the phase-space topology of the background, reduce spurious reconstruction errors and give a small but consistent anomaly-detection gain ov...
-
Look everywhere effects in anomaly detection
Weakly supervised anomaly detectors that train and test on the same data produce badly miscalibrated p-values; independent test sets are calibrated but insensitive, while k-fold cross-validation is a workable middle ground.
-
FlexCAST: Enabling Flexible Scientific Data Analyses
FlexCAST preserves the design of a scientific analysis as a reusable functional, enabling reinterpretation with changed input data and parameters, demonstrated on a machine-learning anomaly detection analysis.
-
Generator Based Inference (GBI)
Generator Based Inference uses data-derived background generators to turn resonant anomaly detection into parameter estimation, reaching 0.1 sigma signal sensitivity on the LHCO benchmark.
Discussion (0). Continue with ORCID to comment.