REVIEW 3 cited by
Anomaly Detection in Presence of Irrelevant Features
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
Anomaly Detection in Presence of Irrelevant Features
read the original abstract
Experiments at particle colliders are the primary source of insight into physics at microscopic scales. Searches at these facilities often rely on optimization of analyses targeting specific models of new physics. Increasingly, however, data-driven model-agnostic approaches based on machine learning are also being explored. A major challenge is that such methods can be highly sensitive to the presence of many irrelevant features in the data. This paper presents Boosted Decision Tree (BDT)-based techniques to improve anomaly detection in the presence of many irrelevant features. First, a BDT classifier is shown to be more robust than neural networks for the Classification Without Labels approach to finding resonant excesses assuming independence of resonant and non-resonant observables. Next, a tree-based probability density estimator using copula transformations demonstrates significant stability and improved performance over normalizing flows as irrelevant features are added. The results make a compelling case for further development of tree-based algorithms for more robust resonant anomaly detection in high energy physics.
Forward citations
Cited by 3 Pith papers
-
Towards anomaly detection searches for new physics signatures including Higgs bosons with weakly supervised machine learning
HAXAD, a weakly supervised anomaly-detection search for Higgs-plus-X new physics, is extended with new embeddings and limit-setting, and on 470 fb^-1 of pseudo-data it matches or exceeds the best single cut-based limi...
-
Look everywhere effects in anomaly detection
Weakly supervised anomaly detectors that train and test on the same data produce badly miscalibrated p-values; independent test sets are calibrated but insensitive, while k-fold cross-validation is a workable middle ground.
-
Kitchen Sink Anomaly Detection
A combined kitchen sink observable set of Energy Flow Polynomials and subjettiness variables outperforms standard baselines in sensitivity to a wide range of resonant signals, with new public benchmarks released and a...
discussion (0)
Sign in with ORCID, Apple, or X to comment. Anyone can read and Pith papers without signing in.