REVIEW 5 minor 1 cited by
Are We Ready for AI-Driven Discovery? AI Verification Before the Next Fundamental Physics Breakthrough
T0 review · 0 major / 5 minor · reviewed 2026-07-14 · grok-4.5
Pith's one-line read Verification of machine learning is essential only when its outputs enter statistical modeling, inference, or hypothesis testing for discovery claims.
desk verdict Solid community synthesis that maps when ML verification is load-bearing for discovery claims; useful reference, not a new result. read the letter →
The pith
A machine-rendered reading of the paper's core claim, the machinery that carries it, and where it could break.
The reading
What carries the argument
The four-stage statistical workflow (data collection, summarization, modeling, inference) together with the aleatoric/epistemic uncertainty distinction; these locate every ML tool and dictate whether, and which, verification is required.
What would settle it
A concrete ML application whose outputs enter a discovery claim yet cannot be classified as either (a) a summarization/exploratory step whose imperfections only reduce power or (b) a modeling/surrogate step whose residual bias and epistemic uncertainty can be quantified and propagated, thereby leaving the paper’s decision rules incomplete.
Extended reading notes
Core claim
Verification of machine learning is essential precisely when its outputs form part of the statistical model used for inference or hypothesis testing; at other stages of the discovery workflow imperfect models are acceptable provided residual uncertainties are quantified and systematic biases are either calibrated out or demonstrably absent.
Load-bearing premise
That the four-stage workflow and the aleatoric/epistemic split cleanly cover every present and future machine-learning use in fundamental physics, so the verification rules derived from them stay complete.
Signed reviews
Editorial analysis
A structured set of objections, weighed in public.
Referee Report
Summary. This VERaiPHY community review argues that ML verification in fundamental physics is essential precisely when model outputs enter statistical modeling, inference, or hypothesis testing, while imperfect models remain tolerable in summarization, exploratory analysis, and calibrated surrogates provided residual uncertainties are quantified and unmodeled systematic bias is avoided. It situates ML within a four-stage discovery workflow (data collection, summarization, modeling, inference), surveys computational bottlenecks and emerging paradigms (differentiable design, foundation models, anomaly detection, agentic AI), and articulates irreducible limits (inductive bias, observational constraints, computational bounds, verification incompleteness). The closing sections discuss the physicist’s evolving role as designer, evaluator, and teacher of AI systems and offer high-level guidelines for responsible deployment.
Significance. As a synthesis paper rather than a primary-result claim, its value lies in organizing a fragmented literature into a coherent, workflow-based verification framework that spans particle physics, astrophysics, and cosmology. The contextual distinction between performance degradation and statistical invalidity (Sections 3.1–3.3), the explicit treatment of agentic systems and verification limits (2.3, 4.4), and the reflection on human oversight (Section 5) are timely contributions for a community facing increasingly autonomous ML. The paper correctly grounds its arguments in standard statistical practice (look-elsewhere effects, calibration, coverage) and citable results (No Free Lunch, data-processing inequality, SBI surveys). It does not overclaim completeness and is well positioned as an entry point to the broader VERaiPHY series.
minor comments (5)
- Several companion VERaiPHY reviews are cited as “in preparation” (e.g., Refs. [36], [61], [103], [114], [118]). For archival permanence, either update with arXiv identifiers where available or flag more clearly which claims rest only on forthcoming companion pieces.
- Figure 1 and Figure 3 are conceptually clear but would benefit from slightly more explicit captions linking each panel to the corresponding workflow stage or uncertainty type discussed in the text.
- Section 2.3 on agentic AI is appropriately cautious; a short forward pointer to concrete verification protocols (even if only as open problems) would strengthen the bridge to Section 4.4.
- Minor typographical and formatting inconsistencies appear (e.g., spacing around citations, occasional hyphenation of “black-box” / “black boxes”). A light copy-edit pass would polish the manuscript.
- The abstract and concluding guidelines are strong; ensuring the five bullet guidelines in Section 6 map one-to-one onto the section structure would improve navigability for practitioners.
Circularity Check
No circularity: community review with normative framing, no fitted predictions or self-definitional reductions.
full rationale
This is a VERaiPHY community review that organizes existing statistical practice around a four-stage workflow (data collection, summarization, modeling, inference) and an aleatoric/epistemic taxonomy. It does not claim to derive quantitative predictions, uniqueness theorems, or first-principles results from fitted parameters or self-defined quantities. Load-bearing external anchors (Wolpert No Free Lunch, Cover data-processing inequality, Cranmer SBI survey, standard frequentist/Bayesian practice) are independent of the authors. Companion VERaiPHY citations supply depth on subtopics but are not required to force the central normative claim that verification is essential precisely when ML outputs enter statistical modeling, inference, or hypothesis testing. No equation reduces to its own input by construction; no ansatz is smuggled via self-citation; no known empirical pattern is merely renamed. Score 0 is the correct honest finding.
Assumptions & free parameters
assumptions (4)
- standard math No Free Lunch theorems: no learning algorithm is universally superior; inductive bias is unavoidable.
- standard math Data Processing Inequality: deterministic or stochastic transformations cannot increase mutual information with the quantity of interest.
- domain assumption Discovery in fundamental physics proceeds via statistical inference on noisy, incomplete observables rather than direct observation of the target entities.
- domain assumption The four-stage workflow (collection, summarization, modeling, inference) plus the aleatoric/epistemic distinction covers the relevant ML insertion points.
Cite this review
Pith. "Pith review of Are We Ready for AI-Driven Discovery? AI Verification Before the Next Fundamental Physics Breakthrough." pith.science (2026). https://pith.science/paper/FKYLUFS7
@misc{pith2026260710039,
author = {Pith},
title = {Pith review of: Are We Ready for AI-Driven Discovery? AI Verification Before the Next Fundamental Physics Breakthrough},
year = {2026},
howpublished = {\url{https://pith.science/paper/FKYLUFS7}},
note = {Machine review of arXiv:2607.10039}
}
read the original abstract
Machine learning (ML) has become integral to fundamental physics, accelerating statistical workflows from data acquisition through inference and hypothesis testing. As ML systems grow increasingly autonomous, ensuring their reliability for discovery claims becomes critical. This review synthesizes the VERaiPHY (Validation & Evaluation for Robust AI in PHYsics) initiative's frameworks for rigorous ML assessment across particle physics, astrophysics, and cosmology. We establish when verification is essential by contextualizing ML within the statistical discovery workflow. We emphasize fundamental limitations: inductive bias is unavoidable, sample complexity bounds learning, and experimental constraints limit discovery. We reflect on physicists' evolving role as both experimental designers and evaluators whose judgments encode scientific rigor into AI systems. Responsible integration requires understanding ML's transformative potential alongside its intrinsic boundaries.
Figures
Figures from the paper (1 more)
Forward citations
Cited by 1 Pith paper
-
Machine Learning is Good for Physics - and Vice Versa
A perspective essay arguing that AI should be integrated into fundamental physics while preserving the field's statistical and theory-based standards, and that physics can enrich machine learning.
Reference graph
Works this paper leans on
-
[1]
2025 , month =
Cho, Kyunghyun , title =. 2025 , month =
2025
-
[2]
Coffea: Columnar Object Framework For Effective Analysis
Smith, Nicholas and others. Coffea: Columnar Object Framework For Effective Analysis. EPJ Web Conf. 2020. doi:10.1051/epjconf/202024506012. arXiv:2008.12712
-
[3]
Proceedings of the 24th International Conference on Computing in High Energy and Nuclear Physics (CHEP 2019) , journal =
Awkward Arrays in Python, C++, and Numba , author =. Proceedings of the 24th International Conference on Computing in High Energy and Nuclear Physics (CHEP 2019) , journal =. 2020 , doi =
2019
-
[4]
Brun, R. and Rademakers, F. ROOT: An object oriented data analysis framework. Nucl. Instrum. Meth. A. 1997. doi:10.1016/S0168-9002(97)00048-X
-
[5]
Active Learning reinterpretation of an ATLAS Dark Matter search constraining a model of a dark Higgs boson decaying to two b-quarks. 2022
2022
-
[6]
Observation of Gravitational Waves from a Binary Black Hole Merger , author =. Phys. Rev. Lett. , volume =. 2016 , month =. doi:10.1103/PhysRevLett.116.061102 , url =
-
[7]
The Astronomical Journal , volume=
Data release 1 of the dark energy spectroscopic instrument , author=. The Astronomical Journal , volume=. 2026 , publisher=
2026
-
[8]
Monthly Notices of the Royal Astronomical Society , volume=
Cosmic confusion: degeneracies among cosmological parameters derived from measurements of microwave background anisotropies , author=. Monthly Notices of the Royal Astronomical Society , volume=. 1999 , publisher=
1999
Show all 142 references
-
[9]
Monthly Notices of the Royal Astronomical Society , volume=
Massively parallel Bayesian inference for transient gravitational-wave astronomy , author=. Monthly Notices of the Royal Astronomical Society , volume=. 2020 , publisher=
2020
- [10]
-
[11]
Observation of a new particle in the search for the Standard Model Higgs boson with the ATLAS detector at the LHC
Aad, Georges and others. Observation of a new particle in the search for the Standard Model Higgs boson with the ATLAS detector at the LHC. Phys. Lett. B. 2012. doi:10.1016/j.physletb.2012.08.020. arXiv:1207.7214
- [12]
- [13]
-
[14]
Journal of Instrumentation , volume=
The CMS trigger system , author=. Journal of Instrumentation , volume=
-
[15]
Journal of Instrumentation , volume=
Operation of the ATLAS trigger system in Run 2 , author=. Journal of Instrumentation , volume=
-
[16]
arXiv preprint arXiv:1902.08960 , year=
Precision laser-based measurements of the single electron response of SPCs for the NEWS-G light dark matter search experiment , author=. arXiv preprint arXiv:1902.08960 , year=
1902 arXiv
-
[17]
Physical Review D , volume=
First results from the CRESST-III low-mass dark matter program , author=. Physical Review D , volume=. 2019 , publisher=
2019
-
[18]
Fast neural-net based fake track rejection in the LHCb reconstruction , author=
-
[19]
The Phase-2 Upgrade of the CMS Level-1 Trigger
Zabi, Alexandre and Berryhill, Jeffrey Wayne and Perez, Emmanuelle and Tapper, Alexander D. The Phase-2 Upgrade of the CMS Level-1 Trigger. 2020
2020
-
[20]
Journal of instrumentation , volume=
Fast inference of deep neural networks in FPGAs for particle physics , author=. Journal of instrumentation , volume=. 2018 , publisher=
2018
-
[21]
arXiv preprint arXiv:2509.24371 , year=
Advancing the CMS Level-1 Trigger: Jet Tagging with DeepSets at the HL-LHC , author=. arXiv preprint arXiv:2509.24371 , year=
-
[22]
Autoencoders on field-programmable gate arrays for real-time, unsupervised new physics detection at 40 MHz at the Large Hadron Collider
Govorkova, Ekaterina and others. Autoencoders on field-programmable gate arrays for real-time, unsupervised new physics detection at 40 MHz at the Large Hadron Collider. Nature Mach. Intell. 2022. doi:10.1038/s42256-022-00441-3. arXiv:2108.03986
-
[23]
Matching Matched Filtering with Deep Networks for Gravitational-Wave Astronomy , author =. Phys. Rev. Lett. , volume =. 2018 , month =. doi:10.1103/PhysRevLett.120.141103 , url =
2018 doi
-
[24]
Physical Review D , volume=
Generalized approach to matched filtering using neural networks , author=. Physical Review D , volume=. 2022 , publisher=
2022
-
[25]
Physical Review D , volume=
Deep neural networks to enable real-time multimessenger astrophysics , author=. Physical Review D , volume=. 2018 , publisher=
2018
-
[26]
Machine Learning: Science and Technology , volume=
GWAK: gravitational-wave anomalous knowledge with recurrent autoencoders , author=. Machine Learning: Science and Technology , volume=. 2024 , publisher=
2024
-
[27]
SEAL - A Symmetry EncourAging Loss for High Energy Physics
Hebbar, Pradyun and Madula, Thandikire and Mikuni, Vinicius and Nachman, Benjamin and Outmezguine, Nadav and Savoray, Inbar. SEAL - A Symmetry EncourAging Loss for High Energy Physics. 2025. arXiv:2511.01982
2025
-
[28]
Machine Learning: Science and Technology , volume=
Probing the effects of broken symmetries in machine learning , author=. Machine Learning: Science and Technology , volume=. 2024 , publisher=
2024
-
[29]
Physical Review D , volume=
Learning broken symmetries with approximate invariance , author=. Physical Review D , volume=. 2025 , publisher=
2025
-
[30]
and Ipp, Andreas and M
Aarts, Gert and Habibi, Diaa E. and Ipp, Andreas and M. Generalizable Equivariant Diffusion Models for Non-Abelian Lattice Gauge Theory. 2026. arXiv:2601.19552
2026
-
[31]
A Living Review of Machine Learning for Particle Physics
Feickert, Matthew and Nachman, Benjamin. A Living Review of Machine Learning for Particle Physics. 2021. arXiv:2102.02770
2021 arXiv
-
[32]
OmniJet- : The first cross-task foundation model for particle physics
Birk, Joschka and Hallin, Anna and Kasieczka, Gregor. OmniJet- : The first cross-task foundation model for particle physics. 2024. arXiv:2403.05618
2024 arXiv
-
[33]
Foundation models for high-energy physics
Hallin, Anna. Foundation models for high-energy physics. 2nd European AI for Fundamental Physics Conference. 2025. arXiv:2509.21434
2025
-
[34]
Re-Simulation-based Self-Supervised Learning for Pre-Training Foundation Models
Harris, Philip and Kagan, Michael and Krupa, Jeffrey and Maier, Benedikt and Woodward, Nathaniel. Re-Simulation-based Self-Supervised Learning for Pre-Training Foundation Models. 2024. arXiv:2403.07066
2024 arXiv
-
[35]
Masked particle modeling on sets: towards self-supervised high energy physics foundation models
Golling, Tobias and Heinrich, Lukas and Kagan, Michael and Klein, Samuel and Leigh, Matthew and Osadchy, Margarita and Raine, John Andrew. Masked particle modeling on sets: towards self-supervised high energy physics foundation models. Mach. Learn. Sci. Tech. 2024. doi:10.1088...
-
[36]
Finetuning foundation models for joint analysis optimization in High Energy Physics
Vigl, Matthias and Hartman, Nicole and Heinrich, Lukas. Finetuning foundation models for joint analysis optimization in High Energy Physics. Mach. Learn. Sci. Tech. 2024. doi:10.1088/2632-2153/ad55a3. arXiv:2401.13536
-
[37]
Method to simultaneously facilitate all jet physics tasks
Mikuni, Vinicius and Nachman, Benjamin. Method to simultaneously facilitate all jet physics tasks. Phys. Rev. D. 2025. doi:10.1103/PhysRevD.111.054015. arXiv:2502.14652
2025 doi
-
[38]
Solving key challenges in collider physics with foundation models
Mikuni, Vinicius and Nachman, Benjamin. Solving key challenges in collider physics with foundation models. Phys. Rev. D. 2025. doi:10.1103/PhysRevD.111.L051504. arXiv:2404.16091
2025 doi
-
[39]
Is Tokenization Needed for Masked Particle Modelling?
Leigh, Matthew and Klein, Samuel and Charton, Fran c ois and Golling, Tobias and Heinrich, Lukas and Kagan, Michael and Ochoa, In \^e s and Osadchy, Margarita. Is Tokenization Needed for Masked Particle Modelling?. 2024. arXiv:2409.12589
2024 arXiv
-
[40]
HEP-JEPA: A foundation model for collider physics using joint embedding predictive architecture
Bardhan, Jai and Agrawal, Radhikesh and Tilak, Abhiram and Neeraj, Cyrin and Mitra, Subhadip. HEP-JEPA: A foundation model for collider physics using joint embedding predictive architecture. 2025. arXiv:2502.03933
2025 arXiv
-
[41]
Particle transformers for identifying Lorentz-boosted Higgs bosons decaying to a pair of W bosons
Hayrapetyan, Aram and others. Particle transformers for identifying Lorentz-boosted Higgs bosons decaying to a pair of W bosons. 2026. arXiv:2604.09809
2026 arXiv
-
[42]
OmniCosmos: Transferring Particle Physics Knowledge Across the Cosmos
Mikuni, Vinicius and Elsharkawy, Ibrahim and Nachman, Benjamin. OmniCosmos: Transferring Particle Physics Knowledge Across the Cosmos. 2025. arXiv:2512.24422
2025
-
[43]
OmniMol: Transferring Particle Physics Knowledge to Molecular Dynamics with Point-Edge Transformers
Elsharkawy, Ibrahim and Mikuni, Vinicius and Bhimji, Wahid and Nachman, Benjamin. OmniMol: Transferring Particle Physics Knowledge to Molecular Dynamics with Point-Edge Transformers. 2026. arXiv:2601.10791
2026 arXiv
-
[44]
PASCL: supervised contrastive learning with perturbative augmentation for particle decay reconstruction
Lu, Junjian and Liu, Siwei and Kobylianskii, Dmitrii and Dreyer, Etienne and Gross, Eilam and Liang, Shangsong. PASCL: supervised contrastive learning with perturbative augmentation for particle decay reconstruction. Mach. Learn. Sci. Tech. 2024. doi:10.1088/2632-2153/ad8060. ...
-
[46]
Advances in Neural Information Processing Systems , volume=
Autoscidact: Automated scientific discovery through contrastive embedding and hypothesis testing , author=. Advances in Neural Information Processing Systems , volume=
- [47]
- [48]
-
[49]
A unified approach for jet tagging in Run 3 at s =13.6 TeV in CMS. 2024
2024
-
[50]
Transforming jet flavour tagging at ATLAS
Aad, Georges and others. Transforming jet flavour tagging at ATLAS. Nature Commun. 2026. doi:10.1038/s41467-025-65059-6. arXiv:2505.19689
2026 doi
-
[51]
SciPost Physics , volume=
Machine learning and LHC event generation , author=. SciPost Physics , volume=
-
[52]
CaloChallenge 2022: a community challenge for fast calorimeter simulation
Amram, Oz and others. CaloChallenge 2022: a community challenge for fast calorimeter simulation. Rept. Prog. Phys. 2025. doi:10.1088/1361-6633/ae1304. arXiv:2410.21611
2022 doi
-
[53]
arXiv preprint arXiv:2110.07925 , year=
Machine learning for the LHCb simulation , author=. arXiv preprint arXiv:2110.07925 , year=
-
[54]
Machine Learning: Science and Technology , volume=
Particle-based fast jet simulation at the LHC with variational autoencoders , author=. Machine Learning: Science and Technology , volume=. 2022 , publisher=
2022
-
[55]
2024 , institution=
The Fast Simulation Program of ATLAS at the LHC , author=. 2024 , institution=
2024
-
[56]
arXiv preprint arXiv:2511.02020 , year=
Machine learning in LHCb Simulation: From fast to flash , author=. arXiv preprint arXiv:2511.02020 , year=
-
[57]
Proceedings of the National Academy of Sciences , volume =
Siyu He and Yin Li and Yu Feng and Shirley Ho and Siamak Ravanbakhsh and Wei Chen and Barnabás Póczos , title =. Proceedings of the National Academy of Sciences , volume =. 2019 , doi =
2019
-
[58]
Computational Astrophysics and Cosmology , volume=
Fast cosmic web simulations with generative adversarial networks , author=. Computational Astrophysics and Cosmology , volume=. 2018 , publisher=
2018
-
[59]
Machine Learning: Science and Technology , url=
Cappelli, Pietro and Grosso, Gaia and Letizia, Marco and Reyes-González, Humberto and Zanetti, Marco , title=. Machine Learning: Science and Technology , url=
-
[60]
Evaluating generative models in high energy physics
Kansal, Raghav and Li, Anni and Duarte, Javier and Chernyavskaya, Nadezda and Pierini, Maurizio and Orzari, Breno and Tomei, Thiago. Evaluating generative models in high energy physics. Phys. Rev. D. 2023. doi:10.1103/PhysRevD.107.076017. arXiv:2211.10295
-
[61]
hydrodynamical simulations , author=
Machine learning and cosmological simulations--ii. hydrodynamical simulations , author=. Monthly Notices of the Royal Astronomical Society , volume=. 2016 , publisher=
2016
-
[62]
Proceedings of the National Academy of Sciences , volume=
The frontier of simulation-based inference , author=. Proceedings of the National Academy of Sciences , volume=. 2020 , publisher=
2020
-
[63]
Artificial Intelligence for High Energy Physics , pages=
Simulation-based inference methods for particle physics , author=. Artificial Intelligence for High Energy Physics , pages=. 2022 , publisher=
2022
-
[64]
Reports on Progress in Physics , volume=
An implementation of neural simulation-based inference for parameter estimation in ATLAS , author=. Reports on Progress in Physics , volume=. 2025 , publisher=
2025
-
[65]
Machine Learning: Science and Technology , volume=
Robust simulation-based inference in cosmology with Bayesian neural networks , author=. Machine Learning: Science and Technology , volume=. 2023 , publisher=
2023
-
[66]
Nature Astronomy , volume=
A deep-learning algorithm to disentangle self-interacting dark matter and AGN feedback models , author=. Nature Astronomy , volume=. 2024 , publisher=
2024
-
[67]
Machine Learning: Science and Technology , abstract =
Astrand, S and Boggia, L and Borsato, M and Bozianu, L and Cocha Toapaxi, C E and Giasemis, F I and Hansen, J and Inkaew, P and Iversen, K E and Jawahar, P and Pineiro Monteagudo, H and Olocco, M and Schramm, S , title =. Machine Learning: Science and Technology , abstract =. ...
2026 doi
-
[68]
arXiv preprint arXiv:1702.08608 , year=
Towards a rigorous science of interpretable machine learning , author=. arXiv preprint arXiv:1702.08608 , year=
-
[69]
Nature machine intelligence , volume=
Stop explaining black box machine learning models for high stakes decisions and use interpretable models instead , author=. Nature machine intelligence , volume=. 2019 , publisher=
2019
-
[70]
, author=
The mythos of model interpretability: In machine learning, the concept of interpretability is both important and slippery. , author=. Queue , volume=. 2018 , publisher=
2018
-
[71]
arXiv preprint arXiv:2601.03220 , year=
From Entropy to Epiplexity: Rethinking Information for Computationally Bounded Intelligence , author=. arXiv preprint arXiv:2601.03220 , year=
-
[72]
arXiv preprint arXiv:2503.02113 , year=
Deep learning is not so mysterious or different , author=. arXiv preprint arXiv:2503.02113 , year=
-
[73]
arXiv preprint arXiv:2304.05366 , year=
The no free lunch theorem, kolmogorov complexity, and the role of inductive biases in machine learning , author=. arXiv preprint arXiv:2304.05366 , year=
-
[74]
Neural computation , volume=
The lack of a priori distinctions between learning algorithms , author=. Neural computation , volume=. 1996 , publisher=
1996
-
[75]
1995 , institution=
No free lunch theorems for search , author=. 1995 , institution=
1995
-
[76]
IEEE transactions on evolutionary computation , volume=
No free lunch theorems for optimization , author=. IEEE transactions on evolutionary computation , volume=. 2002 , publisher=
2002
-
[77]
arXiv preprint arXiv:1805.08522 , year=
Deep learning generalizes because the parameter-function map is biased towards simple functions , author=. arXiv preprint arXiv:1805.08522 , year=
-
[78]
arXiv preprint arXiv:2103.10427 , year=
The low-rank simplicity bias in deep networks , author=. arXiv preprint arXiv:2103.10427 , year=
-
[79]
arXiv preprint arXiv:2509.22445 , year=
Bridging Kolmogorov Complexity and Deep Learning: Asymptotically Optimal Description Length Objectives for Transformers , author=. arXiv preprint arXiv:2509.22445 , year=
-
[80]
Resimulation-based self-supervised learning for pretraining physics foundation models , author =. Phys. Rev. D , volume =. 2025 , month =. doi:10.1103/PhysRevD.111.032010 , url =
2025 doi
-
[81]
Anomaly-preserving contrastive neural embeddings for end-to-end model-independent searches at the LHC , author =. Phys. Rev. D , volume =. 2025 , month =. doi:10.1103/5n77-ynsp , url =
2025 doi
-
[82]
arXiv preprint arXiv:2002.10689 , year=
A theory of usable information under computational constraints , author=. arXiv preprint arXiv:2002.10689 , year=
2002 arXiv
-
[83]
1986 , publisher=
RESOURCE BOUNDED KOLMOGOROV COMPLEXITY, A LINK BETWEEN COMPUTATIONAL COMPLEXITY AND INFORMATION THEORY (RANDOMNESS) , author=. 1986 , publisher=
1986
-
[84]
SIAM Journal on Computing , volume=
Power from random strings , author=. SIAM Journal on Computing , volume=. 2006 , publisher=
2006
-
[85]
2026 , url =
Get Physics Done (GPD) , version =. 2026 , url =
2026
-
[86]
Moreno, Eric and others , title =
-
[87]
Karpathy, Andrej , title =
-
[88]
Reviews in Physics , volume=
Toward the end-to-end optimization of particle physics instruments with differentiable programming , author=. Reviews in Physics , volume=. 2023 , publisher=
2023
-
[89]
Reviews in physics , volume=
Progress in end-to-end optimization of fundamental physics experimental apparata with differentiable programming , author=. Reviews in physics , volume=. 2025 , publisher=
2025
-
[90]
arXiv preprint arXiv:2603.15726 , year=
MiroThinker-1.7 & H1: Towards Heavy-Duty Research Agents via Verification , author=. arXiv preprint arXiv:2603.15726 , year=
-
[91]
and Bright-Thonney, Samuel and Novak, Andrzej and Garcia, Dolores and Harris, Philip
Moreno, Eric A. and Bright-Thonney, Samuel and Novak, Andrzej and Garcia, Dolores and Harris, Philip. AI Agents Can Already Autonomously Perform Experimental High Energy Physics. 2026. arXiv:2603.20179
2026 arXiv
-
[92]
arXiv preprint arXiv:2603.05735 , year=
Agentic AI--Physicist Collaboration in Experimental Particle Physics: A Proof-of-Concept Measurement with LEP Open Data , author=. arXiv preprint arXiv:2603.05735 , year=
-
[93]
arXiv preprint arXiv:2408.06292 , year=
The ai scientist: Towards fully automated open-ended scientific discovery , author=. arXiv preprint arXiv:2408.06292 , year=
-
[94]
arXiv preprint arXiv:2509.06855 , year=
Seeing the Forest Through the Trees: Knowledge Retrieval for Streamlining Particle Physics Analysis , author=. arXiv preprint arXiv:2509.06855 , year=
-
[95]
arXiv preprint arXiv:2512.15867 , year=
HEPTAPOD: Orchestrating High Energy Physics Workflows Towards Autonomous Agency , author=. arXiv preprint arXiv:2512.15867 , year=
-
[96]
arXiv preprint arXiv:2404.08001 , year=
Xiwu: A basis flexible and learnable llm for high energy physics , author=. arXiv preprint arXiv:2404.08001 , year=
-
[97]
arXiv preprint arXiv:2509.08535 , year=
Agents of Discovery , author=. arXiv preprint arXiv:2509.08535 , year=
-
[98]
arXiv preprint arXiv:1412.6572 , year=
Explaining and harnessing adversarial examples , author=. arXiv preprint arXiv:1412.6572 , year=
-
[99]
International Conference on Automated Deduction , pages=
The lean 4 theorem prover and programming language , author=. International Conference on Automated Deduction , pages=. 2021 , organization=
2021
-
[100]
Advances in neural information processing systems , volume=
Deep reinforcement learning from human preferences , author=. Advances in neural information processing systems , volume=
-
[101]
Advances in neural information processing systems , volume=
Training language models to follow instructions with human feedback , author=. Advances in neural information processing systems , volume=
-
[102]
The twelfth international conference on learning representations , year=
Let's verify step by step , author=. The twelfth international conference on learning representations , year=
-
[103]
arXiv preprint arXiv:2212.08073 , year=
Constitutional ai: Harmlessness from ai feedback , author=. arXiv preprint arXiv:2212.08073 , year=
-
[104]
arXiv preprint arXiv:2211.14275 , year=
Solving math word problems with process-and outcome-based feedback , author=. arXiv preprint arXiv:2211.14275 , year=
-
[105]
1999 , publisher=
Elements of information theory , author=. 1999 , publisher=
1999
-
[106]
Savannah Thais and Roberto Trotta and Nathan Suri and Emily Sullivan and Viyan Poonamallee and Tanaporn Na Narong and Rupert Croft and Nicole Hartman , title =
-
[107]
Nature Reviews Physics , volume =
Thais, Savannah , title =. Nature Reviews Physics , volume =. 2024 , month = nov, doi =
2024
-
[108]
arXiv preprint arXiv:2604.21691 , year=
There will be a scientific theory of deep learning , author=. arXiv preprint arXiv:2604.21691 , year=
-
[109]
arXiv preprint arXiv:2605.10378 , year=
Uncertainty in Physics and AI: Taxonomy, Quantification, and Validation , author=. arXiv preprint arXiv:2605.10378 , year=
-
[110]
2026 , eprint =
Gambhir, Rikab and Lucie-Smith, Luisa and Thaler, Jesse , title =. 2026 , eprint =
2026
-
[111]
in preparation , year=
Symmetry-Informed Machine Learning for Fundamental Physics , author=. in preparation , year=
-
[112]
in preparation , year=
Unknown Unknowns in Machine Learning for Physics , author=. in preparation , year=
-
[113]
in preparation , year=
Representation Learning in Fundamental Physics , author=. in preparation , year=
-
[114]
arXiv preprint arXiv:2606.20299 , year=
Statistical Properties of Training & Generalization , author=. arXiv preprint arXiv:2606.20299 , year=
-
[115]
arXiv preprint arXiv:2605.30453 , year=
Generative Models and Statistical Validation , author=. arXiv preprint arXiv:2605.30453 , year=
-
[116]
arXiv preprint arXiv:2605.31103 , year=
Model-Agnostic Signal Discovery with Machine Learning: Bridging the Gap Between Theory and Practice , author=. arXiv preprint arXiv:2605.31103 , year=
-
[117]
Blind analysis in nuclear and particle physics , author=. Annu. Rev. Nucl. Part. Sci. , volume=. 2005 , publisher=
2005
-
[118]
Nuclear instruments and methods in physics research section A: Accelerators, Spectrometers, Detectors and Associated Equipment , volume=
Geant4—a simulation toolkit , author=. Nuclear instruments and methods in physics research section A: Accelerators, Spectrometers, Detectors and Associated Equipment , volume=. 2003 , publisher=
2003
-
[119]
Physical Review X , volume=
GWTC-3: Compact binary coalescences observed by LIGO and Virgo during the second part of the third observing run , author=. Physical Review X , volume=. 2023 , publisher=
2023
-
[120]
The Astrophysical Journal Letters , volume=
GWTC-4.0: Updating the gravitational-wave transient catalog with observations from the first part of the fourth LIGO--Virgo--KAGRA observing run , author=. The Astrophysical Journal Letters , volume=. 2026 , publisher=
2026
-
[121]
Journal of instrumentation , volume=
LHC machine , author=. Journal of instrumentation , volume=
-
[122]
Physical review letters , volume=
Dark matter search results from a one ton-year exposure of XENON1T , author=. Physical review letters , volume=. 2018 , publisher=
2018
-
[123]
Biometrika , volume=
Hypothesis testing when a nuisance parameter is present only under the alternative , author=. Biometrika , volume=. 1977 , publisher=
1977
-
[124]
Davies , journal =
Robert B. Davies , journal =. Hypothesis Testing when a Nuisance Parameter is Present Only Under the Alternatives , urldate =
- [125]
- [126]
-
[127]
The European Physical Journal C , volume=
Multiple testing for signal-agnostic searches for new physics with machine learning , author=. The European Physical Journal C , volume=. 2025 , publisher=
2025
-
[128]
2025 , eprint =
Hein, Marie and Nachman, Benjamin and Shih, David , title =. 2025 , eprint =
2025
- [129]
-
[130]
2026 , url =
Tong, Shelley and Grosso, Gaia and Harris, Philip , title =. 2026 , url =
2026
-
[131]
The Annals of Statistics , pages=
Valid post-selection inference , author=. The Annals of Statistics , pages=. 2013 , publisher=
2013
-
[132]
and Kolassa, John E
Kuchibhotla, Arun K. and Kolassa, John E. and Kuffner, Todd A. , title =. Annual Review of Statistics and Its Application , volume =. 2022 , doi =
2022
-
[133]
Nested sampling for general Bayesian computation , author=
-
[134]
International conference on machine learning , pages=
Group equivariant convolutional networks , author=. International conference on machine learning , pages=. 2016 , organization=
2016
-
[135]
arXiv preprint arXiv:2111.00899 , year=
Equivariant contrastive learning , author=. arXiv preprint arXiv:2111.00899 , year=
-
[136]
The journal of chemical physics , volume=
Equation of state calculations by fast computing machines , author=. The journal of chemical physics , volume=. 1953 , publisher=
1953
-
[137]
1970 , publisher=
Monte Carlo sampling methods using Markov chains and their applications , author=. 1970 , publisher=
1970
-
[138]
arXiv preprint arXiv:2108.07258 , year=
On the opportunities and risks of foundation models , author=. arXiv preprint arXiv:2108.07258 , year=
-
[139]
ACM computing surveys , volume=
Diffusion models: A comprehensive survey of methods and applications , author=. ACM computing surveys , volume=. 2023 , publisher=
2023
-
[140]
arXiv preprint arXiv:2001.08361 , year=
Scaling laws for neural language models , author=. arXiv preprint arXiv:2001.08361 , year=
2001 arXiv
-
[141]
arXiv preprint arXiv:2203.15556 , volume=
Training compute-optimal large language models , author=. arXiv preprint arXiv:2203.15556 , volume=
-
[142]
Annual review of condensed matter physics , volume=
Statistical mechanics of deep learning , author=. Annual review of condensed matter physics , volume=. 2020 , publisher=
2020
-
[143]
Proceedings of the National Academy of Sciences , volume =
Yasaman Bahri and Ethan Dyer and Jared Kaplan and Jaehoon Lee and Utkarsh Sharma , title =. Proceedings of the National Academy of Sciences , volume =. 2024 , doi =. https://www.pnas.org/doi/pdf/10.1073/pnas.2311878121 , abstract =
2024 doi
Reviewed July 14, 2026 · model on record in the stance chip above.
Discussion (0). Continue with ORCID to comment.