REVIEW 3 major objections 6 minor 21 references
Deep Learning Models for ADITYA-U MHD Equilibrium
T0 review · 3 major / 6 minor · reviewed 2026-07-11 · grok-4.5
Pith's one-line read Deep-learning surrogates predict ADITYA-U MHD equilibrium parameters and profiles from a large synthetic free-boundary dataset within the machine’s circular-limiter flat-top domain.
desk verdict Solid first-for-ADITYA-U equilibrium surrogate suite: careful methods and honest scoping, limited mainly by synthetic-only evaluation. read the letter →
The pith
A machine-rendered reading of the paper's core claim, the machinery that carries it, and where it could break.
The reading
What carries the argument
A large, physics-filtered synthetic free-boundary equilibrium library (pyIPREQ solutions of the Grad-Shafranov equation with a three-parameter Jφ profile, Bayesian-optimized γ, and a probabilistic βp prior) that supplies training targets for dense, PCA, CNN and Grad-Shafranov-residual PINN models.
What would settle it
Apply the trained forward models to a set of real ADITYA-U flat-top discharges that have independent equilibrium reconstructions (or diamagnetic βp and magnetic-axis measurements) never used in the synthetic generation pipeline; systematic errors larger than the reported synthetic percentiles would falsify the claimed transferability.
Extended reading notes
Core claim
Key ADITYA-U MHD equilibrium parameters and profiles—Rax, Zax, βp, ℓi, q1, ψaxs/ψlim, the full q(ρ) profile, the full ψ(R,Z) map, and selected PF coil currents—can be accurately estimated by deep-learning surrogates trained on a 100,760-case pyIPREQ free-boundary synthetic dataset, with inference times of order 1 ms, inside the circular-limiter flat-top operational domain represented by that dataset.
Load-bearing premise
The synthetic library—built with a fixed parametric current-density form, a linear βp model fitted on only a few dozen measured discharges, and hard filters on axis shift, inductance and safety factor—must faithfully cover the equilibria that real ADITYA-U flat-top plasmas actually produce, so that synthetic test errors transfer to experiment.
Editorial extensions
If this is right
- Magnetic-axis position, βp, ℓi and edge q can be obtained in ~1 ms from magnetic diagnostics and coil currents, enabling real-time equilibrium feedback on ADITYA-U.
- Full q(ρ) and ψ(R,Z) maps become available for rapid post-shot analysis without repeated free-boundary solves.
- The inverse coil-current model supplies candidate PF actuator settings for desired plasma parameters inside the trained domain, supporting experimental planning.
- Physics-informed residual losses keep the predicted flux maps approximately consistent with the Grad-Shafranov equation, reducing the chance of non-physical reconstructions.
- The same library-plus-surrogate pattern can be reused for other circular-limiter machines once an analogous synthetic database is generated.
Reading between the lines
- Because the models rely on magnetic-probe and loop-voltage inputs, they cannot yet replace free-boundary solvers that operate from coil currents alone; a pure-actuator forward map would be a natural next library.
- The low-dimensional PCA manifolds for both q and ψ suggest that simple parametric families already capture most ADITYA-U flat-top equilibria, so uncertainty-aware or active-learning extensions could focus data collection on the residual high-order modes.
- Localized CNN distortions versus globally smooth PCA errors imply that hybrid PCA–CNN or ensemble predictors may be needed before the maps are trusted for stability calculations that depend on local shear or curvature.
Signed reviews
Editorial analysis
A structured set of objections, weighed in public.
Referee Report
Summary. The manuscript develops deep-learning surrogates for ADITYA-U free-boundary MHD equilibria. A 100,760-case synthetic library is generated with pyIPREQ from 766 experimental discharges (circular limiter, near flat-top, Ip>100 kA), using a fixed parametric Jφ form (Eq. 2), Bayesian-optimized γ, a linear probabilistic βp prior fitted on 39 diamagnetic shots plus noise, and physics-informed filters on ΔR, Zax, ℓi, q0, q1. Dense networks predict scalars (Rax, Zax, βp, ℓi, q1, ψaxs/ψlim) and an inverse map from desired equilibrium parameters to selected PF coil currents. PCA and 1d-CNN models reconstruct q(ρ) (the latter via softplus-constrained q′); PCA and 2d-CNN PINNs reconstruct ψ(R,Z) with Grad–Shafranov residual losses and simultaneous prediction of α,β,γ. Shot-wise 80/10/10 splits, Hyperband search, and 95th/99th-percentile absolute errors plus median/worst-case profile figures are reported. Inference of the largest model is ~1 ms on CPU. Claims are scoped to accuracy inside the synthetic operational domain.
Significance. If the reported synthetic accuracy transfers, the work supplies the first large-scale, multi-architecture ML equilibrium framework for ADITYA-U, with practical utility for real-time control, rapid discharge analysis, and actuator planning. Strengths include the experimentally motivated input ranges, shot-wise splits that avoid temporal leakage, systematic PCA-vs-CNN comparison, physics-informed constraints (monotonicity of q, progressive GS residual with uncertainty weighting), and explicit quantitative error distributions rather than only mean metrics. The inverse coil-current model and cascade-ready scalar predictors are useful engineering contributions within the stated domain. The principal limitation is that all validation remains synthetic; experimental transfer is left for future work and is already acknowledged.
major comments (3)
- The central claim is carefully scoped to the synthetic domain, yet the abstract, introduction and conclusion repeatedly advertise utility for real-time control and experimental planning. No experimental reconstruction or even a single pyIPREQ-vs-diagnostic comparison on held-out ADITYA-U shots is shown. A minimal experimental sanity check (or a clearly labeled “synthetic-only” caveat in the abstract) is needed so that the transfer premise of §2.1–2.2 does not over-extend the reported numbers.
- §2.2: the βp prior is a linear model trained on only 39 diamagnetic discharges; Fig. 3 shows clear bias at high/low βp and the residual noise N(0,0.04) is then injected into all 100k cases. Because βp directly sets the Jφ parameter β (Eq. 2) and therefore shapes ℓi and q, the sensitivity of the downstream scalar and profile errors to this prior should be quantified (e.g., by re-training with a wider or alternative βp distribution).
- §4.2 and cascade discussion: Rax/Zax (and later q1, ψaxs/ψlim) are used as inputs to subsequent models, yet training uses ground-truth values; cascading error is never measured. Because the intended real-time pipeline is cascaded, the reported test-set errors for βp/ℓi, q-profile and ψ-profile are optimistic. A short end-to-end cascade evaluation on the test shots is required.
minor comments (6)
- Table 1 lists coil geometry but omits the actual current ranges present in the dataset; a short summary row would help readers judge the inverse-model domain.
- Figs. 9–22 report absolute errors with medians and 95th percentiles in the captions; adding the same numbers to the main text or a summary table would improve readability.
- Eq. (3) and the progressive factor f are introduced without a short derivation or reference to the uncertainty-weighting paper beyond [21]; a one-sentence justification of the numerical constants (−50, 0.15) would help reproducibility.
- The 1d-CNN predicts 100 values of q′ and reconstructs q by integrating from the separately predicted q1; the accumulation of integration error should be stated explicitly when comparing core-region accuracy with the PCA model.
- Minor typographical issues: “two-dimensionalpoloidalfluxprofiles” (abstract), “physics-informedneuralnetworks” (abstract), and occasional missing spaces after commas in the introduction.
- Acknowledgment notes LLM rephrasing; a brief statement that scientific content and all numerical results were verified by the authors would be appropriate.
Circularity Check
No significant circularity: models are intentional surrogates of pyIPREQ on a filtered synthetic library; accuracy claims are scoped to that domain and do not reduce by construction to inputs.
full rationale
The paper's derivation chain is standard surrogate modeling: experimental coil/Ip/Bphi ranges (plus a linear beta_p prior fitted on 39 diamagnetic shots) seed pyIPREQ free-boundary solutions under a fixed J_phi ansatz (Eq. 2) and hard filters; dense/PCA/CNN/PINN networks then learn the resulting input-to-output map and are evaluated on shot-wise held-out synthetic cases. Reported errors (Tables 9/11/13/15, Figs. 9-42) therefore measure fidelity to the solver library, not an independent physical derivation. The beta_p linear model (Sec. 2.2) is an acknowledged stochastic prior used only for dataset generation, not a fitted parameter later re-labeled as a first-principles prediction. Self-citations to pyIPREQ [13,14] simply identify the data-generation tool; no uniqueness theorem or load-bearing external result is imported from the authors. Inverse coil-current mapping is explicitly noted as non-unique and statistical. PINN GS residuals act as regularization, not as a circular definition of psi. All central claims are carefully limited to 'within the operational domain represented by the dataset' (Abstract, Sec. 7). No step reduces a claimed prediction to its own inputs by construction.
Assumptions & free parameters
free parameters (7)
- βp linear-model residual noise =
N(0, 0.04), clip (0.05, 0.4)
- α sampling distribution for Jφ =
N(4.5, 0.5)
- Rax target sampling distribution =
N(0.7525, 0.0075) m
- Equilibrium filtering thresholds =
as listed in §2.1
- 2% uniform noise on coil currents, Ip, Bφ =
2% uniform
- PINN loss weights / uncertainty sigmas and progressive factor f =
f=sigmoid(-50(MSEψ+MSEJ-0.15)); σ trainable
- Network architectures and hyperparameters =
Tables 3–14
assumptions (6)
- domain assumption Axisymmetric ideal MHD equilibrium is described by the Grad–Shafranov equation (Eq. 1).
- domain assumption Toroidal current density follows the three-parameter form Jφ ∝ (β R/R0 + (1-β) R0/R) (1-(1-ψ̄)^α)^γ (Eq. 2).
- domain assumption Eddy currents can be neglected for the selected flat-top windows because the vessel is toroidally discontinuous and temporal variations are weak.
- ad hoc to paper Circular limiter plasmas near flat-top with Ip>100 kA for ≥80 ms adequately represent the operational domain of interest.
- standard math Shot-wise train/val/test splits prevent temporal leakage and give a realistic generalization estimate to unseen discharges.
- ad hoc to paper A linear regression on transformed Ip, coil currents, Vloop and timing features plus Gaussian noise is a sufficient stochastic prior for βp when diamagnetic data are sparse.
Cite this review
Pith. "Pith review of Deep Learning Models for ADITYA-U MHD Equilibrium." pith.science (2026). https://pith.science/paper/6WIRAPOV
@misc{pith2026260704865,
author = {Pith},
title = {Pith review of: Deep Learning Models for ADITYA-U MHD Equilibrium},
year = {2026},
howpublished = {\url{https://pith.science/paper/6WIRAPOV}},
note = {Machine review of arXiv:2607.04865}
}
read the original abstract
This work presents deep learning models to predict magnetohydrodynamic equilibrium parameters and profiles for the ADITYA-U tokamak. A synthetic free-boundary equilibrium dataset consisting of 100,760 cases was generated using the pyIPREQ Grad-Shafranov solver, with inputs derived from 766 ADITYA-U plasma discharges and constrained to experimentally relevant circular limiter plasmas near the flat-top phase. Several deep learning approaches were investigated for predicting scalar equilibrium quantities, one-dimensional safety factor profiles and two-dimensional poloidal flux profiles. These approaches included Dense neural networks, principal component analysis based reduced-order models, one-dimensional and two-dimensional convolutional neural networks, and physics-informed neural networks incorporating Grad-Shafranov residual constraints. In addition, an inverse model was developed to estimate poloidal field coil currents from desired plasma equilibrium conditions. The results demonstrate that key equilibrium parameters and profiles can be accurately estimated within the operational domain represented by the dataset. The developed models provide a computationally efficient alternative to conventional equilibrium estimation and can be useful for real-time plasma control, rapid equilibrium analysis, and experimental planning in ADITYA-U operations.
Figures
Figures from the paper (31 more)
Reference graph
Works this paper leans on
-
[1]
L.L. Lao, H. St. John, R.D. Stambaugh, A.G. Kellman, andW.Pfeiffer. Reconstructionofcurrentprofileparam- eters and plasma shapes in tokamaks.Nuclear Fusion, 25(11):1611, Nov 1985
1985
-
[2]
Bak, S.G
Semin Joung, Jaewook Kim, Sehyun Kwak, J.G. Bak, S.G. Lee, H.S. Han, H.S. Kim, Geunho Lee, Daeho Kwon, and Y.-C. Ghim. Deep neural network Grad–Shafranov solver constrained with measured mag- netic signals.Nuclear Fusion, 60(1):016034, Dec 2019
2019
-
[3]
Wai, M.D
J.T. Wai, M.D. Boyer, and E. Kolemen. Neural net modeling of equilibria in NSTX-U.Nuclear Fusion, 62(8):086042, Jul 2022
2022
-
[4]
EAST discharge prediction with- out integrating simulation results.Nuclear Fusion, 62(12):126060, Nov 2022
Chenguang Wan, Zhi Yu, Alessandro Pau, Xiaojuan Liu, and Jiangang Li. EAST discharge prediction with- out integrating simulation results.Nuclear Fusion, 62(12):126060, Nov 2022
2022
-
[5]
A machine-learning-based tool for last closed-flux surface reconstruction on tokamaks.Nuclear Fusion, 63(5):056019, Apr 2023
Chenguang Wan, Zhi Yu, Alessandro Pau, Olivier Sauter, Xiaojuan Liu, Qiping Yuan, and Jiangang Li. A machine-learning-based tool for last closed-flux surface reconstruction on tokamaks.Nuclear Fusion, 63(5):056019, Apr 2023
2023
-
[6]
Fast equilibrium reconstruction by deep learning on EAST tokamak.AIP Advances, 13(7):075007, Jul 2023
Jingjing Lu, Youjun Hu, Nong Xiang, and Youwen Sun. Fast equilibrium reconstruction by deep learning on EAST tokamak.AIP Advances, 13(7):075007, Jul 2023
2023
-
[7]
Reconstruction of tokamak plasma safety factor profile using deep learning.Nuclear Fusion, 63(8):086020, Jun 2023
Xishuo Wei, Shuying Sun, William Tang, Zhihong Lin, Hongfei Du, and Ge Dong. Reconstruction of tokamak plasma safety factor profile using deep learning.Nuclear Fusion, 63(8):086020, Jun 2023
2023
-
[8]
Madireddy, C
S. Madireddy, C. Akçay, S. E. Kruger, T. Bechtel Amara, X. Sun, J. McClenaghan, J. Koo, A. Samad- dar, Y. Liu, P. Balaprakash, and L. L. Lao. EFIT- Prime: Probabilistic and physics-constrained reduced- order neural network model for equilibrium reconstruc- tion in DIII-D.Physics of Plasmas, 31(9):092505, Sep 2024
2024
Show all 21 references
-
[9]
Novella Rutigliano, Andrea Murari, Pasquale Gaudio, Michela Gelfusa, Riccardo Rossi, on behalf of JET Con- tributors, and the EUROfusion Tokamak Exploita- tion Team. Multi-diagnostics reconstruction of magnetic equilibrium and kinetic profiles using physics-informed neural net...
2026
-
[10]
Machine learning prediction of plasma behav- ior from discharge configurations on WEST.Nuclear Fusion, 66(4):044001, Mar 2026
Chenguang Wan, Feda Almuhisen, Philippe Moreau, Rémy Nouailletas, Zhisong Qu, Youngwoo Cho, Robin Varennes, Kyungtak Lim, Kunpeng Li, Jia Huang, Wei- dong Chen, Jiangang Li, Xavier Garbet, and WEST Team. Machine learning prediction of plasma behav- ior from discharge configura...
2026
-
[11]
Zheng, S.F
G.H. Zheng, S.F. Liu, H.S. Xie, X. Gu, Z.Y. Chen, X.C. Lun, Y. Liu, J. Li, D. Guo, R.Y. Tao, et al. EFIT-mini: an embedded, multi-task neural network- driven equilibrium inversion algorithm.Nuclear Fusion, 65(10):106008, Sep 2025
2025
-
[12]
Tanna, J
R.L. Tanna, J. Ghosh, K.A. Jadeja, Rohit Kumar, Suman Aich, K.M. Patel, Harshita Raj, Kaushlender Singh, Suman Dolui, Kajal Shah, et al. Overview of physics results from the ADITYA-U tokamak and fu- ture experiments.Nuclear Fusion, 64(11):112011, Aug 2024
2024
-
[13]
Singh, Suman Aich, Jaga- bandhu Kumar, Rohit Kumar, and Daniel Raju
Udaya Maurya, Amit K. Singh, Suman Aich, Jaga- bandhu Kumar, Rohit Kumar, and Daniel Raju. Plasma position constrained free-boundary MHD equilibrium in Tokamaks using pyIPREQ.Physics of Plasmas, 32(10):102501, Oct 2025
2025
-
[14]
pyIPREQ: Free Boundary MHD Equiliibrium Code.https://github.com/udy11/ pyIPREQ,https://gitlab.com/udy11/pyIPREQ, 2025
Udaya Maurya. pyIPREQ: Free Boundary MHD Equiliibrium Code.https://github.com/udy11/ pyIPREQ,https://gitlab.com/udy11/pyIPREQ, 2025
2025
-
[15]
Plasma column position measurements using magnetic diagnostics in ADITYA-U tokamak.Plasma Research Express, 3(3):035005, Sep 2021
S Aich, R Kumar, T M Macwan, D Kumavat, S Jha, R L Tanna, K Sathyanarayana, J Ghosh, K A Jadeja, K Pa- tel, et al. Plasma column position measurements using magnetic diagnostics in ADITYA-U tokamak.Plasma Research Express, 3(3):035005, Sep 2021
2021
-
[16]
S. Aich, T. M. Macwan, K. Galodiya, K. Singh, S. Dolui, J. Ghosh, R. L. Tanna, Abhijeet Kumar, H. Mandliya, E. V. Praveenlal, et al. Design and measurements of the diamagnetic loop in Aditya-U tokamak.Radiation Effects and Defects in Solids, 180(3-4):435–446, 2024
2024
-
[17]
Nu- merical determination of axisymmetric toroidal magne- tohydrodynamic equilibria.Journal of Computational Physics, 32(2):212–234, 1979
J.LJohnson, H.EDalhed, J.MGreene, R.CGrimm, Y.Y Hsieh, S.C Jardin, J Manickam, M Okabayashi, R.G Storer, A.M.M Todd, D.E Voss, and K.E Weimer. Nu- merical determination of axisymmetric toroidal magne- tohydrodynamic equilibria.Journal of Computational Physics, 32(2):212–234, 1979
1979
-
[18]
Realtimeverticalpo- sition estimation of plasma column using fast imaging in Aditya-U tokamak.Nuclear Fusion, Jul 2025
Suman Aich, Sharvil Patel, Laxmikanta Pradhan, Ashok Kumar Kumawat, Bharat Hegde, Kalpesh Ga- lodiya, RLTanna, KumarpalsinhAJadeja, MalayBikas Chowdhuri, NandiniYadava, etal. Realtimeverticalpo- sition estimation of plasma column using fast imaging in Aditya-U tokamak.Nuclear ...
2025
-
[19]
Corrado, Andy Davis, Jeffrey Dean, Matthieu Devin, et al
Martín Abadi, Ashish Agarwal, Paul Barham, Eugene Brevdo, Zhifeng Chen, Craig Citro, Greg S. Corrado, Andy Davis, Jeffrey Dean, Matthieu Devin, et al. Ten- sorFlow: Large-Scale Machine Learning on Heteroge- neous Systems, 2015. Software available from tensor- flow.org
2015
-
[20]
KerasTuner
Tom O’Malley, Elie Bursztein, James Long, François Chollet, Haifeng Jin, Luca Invernizzi, et al. KerasTuner. https://github.com/keras-team/keras-tuner, 2019
2019
-
[21]
Multi- Task Learning Using Uncertainty to Weigh Losses for Scene Geometry and Semantics
Alex Kendall, Yarin Gal, and Roberto Cipolla. Multi- Task Learning Using Uncertainty to Weigh Losses for Scene Geometry and Semantics. InProceedings of the IEEE Conference on Computer Vision and Pattern Recognition (CVPR), Jun 2018. 18
2018
Reviewed July 11, 2026 · model on record in the stance chip above.
Discussion (0). Continue with ORCID to comment.