Pith. sign in

REVIEW 7 cited by

PDEBENCH: An Extensive Benchmark for Scientific Machine Learning

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2210.07182 v7 pith:33FY6EGQ submitted 2022-10-13 cs.LG cs.CVphysics.flu-dynphysics.geo-ph

PDEBENCH: An Extensive Benchmark for Scientific Machine Learning

classification cs.LG cs.CVphysics.flu-dynphysics.geo-ph
keywords pdebenchbenchmarklearningmachinemethodsmodelsproblemsscientific
verification ladder T0 review T1 audit T2 compute T3 formal T4 reserved
0 comments
read the original abstract

Machine learning-based modeling of physical systems has experienced increased interest in recent years. Despite some impressive progress, there is still a lack of benchmarks for Scientific ML that are easy to use but still challenging and representative of a wide range of problems. We introduce PDEBench, a benchmark suite of time-dependent simulation tasks based on Partial Differential Equations (PDEs). PDEBench comprises both code and data to benchmark the performance of novel machine learning models against both classical numerical simulations and machine learning baselines. Our proposed set of benchmark problems contribute the following unique features: (1) A much wider range of PDEs compared to existing benchmarks, ranging from relatively common examples to more realistic and difficult problems; (2) much larger ready-to-use datasets compared to prior work, comprising multiple simulation runs across a larger number of initial and boundary conditions and PDE parameters; (3) more extensible source codes with user-friendly APIs for data generation and baseline results with popular machine learning models (FNO, U-Net, PINN, Gradient-Based Inverse Method). PDEBench allows researchers to extend the benchmark freely for their own purposes using a standardized API and to compare the performance of new models to existing baseline methods. We also propose new evaluation metrics with the aim to provide a more holistic understanding of learning methods in the context of Scientific ML. With those metrics we identify tasks which are challenging for recent ML methods and propose these tasks as future challenges for the community. The code is available at https://github.com/pdebench/PDEBench.

discussion (0)

Sign in with ORCID, Apple, or X to comment. Anyone can read and Pith papers without signing in.

Forward citations

Cited by 7 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score.

  1. ELADO: Elliptic PDE Assessment Datasets for Operator Learning

    cs.LG 2026-06 unverdicted novelty 7.0

    ELADO provides a benchmark suite of elliptic PDE datasets designed to isolate and quantify failure modes in neural operator architectures.

  2. Walrus: A Cross-Domain Foundation Model for Continuum Dynamics

    cs.LG 2025-11 conditional novelty 7.0

    Walrus, a transformer pretrained on diverse 2D/3D continuum-dynamics data, outperforms prior physics foundation models across most downstream emulation tasks.

  3. Mechanism Learning: Prototype-Anchored Mechanism Inference for Scientific Forecasting

    cs.LG 2026-05 unverdicted novelty 6.0

    Mechanism learning infers active local evolution rules via prototype-anchored descriptors to achieve more robust forecasting than direct state prediction on benchmarks like Burgers, WeatherBench2, and Lorenz96.

  4. Late Fusion Neural Operators for Extrapolation Across Parameter Space in Partial Differential Equations

    cs.LG 2026-04 unverdicted novelty 6.0

    Late Fusion Neural Operators disentangle state and parameter learning to outperform FNO and CAPE-FNO on advection, Burgers, and reaction-diffusion PDEs with 72% average RMSE reduction in and out of domain.

  5. Flow marching for a generative PDE foundation model

    cs.LG 2025-09 unverdicted novelty 6.0

    Flow Marching jointly samples noise and physical time to learn a velocity field for generative PDE modeling, paired with a latent autoencoder and efficient transformer for large-scale pretraining on 2.5M trajectories.

  6. DRIFT: Direct Reduced Fourier Transforms for Distributed Spectral Neural Operators

    cs.DC 2026-07 conditional novelty 5.0

    Distributed Fourier Neural Operators can compute their truncated spectra with local partial DFTs and two collectives on the kept modes, giving exact results with communication independent of grid resolution.

  7. Data assimilation via model reference adaptation for linear and nonlinear dynamical systems

    math.OC 2026-02 conditional novelty 5.0

    A model reference adaptive system recovers unknown coefficients in nonlinear parabolic PDEs from time-series data in four benchmark tests.