Pith. sign in

REVIEW 5 cited by

PDEBENCH: An Extensive Benchmark for Scientific Machine Learning

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2210.07182 v7 pith:33FY6EGQ submitted 2022-10-13 cs.LG cs.CVphysics.flu-dynphysics.geo-ph

classification cs.LGcs.CVphysics.flu-dynphysics.geo-ph
keywords pdebenchbenchmarklearningmachinemethodsmodelsproblemsscientific
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Machine learning-based modeling of physical systems has experienced increased interest in recent years. Despite some impressive progress, there is still a lack of benchmarks for Scientific ML that are easy to use but still challenging and representative of a wide range of problems. We introduce PDEBench, a benchmark suite of time-dependent simulation tasks based on Partial Differential Equations (PDEs). PDEBench comprises both code and data to benchmark the performance of novel machine learning models against both classical numerical simulations and machine learning baselines. Our proposed set of benchmark problems contribute the following unique features: (1) A much wider range of PDEs compared to existing benchmarks, ranging from relatively common examples to more realistic and difficult problems; (2) much larger ready-to-use datasets compared to prior work, comprising multiple simulation runs across a larger number of initial and boundary conditions and PDE parameters; (3) more extensible source codes with user-friendly APIs for data generation and baseline results with popular machine learning models (FNO, U-Net, PINN, Gradient-Based Inverse Method). PDEBench allows researchers to extend the benchmark freely for their own purposes using a standardized API and to compare the performance of new models to existing baseline methods. We also propose new evaluation metrics with the aim to provide a more holistic understanding of learning methods in the context of Scientific ML. With those metrics we identify tasks which are challenging for recent ML methods and propose these tasks as future challenges for the community. The code is available at https://github.com/pdebench/PDEBench.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 5 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Walrus: A Cross-Domain Foundation Model for Continuum Dynamics

    cs.LG 2025-11 conditional novelty 7.0 of 10

    Walrus, a transformer pretrained on diverse 2D/3D continuum-dynamics data, outperforms prior physics foundation models across most downstream emulation tasks.

  2. Modeling turbulent and self-gravitating fluids with Fourier neural operators

    astro-ph.GA 2025-07 conditional novelty 6.0 of 10

    Fourier neural operators trained on 2D projected views of 3D astrophysical simulations can forecast subsequent projected density and velocity snapshots with 5-25% RMS error, though a constant hidden magnetic field mea...

  3. Hierarchical Implicit Neural Emulators

    cs.LG 2025-06 conditional novelty 6.0 of 10

    Feeding a hierarchy of predicted coarse-grained future states into an autoregressive neural emulator greatly improves long-term stability for 2D turbulent flow forecasting.

  4. DRIFT: Direct Reduced Fourier Transforms for Distributed Spectral Neural Operators

    cs.DC 2026-07 conditional novelty 5.0 of 10

    Distributed Fourier Neural Operators can compute their truncated spectra with local partial DFTs and two collectives on the kept modes, giving exact results with communication independent of grid resolution.

  5. Data assimilation via model reference adaptation for linear and nonlinear dynamical systems

    math.OC 2026-02 conditional novelty 5.0 of 10

    A model reference adaptive system recovers unknown coefficients in nonlinear parabolic PDEs from time-series data in four benchmark tests.

Pith tools