REVIEW 17 cited by
GraphCast: Learning skillful medium-range global weather forecasting
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
read the original abstract
Global medium-range weather forecasting is critical to decision-making across many social and economic domains. Traditional numerical weather prediction uses increased compute resources to improve forecast accuracy, but cannot directly use historical weather data to improve the underlying model. We introduce a machine learning-based method called "GraphCast", which can be trained directly from reanalysis data. It predicts hundreds of weather variables, over 10 days at 0.25 degree resolution globally, in under one minute. We show that GraphCast significantly outperforms the most accurate operational deterministic systems on 90% of 1380 verification targets, and its forecasts support better severe event prediction, including tropical cyclones, atmospheric rivers, and extreme temperatures. GraphCast is a key advance in accurate and efficient weather forecasting, and helps realize the promise of machine learning for modeling complex dynamical systems.
Forward citations
Cited by 17 Pith papers
-
FourCastNet 3: A geometric approach to probabilistic machine-learning weather forecasting at scale
A purely convolutional, spherical-geometry weather model trained with a combined spatial and spectral CRPS loss delivers GenCast-level skill, IFS-beating accuracy, and stable spectra out to 60 days.
-
Skillful joint probabilistic weather forecasting from marginals
FGN, a neural weather model trained only on per-location forecast scores, produces more accurate global ensemble forecasts than GenCast and captures realistic spatial correlations.
-
HourGlass: A probabilistic data-driven temporal downscaler for global and regional weather forecasting
HourGlass probabilistically reconstructs hourly weather evolution between 6-hourly forecast states using CRPS training on NWP trajectories, preserving skill and small-scale variability better than deterministic downscalers.
-
A PMP-inspired Evaluation Framework for Assessing Deep-Learning Earth System Models
The paper presents a PMP-based evaluation framework to test deep-learning Earth system models on climatology and modes of variability using observational data.
-
On the under-reaching phenomenon in message-passing neural PDE solvers: revisiting the CFL condition
Graph neural network PDE solvers require at least a CFL-like number of message passes for hyperbolic problems and domain-spanning passes for parabolic and elliptic problems.
-
Generative Lagrangian data assimilation for ocean dynamics under extreme sparsity
A neural operator-conditioned diffusion model reconstructs ocean surface states with high-wavenumber fidelity from 99% to 99.9% sparse observations, outperforming standard UNET and FNO baselines.
-
LaDCast: A Latent Diffusion Model for Medium-Range Ensemble Weather Forecasting
A latent diffusion model generates medium-range global weather ensembles at 1.5 degrees that match ECMWF IFS-ENS deterministic skill at lower compute, with weaker probabilistic spread and anecdotal cyclone advantages.
-
DEF: Diffusion-augmented Ensemble Forecasting
A conditional diffusion model that generates perturbed initial states can turn any deterministic neural weather forecast model into an ensemble, with measured error reduction on a single ERA5 case study.
-
PEAR: Equal Area Weather Forecasting on the Sphere
A transformer weather model operating natively on the equal-area HEALPix grid beats an equiangular-grid counterpart at longer lead times with 2.6x fewer parameters.
-
A robust and stable hybrid neural network/finite element method for 2D flows that generalizes to different geometries
Replay training on the DNN-MG's own perturbed trajectories removes the long-time instability of the hybrid Navier-Stokes solver, while Transformers or larger patches improve accuracy on unseen geometries.
-
Finetuning AI Foundation Models to Develop Subgrid-Scale Parameterizations: A Case Study on Atmospheric Gravity Waves
Fine-tuning Prithvi WxC to predict gravity wave fluxes beats an Attention U-Net baseline on one month of ERA5 validation, including in upper stratospheric levels absent from the pre-training data.
-
Physics-Guided Learning of Meteorological Dynamics for Weather Downscaling and Forecasting
PhyDL-NWP trains neural surrogates with a fitted PDE regularizer and a latent force term, improving weather downscaling and fine-tuning forecasts over 17 baselines.
-
Enhancing a high resolution data-driven weather prediction model with surface descriptors
Surface descriptors cut 2 m temperature and 10 m wind MAE by 1.9% and 3.0% domain-wide (about 12% for urban temperature) in a stretched-grid data-driven weather model, and glacier removal raises temperature without re...
-
The Virtuous Cycle of Quantum-Classical Machine Learning
Classical ML and quantum computing mutually accelerate each other through error correction, control, simulation data, and quantum-native learning, forming a virtuous cycle toward quantum intelligence.
-
Modernizing CNN-based Weather Forecast Model towards Higher Computational Efficiency
A 7-million-parameter convolutional weather model trains in 12 hours on one GPU and is reported to match or beat much larger AI and numerical weather models in medium-range forecasts.
-
Structure of System-Specific Hazard Modeling for Physical Infrastructures in a Climate Change Context
A hazard is defined as any intensity exceeding a system-specific threshold, and the paper uses this to generate wind and temperature hazard maps for different infrastructure types.
-
Exploratory Analysis of Deep Learning Models for Forecasting Meteorological Parameters in the Agricultural Sector
On Ioannina ERA5 hourly data, hybrid 1D-CNN–RNN models raise a composite WQS by 1.22–1.63% at 24 h and 0.44–0.45% at 168 h over the best single-layer GRU/LSTM baselines.
Discussion (0). Continue with ORCID to comment.