Medium-range weather forecasts from a graph neural network beat the operational physics model
Medium-range weather forecasts from a graph neural network beat the operational physics model is graded peer reviewed on whataifound.org, with the AI's role graded ai-led.
GraphCast produced more accurate 10-day forecasts than ECMWF's HRES on roughly 90% of the evaluated variable and lead-time combinations, in under a minute on a single TPU against hours on a supercomputer.
- Verification
- Peer reviewed
- Autonomy
- AI-led
- Lab
- Google DeepMind
- Model
- GraphCast
- Field
- Climate science
- Date
- 2023-11-14
- Human collaborators
- Remi Lam, Peter Battaglia
What was found
GraphCast takes the two most recent global atmospheric states and predicts the next state six hours ahead on a roughly 0.25° grid, applied repeatedly to reach ten days. It is trained on ERA5 reanalysis rather than on the equations of motion, and it beat HRES — the operational gold standard — on the large majority of evaluated variables, rising to almost all of them within the troposphere. It also flagged severe weather events, including cyclone tracks, earlier than the physics model despite never being trained to look for them. Published in Science, with code and weights released.
Novelty check
Machine-learning weather models existed before (FourCastNet, Pangu-Weather), and Pangu had already reported beating HRES on some measures. GraphCast's contribution is the breadth of the win across variables and lead times, at operational resolution, with public code — which is what moved national weather services to adopt learned models.
Caveats
GraphCast is trained on ECMWF reanalysis and initialised from ECMWF analyses, so it depends on the physics-based system it is measured against rather than replacing it. Deterministic models trained on a mean-squared-error objective blur fine structure and underestimate extremes such as peak cyclone intensity. It has no conservation guarantees and cannot forecast anything outside its trained variables. Later ensemble models such as GenCast address part of this.
Sources
How this is graded
whataifound.org grades every entry on two axes: verification (how solid the result is, from a machine-checked proof down to refuted) and autonomy (how much the AI did versus its human collaborators). This finding is peer reviewed and ai-led. Full definitions are in the methodology.
Cite this entry
whataifound.org (2023). Medium-range weather forecasts from a graph neural network beat the operational physics model. whataifound.org: A Registry of AI Scientific and Mathematical Discoveries. https://whataifound.org/finding/2023-11-14-graphcast