AIIC AI Intelligence Centre

SOURCE-LINKED INTELLIGENCE

Online Reinforcement Learning in the Met Office Unified Model through Distributed Model-Agent Coupling

arXiv · AI, language, vision and robotics · article · Sep 2, 2026 · UTC

Machine-learnt corrections can complement numerical weather prediction provided that they operate stably within an evolving numerical model. In this study, we couple the Met Office (UKMO) Unified Model (UM) with distributed reinforcement-learning agents through rank-local tensors. A column-aware deep deterministic policy gradient (DDPG) actor uses local vertical structure together with full-column context to apply bounded corrections to potential temperature and horizontal wind. During training, we perform ten nudged 6-hr 12-min forecasts, with nudging towards the UKMO operational analysis pro

Read original source ↗ Open in workspace

recordType
paper
region
Global

Evidence & attribution

First collected: 2026-09-21T05:32:15.665Z. This is not the publication date.