TY - RPRT TI - Online Reinforcement Learning in the Met Office Unified Model through Distributed Model-Agent Coupling AU - Pritthijit Nath AU - Sebastian Schemm AU - Peter Haynes AU - Emily Shuckburgh AU - Mark Webb PY - 2026 UR - https://arxiv.org/abs/2609.02566 ID - 2609.02566 ER -