Search | arXiv e-print repository

arXiv:2412.15832 [pdf, other]

AIFS-CRPS: Ensemble forecasting using a model trained with a loss function based on the Continuous Ranked Probability Score

Authors: Simon Lang, Mihai Alexe, Mariana C. A. Clare, Christopher Roberts, Rilwan Adewoyin, Zied Ben Bouallègue, Matthew Chantry, Jesper Dramsch, Peter D. Dueben, Sara Hahner, Pedro Maciel, Ana Prieto-Nemesio, Cathal O'Brien, Florian Pinault, Jan Polster, Baudouin Raoult, Steffen Tietsche, Martin Leutbecher

Abstract: Over the last three decades, ensemble forecasts have become an integral part of forecasting the weather. They provide users with more complete information than single forecasts as they permit to estimate the probability of weather events by representing the sources of uncertainties and accounting for the day-to-day variability of error growth in the atmosphere. This paper presents a novel approach… ▽ More Over the last three decades, ensemble forecasts have become an integral part of forecasting the weather. They provide users with more complete information than single forecasts as they permit to estimate the probability of weather events by representing the sources of uncertainties and accounting for the day-to-day variability of error growth in the atmosphere. This paper presents a novel approach to obtain a weather forecast model for ensemble forecasting with machine-learning. AIFS-CRPS is a variant of the Artificial Intelligence Forecasting System (AIFS) developed at ECMWF. Its loss function is based on a proper score, the Continuous Ranked Probability Score (CRPS). For the loss, the almost fair CRPS is introduced because it approximately removes the bias in the score due to finite ensemble size yet avoids a degeneracy of the fair CRPS. The trained model is stochastic and can generate as many exchangeable members as desired and computationally feasible in inference. For medium-range forecasts AIFS-CRPS outperforms the physics-based Integrated Forecasting System (IFS) ensemble for the majority of variables and lead times. For subseasonal forecasts, AIFS-CRPS outperforms the IFS ensemble before calibration and is competitive with the IFS ensemble when forecasts are evaluated as anomalies to remove the influence of model biases. △ Less

Submitted 20 December, 2024; originally announced December 2024.

arXiv:2412.15687 [pdf, other]

GraphDOP: Towards skilful data-driven medium-range weather forecasts learnt and initialised directly from observations

Authors: Mihai Alexe, Eulalie Boucher, Peter Lean, Ewan Pinnington, Patrick Laloyaux, Anthony McNally, Simon Lang, Matthew Chantry, Chris Burrows, Marcin Chrust, Florian Pinault, Ethel Villeneuve, Niels Bormann, Sean Healy

Abstract: We introduce GraphDOP, a new data-driven, end-to-end forecast system developed at the European Centre for Medium-Range Weather Forecasts (ECMWF) that is trained and initialised exclusively from Earth System observations, with no physics-based (re)analysis inputs or feedbacks. GraphDOP learns the correlations between observed quantities - such as brightness temperatures from polar orbiters and geos… ▽ More We introduce GraphDOP, a new data-driven, end-to-end forecast system developed at the European Centre for Medium-Range Weather Forecasts (ECMWF) that is trained and initialised exclusively from Earth System observations, with no physics-based (re)analysis inputs or feedbacks. GraphDOP learns the correlations between observed quantities - such as brightness temperatures from polar orbiters and geostationary satellites - and geophysical quantities of interest (that are measured by conventional observations), to form a coherent latent representation of Earth System state dynamics and physical processes, and is capable of producing skilful predictions of relevant weather parameters up to five days into the future. △ Less

Submitted 20 December, 2024; originally announced December 2024.

Comments: 23 pages, 15 figures

arXiv:2410.16343 [pdf, other]

Hydra-LSTM: A semi-shared Machine Learning architecture for prediction across Watersheds

Authors: Karan Ruparell, Robert J. Marks, Andy Wood, Kieran M. R. Hunt, Hannah L. Cloke, Christel Prudhomme, Florian Pappenberger, Matthew Chantry

Abstract: Long Short Term Memory networks (LSTMs) are used to build single models that predict river discharge across many catchments. These models offer greater accuracy than models trained on each catchment independently if using the same data. However, the same data is rarely available for all catchments. This prevents the use of variables available only in some catchments, such as historic river dischar… ▽ More Long Short Term Memory networks (LSTMs) are used to build single models that predict river discharge across many catchments. These models offer greater accuracy than models trained on each catchment independently if using the same data. However, the same data is rarely available for all catchments. This prevents the use of variables available only in some catchments, such as historic river discharge or upstream discharge. The only existing method that allows for optional variables requires all variables to be considered in the initial training of the model, limiting its transferability to new catchments. To address this limitation, we develop the Hydra-LSTM. The Hydra-LSTM processes variables used across all catchments and variables used in only some catchments separately to allow general training and use of catchment-specific data in individual catchments. The bulk of the model can be shared across catchments, maintaining the benefits of multi-catchment models to generalise, while also benefitting from the advantages of using bespoke data. We apply this methodology to 1 day-ahead river discharge prediction in the Western US, as next-day river discharge prediction is the first step towards prediction across longer time scales. We obtain state-of-the-art performance, generating more accurate median and quantile predictions than Multi-Catchment and Single-Catchment LSTMs while allowing local forecasters to easily introduce and remove variables from their prediction set. We test the ability of the Hydra-LSTM to incorporate catchment-specific data by introducing historical river discharge as a catchment-specific input, outperforming state-of-the-art models without needing to train an entirely new model. △ Less

Submitted 21 October, 2024; originally announced October 2024.

arXiv:2409.18529 [pdf, other]

Robustness of AI-based weather forecasts in a changing climate

Authors: Thomas Rackow, Nikolay Koldunov, Christian Lessig, Irina Sandu, Mihai Alexe, Matthew Chantry, Mariana Clare, Jesper Dramsch, Florian Pappenberger, Xabier Pedruzo-Bagazgoitia, Steffen Tietsche, Thomas Jung

Abstract: Data-driven machine learning models for weather forecasting have made transformational progress in the last 1-2 years, with state-of-the-art ones now outperforming the best physics-based models for a wide range of skill scores. Given the strong links between weather and climate modelling, this raises the question whether machine learning models could also revolutionize climate science, for example… ▽ More Data-driven machine learning models for weather forecasting have made transformational progress in the last 1-2 years, with state-of-the-art ones now outperforming the best physics-based models for a wide range of skill scores. Given the strong links between weather and climate modelling, this raises the question whether machine learning models could also revolutionize climate science, for example by informing mitigation and adaptation to climate change or to generate larger ensembles for more robust uncertainty estimates. Here, we show that current state-of-the-art machine learning models trained for weather forecasting in present-day climate produce skillful forecasts across different climate states corresponding to pre-industrial, present-day, and future 2.9K warmer climates. This indicates that the dynamics shaping the weather on short timescales may not differ fundamentally in a changing climate. It also demonstrates out-of-distribution generalization capabilities of the machine learning models that are a critical prerequisite for climate applications. Nonetheless, two of the models show a global-mean cold bias in the forecasts for the future warmer climate state, i.e. they drift towards the colder present-day climate they have been trained for. A similar result is obtained for the pre-industrial case where two out of three models show a warming. We discuss possible remedies for these biases and analyze their spatial distribution, revealing complex warming and cooling patterns that are partly related to missing ocean-sea ice and land surface information in the training data. Despite these current limitations, our results suggest that data-driven machine learning models will provide powerful tools for climate science and transform established approaches by complementing conventional physics-based models. △ Less

Submitted 27 September, 2024; originally announced September 2024.

Comments: 14 pages, 4 figures

arXiv:2409.02891 [pdf, other]

Regional data-driven weather modeling with a global stretched-grid

Authors: Thomas Nils Nipen, Håvard Homleid Haugen, Magnus Sikora Ingstad, Even Marius Nordhagen, Aram Farhad Shafiq Salihi, Paulina Tedesco, Ivar Ambjørn Seierstad, Jørn Kristiansen, Simon Lang, Mihai Alexe, Jesper Dramsch, Baudouin Raoult, Gert Mertes, Matthew Chantry

Abstract: A data-driven model (DDM) suitable for regional weather forecasting applications is presented. The model extends the Artificial Intelligence Forecasting System by introducing a stretched-grid architecture that dedicates higher resolution over a regional area of interest and maintains a lower resolution elsewhere on the globe. The model is based on graph neural networks, which naturally affords arb… ▽ More A data-driven model (DDM) suitable for regional weather forecasting applications is presented. The model extends the Artificial Intelligence Forecasting System by introducing a stretched-grid architecture that dedicates higher resolution over a regional area of interest and maintains a lower resolution elsewhere on the globe. The model is based on graph neural networks, which naturally affords arbitrary multi-resolution grid configurations. The model is applied to short-range weather prediction for the Nordics, producing forecasts at 2.5 km spatial and 6 h temporal resolution. The model is pre-trained on 43 years of global ERA5 data at 31 km resolution and is further refined using 3.3 years of 2.5 km resolution operational analyses from the MetCoOp Ensemble Prediction System (MEPS). The performance of the model is evaluated using surface observations from measurement stations across Norway and is compared to short-range weather forecasts from MEPS. The DDM outperforms both the control run and the ensemble mean of MEPS for 2 m temperature. The model also produces competitive precipitation and wind speed forecasts, but is shown to underestimate extreme events. △ Less

Submitted 4 September, 2024; originally announced September 2024.

arXiv:2408.02161 [pdf, other]

Distilling Machine Learning's Added Value: Pareto Fronts in Atmospheric Applications

Authors: Tom Beucler, Arthur Grundner, Sara Shamekh, Peter Ukkonen, Matthew Chantry, Ryan Lagerquist

Abstract: The added value of machine learning for weather and climate applications is measurable through performance metrics, but explaining it remains challenging, particularly for large deep learning models. Inspired by climate model hierarchies, we propose that a full hierarchy of Pareto-optimal models, defined within an appropriately determined error-complexity plane, can guide model development and hel… ▽ More The added value of machine learning for weather and climate applications is measurable through performance metrics, but explaining it remains challenging, particularly for large deep learning models. Inspired by climate model hierarchies, we propose that a full hierarchy of Pareto-optimal models, defined within an appropriately determined error-complexity plane, can guide model development and help understand the models' added value. We demonstrate the use of Pareto fronts in atmospheric physics through three sample applications, with hierarchies ranging from semi-empirical models with minimal parameters to deep learning algorithms. First, in cloud cover parameterization, we find that neural networks identify nonlinear relationships between cloud cover and its thermodynamic environment, and assimilate previously neglected features such as vertical gradients in relative humidity that improve the representation of low cloud cover. This added value is condensed into a ten-parameter equation that rivals deep learning models. Second, we establish a machine learning model hierarchy for emulating shortwave radiative transfer, distilling the importance of bidirectional vertical connectivity for accurately representing absorption and scattering, especially for multiple cloud layers. Third, we emphasize the importance of convective organization information when modeling the relationship between tropical precipitation and its surrounding environment. We discuss the added value of temporal memory when high-resolution spatial information is unavailable, with implications for precipitation parameterization. Therefore, by comparing data-driven models directly with existing schemes using Pareto optimality, we promote process understanding by hierarchically unveiling system complexity, with the hope of improving the trustworthiness of machine learning models in atmospheric applications. △ Less

Submitted 18 January, 2025; v1 submitted 4 August, 2024; originally announced August 2024.

Comments: 18 pages, 4 figures, submitted to AMS Artificial Intelligence for the Earth Systems (AIES)

arXiv:2407.16463 [pdf, other]

Advances in Land Surface Model-based Forecasting: A comparative study of LSTM, Gradient Boosting, and Feedforward Neural Network Models as prognostic state emulators

Authors: Marieke Wesselkamp, Matthew Chantry, Ewan Pinnington, Margarita Choulga, Souhail Boussetta, Maria Kalweit, Joschka Boedecker, Carsten F. Dormann, Florian Pappenberger, Gianpaolo Balsamo

Abstract: Most useful weather prediction for the public is near the surface. The processes that are most relevant for near-surface weather prediction are also those that are most interactive and exhibit positive feedback or have key role in energy partitioning. Land surface models (LSMs) consider these processes together with surface heterogeneity and forecast water, carbon and energy fluxes, and coupled wi… ▽ More Most useful weather prediction for the public is near the surface. The processes that are most relevant for near-surface weather prediction are also those that are most interactive and exhibit positive feedback or have key role in energy partitioning. Land surface models (LSMs) consider these processes together with surface heterogeneity and forecast water, carbon and energy fluxes, and coupled with an atmospheric model provide boundary and initial conditions. This numerical parametrization of atmospheric boundaries being computationally expensive, statistical surrogate models are increasingly used to accelerated progress in experimental research. We evaluated the efficiency of three surrogate models in speeding up experimental research by simulating land surface processes, which are integral to forecasting water, carbon, and energy fluxes in coupled atmospheric models. Specifically, we compared the performance of a Long-Short Term Memory (LSTM) encoder-decoder network, extreme gradient boosting, and a feed-forward neural network within a physics-informed multi-objective framework. This framework emulates key states of the ECMWF's Integrated Forecasting System (IFS) land surface scheme, ECLand, across continental and global scales. Our findings indicate that while all models on average demonstrate high accuracy over the forecast period, the LSTM network excels in continental long-range predictions when carefully tuned, the XGB scores consistently high across tasks and the MLP provides an excellent implementation-time-accuracy trade-off. The runtime reduction achieved by the emulators in comparison to the full numerical models are significant, offering a faster, yet reliable alternative for conducting numerical experiments on land surfaces. △ Less

Submitted 23 July, 2024; originally announced July 2024.

arXiv:2407.15586 [pdf, other]

Data driven weather forecasts trained and initialised directly from observations

Authors: Anthony McNally, Christian Lessig, Peter Lean, Eulalie Boucher, Mihai Alexe, Ewan Pinnington, Matthew Chantry, Simon Lang, Chris Burrows, Marcin Chrust, Florian Pinault, Ethel Villeneuve, Niels Bormann, Sean Healy

Abstract: Skilful Machine Learned weather forecasts have challenged our approach to numerical weather prediction, demonstrating competitive performance compared to traditional physics-based approaches. Data-driven systems have been trained to forecast future weather by learning from long historical records of past weather such as the ECMWF ERA5. These datasets have been made freely available to the wider re… ▽ More Skilful Machine Learned weather forecasts have challenged our approach to numerical weather prediction, demonstrating competitive performance compared to traditional physics-based approaches. Data-driven systems have been trained to forecast future weather by learning from long historical records of past weather such as the ECMWF ERA5. These datasets have been made freely available to the wider research community, including the commercial sector, which has been a major factor in the rapid rise of ML forecast systems and the levels of accuracy they have achieved. However, historical reanalyses used for training and real-time analyses used for initial conditions are produced by data assimilation, an optimal blending of observations with a physics-based forecast model. As such, many ML forecast systems have an implicit and unquantified dependence on the physics-based models they seek to challenge. Here we propose a new approach, training a neural network to predict future weather purely from historical observations with no dependence on reanalyses. We use raw observations to initialise a model of the atmosphere (in observation space) learned directly from the observations themselves. Forecasts of crucial weather parameters (such as surface temperature and wind) are obtained by predicting weather parameter observations (e.g. SYNOP surface data) at future times and arbitrary locations. We present preliminary results on forecasting observations 12-hours into the future. These already demonstrate successful learning of time evolutions of the physical processes captured in real observations. We argue that this new approach, by staying purely in observation space, avoids many of the challenges of traditional data assimilation, can exploit a wider range of observations and is readily expanded to simultaneous forecasting of the full Earth system (atmosphere, land, ocean and composition). △ Less

Submitted 22 July, 2024; originally announced July 2024.

arXiv:2406.01465 [pdf, other]

AIFS -- ECMWF's data-driven forecasting system

Authors: Simon Lang, Mihai Alexe, Matthew Chantry, Jesper Dramsch, Florian Pinault, Baudouin Raoult, Mariana C. A. Clare, Christian Lessig, Michael Maier-Gerber, Linus Magnusson, Zied Ben Bouallègue, Ana Prieto Nemesio, Peter D. Dueben, Andrew Brown, Florian Pappenberger, Florence Rabier

Abstract: Machine learning-based weather forecasting models have quickly emerged as a promising methodology for accurate medium-range global weather forecasting. Here, we introduce the Artificial Intelligence Forecasting System (AIFS), a data driven forecast model developed by the European Centre for Medium-Range Weather Forecasts (ECMWF). AIFS is based on a graph neural network (GNN) encoder and decoder, a… ▽ More Machine learning-based weather forecasting models have quickly emerged as a promising methodology for accurate medium-range global weather forecasting. Here, we introduce the Artificial Intelligence Forecasting System (AIFS), a data driven forecast model developed by the European Centre for Medium-Range Weather Forecasts (ECMWF). AIFS is based on a graph neural network (GNN) encoder and decoder, and a sliding window transformer processor, and is trained on ECMWF's ERA5 re-analysis and ECMWF's operational numerical weather prediction (NWP) analyses. It has a flexible and modular design and supports several levels of parallelism to enable training on high-resolution input data. AIFS forecast skill is assessed by comparing its forecasts to NWP analyses and direct observational data. We show that AIFS produces highly skilled forecasts for upper-air variables, surface weather parameters and tropical cyclone tracks. AIFS is run four times daily alongside ECMWF's physics-based NWP model and forecasts are available to the public under ECMWF's open data policy. △ Less

Submitted 7 August, 2024; v1 submitted 3 June, 2024; originally announced June 2024.

arXiv:2404.00411 [pdf, other]

Aardvark weather: end-to-end data-driven weather forecasting

Authors: Anna Vaughan, Stratis Markou, Will Tebbutt, James Requeima, Wessel P. Bruinsma, Tom R. Andersson, Michael Herzog, Nicholas D. Lane, Matthew Chantry, J. Scott Hosking, Richard E. Turner

Abstract: Weather forecasting is critical for a range of human activities including transportation, agriculture, industry, as well as the safety of the general public. Machine learning models have the potential to transform the complex weather prediction pipeline, but current approaches still rely on numerical weather prediction (NWP) systems, limiting forecast speed and accuracy. Here we demonstrate that a… ▽ More Weather forecasting is critical for a range of human activities including transportation, agriculture, industry, as well as the safety of the general public. Machine learning models have the potential to transform the complex weather prediction pipeline, but current approaches still rely on numerical weather prediction (NWP) systems, limiting forecast speed and accuracy. Here we demonstrate that a machine learning model can replace the entire operational NWP pipeline. Aardvark Weather, an end-to-end data-driven weather prediction system, ingests raw observations and outputs global gridded forecasts and local station forecasts. Further, it can be optimised end-to-end to maximise performance over quantities of interest. Global forecasts outperform an operational NWP baseline for multiple variables and lead times. Local station forecasts are skillful up to ten days lead time and achieve comparable and often lower errors than a post-processed global NWP baseline and a state-of-the-art end-to-end forecasting system with input from human forecasters. These forecasts are produced with a remarkably simple neural process model using just 8% of the input data and three orders of magnitude less compute than existing NWP and hybrid AI-NWP methods. We anticipate that Aardvark Weather will be the starting point for a new generation of end-to-end machine learning models for medium-range forecasting that will reduce computational costs by orders of magnitude and enable the rapid and cheap creation of bespoke models for users in a variety of fields, including for the developing world where state-of-the-art local models are not currently available. △ Less

Submitted 13 July, 2024; v1 submitted 30 March, 2024; originally announced April 2024.

arXiv:2309.15689 [pdf, other]

Further analysis of cGAN: A system for Generative Deep Learning Post-processing of Precipitation

Authors: Fenwick C. Cooper, Andrew T. T. McRae, Matthew Chantry, Bobby Antonio, Tim N. Palmer

Abstract: The conditional generative adversarial rainfall model "cGAN" developed for the UK \cite{Harris22} was trained to post-process into an ensemble and downscale ERA5 rainfall to 1km resolution over three regions of the USA and the UK. Relative to radar data (stage IV and NIMROD), the quality of the forecast rainfall distribution was quantified locally at each grid point and between grid points using t… ▽ More The conditional generative adversarial rainfall model "cGAN" developed for the UK \cite{Harris22} was trained to post-process into an ensemble and downscale ERA5 rainfall to 1km resolution over three regions of the USA and the UK. Relative to radar data (stage IV and NIMROD), the quality of the forecast rainfall distribution was quantified locally at each grid point and between grid points using the spatial correlation structure. Despite only having information from a single lower quality analysis, the ensembles of post processed rainfall produced were found to be competitive with IFS ensemble forecasts with lead times of between 8 and 16 hours. Comparison to the original cGAN trained on the UK using the IFS HRES forecast indicates that improved training forecasts result in improved post-processing. The cGAN models were additionally applied to the regions that they were not trained on. Each model performed well in their own region indicating that each model is somewhat region specific. However the model trained on the Washington DC, Atlantic coast, region achieved good scores across the USA and was competitive over the UK. There are more overall rainfall events spread over the whole region so the improved scores might be simply due to increased data. A model was therefore trained using data from all four regions which then outperformed the models trained locally. △ Less

Submitted 27 September, 2023; originally announced September 2023.

arXiv:2308.15560 [pdf, other]

WeatherBench 2: A benchmark for the next generation of data-driven global weather models

Authors: Stephan Rasp, Stephan Hoyer, Alexander Merose, Ian Langmore, Peter Battaglia, Tyler Russel, Alvaro Sanchez-Gonzalez, Vivian Yang, Rob Carver, Shreya Agrawal, Matthew Chantry, Zied Ben Bouallegue, Peter Dueben, Carla Bromberg, Jared Sisk, Luke Barrington, Aaron Bell, Fei Sha

Abstract: WeatherBench 2 is an update to the global, medium-range (1-14 day) weather forecasting benchmark proposed by Rasp et al. (2020), designed with the aim to accelerate progress in data-driven weather modeling. WeatherBench 2 consists of an open-source evaluation framework, publicly available training, ground truth and baseline data as well as a continuously updated website with the latest metrics and… ▽ More WeatherBench 2 is an update to the global, medium-range (1-14 day) weather forecasting benchmark proposed by Rasp et al. (2020), designed with the aim to accelerate progress in data-driven weather modeling. WeatherBench 2 consists of an open-source evaluation framework, publicly available training, ground truth and baseline data as well as a continuously updated website with the latest metrics and state-of-the-art models: https://sites.research.google/weatherbench. This paper describes the design principles of the evaluation framework and presents results for current state-of-the-art physical and data-driven weather models. The metrics are based on established practices for evaluating weather forecasts at leading operational weather centers. We define a set of headline scores to provide an overview of model performance. In addition, we also discuss caveats in the current evaluation setup and challenges for the future of data-driven weather forecasting. △ Less

Submitted 26 January, 2024; v1 submitted 29 August, 2023; originally announced August 2023.

arXiv:2307.10128 [pdf, other]

doi 10.1175/BAMS-D-23-0162.1

The rise of data-driven weather forecasting

Authors: Zied Ben-Bouallegue, Mariana C A Clare, Linus Magnusson, Estibaliz Gascon, Michael Maier-Gerber, Martin Janousek, Mark Rodwell, Florian Pinault, Jesper S Dramsch, Simon T K Lang, Baudouin Raoult, Florence Rabier, Matthieu Chevallier, Irina Sandu, Peter Dueben, Matthew Chantry, Florian Pappenberger

Abstract: Data-driven modeling based on machine learning (ML) is showing enormous potential for weather forecasting. Rapid progress has been made with impressive results for some applications. The uptake of ML methods could be a game-changer for the incremental progress in traditional numerical weather prediction (NWP) known as the 'quiet revolution' of weather forecasting. The computational cost of running… ▽ More Data-driven modeling based on machine learning (ML) is showing enormous potential for weather forecasting. Rapid progress has been made with impressive results for some applications. The uptake of ML methods could be a game-changer for the incremental progress in traditional numerical weather prediction (NWP) known as the 'quiet revolution' of weather forecasting. The computational cost of running a forecast with standard NWP systems greatly hinders the improvements that can be made from increasing model resolution and ensemble sizes. An emerging new generation of ML models, developed using high-quality reanalysis datasets like ERA5 for training, allow forecasts that require much lower computational costs and that are highly-competitive in terms of accuracy. Here, we compare for the first time ML-generated forecasts with standard NWP-based forecasts in an operational-like context, initialized from the same initial conditions. Focusing on deterministic forecasts, we apply common forecast verification tools to assess to what extent a data-driven forecast produced with one of the recently developed ML models (PanguWeather) matches the quality and attributes of a forecast from one of the leading global NWP systems (the ECMWF IFS). The results are very promising, with comparable skill for both global metrics and extreme events, when verified against both the operational analysis and synoptic observations. Increasing forecast smoothness and bias drift with forecast lead time are identified as current drawbacks of ML-based forecasts. A new NWP paradigm is emerging relying on inference from ML models and state-of-the-art analysis and reanalysis datasets for forecast initialization and model training. △ Less

Submitted 3 November, 2023; v1 submitted 19 July, 2023; originally announced July 2023.

arXiv:2303.17195 [pdf, other]

Improving medium-range ensemble weather forecasts with hierarchical ensemble transformers

Authors: Zied Ben-Bouallegue, Jonathan A Weyn, Mariana C A Clare, Jesper Dramsch, Peter Dueben, Matthew Chantry

Abstract: Statistical post-processing of global ensemble weather forecasts is revisited by leveraging recent developments in machine learning. Verification of past forecasts is exploited to learn systematic deficiencies of numerical weather predictions in order to boost post-processed forecast performance. Here, we introduce PoET, a post-processing approach based on hierarchical transformers. PoET has 2 maj… ▽ More Statistical post-processing of global ensemble weather forecasts is revisited by leveraging recent developments in machine learning. Verification of past forecasts is exploited to learn systematic deficiencies of numerical weather predictions in order to boost post-processed forecast performance. Here, we introduce PoET, a post-processing approach based on hierarchical transformers. PoET has 2 major characteristics: 1) the post-processing is applied directly to the ensemble members rather than to a predictive distribution or a functional of it, and 2) the method is ensemble-size agnostic in the sense that the number of ensemble members in training and inference mode can differ. The PoET output is a set of calibrated members that has the same size as the original ensemble but with improved reliability. Performance assessments show that PoET can bring up to 20% improvement in skill globally for 2m temperature and 2% for precipitation forecasts and outperforms the simpler statistical member-by-member method, used here as a competitive benchmark. PoET is also applied to the ENS10 benchmark dataset for ensemble post-processing and provides better results when compared to other deep learning solutions that are evaluated for most parameters. Furthermore, because each ensemble member is calibrated separately, downstream applications should directly benefit from the improvement made on the ensemble forecast with post-processing. △ Less

Submitted 20 October, 2023; v1 submitted 30 March, 2023; originally announced March 2023.

arXiv:2210.16746 [pdf, other]

doi 10.5194/egusphere-2022-1177

Deep learning for quality control of surface physiographic fields using satellite Earth observations

Authors: Tom Kimpson, Margarita Choulga, Matthew Chantry, Gianpaolo Balsamo, Souhail Boussetta, Peter Dueben, Tim Palmer

Abstract: A purposely built deep learning algorithm for the Verification of Earth-System ParametERisation (VESPER) is used to assess recent upgrades of the global physiographic datasets underpinning the quality of the Integrated Forecasting System (IFS) of the European Centre for Medium-Range Weather Forecasts (ECMWF), which is used both in numerical weather prediction and climate reanalyses. A neural netwo… ▽ More A purposely built deep learning algorithm for the Verification of Earth-System ParametERisation (VESPER) is used to assess recent upgrades of the global physiographic datasets underpinning the quality of the Integrated Forecasting System (IFS) of the European Centre for Medium-Range Weather Forecasts (ECMWF), which is used both in numerical weather prediction and climate reanalyses. A neural network regression model is trained to learn the mapping between the surface physiographic dataset plus the meteorology from ERA5, and the MODIS satellite skin temperature observations. Once trained, this tool is applied to rapidly assess the quality of upgrades of the land-surface scheme. Upgrades which improve the prediction accuracy of the machine learning tool indicate a reduction of the errors in the surface fields used as input to the surface parametrisation schemes. Conversely, incorrect specifications of the surface fields decrease the accuracy with which VESPER can make predictions. We apply VESPER to assess the accuracy of recent upgrades of the permanent lake and glaciers covers as well as planned upgrades to represent seasonally varying water bodies (i.e. ephemeral lakes). We show that for grid-cells where the lake fields have been updated, the prediction accuracy in the land surface temperature (i.e mean absolute error difference between updated and original physiographic datasets) improves by 0.37 K on average, whilst for the subset of points where the lakes have been exchanged for bare ground (or vice versa) the improvement is 0.83 K. We also show that updates to the glacier cover improve the prediction accuracy by 0.22 K. We highlight how neural networks such as VESPER can assist the research and development of surface parameterizations and their input physiography to better represent Earth's surface couples processes in weather and climate models. △ Less

Submitted 24 October, 2023; v1 submitted 30 October, 2022; originally announced October 2022.

Comments: 26 pages, 16 figures. Accepted for publication in Hydrology and Earth System Sciences (HESS)

arXiv:2207.14598 [pdf, other]

doi 10.1002/qj.4435

Climate Change Modelling at Reduced Float Precision with Stochastic Rounding

Authors: Tom Kimpson, E. Adam Paxton, Matthew Chantry, Tim Palmer

Abstract: Reduced precision floating point arithmetic is now routinely deployed in numerical weather forecasting over short timescales. However the applicability of these reduced precision techniques to longer timescale climate simulations - especially those which seek to describe a dynamical, changing climate - remains unclear. We investigate this question by deploying a global atmospheric, coarse resoluti… ▽ More Reduced precision floating point arithmetic is now routinely deployed in numerical weather forecasting over short timescales. However the applicability of these reduced precision techniques to longer timescale climate simulations - especially those which seek to describe a dynamical, changing climate - remains unclear. We investigate this question by deploying a global atmospheric, coarse resolution model known as SPEEDY to simulate a changing climate system subject to increased $\text{CO}_2$ concentrations, over a 100 year timescale. Whilst double precision is typically the operational standard for climate modelling, we find that reduced precision solutions (Float32, Float16) are sufficiently accurate. Rounding the finite precision floats stochastically, rather than using the more common ``round-to-nearest" technique, notably improves the performance of the reduced precision solutions. Over 100 years the mean bias error (MBE) in the global mean surface temperature (precipitation) relative to the double precision solution is $+2 \times 10^{-4}$K ($-8 \times 10^{-5}$ mm/6hr) at single precision and $-3.5\times 10^{-2}$ K($-1 \times 10^{-2}$ mm/6hr) at half precision, whilst the inclusion of stochastic rounding reduced the half precision error to +1.8 $\times 10^{-2}$ K ($-8 \times10^{-4}$ mm/6hr). By examining the resultant climatic distributions that arise after 100 years, the difference in the expected value of the global surface temperature, relative to the double precision solution is $\leq 5 \times 10^{-3}$ K and for precipitation $8 \times 10^{-4}$ mm/6h when numerically integrating at half precision with stochastic rounding. Areas of the model which notably improve due to the inclusion of stochastic over deterministic rounding are also explored and discussed. [abridged] △ Less

Submitted 29 July, 2022; originally announced July 2022.

Comments: 13 pages, 9 figures. Submitted to QJRMS. Abstract abridged to meet arxiv requirements

arXiv:2204.02028 [pdf, other]

doi 10.1029/2022MS003120

A Generative Deep Learning Approach to Stochastic Downscaling of Precipitation Forecasts

Authors: Lucy Harris, Andrew T. T. McRae, Matthew Chantry, Peter D. Dueben, Tim N. Palmer

Abstract: Despite continuous improvements, precipitation forecasts are still not as accurate and reliable as those of other meteorological variables. A major contributing factor to this is that several key processes affecting precipitation distribution and intensity occur below the resolved scale of global weather models. Generative adversarial networks (GANs) have been demonstrated by the computer vision c… ▽ More Despite continuous improvements, precipitation forecasts are still not as accurate and reliable as those of other meteorological variables. A major contributing factor to this is that several key processes affecting precipitation distribution and intensity occur below the resolved scale of global weather models. Generative adversarial networks (GANs) have been demonstrated by the computer vision community to be successful at super-resolution problems, i.e., learning to add fine-scale structure to coarse images. Leinonen et al. (2020) previously applied a GAN to produce ensembles of reconstructed high-resolution atmospheric fields, given coarsened input data. In this paper, we demonstrate this approach can be extended to the more challenging problem of increasing the accuracy and resolution of comparatively low-resolution input from a weather forecasting model, using high-resolution radar measurements as a "ground truth". The neural network must learn to add resolution and structure whilst accounting for non-negligible forecast error. We show that GANs and VAE-GANs can match the statistical properties of state-of-the-art pointwise post-processing methods whilst creating high-resolution, spatially coherent precipitation maps. Our model compares favourably to the best existing downscaling methods in both pixel-wise and pooled CRPS scores, power spectrum information and rank histograms (used to assess calibration). We test our models and show that they perform in a range of scenarios, including heavy rainfall. △ Less

Submitted 28 July, 2022; v1 submitted 5 April, 2022; originally announced April 2022.

Comments: Revised version 28/7/22

arXiv:2104.15076 [pdf, other]

doi 10.1175/JCLI-D-21-0343.1

Climate Modelling in Low-Precision: Effects of both Deterministic & Stochastic Rounding

Authors: E. Adam Paxton, Matthew Chantry, Milan Klöwer, Leo Saffin, Tim Palmer

Abstract: Motivated by recent advances in operational weather forecasting, we study the efficacy of low-precision arithmetic for climate simulations. We develop a framework to measure rounding error in a climate model which provides a stress-test for a low-precision version of the model, and we apply our method to a variety of models including the Lorenz system; a shallow water approximation for flow over a… ▽ More Motivated by recent advances in operational weather forecasting, we study the efficacy of low-precision arithmetic for climate simulations. We develop a framework to measure rounding error in a climate model which provides a stress-test for a low-precision version of the model, and we apply our method to a variety of models including the Lorenz system; a shallow water approximation for flow over a ridge; and a coarse resolution global atmospheric model with simplified parameterisations (SPEEDY). Although double precision (52 significant bits) is standard across operational climate models, in our experiments we find that single precision (23 sbits) is more than enough and that as low as half precision (10 sbits) is often sufficient. For example, SPEEDY can be run with 12 sbits across the entire code with negligible rounding error and this can be lowered to 10 sbits if very minor errors are accepted, amounting to less than 0.1 mm/6hr for the average grid-point precipitation, for example. Our test is based on the Wasserstein metric and this provides stringent non-parametric bounds on rounding error accounting for annual means as well as extreme weather events. In addition, by testing models using both round-to-nearest (RN) and stochastic rounding (SR) we find that SR can mitigate rounding error across a range of applications. Thus our results also provide evidence that SR could be relevant to next-generation climate models. While many studies have shown that low-precision arithmetic can be suitable on short-term weather forecasting timescales, our results give the first evidence that a similar low precision level can be suitable for climate. △ Less

Submitted 30 April, 2021; originally announced April 2021.

arXiv:2104.03196 [pdf, other]

doi 10.1007/s00382-022-06395-x

A topological perspective on weather regimes

Authors: Kristian Strommen, Matthew Chantry, Joshua Dorrington, Nina Otter

Abstract: It has long been suggested that the mid-latitude atmospheric circulation possesses what has come to be known as `weather regimes', loosely categorised as regions of phase space with above-average density and/or extended persistence. Their existence and behaviour has been extensively studied in meteorology and climate science, due to their potential for drastically simplifying the complex and chaot… ▽ More It has long been suggested that the mid-latitude atmospheric circulation possesses what has come to be known as `weather regimes', loosely categorised as regions of phase space with above-average density and/or extended persistence. Their existence and behaviour has been extensively studied in meteorology and climate science, due to their potential for drastically simplifying the complex and chaotic mid-latitude dynamics. Several well-known, simple non-linear dynamical systems have been used as toy-models of the atmosphere in order to understand and exemplify such regime behaviour. Nevertheless, no agreed-upon and clear-cut definition of a `regime' exists in the literature. We argue here for an approach which equates the existence of regimes in a dynamical system with the existence of non-trivial topological structure of the system's attractor. We show using persistent homology, an algorithmic tool in topological data analysis, that this approach is computationally tractable, practically informative, and identifies the relevant regime structure across a range of examples. △ Less

Submitted 6 September, 2021; v1 submitted 7 April, 2021; originally announced April 2021.

Comments: Major revisions and improvements to exposition. More mathematical details in the Appendix. New title

arXiv:2101.08195 [pdf, other]

doi 10.1029/2021MS002477

Machine learning emulation of gravity wave drag in numerical weather forecasting

Authors: Matthew Chantry, Sam Hatfield, Peter Duben, Inna Polichtchouk, Tim Palmer

Abstract: We assess the value of machine learning as an accelerator for the parameterisation schemes of operational weather forecasting systems, specifically the parameterisation of non-orographic gravity wave drag. Emulators of this scheme can be trained to produce stable and accurate results up to seasonal forecasting timescales. Generally, more complex networks produce more accurate emulators. By trainin… ▽ More We assess the value of machine learning as an accelerator for the parameterisation schemes of operational weather forecasting systems, specifically the parameterisation of non-orographic gravity wave drag. Emulators of this scheme can be trained to produce stable and accurate results up to seasonal forecasting timescales. Generally, more complex networks produce more accurate emulators. By training on an increased complexity version of the existing parameterisation scheme we build emulators that produce more accurate forecasts. {For medium range forecasting we find evidence our emulators are more accurate} than the version of the parametrisation scheme that is used for operational predictions. Using the current operational CPU hardware our emulators have a similar computational cost to the existing scheme, but are heavily limited by data movement. On GPU hardware our emulators perform ten times faster than the existing scheme on a CPU. △ Less

Submitted 7 May, 2021; v1 submitted 20 January, 2021; originally announced January 2021.

arXiv:2012.09670 [pdf, other]

RainBench: Towards Global Precipitation Forecasting from Satellite Imagery

Authors: Christian Schroeder de Witt, Catherine Tong, Valentina Zantedeschi, Daniele De Martini, Freddie Kalaitzis, Matthew Chantry, Duncan Watson-Parris, Piotr Bilinski

Abstract: Extreme precipitation events, such as violent rainfall and hail storms, routinely ravage economies and livelihoods around the developing world. Climate change further aggravates this issue. Data-driven deep learning approaches could widen the access to accurate multi-day forecasts, to mitigate against such events. However, there is currently no benchmark dataset dedicated to the study of global pr… ▽ More Extreme precipitation events, such as violent rainfall and hail storms, routinely ravage economies and livelihoods around the developing world. Climate change further aggravates this issue. Data-driven deep learning approaches could widen the access to accurate multi-day forecasts, to mitigate against such events. However, there is currently no benchmark dataset dedicated to the study of global precipitation forecasts. In this paper, we introduce \textbf{RainBench}, a new multi-modal benchmark dataset for data-driven precipitation forecasting. It includes simulated satellite data, a selection of relevant meteorological data from the ERA5 reanalysis product, and IMERG precipitation data. We also release \textbf{PyRain}, a library to process large precipitation datasets efficiently. We present an extensive analysis of our novel dataset and establish baseline results for two benchmark medium-range precipitation forecasting tasks. Finally, we discuss existing data-driven weather forecasting methodologies and suggest future research avenues. △ Less

Submitted 17 December, 2020; originally announced December 2020.

Comments: Work completed during the 2020 Frontier Development Lab research accelerator, a private-public partnership with NASA in the US, and ESA in Europe. Accepted as a spotlight/long oral talk at both Climate Change and AI, as well as AI for Earth Sciences Workshops at NeurIPS 2020

arXiv:2009.02108 [pdf, other]

doi 10.1103/PhysRevFluids.5.103902

Far field of turbulent spots

Authors: Pavan V. Kashyap, Yohann Duguet, Matthew Chantry

Abstract: The proliferation of turbulence in subcritical wall-bounded shear flows involves spatially localised coherent structures. Turbulent spots correspond to finite-time nonlinear responses to pointwise disturbances and are regarded as seeds of turbulence during transition. The rapid spatial decay of the turbulent fluctuations away from a spot is accompanied by large-scale flows with a robust structurat… ▽ More The proliferation of turbulence in subcritical wall-bounded shear flows involves spatially localised coherent structures. Turbulent spots correspond to finite-time nonlinear responses to pointwise disturbances and are regarded as seeds of turbulence during transition. The rapid spatial decay of the turbulent fluctuations away from a spot is accompanied by large-scale flows with a robust structuration. The far field velocity field of these spots is investigated numerically using spectral methods in large domains in four different flow scenarios (plane Couette, plane Poiseuille, Couette-Poiseuille and a sinusoidal shear flow). At odds with former expectations, the planar components of the velocity field decay algebraically. These decay exponents depend only on the symmetries of the system, which here depend on the presence of an applied gradient, and not on the Reynolds number. This suggests an effective two-dimensional multipolar expansion for the far field, dominated by a quadrupolar flow component or, for asymmetric flow fields, by a dipolar flow component. △ Less

Submitted 4 September, 2020; originally announced September 2020.

Comments: Article accepted for publication in Physical Review Fluids

Journal ref: Phys. Rev. Fluids 5, 103902 (2020)

arXiv:1704.03567 [pdf, other]

doi 10.1017/jfm.2017.405

Universal continuous transition to turbulence in a planar shear flow

Authors: Matthew Chantry, Laurette S. Tuckerman, Dwight Barkley

Abstract: We examine the onset of turbulence in Waleffe flow -- the planar shear flow between stress-free boundaries driven by a sinusoidal body force. By truncating the wall-normal representation to four modes, we are able to simulate system sizes an order of magnitude larger than any previously simulated, and thereby to attack the question of universality for a planar shear flow. We demonstrate that the e… ▽ More We examine the onset of turbulence in Waleffe flow -- the planar shear flow between stress-free boundaries driven by a sinusoidal body force. By truncating the wall-normal representation to four modes, we are able to simulate system sizes an order of magnitude larger than any previously simulated, and thereby to attack the question of universality for a planar shear flow. We demonstrate that the equilibrium turbulence fraction increases continuously from zero above a critical Reynolds number and that statistics of the turbulent structures exhibit the power-law scalings of the (2+1)-D directed percolation universality class. △ Less

Submitted 15 August, 2017; v1 submitted 11 April, 2017; originally announced April 2017.

Journal ref: J. Fluid Mech. 824, R1 (2017)

arXiv:1506.05002 [pdf, other]

doi 10.1017/jfm.2016.92

Turbulent-laminar patterns in shear flows without walls

Authors: Matthew Chantry, Laurette S. Tuckerman, Dwight Barkley

Abstract: Turbulent-laminar intermittency, typically in the form of bands and spots, is a ubiquitous feature of the route to turbulence in wall-bounded shear flows. Here we study the idealised shear between stress-free boundaries driven by a sinusoidal body force and demonstrate quantitative agreement between turbulence in this flow and that found in the interior of plane Couette flow -- the region excludin… ▽ More Turbulent-laminar intermittency, typically in the form of bands and spots, is a ubiquitous feature of the route to turbulence in wall-bounded shear flows. Here we study the idealised shear between stress-free boundaries driven by a sinusoidal body force and demonstrate quantitative agreement between turbulence in this flow and that found in the interior of plane Couette flow -- the region excluding the boundary layers. Exploiting the absence of boundary layers, we construct a model flow that uses only four Fourier modes in the shear direction and yet robustly captures the range of spatiotemporal phenomena observed in transition, from spot growth to turbulent bands and uniform turbulence. The model substantially reduces the cost of simulating intermittent turbulent structures while maintaining the essential physics and a direct connection to the Navier-Stokes equations. We demonstrate the generic nature of this process by introducing stress-free equivalent flows for plane Poiseuille and pipe flows which again capture the turbulent-laminar structures seen in transition. △ Less

Submitted 17 February, 2016; v1 submitted 16 June, 2015; originally announced June 2015.

Comments: 13 pages, 9 figures

arXiv:1410.2035 [pdf, ps, other]

doi 10.1103/PhysRevE.91.043005

Localization in a spanwise-extended model of plane Couette flow

Authors: Matthew Chantry, Rich R Kerswell

Abstract: We consider a 9-PDE (1-space and 1-time) model of plane Couette flow in which the degrees of freedom are severely restricted in the streamwise and cross-stream directions to study spanwise localisation in detail. Of the many steady Eckhaus (spanwise modulational) instabilities identified of global steady states, none lead to a localized state. Localized periodic solutions were found instead which… ▽ More We consider a 9-PDE (1-space and 1-time) model of plane Couette flow in which the degrees of freedom are severely restricted in the streamwise and cross-stream directions to study spanwise localisation in detail. Of the many steady Eckhaus (spanwise modulational) instabilities identified of global steady states, none lead to a localized state. Localized periodic solutions were found instead which arise in saddle node bifurcations in the Reynolds number. These solutions appear global (domain filling) in narrow (small spanwise) domains yet can be smoothly continued out to fully spanwise-localised states in very wide domains. This smooth localisation behaviour, which has also been seen in fully-resolved duct flow (Okino 2011), indicates that an apparently global flow structure needn't have to suffer a modulational instability to localize in wide domains. △ Less

Submitted 8 October, 2014; originally announced October 2014.

Comments: 14 pages, 18 figures

Journal ref: Phys. Rev. E 91, 043005 (2015)

arXiv:1308.6224 [pdf, ps, other]

doi 10.1103/PhysRevLett.112.164501

The genesis of streamwise-localized solutions from globally periodic travelling waves in pipe flow

Authors: Matthew Chantry, Ashley P. Willis, Rich R. Kerswell

Abstract: The aim in the dynamical systems approach to transitional turbulence is to construct a scaffold in phase space for the dynamics using simple invariant sets (exact solutions) and their stable and unstable manifolds. In large (realistic) domains where turbulence can co-exist with laminar flow, this requires identifying exact localized solutions. In wall-bounded shear flows the first of these has rec… ▽ More The aim in the dynamical systems approach to transitional turbulence is to construct a scaffold in phase space for the dynamics using simple invariant sets (exact solutions) and their stable and unstable manifolds. In large (realistic) domains where turbulence can co-exist with laminar flow, this requires identifying exact localized solutions. In wall-bounded shear flows the first of these has recently been found in pipe flow, but questions remain as to how they are connected to the many known streamwise-periodic solutions. Here we demonstrate the origin of the first localized solution in a modulational symmetry-breaking Hopf bifurcation from a known global travelling wave that has 2-fold rotational symmetry about the pipe axis. Similar behaviour is found for a global wave of 3-fold rotational symmetry, this time leading to two localized relative periodic orbits. The clear implication is that all global solutions should be expected to lead to more realistic localised counterparts through such bifurcations, which provides a constructive route for their generation. △ Less

Submitted 22 April, 2014; v1 submitted 28 August, 2013; originally announced August 2013.

Comments: 5 pages, 6 figures

Journal ref: Phys. Rev. Lett. 112, 164501 (2014)

arXiv:1308.5902 [pdf, ps, other]

doi 10.1017/jfm.2014.150

Studying edge geometry in transiently turbulent shear flows

Authors: Matthew Chantry, Tobias M. Schneider

Abstract: In linearly stable shear flows at moderate Re, turbulence spontaneously decays despite the existence of a codimension-one manifold, termed the edge of chaos, which separates decaying perturbations from those triggering turbulence. We statistically analyse the decay in plane Couette flow, quantify the breaking of self-sustaining feedback loops and demonstrate the existence of a whole continuum of p… ▽ More In linearly stable shear flows at moderate Re, turbulence spontaneously decays despite the existence of a codimension-one manifold, termed the edge of chaos, which separates decaying perturbations from those triggering turbulence. We statistically analyse the decay in plane Couette flow, quantify the breaking of self-sustaining feedback loops and demonstrate the existence of a whole continuum of possible decay paths. Drawing parallels with low-dimensional models and monitoring the location of the edge relative to decaying trajectories we provide evidence, that the edge of chaos separates state space not globally. It is instead wrapped around the turbulence generating structures and not an independent dynamical structure but part of the chaotic saddle. Thereby, decaying trajectories need not cross the edge, but circumnavigate it while unwrapping from the turbulent saddle. △ Less

Submitted 10 December, 2013; v1 submitted 27 August, 2013; originally announced August 2013.

Comments: 11 pages, 6 figures

Showing 1–27 of 27 results for author: Chantry, M