Generative Atmospheric Super-Resolution from Heterogeneous In Situ Observations through Composable Interfaces

📅 2026-09-24
📈 Citations: 0
✨ Influential: 0
📄 PDF
🤖 AI Summary
This study addresses the structural mismatch between sparse, heterogeneous atmospheric observations and gridded generative models by proposing a modular observation interface based on composable likelihood factors. This approach unifies multi-source heterogeneous data—including radiosonde, aircraft, and surface station measurements—into likelihood factors to condition a pretrained 13-variable atmospheric diffusion model via Bayesian inference, enabling training-free data assimilation and atmospheric super-resolution reconstruction. The framework effectively bridges geometric and sampling discrepancies across diverse data sources. Compared to a baseline using only radiosonde data, it reduces RMSE by 9.24% and yields significant improvements in CRPS, substantially enhancing global reconstruction accuracy.
📝 Abstract
Atmospheric observations are sparse, heterogeneous, and unevenly distributed, whereas many generative atmospheric models learn distributions over regularly gridded multivariate states. Once pretrained, diffusion models can supply atmospheric priors that can be combined with observation-derived likelihood factors in a Bayesian formulation. However, these observation sources differ substantially in geometry and sampling density, complicating the consistent use of their observations within a common inference framework. Here, we formulate this reconstruction problem as generative atmospheric super-resolution and introduce composable observation interfaces for conditioning a single pretrained 13-variable atmospheric diffusion model. The interfaces convert sparse radiosonde (R), clustered aircraft (A), and dense irregular surface-station (S) observations into source-specific likelihood factors that specify where observations constrain the gridded state, how residuals are counted under uneven sampling, and how strongly each source guides posterior sampling. We developed the aircraft and surface observation interfaces using 2019 observations and evaluated the selected interfaces throughout 2020 without further tuning. Compared with reconstructions conditioned only on radiosonde observations, the composed R+A+S interface reduces RMSE evaluated against ERA5 by $9.24\%$ across all 13 state variables over the CONUS domain. The aircraft and surface factors provide complementary improvements in upper-air and surface variables. The R+A+S combination also lowers the Continuous Ranked Probability Score (CRPS), while evaluations at held-out aircraft and surface-station observations show reduced prediction errors. Together, these results demonstrate a modular route for conditioning a pretrained atmospheric generative prior on heterogeneous in situ observations without retraining the underlying model.
Problem

Research questions and friction points this paper is trying to address.

atmospheric super-resolution
heterogeneous observations
diffusion models
data assimilation
in situ observations
Innovation

Methods, ideas, or system contributions that make the work stand out.

Generative Atmospheric Super-Resolution
Composable Observation Interfaces
Diffusion Models
Heterogeneous In Situ Observations
Bayesian Conditioning
🔎 Similar Papers
No similar papers found.
Yang Xu
Yang Xu
Purdue University
Reinforcement LearningStatistical Learning
D
Dibyajyoti Chakraborty
College of Information Sciences and Technology, The Pennsylvania State University, University Park, Pennsylvania, USA
H
Haiwen Guan
College of Information Sciences and Technology, The Pennsylvania State University, University Park, Pennsylvania, USA
S
Sen Wang
School of Mechanical Engineering, Purdue University, West Lafayette, Indiana, USA
Romit Maulik
Romit Maulik
Assistant Professor and ICDS Co-Hire: Pennsylvania State University
Scientific Machine LearningComputational Fluid Dynamics