arXiv AI

Data assimilation for subsurface flow using latent diffusion model parameterization: performance of ensemble-Kalman and Monte Carlo techniques

arXiv:2606. 11140v1 Announce Type: cross Abstract: Data assimilation (DA) in subsurface flow entails calibrating model parameters to match observed data, typically at wells, while preserving geological realism.

arXiv Machine Learning
Sep 23

FREESIA: Covariance-Aware Posterior Transport for Expressive and Scalable Data Assimilation

The paper introduces FREESIA, a training‑free, covariance‑aware posterior transport method for data assimilation that embeds forecast cross‑covariance into a flow‑based transport to recover unobserved states while preserving non‑Gaussian posterior structure. It combines an observation‑adaptive proposal with posterior correction, providing an asymptotically exact approximation of nonlinear posteriors and a Wasserstein error bound. Experiments on Double‑Well, Lorenz‑96, and Kolmogorov flow demonstrate that FREESIA captures complex posterior structures and achieves up to a 56% reduction in RMSE compared to the best baseline in sparse, nonlinear, non‑injective observation scenarios.

By Shiwei Ni, Yangwen Zhang, Hang Qi, Xiaofei Guan, Lili Ju
arXiv AI
Jul 20

Energy-based Transport for Amortized Bayesian Inference

arXiv:2605. 15407v3 Announce Type: replace-cross Abstract: We consider amortized Bayesian inference for nonlinear inverse problems using only samples from the joint distribution of parameters and observations, including problems with unknown functions in a Banach space.

By Ricardo Baptista, Hojjat Kaveh, Andrew M. Stuart
arXiv AI
2d ago

Benchmarking Generative Models for Weather Data Assimilation on Real Station Observations

This study introduces the first controlled benchmark of generative models for weather data assimilation using real station observations from 11,849 NOAA MADIS stations across the U.S. It evaluates key design choices—diffusion vs. flow matching, pixel vs. latent-space formulations, and inference-time conditioning strategies—against a classical 3D-Var baseline. The benchmark finds that learned generative priors and full-gradient guidance improve RMSE over ERA5, while other design variations offer minimal benefit, especially under sparse observation conditions.

By Ruizhe Huang, Qidong Yang, Jonathan Giezendanner, Sherrie Wang