PosteriorBench: From Point Estimates to Posterior Matching in Evaluating Generative Inverse Solvers
Read the original on arXiv Machine Learning →PosteriorBench is a new benchmark that evaluates how well generative inverse solvers recover full posterior distributions rather than just a single reconstruction. It tests four physics-based inverse problems—Darcy flow inversion, Poisson source recovery, carbon capture and storage, and light transport material inference—using high-fidelity reference posteriors generated by rejection sampling and MCMC. The benchmark employs five metrics (posterior-mean error, posterior-standard-deviation error, maximum mean discrepancy, sliced Wasserstein distance, and radially averaged power-spectrum error) to assess pointwise accuracy, uncertainty, distributional alignment, and global frequency fidelity, revealing significant distribution-matching gaps in current solvers and highlighting the importance of neural operators, guidance weights, and generation noise for posterior-variance calibration.
Machine-generated by The Flow from the publisher's headline and feed description — not written or checked by a human. The full article lives at arXiv Machine Learning.