arXiv Machine Learning

A Ground-Truth Framework for Uncertainty Disentanglement with Posterior Risk

The paper introduces a ground‑truth framework for disentangling uncertainty into epistemic and aleatoric components using sample‑conditional pointwise posterior risk. It evaluates current methods, finding that Spectral‑normalized Neural Gaussian Processes and Variational Latent Gaussian Processes best recover the ground‑truth uncertainty, while most methods align more closely with posterior variance and miss predictor bias. The study also explores the entanglement of estimated uncertainties and the impact of modeling choices, providing practical guidance and releasing 13 semi‑synthetic datasets for further validation.

arXiv Machine Learning
Jun 19

Quantifying Aleatoric Uncertainty of In-Context Learning for Robust Measure of LLM Prediction Confidence

arXiv:2606. 19353v1 Announce Type: cross Abstract: In-Context Learning (ICL) allows LLMs to adapt to new tasks from a few demonstrations, but its reliability remains a concern: predictions are highly sensitive to both prompt design and the model's ability to understand the context, obscuring whether failures arise from data properties or model limitations.

By Jinseok Chung, Minkyoung Song, Hyunji Jung, Namhoon Lee
arXiv AI
Jun 16

Bayesian 3D Steerable CNNs: Enabling Equivariance and Uncertainty Quantification Simultaneously

arXiv:2606. 15479v1 Announce Type: cross Abstract: Steerable convolutional neural networks (Steerable-CNNs) guarantee SE(3)-equivariance by parameterizing kernels as linear combinations of steerable basis functions, but their deterministic nature precludes uncertainty quantification - limiting their use in settings where confidence estimates are essential.

By Abhishek Keripale, Ponkrshnan Thiagarajan, Susanta Ghosh
arXiv Machine Learning
Aug 26

It depends: Incorporating correlations for joint aleatoric and epistemic uncertainties of high-dimensional output spaces

arXiv:2608.24518v1 Announce Type: new Abstract: Uncertainty Quantification (UQ) plays a vital role in enhancing the reliability of deep learning model predictions, especially in scenarios with high-d...

By Leonhard F. Feiner, Manuel Nickel, Martin Menten, Laurin Lux, Rickmer Braren, Daniel Rueckert, Georgios Kaissis, Raphael Rehms, Johannes Paetzold