arXiv AI

Information Geometry of Message Passing

arXiv:2608. 15922v1 Announce Type: cross Abstract: We show that the natural-gradient stationary condition of variational inference has an edge-local form on a Forney-style factor graph.

arXiv AI
Sep 25

Direct Message Approximation (DMA): A Consistency-Based Framework for Tractable Approximate Inference on Factor Graphs

arXiv:2609. 29466v1 Announce Type: cross Abstract: Approximate message passing on factor graphs underlies two dominant families of probabilistic inference algorithms: expectation propagation (EP) and variational message passing (VMP).

By Ralf Herbrich, Rainer Schlosser, Jan Lemcke, Johann Ukrow, Anna Kazachkova, Nicolas Alder, Leonhard Hennicke, Theo Bardey, Nico Grimm, Luca Kleinschmidt, Philipp Kolbe, Cezary Kujath, Johanna Schlimme, Karl Matti Sch\"utz
arXiv AI
Aug 19

Expected free energy as an information constraint on the Bethe Lagrangian

The paper introduces a Bethe Lagrangian formulation of expected free energy (EFE) that preserves a Kullback–Leibler structure, enabling message‑passing inference. By imposing an information constraint—requiring the mutual information between future observations, states, and parameters given actions to be at least the entropy of the goal prior—the authors recover the standard EFE solution at a specific Karush‑Kuhn‑Tucker multiplier. They analyze how varying this multiplier transitions the agent’s epistemic drive through inactive, interior, and saturated regimes, and benchmark the constrained Bethe agent against EFE and Q‑MDP on three tasks.

By Wouter M. Kouw
arXiv Machine Learning
Jun 5

Equivariant Neural Belief Propagation

arXiv:2606. 06344v1 Announce Type: new Abstract: Probabilistic inference over spatially embedded variables requires beliefs that respect $SE(3)$ symmetry, yet existing equivariant networks produce only scalars and vectors -- not the rank-2 precision tensors needed for anisotropic uncertainty, and single-component messages collapse multi-modal energy landscapes to physically meaningless averages.

By Zehua Cheng, Wei Dai, Jiahao Sun
arXiv AI
Sep 25

Generalized Graph Variational Autoencoders: Bounded Divergences Control Posterior Collapse

The paper introduces the Generalized Graph Variational Autoencoder (GGVA), which replaces the Kullback–Leibler divergence in the standard variational graph autoencoder with any member of the Rényi–Tsallis family of order $q$. The authors show that for $q<1$ the Tsallis divergence is bounded, whereas the KL and Rényi divergences are unbounded, and that this boundedness can significantly increase the amount of posterior information retained—up to 49× more than the VGAE on several benchmark graphs. Experiments demonstrate that the GGVA’s retained information improves node classification performance, though it does not improve link‑prediction accuracy and only delays, rather than prevents, posterior collapse.

By Kleyton da Costa, Bernardo Modenesi, Ivan F. M. Menezes, Helio Lopes
arXiv Machine Learning
Sep 14

A Generalized Tangent Approximation based Variational Inference Framework for Strongly Super-Gaussian Likelihoods

The paper introduces a new variational inference framework that uses tangent transformations to handle strongly super‑Gaussian likelihoods across a wide range of probability models. By constructing tangent minorants of the log‑likelihood through convex duality, the method achieves conjugacy with Gaussian priors, enabling tractable inference where traditional approaches struggle. The authors provide algorithmic convergence guarantees and near‑parametric risk bounds, and demonstrate superior scalability and accuracy on both simulated and real‑world datasets compared to existing variational algorithms.

By Somjit Roy, Pritam Dey, Debdeep Pati, Bani K. Mallick
arXiv Machine Learning
Aug 5

Information-Geometric Forward Policy Training in GFlowNets

arXiv:2608. 03967v1 Announce Type: cross Abstract: Generative Flow Networks (GFlowNets) have emerged as a flexible framework for amortised inference over discrete and mixed discrete-continuous objects, requiring only an unnormalised target density specified through a reward.

By Yordan Raykov, Rodrigo Veiga
arXiv Statistics ML
2d ago

Exact information accounting for SGD methods

arXiv:2610.00446v1 Announce Type: cross Abstract: As an alternative to the standard geometric analyses, we give an exact, information-theoretic analysis of stochastic gradient descent (SGD) and its v...

By Akshay Balsubramani