arXiv Statistics ML

Repulsive normalizing flow mixtures for adaptive importance sampling: reliability analysis of complex systems

The paper introduces FAMIS, a flow-based multiple importance sampling framework that learns a nonuniform mixture of normalizing flow proposals for rare‑event estimation. It does not need presampled failure data or prior knowledge of failure modes, instead adapting the mixture through sequential evaluations of the limit state function. The method employs a smooth rare‑event surrogate, a tempered target sequence, defensive exploration, Rao‑Blackwellized weight updates, and a Jensen‑Shannon repulsion term to promote diversity, achieving accurate failure probability estimates with fewer training samples and stable variance reduction in complex reliability problems.

arXiv Machine Learning
Aug 5

Information-Geometric Forward Policy Training in GFlowNets

arXiv:2608. 03967v1 Announce Type: cross Abstract: Generative Flow Networks (GFlowNets) have emerged as a flexible framework for amortised inference over discrete and mixed discrete-continuous objects, requiring only an unnormalised target density specified through a reward.

By Yordan Raykov, Rodrigo Veiga
arXiv Machine Learning
Jul 13

SafeExplorer: An Unbiased Policy Gradient for Reinforcement Learning with Recovery Interventions

arXiv:2607. 08925v1 Announce Type: new Abstract: Training reinforcement-learning agents directly on physical robots makes every fall costly, since a fall can damage the platform and cannot be undone like a simulator reset; the goal is therefore to minimize falls during training rather than trade them off against return, as constrained Markov decision process (MDP) formulations do.

By Elham Daneshmand, Majid Khadiv, Glen Berseth, Hsiu-Chin Lin