arXiv Machine Learning

Three-Body Scattering for Generative Modeling

arXiv:2607. 18198v1 Announce Type: new Abstract: Modern generative models typically rely on an adversarial critic, a prescribed noise-to-data path, or an autoregressive factorization.

Hugging Face Trending Papers
Jul 20

Three-Body Scattering for Generative Modeling

Modern generative models typically rely on an adversarial critic, a prescribed noise-to-data path, or an autoregressive factorization. Instead, we show that a proper distributional energy can induce sample-level motion and provide direct regression supervision for a one-step generator.

arXiv Machine Learning
1d ago

One-Step Generative Modeling via Training Dynamics Action

The paper introduces TDAction, a method for one‑step generative modeling that selects transport targets during training based on a cost reflecting shared‑parameter effort and terminal mismatch. By formulating this as a soft‑terminal control problem, the authors derive a closed‑form Batch Tangent Action‑to‑Go value that captures cross‑sample interactions and can be efficiently implemented with randomized tangent probes. Experiments on ImageNet 256×256 demonstrate that TDAction achieves an FID below 1.1 without distillation.

By Zhangyong Liang, Ying Huang, Haibin Ling
arXiv Machine Learning
Aug 13

Fine-Tuning Generative Models for Extreme Events via CVaR-Penalized Wasserstein Gradient Flows

arXiv:2608. 11544v1 Announce Type: cross Abstract: We propose CVaR-penalized Generative Particle Algorithm (CVaR-GPA), a robust, tail-agnostic algorithm for fine-tuning generative models to learn heavy-tailed distributions and capture extreme events, requiring no prior knowledge or estimation of the target's tail characteristics.

By Thejani Gamage, Hyemin Gu, Zhizhen Zhang, Ziyu Chen, Markos Katsoulakis, Luc Rey-Bellet
Hugging Face Trending Papers
Aug 20

Continuous Adversarial MeanFlow Transfer

Training fast generators on new domains with limited data remains challenging for two reasons. First, adapting a pretrained diffusion or flow model to a new domain leaves its costly multi-step sampling unaddressed, and existing acceleration methods are tied to the source parameterization--$ε$, $x$, $v$, or $u$--leaving heterogeneous pretrained models with no common acceleration target.

arXiv AI
2d ago

Discrete Wasserstein Flows for One-Step Generative Modeling

The paper presents a new one‑step generative modeling framework for finite state spaces, leveraging discrete Wasserstein geometry to define a target‑relative KL gradient flow over a reversible Markov kernel. The authors implement this flow at the particle level using Markov jumps and encode the resulting transport updates into a latent‑conditioned generator, enabling one‑step inference after training. Experiments on a controlled setting confirm KL dissipation, consistency between particle dynamics and probability flow, and accurate numerical scaling, while a finite‑capacity neural generator successfully tracks the exact transport targets.

By Alessandro Micheli, Andrea Zerio, Samir Bhatt
arXiv Machine Learning
Sep 22

When Does Adversarial Refinement Help? A Negative Result and Open Problem in Adapting R3GAN to Time Series Imputation

The paper investigates whether the stable GAN architecture R3GAN can improve time‑series imputation when adapted to 1‑D temporal data. Using a coarse‑to‑fine refinement framework and a frequency‑domain discriminator, the authors evaluate 14 saved configurations across three datasets and find a negative result: most configurations either show negligible improvement or degrade performance compared to baseline methods. The study highlights that the usual argument—GANs optimize distributional objectives rather than point‑wise ones—does not fully explain the lack of benefit, and it poses an open problem regarding why a learned discriminator fails to provide useful refinement gradients while diffusion denoisers succeed, offering practical guidance on when adversarial refinement may be worthwhile.

By Yufeng He
arXiv Machine Learning
4d ago

Improved Distributional Diffusion Models

arXiv:2609.37147v1 Announce Type: cross Abstract: Distributional Diffusion Models (DDMs) replace the standard mean-prediction denoiser with a \emph{distributional} denoiser trained via a scoring rule...

By Tommaso Martorella, Alexandre Galashov, Felix Krause, Stefan Andreas Baumann, Valentin De Bortoli, Arthur Gretton, Bj\"orn Ommer