← Back to all news
arXiv AI October 1, 2026 By Yilie Huang, Wenpin Tang, Xunyu Zhou

ART for Diffusion Sampling: A Reinforcement Learning Approach to Timestep Schedule

Read the original on arXiv AI →

The Flow has not summarised this story yet — read it at arXiv AI.

  • diffusion
  • reinforcement-learning

One email a morning, machine-written

One email a day, machine-written, one click to leave. We never share your address.

Related stories

arXiv AI
Jul 3

ART for Diffusion Sampling: Continuous-Time Control and Actor-Critic Learning

arXiv:2607. 02137v1 Announce Type: cross Abstract: We study timestep allocation for score-based diffusion sampling, where a learned reverse-time dynamics is discretized on a finite grid.

By Yilie Huang, Wenpin Tang, Xun Yu Zhou
diffusionreinforcement-learning
More like this →
Hugging Face Trending Papers
Jul 2

ART for Diffusion Sampling: Continuous-Time Control and Actor-Critic Learning

We study timestep allocation for score-based diffusion sampling, where a learned reverse-time dynamics is discretized on a finite grid. Uniform and hand-crafted schedules are standard choices, but they rely on fixed prescriptions and can therefore be suboptimal.

diffusionreinforcement-learning
More like this →
arXiv AI
2d ago

Adaptive Reparametrized Time for Score-Based Diffusion Sampling

arXiv:2607.02137v3 Announce Type: replace-cross Abstract: We study timestep allocation for score-based diffusion sampling, where a learned reverse-time dynamics is discretized on a finite grid. Unifo...

By Yilie Huang, Wenpin Tang, Xun Yu Zhou
diffusionreinforcement-learning
More like this →
arXiv Machine Learning
Jun 2

Learning To Sample From Diffusion Models Via Inverse Reinforcement Learning

arXiv:2602. 08689v2 Announce Type: replace Abstract: Diffusion models generate samples through an iterative denoising process guided by a pretrained neural network.

By Constant Bourdrez, Alexandre V\'erine, Olivier Capp\'e
diffusionreinforcement-learningfine-tuning
More like this →
arXiv Machine Learning
Jun 16

Temporal Difference Learning for Diffusion Models

arXiv:2606. 15048v1 Announce Type: new Abstract: Diffusion models are typically trained with objectives that focus on local denoising targets at individual time steps (or adjacent pairs), which do not enforce consistency between predictions along the denoising trajectory.

By Qizhen Ying, Yangchen Pan, Victor Adrian Prisacariu, Junfeng Wen
diffusionreinforcement-learningbenchmarks
More like this →
arXiv AI
Jul 9

Selective Timestep Weighting and Advantage-Based Replay for Sample-Efficient Diffusion RLHF

arXiv:2607. 07693v1 Announce Type: cross Abstract: Reinforcement learning from human feedback (RLHF) has emerged as a powerful paradigm for aligning generative models with human preferences.

By Eric Zhu, Abhinav Shrivastava, Soumik Mukhopadhyay
diffusionreinforcement-learning
More like this →
About Pricing API Newsletter Sources Privacy Terms Refunds Accessibility Provider info Contact RSS

The Flow links to publishers and never republishes their articles. Summaries are machine-generated.

v1.1.0 · 5f852ea