arXiv Machine Learning By Jaeyeon Kim, Jonathan Geuter, David Alvarez-Melis, Sham Kakade, Sitan Chen

Stop Training for the Worst: Progressive Unmasking Accelerates Masked Diffusion Training

Read the original on arXiv Machine Learning →

arXiv:2602. 10314v2 Announce Type: replace Abstract: Masked Diffusion Models (MDMs) have emerged as a promising approach for generative modeling in discrete spaces.

Summary generated by The Flow from the publisher's feed. The full article lives at arXiv Machine Learning.