arXiv Machine Learning

On the Redundancy of Timestep Embeddings in Diffusion Models

arXiv:2606. 20416v1 Announce Type: new Abstract: Diffusion models rely heavily on explicit timestep embeddings to modulate the denoising process across various noise scales.

arXiv AI
Jul 3

Subliminal Clocks: Latent Time Modelling in Diffusion Language Models

arXiv:2607. 01774v1 Announce Type: new Abstract: Diffusion Language Models (DLMs) have recently emerged as a promising alternative to autoregressive models.

By Maximo Rulli (Sapienza University of Rome), Thomas Fontanari (Sapienza University of Rome), Simone Petruzzi (Sapienza University of Rome), Federico Alvetreti (Sapienza University of Rome), Giorgio Strano (Sapienza University of Rome), Donato Crisostomi (Sapienza University of Rome), Giorgos Nikolaou (EPFL), Tommaso Mencattini (EPFL), Andrea Santilli (Independent researcher), Emanuele Rodol\`a (Sapienza University of Rome), Simone Scardapane (Sapienza University of Rome), Alessio Devoto (Independent researcher)
arXiv AI
6d ago

Does Uniform Discrete Diffusion Need Time?

Uniform discrete diffusion models (UDMs) typically rely on explicit time conditioning, yet this study finds that such conditioning is often unnecessary in practice. While the population‑optimal UDM predictor generally depends on time—controlling how much the model should trust the observed context—the dependence becomes negligible in finite‑data language settings. Empirical results show that trained language UDMs exhibit limited time sensitivity across most of the diffusion trajectory, and time‑agnostic predictors can match or outperform time‑conditioned models on various datasets and training objectives.

By Chunsan Hong, Chieh-Hsin Lai, Satoshi Hayakawa, Yuhta Takida, Jong Chul Ye, Yuki Mitsufuji
arXiv Machine Learning
Sep 14

DCRA: Diffusion-Conditioned Representation Alignment for Robust Time-Series Learning

The paper introduces Diffusion-Conditioned Representation Alignment (DCRA), a training framework that uses the forward diffusion process as a structured corruption scheduler for time‑series representation learning. DCRA aligns representations across noise levels with a feature‑level consistency objective, preserving class‑discriminative structure and enabling smooth, semantically coherent trajectories in latent space. Experiments on the CHB‑MIT EEG dataset demonstrate that DCRA improves seizure detection performance under various noise conditions, achieving higher sensitivity at low false‑positive rates and producing more balanced, structured representations than baseline methods.

By Wenrui Xu, Anas Enanaa, Keshab K. Parhi
arXiv AI
Jun 17

Rethinking Cross-Layer Information Routing in Diffusion Transformers

arXiv:2605. 20708v2 Announce Type: replace-cross Abstract: Diffusion Transformers (DiTs) have become a de facto backbone of modern visual generation, and nearly every major axis of their design -- tokenization, attention, conditioning, objectives, and latent autoencoders -- has been extensively revisited.

By Chao Xu, Maohua Li, Qirui Li, Yixuan Xu, Yanke Zhou, Yunhe Li, Cuifeng Shen, Hanlin Tang, Kan Liu, Tao Lan, Lin Qu, Shao-Qun Zhang