arXiv Machine Learning By Xing Cong, Hanlin Tang, Kan Liu, Lan Tao, Lin Qu, Chenhao Xie

RT-Lynx: Putting the GEMM Sparsity In a Right Way for Diffusion Models

Read the original on arXiv Machine Learning →

arXiv:2605. 26632v2 Announce Type: replace Abstract: Diffusion Transformers (DiT) achieve strong performance in image generation but incur substantial inference costs.

Machine-generated by The Flow from the publisher's headline and feed description — not written or checked by a human. The full article lives at arXiv Machine Learning.

arXiv Computer Vision
Sep 7

Importance-Aware Low-Rank Distillation of Diffusion Transformers

The paper introduces SVDtrunc, a two‑step block‑level compression method for Diffusion Transformers (DiTs) that allocates ranks across blocks, applies truncated SVD to the least important ones, and then fine‑tunes all blocks with modular knowledge distillation and a rectified‑flow objective. Experiments on FLUX.dev show that SVDtrunc achieves near‑full performance at 68% of the original parameters and remains competitive even at 57%, outperforming all competing approaches on GenEval, HPSv2, and DPG benchmarks. The method also works well without fine‑tuning, complementing step distillation and offering a practical path to efficient large‑scale generative models.

By Denis Zavadski, Sebastian Heid, Damjan Kal\v{s}an, Stefan Roth, Carsten Rother