arXiv AI By Jiacheng Liu, Jason Liu

MegaSlide-DiT: Memory-Centric Adaptation and Deformable Local Attention for Efficient Video Diffusion

Read the original on arXiv AI →

arXiv:2607. 22696v1 Announce Type: cross Abstract: High-resolution video diffusion models built on Diffusion Transformers (DiTs) deliver strong fidelity but quickly exhaust the memory budget of a single workstation.

Summary generated by The Flow from the publisher's feed. The full article lives at arXiv AI.