arXiv Machine Learning By Parsa Esmati, Somjit Nath, Katja Hofmann, Derek Nowrouzezahrai, Samira Ebrahimi Kahou, Majid Mirmehdi

The Invisible Hand of Physics: When Video Diffusion Models Know More Than They Show

Read the original on arXiv Machine Learning →

arXiv:2606. 05328v1 Announce Type: cross Abstract: Modern video diffusion models generate increasingly realistic and temporally coherent videos, motivating their use as candidate world simulators.

Summary generated by The Flow from the publisher's feed. The full article lives at arXiv Machine Learning.