Our latest Veo update generates lively, dynamic clips that feel natural and engaging — and supports vertical video generation.
SkillPE is a prompt‑engineering framework that evolves reusable cinematic skills from expert‑authored seeds to improve text‑to‑video generation for non‑experts. It encodes shot logic, composition, lighting, sound design, and other filmmaking cues in a fine‑grained format, and uses movie references classified as resonators, dissonants, and divergents to refine skill application and inspire creative alternatives. Experiments on StoryEval and VBench demonstrate up to 1.40‑point gains over the strongest baseline and 0.51 points over seed skills on a 7‑point four‑dimensional evaluation, while remaining competitive on benchmark‑native metrics.
Despite remarkable progress in text-guided image editing, generative models frequently fail to preserve visual object consistency, defined as the preservation of a subject's key attributes throughout the editing process. We address this limitation through three contributions.
The paper "Iterative Flow Matching: Path Correction and Gradual Refinement for Enhanced Generative Modeling" investigates the use of flow matching for image generation and identifies that this approach can produce hallucinations—unrealistic images. It proposes an iterative refinement process that can be incorporated into virtually any generative modeling technique to improve performance and robustness. The authors demonstrate how their method corrects the generation path and gradually refines outputs to mitigate hallucinations.
By Eldad Haber, Shadab Ahamed, Md. Shahriar Rahim Siddiqui, Niloufar Zakariaei, Moshe Eliasof
We partnered with Darren Aronofsky, Eliza McNitt and a team of more than 200 people to make a film using Veo and live-action filmmaking.
arXiv:2608. 08101v1 Announce Type: new Abstract: Generative AI has emerged as one of the most transformative forces in modern artificial intelligence, reshaping how we create, imagine, and interact with digital content.
By Jun Lu