The paper introduces Spectral Correction Guidance, a training‑free method that uses spectral alignment to detect and correct deviations in guided diffusion trajectories. By comparing intermediate states to an analytic reference spectrum, the approach improves consistency with the forward process and enhances image generation quality. Experiments show consistent gains over baseline guidance in text‑to‑image tasks and on ImageNet, with benefits across guidance scales and fewer denoising steps.
By Gihoon Kim, Taesup Kim
Steering Fields introduce adaptive vector fields that re-estimate steering directions at each step of a flow-based text-to-image generation process, replacing the fixed global steering vectors traditionally used. By operating on noisy states, they provide a continuous trade-off between steering strength and content preservation, and allow simultaneous induction and inhibition of concepts without explicit spatial masks or object priors. The method achieves state-of-the-art safety steering benchmarks and can also function as a structure-preserving image-editing technique, delivering high semantic fidelity while remaining model-agnostic and inversion-free.
By Simone Facchiano, Jan Eric Lenssen, Bernt Schiele, Wolfgang Stammer, Fabio Galasso, Jonas Fischer
arXiv:2510.03075v4 Announce Type: replace-cross
Abstract: Compositional generalization, the ability to generate novel combinations of known concepts, is a key ingredient for visual generative models....
By Karim Farid, Rajat Sahay, Yumna Ali Alnaggar, Simon Schrodi, Volker Fischer, Cordelia Schmid, Thomas Brox
The paper introduces Global Transport (GT), a class‑agnostic optimal‑transport coupling that can be computed without class labels. GT associates different conditions with distinct regions of the source noise, which degrades performance when used without guidance but consistently improves generation when combined with classifier‑free guidance across various domains, model scales, and sampling budgets. The authors argue that coupling design should be evaluated under guided flow conditions rather than unguided generation, and demonstrate GT’s benefits on both discrete class‑conditioned and continuous text‑conditioned image generation.
By Katarina Petrovi\'c, Zander W. Blasingame, Danyal Rehman, \.Ismail \.Ilkan Ceylan, Michael Bronstein, Stephen Y. Zhang, Lazar Atanackovic, Alexander Tong
arXiv:2605. 31162v1 Announce Type: cross Abstract: Unconditional diffusion models offer powerful generative priors, yet steering them toward aesthetically enhanced outputs remains largely unexplored.
By Shreyansh Modi, Akshat Tomar, Aarush Aggarwal
arXiv:2602. 20360v2 Announce Type: replace Abstract: Flow-based generative methods offer a simple and effective framework for high-fidelity generation, yet pretrained flow models are rarely used in their vanilla conditional form: in image generation, samples without guidance often appear diffuse and lack fine-grained detail.
By Runlong Liao, Jian Yu, Baiyu Su, Chi Zhang, Lizhang Chen, Qiang Liu