← Back to all news
arXiv Computer Vision October 1, 2026 By Taewon Kang, Yu Shen, Ming C. Lin

Dynamics-Inspired Diffusion for Foreground-Preserving Document Background Editing

Read the original on arXiv Computer Vision →

The Flow has not summarised this story yet — read it at arXiv Computer Vision.

  • diffusion
  • multimodal
  • benchmarks

One email a morning, machine-written

One email a day, machine-written, one click to leave. We never share your address.

Related stories

arXiv AI
Jun 15

Conditioning Matters: Stabilizing Inversion and Attention in Diffusion Image Editing

arXiv:2606. 14125v1 Announce Type: cross Abstract: Inversion-based image editing offers flexible and training-free control but still struggles with inversion accuracy and the trade-off between editing fidelity and background preservation.

By Zheyuan Zhan, Hongchen Li, Can Wang, Yinfei Ma, Mingzhen Huang, Ruoshi Bai, Jiawei Chen, Siwei Lyu, Defang Chen
diffusionroboticssafety
More like this →
arXiv AI
Jun 4

On-the-fly Repulsion in the Contextual Space for Rich Diversity in Diffusion Transformers

arXiv:2603. 28762v2 Announce Type: replace-cross Abstract: Modern Text-to-Image (T2I) diffusion models have achieved remarkable semantic alignment, yet they often suffer from a significant lack of variety, converging on a narrow set of visual solutions for any given prompt.

By Omer Dahary, Benaya Koren, Daniel Garibi, Daniel Cohen-Or
llmsdiffusionmultimodalsafety
More like this →
arXiv AI
Jun 17

Detail++: Training-Free Detail Enhancer for Text-to-Image Diffusion Models

arXiv:2507. 17853v2 Announce Type: replace-cross Abstract: Recent advances in text-to-image (T2I) generation have led to impressive visual results.

By Lifeng Chen, Jiner Wang, Zihao Pan, Beier Zhu, Xiaofeng Yang, Chi Zhang
diffusionbenchmarkssafety
More like this →
arXiv Machine Learning
1d ago

ELROND: Exploring and decomposing intrinsic capabilities of diffusion models

arXiv:2602.10216v2 Announce Type: replace Abstract: A single text prompt passed to a diffusion model yields a wide range of visual outputs determined solely by a stochastic process, leaving users wit...

By Pawe{\l} Skier\'s, Emilia Kaczmarczyk, Tomasz Trzci\'nski, Kamil Deja
diffusion
More like this →
arXiv AI
Jun 6

Edit-R2: Context-Aware Reinforcement Learning for Multi-Turn Image Editing

arXiv:2606. 05950v1 Announce Type: new Abstract: Text-guided image editing has advanced rapidly with diffusion models and unified multimodal foundation models.

By Yuxiao Ye, Haoran He, Fangyuan Kong, Xintao Wang, Pengfei Wan, Kun Gai, Ling Pan
diffusionreinforcement-learningmultimodalbenchmarks
More like this →
arXiv Computer Vision
Oct 2

Rethinking Memorization Mitigation in Diffusion Models: Reinforcing Text Conditioning

arXiv:2610.01723v1 Announce Type: new Abstract: Text-to-image diffusion models have achieved remarkable progress in image synthesis, yet can exhibit memorization by closely reproducing individual tra...

By Hyungjun Joo, Sehwan Kim, Hyeonggeun Han, Sangwoo Hong, Jungwoo Lee
diffusionsafety
More like this →
About Pricing API Newsletter Sources Privacy Terms Refunds Accessibility Provider info Contact RSS

The Flow links to publishers and never republishes their articles. Summaries are machine-generated.

v1.1.0 · 5f852ea