← Back to all news
arXiv Computer Vision September 30, 2026 By Xinnan Zhu, Ruijie Xu, Jiayu Ying, Daoguo Dong, Jiachen Xu, Yuan Xie, Xin Tan

JointEdit3D: Feed-Forward 3D Scene Editing in a Unified Latent Space

Read the original on arXiv Computer Vision →

The Flow has not summarised this story yet — read it at arXiv Computer Vision.

  • diffusion
  • benchmarks

One email a morning, machine-written

One email a day, machine-written, one click to leave. We never share your address.

Related stories

arXiv Computer Vision
Sep 1

DisCo3D: Distilling Multi-View Consistency for 3D Scene Editing

arXiv:2508.01684v2 Announce Type: replace Abstract: While diffusion models have demonstrated remarkable progress in 2D image generation and editing, extending these capabilities to 3D editing remains...

By Yufeng Chi, Huimin Ma, Kafeng Wang, Jianmin Li
diffusionefficiencybenchmarks
More like this →
arXiv Computer Vision
2d ago

GeoVerse: World-Consistent Novel View Synthesis in Geometric Latent Space

arXiv:2609.35734v2 Announce Type: replace Abstract: Novel view synthesis from sparse images must reconcile faithful reconstruction of observed regions with plausible completion of unseen content, whi...

By Kerui Ren, Tao Lu, Linning Xu, Changjian Jiang, Mu Huang, Chunhua Shen, Mulin Yu, Bo Dai
diffusion
More like this →
arXiv Computer Vision
Sep 15

What Makes a 3D Scene Editable? A Factorized Benchmark of Fidelity, Locality, Consistency, and Preservation

arXiv:2609.14899v1 Announce Type: new Abstract: Neural 3D scene editing is often evaluated by semantic alignment alone, although a convincing result may alter unrelated content or become inconsistent...

By Sariah Patro, Arjun Mehra, Nikhil Bhatia
roboticsbenchmarkssafety
More like this →
arXiv AI
Jun 8

Native3D: End-to-End 3D Scene Generation via Unified Mesh-Texture Modeling and Semantic Alignment

arXiv:2606. 07117v1 Announce Type: cross Abstract: This paper presents Native3D, the first end-to-end 3D scene generation framework that completely bypasses 2D intermediate representations.

By Yibo Liu, Ziwei Zhang, Haozhou Pang, Menghao Li, Lanshan He, Gan Qi
llmsdiffusionfine-tuningsafety
More like this →
arXiv Computer Vision
Sep 22

Mira-Scene: Pixel-Aligned Layouts for Generative 3D Scene

arXiv:2609.23796v1 Announce Type: new Abstract: Single-image 3D object generation can now produce high-fidelity assets, yet accurately placing them into a coherent scene layout remains an open challe...

By Yang-Tian Sun, Tianjia Liu, Zehuan Huang, Yi-Hua Huang, Xiaoyang Lyu, Ziyi Yang, Zi-Xin Zou, Yuan-Chen Guo, Yan-Pei Cao, Xiaojuan Qi
llmsdiffusionmultimodalsafety
More like this →
arXiv AI
Jun 30

Edit in 2D, Verify in 3D: Reinforcement Learning for Multi-view Consistent Scene Editing

arXiv:2603. 03143v2 Announce Type: replace-cross Abstract: Leveraging the priors of 2D diffusion models for 3D editing has emerged as a promising paradigm.

By Jiyuan Wang, Chunyu Lin, Lei Sun, Zhi Cao, Yuyang Yin, Lang Nie, Zhenlong Yuan, Xiangxiang Chu, Yunchao Wei, Kang Liao, Guosheng Lin
diffusioncomputer-visionreinforcement-learningfine-tuningbenchmarks
More like this →
About Pricing API Newsletter Sources Privacy Terms Refunds Accessibility Provider info Contact RSS

The Flow links to publishers and never republishes their articles. Summaries are machine-generated.

v1.1.0 · 5f852ea