arXiv Computer Vision By Thanh-Khoi Nguyen, Thien-Phuc Tran, Minh-Triet Tran

Query Rewriting for Complex Object Segmentation in 4D Gaussian Representations

Read the original on arXiv Computer Vision →

The paper examines how rewriting verbose, narrative-style queries into concise keyword‑grounded forms improves complex object segmentation in 4D Gaussian representations. By applying a training‑free reinterpretation strategy, the authors reduce linguistic noise while preserving essential semantic anchors. Experiments on HyperNeRF and Neu3D show that rewritten queries boost temporal accuracy from 60.92% to 92.21% and vIoU from 20.08% to 76.94%, with ablation studies confirming the benefits of shorter, keyword‑focused queries.

Machine-generated by The Flow from the publisher's headline and feed description — not written or checked by a human. The full article lives at arXiv Computer Vision.

arXiv AI
Jul 29

Reasoning with Memory: A Temporal Granularity-Adaptive Framework for Training-Free Long Video Understanding

arXiv:2607. 24794v1 Announce Type: new Abstract: While Multimodal Large Language Models (MLLMs) demonstrate superior generalization in fundamental video tasks, restricted context windows limit their long video understanding.

By Linghao Meng, Qiankun Li, Junyuan Mao, Pujin Liao, Zhicheng He, Enbo Zhang, Kun Wang, Yang Liu, Huazhu Fu, Yueming Jin
arXiv AI
Aug 25

ExtrinSplat: Decoupling Geometry and Semantics for Open-Vocabulary Understanding in 3D Gaussian Splatting

ExtrinSplat is a new framework that separates geometry from semantics in 3D Gaussian Splatting scenes. It clusters Gaussians into overlapping 3D object groups and uses a Vision‑Language Model to generate lightweight textual hypotheses, creating an extrinsic index layer that handles complex polysemy. This approach reduces adaptation time from hours to minutes, cuts storage overhead by orders of magnitude, and outperforms existing embedding‑based methods on open‑vocabulary 3D object selection and semantic segmentation benchmarks.

By Jiayu Ding, Xinpeng Liu, Zhiyi Pan, Shiqiang Long, Ge Li