arXiv Computer Vision By Teemu Saukkio, Hashem Haghbayan, Juha Plosila

Overlapping Visual Grouping Without Semantic Priors

Read the original on arXiv Computer Vision →

The paper introduces Domain Parent Grouping (DPG), a sensor‑grounded method that forms perceptual units directly from raw measurements without relying on semantic priors. DPG operates across three domains—local luminance, direct chromatic, and contextual chromatic—creating spatially connected groups that overlap across domains, thus producing a non‑exclusive grouping representation. Experiments on the BSDS500 dataset show that DPG’s groups align with low‑level image structure and correlate with human‑annotated regions and boundaries.

Machine-generated by The Flow from the publisher's headline and feed description — not written or checked by a human. The full article lives at arXiv Computer Vision.

Hugging Face Trending Papers
Aug 6

Prior-SG: Task and Prior Driven Region Segmentation for Scene Graphs in Arbitrarily-Structured Environments

Hierarchical 3D scene graphs are a promising representation for high-level spatial reasoning in autonomous mobile platforms. However, existing extraction frameworks typically rely on purely local visual clustering or strict geometric heuristics, such as wall-separated rooms, which fail in open-plan or arbitrarily-structured environments.

arXiv AI
Oct 2

Geometric Similarity in VLM Low-Level Vision Representations

The paper introduces GeoSim, a four‑level framework for analyzing how vision‑language models (VLMs) represent low‑level vision tasks. It evaluates hidden‑layer representations across 24 tasks and two VLM paradigms—autoregressive models and diffusion transformers—using global similarity, local geometry, sparse feature decomposition, and topological verification. The study uncovers the organizing principles of low‑level visual representations and highlights their limitations in cross‑task and cross‑model agreement, offering an interpretability lens for assessing latent transferability and diagnosing model‑specific issues.

By Shao-Jun Xia, Huixin Zhang, Zhen Lei, Anlan Sun, Yuner Zhang, Xiaoyang Chen