Computer vision

Detection, segmentation, depth and recognition research, plus the vision backbones that keep displacing the last generation.

3,375 stories · RSS feed

arXiv Computer Vision
Sep 30

OmniTaskonomy: When Does Visual Generation Improve Visual Understanding?

arXiv:2609.38079v1 Announce Type: new Abstract: Training a model to generate visual content can encourage it to learn rich perceptual capabilities related to geometry, spatial relationships, and obje...

By Jiaxin Ge, Yiming Qin, Ji Xie, Haozhe Jiang, Xiaochuang Han, Junyi Zhang, Andrew Dai, Yinfei Yang, Jitendra Malik, Ranjay Krishna, Sewon Min, Haiwen Feng, Le Xue, Baifeng Shi, Trevor Darrell, XuDong Wang
arXiv Computer Vision
Sep 30

When to Adapt: Multi-Signal Domain Shift Detection for Efficient Training-Free Adaptation in Open-Vocabulary Segmentation

arXiv:2609.37602v1 Announce Type: cross Abstract: Robust and reliable perception is essential for autonomous robots operating in real-world environments, particularly in long-term missions where envi...

By Michele Antonazzi, Alejandra C. Hernandez, Jos\'e Araujo, Olov Andersson, Patric Jensfelt
arXiv Computer Vision
Sep 30

Achieving detailed medial temporal lobe segmentation with upsampled isotropic training from implicit neural representation

arXiv:2508.17171v3 Announce Type: replace Abstract: Imaging biomarkers in magnetic resonance imaging (MRI) are important tools for diagnosing, tracking and treating Alzheimer's disease (AD). Neurofib...

By Yue Li, Pulkit Khandelwal, Rohit Jena, Long Xie, Michael Duong, Amanda E. Denning, Christopher A. Brown, Laura E. M. Wisse, Sandhitsu R. Das, David A. Wolk, Paul A. Yushkevich