arXiv:2503.10685v3 Announce Type: replace
Abstract: Unsupervised Domain Adaptation (UDA) enables strong generalization from a labeled source domain to an unlabeled target domain, often with limited d...
By Brun\'o B. Englert, Gijs Dubbelman
arXiv:2504.18190v2 Announce Type: replace
Abstract: Unsupervised Domain Adaptation (UDA) can improve a perception model's generalization to an unlabeled target domain starting from a labeled source d...
By Brun\'o B. Englert, Tommie Kerssies, Gijs Dubbelman
arXiv:2609.39681v1 Announce Type: new
Abstract: Unsupervised domain adaptation (UDA) reduces the annotation burden in panoptic segmentation by leveraging a cost-effectively labeled source domain (e.g...
By Ivan Martinovi\'c, Josip \v{S}ari\'c, Yuki M. Asano, Sini\v{s}a \v{S}egvi\'c
The paper introduces VPRef, the first cross‑domain benchmark for Referring Remote Sensing Image Segmentation, containing 46,972 language‑image‑annotation triplets with a three‑tier linguistic hierarchy. It proposes a parameter‑efficient adaptation method based on the Segment Anything Model and Low‑Rank Adaptation, using pseudo‑label self‑training for visual drift and random multi‑granularity prompt mixing for textual drift. Experiments show the approach improves cross‑domain segmentation while altering only 1.08 % of the base model’s parameters, offering a strong baseline for future research.
By Quanwei Liu, Tao Huang, Jiaqi Yang, Wei Xiang
arXiv:2410. 21361v2 Announce Type: replace-cross Abstract: Domain adaptation has been extensively investigated in computer vision but still requires access to target data at the training time, which might be difficult to obtain in real-world autonomous driving scenarios, especially under rare or adverse conditions.
By Mohammad Fahes, Tuan-Hung Vu, Andrei Bursuc, Patrick P\'erez, Raoul de Charette
arXiv:2407. 21311v2 Announce Type: replace-cross Abstract: Unsupervised domain adaptation (UDA) aims to mitigate domain shift, where the distribution of labeled source data differs from that of unlabeled target data.
By Ali Abedi, Q. M. Jonathan Wu, Ning Zhang, Farhad Pourpanah
Rapid advancements in vision-language models have propelled Referring Remote Sensing Image Segmentation (RRSIS) to the forefront of Earth observation. However, practical deployments suffer severe perf...
The paper introduces ICM, an Intra-Class Mixing Consistency framework for unsupervised domain adaptation in semantic segmentation under adverse weather. ICM mixes regions within the same image and semantic class to maintain realistic layouts, contrasting prior methods that combine across images or domains. On the Cityscapes → ACDC benchmark, ICM achieves 75.7% mIoU, surpassing previous state‑of‑the‑art results by 1.9 percentage points.
By Boying Li, Chang Liu, Britta Ayano Wilde, Gy\"orgy Kov\'acs, Tosin Adewumi, Bj\"orn Backe, Hamam Mokayed
arXiv:2606.16996v2 Announce Type: replace-cross
Abstract: Segment Anything Model 3 (SAM 3) provides a strong frozen backbone for concept-prompted segmentation, but applying it directly to open-vocabu...
By Tran Dinh Tien, Zhiqiang Shen
arXiv:2604. 15622v3 Announce Type: replace-cross Abstract: Always-on contextual AI runs language-aligned vision foundation models (VFMs) on edge devices, where the on-device model is the dominant continuous compute cost under strict latency and power limits.
By Yiwei Zhao, Yi Zheng, Huapeng Su, Jieyu Lin, Stefano Ambrogio, Cijo Jose, Michael Ramamonjisoa, Patrick Labatut, Barbara De Salvo, Chiao Liu, Phillip B. Gibbons, Ziyun Li
arXiv:2511.11286v4 Announce Type: replace-cross
Abstract: Out-of-domain (OOD) robustness is challenging to achieve in real-world computer vision, especially in unsupervised domain adaptation scenario...
By Ruoqi Wang, Haitao Wang, Shaojie Guo, Qiong Luo
The paper proposes Difficulty-Aware Sample Allocation (DASA), a framework that assigns stronger data augmentation to training samples deemed more difficult based on a composite difficulty score. This score integrates prediction ambiguity, training loss, class rarity, and boundary complexity, and is used to modulate augmentation strength during iterative training. Experiments on Oxford‑IIIT Pet and binary Pascal VOC with U‑Net, DeepLabV3, and SegFormer‑B0 demonstrate that DASA outperforms standard training and matches or exceeds single‑signal adaptive baselines, notably raising DeepLabV3’s mIoU from 0.633 to 0.740 on Oxford‑IIIT Pet.
By Olasimbo Ayodeji Arigbabu, Abimbola Ismail Arigbabu