Universal Image Segmentation with Mask2Former and OneFormer
Related stories
Introducing Segment Anything: Working toward the first foundation model for image segmentation
LUMA: Benchmarking Segmentation via a Lightweight Universal Mask Adapter
arXiv:2607. 00687v1 Announce Type: cross Abstract: Comparing transformer backbones for image segmentation is confounded: each is paired with a different decoder, recipe, and pretraining, so reported differences rarely reflect the backbone itself.
Mask Proposal Voting Based on Geodesic Framework for Robust Image Segmentation
arXiv:2606. 14912v1 Announce Type: cross Abstract: Despite great advances, finding accurate segmentation remains a challenging task, especially in scenarios with cluttered backgrounds, complex intensity variations and topology appearance.
A Comprehensive Survey of Medical Image Segmentation: Challenges, Benchmarks, and Beyond
arXiv:2606. 16153v1 Announce Type: cross Abstract: Medical image segmentation plays a critical role in clinical diagnostics, treatment planning, disease monitoring, and neurological disorder identification.
Conformal Prediction Sets for Instance Segmentation
arXiv:2602. 10045v2 Announce Type: replace-cross Abstract: Current instance segmentation models achieve high performance on average predictions, but lack principled uncertainty quantification: their outputs are not calibrated, and there is no guarantee that a predicted mask is close to the ground truth.
U-CFR: Uncertainty-Guided Cascade Forward Refinement for Interactive Segmentation
arXiv:2607. 20705v1 Announce Type: cross Abstract: Interactive image segmentation is critical for efficient image annotation; however, existing methods often require many corrective clicks or rely on passive refinement schemes that converge slowly.
Sub-Semantic Image Segmentation
arXiv:2606. 14754v1 Announce Type: cross Abstract: Images can be segmented based on visual cues (i.
++nnU-Net: Scaling nnU-Net with Prefix-Based Data Augmentation
arXiv:2606. 10713v1 Announce Type: cross Abstract: The nnU-Net has demonstrated continuous success in medical segmentation tasks, which heavily rely on the availability and diversity of annotated biomedical data.
Object-centric LeJEPA
arXiv:2607. 02404v1 Announce Type: cross Abstract: Image encoders trained with LeJEPA can deliver strong features for downstream tasks, but, like other image-level self-supervised methods, typically require large training datasets.
AlbumentationsX: One Augmentation Pipeline for Images and Related Annotations
Augmentation can corrupt a training example when an image and its annotations receive different random changes. A crop must use the same coordinates for the image, mask, boxes, keypoints, stereo views, video frames, or volume.
OA-CutMix: Correcting the Label Bias of CutMix
arXiv:2606. 04820v1 Announce Type: cross Abstract: CutMix has become the de facto standard mixing augmentation, yet its label assignment rests on a flawed assumption: The area of the pasted patch faithfully reflects its semantic contribution to the mixed image.