BiCLIP is a bidirectional multimodal framework that enhances medical image segmentation by allowing visual features to iteratively refine textual representations, improving semantic alignment. It incorporates an augmentation consistency objective to stabilize learning against perturbed inputs. Experiments on QaTa-COV19 and MosMedData+ show that BiCLIP outperforms state‑of‑the‑art image‑only and multimodal baselines, achieving strong performance even with only 1% labeled data and resisting common clinical artifacts such as motion blur and low‑dose CT noise.
By Saivan Talaei, Fatemeh Daneshfar, Abdulhady Abas Abdullah, Mourad Oussalah
The paper proposes a two‑stage learning framework for multi‑organ segmentation that handles partially annotated datasets and domain shifts. First, the model learns accurate segmentations from available annotations to build robust feature representations. Second, it introduces learnable organ prototypes and a Sinkhorn‑triplet loss to enforce organ‑wise feature consistency across datasets, keeping embeddings of the same organ close while separating different organs, even when annotations are missing.
By Dakini Mallam Garba, Salim Abdou Daoura
arXiv:2607.12896v3 Announce Type: replace
Abstract: Medical image segmentation foundation models are expected to generalize across diverse clinical scenarios, yet existing universal methods remain fr...
By Yunzhou Li, Jiesi Hu, Yanwu Yang, Hanyang Peng, Chenfei Ye, Jianfeng Cao, Yixuan Yuan, Ting Ma
The paper introduces SSS, a semi‑supervised framework that builds on the Vision Foundation Model SAM‑2 to improve medical image segmentation. It combines a weak‑to‑strong consistency regularization with a Discriminative Feature Enhancement mechanism and a prompt generator that uses Physical Constraints with a Sliding Window to supply prompts for unlabeled data. Experiments on the ACDC and BHSD datasets show that SSS outperforms prior methods, achieving a 53.15 Dice score on BHSD, a +3.65 improvement over the state of the art.
By Hongjie Zhu, Xiwei Liu, Rundong Xue, Zeyu Zhang, Yong Xu, Daji Ergu, Ying Cai, Yang Zhao
arXiv:2609.25850v1 Announce Type: new
Abstract: Deep learning performance generally improves with increasing training data, yet this scaling is fundamentally constrained by annotation cost in large-s...
By Xiaofei Du, Lei Zhang, Shuyu Yan, Manning Wang, Zhijian Song
arXiv:2607.10851v2 Announce Type: replace
Abstract: Medical image classification models are ideally expected to identify diagnostically relevant regions while making predictions, yet standard classif...
By Tonmoy Hossain, Atiqur Rahman, Farhana Hossain Swarnali, Miaomiao Zhang