arXiv Machine Learning

Test-Time Adaptation in Optical Coherence Tomography Using Trajectory-Aligned Time-Independent Flow

arXiv:2606. 18876v1 Announce Type: cross Abstract: Optical coherence tomography (OCT) is essential in ophthalmology, but inconsistent image quality especially in low-cost devices hinders automated analysis.

arXiv Computer Vision
Aug 28

Automated 2D and 3D Segmentation of AMD and DME Lesions in OCT

This study presents four deep‑learning pipelines—two‑dimensional and three‑dimensional—for segmenting age‑related macular degeneration (AMD) and diabetic macular edema (DME) lesions in optical coherence tomography (OCT) images. The models achieve Dice scores between 0.76 and 0.82 and demonstrate strong volumetric and surface calibration (r_vol, r_surf ≥ 0.97) on an in‑domain validation set. Generalization was assessed on the OLIVES clinical cohort using proxy metrics such as biomarker AUROC, central subfield thickness correlation, and longitudinal concordance, showing that the predictions still track clinical biomarkers outside the training distribution, albeit with reduced strength.

By Lucia Sundberg, Zhihao Zhao, M. Ali Nasseri
arXiv Computer Vision
Aug 25

When the Edit Changes the Patient: Measuring Identity Preservation in Counterfactual Retinal Images

arXiv:2608.23024v1 Announce Type: new Abstract: Counterfactual medical image generation aims to modify an existing image to reflect a hypothetical scenario in which certain characteristics of the ima...

By Andrea Posada, Wenke Karbole, Bach Ngoc Doan, Alexander Weers, Solmaz Abdolrahimzadeh, Maria Patsiamanidi, Kahkashan Haider, Vaishali Khare, Daniel Rueckert, Andrew Lotery, Sobha Sivaprasad, Martin J. Menten
arXiv Computer Vision
Sep 14

An Ultra-Widefield Swept-Source OCTA Dataset and a Polar-Gated Mamba Network for Retinal Vessel Segmentation

arXiv:2609.12574v1 Announce Type: new Abstract: Ultra-widefield (UWF) swept-source optical coherence tomography angiography (SS-OCTA) enables large-area retinal vascular imaging, yet vessel segmentat...

By Yang Liu, Yibing Shen, Keming Zhao, Cenk Jiang, Zhenghang Qian, Zhicheng Du, Chen Xiong, Qidong Shao, Zijun Lin, Yunqi Hu, Jingjing Zhou, Lian Zhang, Peter E. Lobie, Peiwu Qin, Chengming Yang
arXiv AI
Jun 16

EyeMVP: OCT-Informed Fundus Representation Learning via Paired CFP--OCT Pretraining

arXiv:2606. 15129v1 Announce Type: cross Abstract: Color fundus photography (CFP) is the mainstay for large-scale retinal screening, yet its diagnostic capacity is constrained by the lack of depth-resolved structural information.

By Zhuo Deng, Ruiheng Zhang, Ziheng Zhang, Weihao Gao, Yitong Li, Qian Wang, Lei Shao, Jiaoyue Dong, Zhixi Zeng, Lijian Fang, Haibo Wang, Xiaobin Lin, Tao Liu, Zhicheng Du, Zhengwei Zhang, Lin Yang, Zheng Gong, Xinyu Zhao, Zhenquan Wu, Fang Li, Zhiguang Zhou, Guoming Zhang, Sun Jing, Han Lv, Wenbin We, Lan Ma
arXiv AI
Aug 3

DualDiT: A Conditional Dual-Output Diffusion Transformer for Joint OCT Image and Segmentation Mask Generation

arXiv:2607. 29337v1 Announce Type: cross Abstract: Background and Objective: Generating realistic medical images with anatomically accurate segmentation masks helps address the shortage of annotated data in medical imaging, particularly in optical coherence tomography (OCT) of mouse eyes, where manual retinal layer delineation is labour-intensive due to tiny structures and required expertise, resulting in scarce datasets.

By Fernando Garc\'ia-Torres, Roc\'io del Amor, Sandra Morales, \'Alvaro Barroso, Peter Heiduschka, Bj\"orn Kemper, Valery Naranjo
arXiv Computer Vision
Sep 3

Evaluating Fundus-Specific Foundation Models for Diabetic Macular Edema Detection

The study evaluates fundus-specific foundation models (FM) for detecting diabetic macular edema (DME) in retinal images. It compares two popular FM—RETFound and FLAIR—against a lightweight EfficientNet-B0 backbone across multiple datasets (IDRiD, MESSIDOR-2, and OCT-and-Eye-FundusImages). Results indicate that FM do not consistently outperform fine‑tuned CNNs; EfficientNet-B0 often matches or exceeds FM performance, with FLAIR being the most competitive FM.

By Franco Javier Arellano, Jos\'e Ignacio Orlando
arXiv AI
Sep 1

Co-Annotator: Expert-Distilled ViT and VLM for Visual and Documentation Guidance in Age-Related Macular Degeneration

Co-Annotator is a clinical AI system that distills expert gaze and dictation into two guidance components: a gaze‑aligned Vision Transformer that highlights fixation‑aligned areas of interest (AOIs) and an ontology‑bounded vision‑language model that pre‑fills editable biomarker summaries for retinal OCT. In controlled studies, each modality independently improved diagnostic accuracy and biomarker generation, and when combined across two academic institutions, the system increased correct diagnoses per minute by 40% and reduced comment editing time by 67% without compromising accuracy.

By Ziheng "Leo" Li, Benjamin Freeman, Akshay Raman, Kavin Aravindhan Rajkumar, Xinxin Fang, Rishabh Srivastava, Steven Feiner, Kaveri A. Thakoor