arXiv AI

EyeMVP: OCT-Informed Fundus Representation Learning via Paired CFP--OCT Pretraining

arXiv:2606. 15129v1 Announce Type: cross Abstract: Color fundus photography (CFP) is the mainstay for large-scale retinal screening, yet its diagnostic capacity is constrained by the lack of depth-resolved structural information.

arXiv AI
Sep 1

Co-Annotator: Expert-Distilled ViT and VLM for Visual and Documentation Guidance in Age-Related Macular Degeneration

Co-Annotator is a clinical AI system that distills expert gaze and dictation into two guidance components: a gaze‑aligned Vision Transformer that highlights fixation‑aligned areas of interest (AOIs) and an ontology‑bounded vision‑language model that pre‑fills editable biomarker summaries for retinal OCT. In controlled studies, each modality independently improved diagnostic accuracy and biomarker generation, and when combined across two academic institutions, the system increased correct diagnoses per minute by 40% and reduced comment editing time by 67% without compromising accuracy.

By Ziheng "Leo" Li, Benjamin Freeman, Akshay Raman, Kavin Aravindhan Rajkumar, Xinxin Fang, Rishabh Srivastava, Steven Feiner, Kaveri A. Thakoor
arXiv Machine Learning
Jul 24

Counterfactual Explainability Framework With CycleGAN And Counterfactual-Classifier Alignnment Score for Retinal Disease Classification

arXiv:2607. 21068v1 Announce Type: new Abstract: Automated detection of vision impairing retina-based ocular conditions from fundus images is important for early screening, timely referral and reducing dependency on specialist-only assessment, for which neural network-based deep learning (DL) models have been widely utilized.

By Kritanu Chattopadhyay, Sayanjit Singha Roy, Soumya Chatterjee
arXiv AI
Sep 10

Clinician-Friendly Foundation Models for Ophthalmic Image Diagnostics without Fine-Tuning or Technical Barriers

The paper introduces GlobeReady, a clinician-friendly platform that leverages the RetiGlobe foundation model for ophthalmic image diagnostics without requiring fine-tuning. RetiGlobe was pretrained in two stages: first with self-supervised learning on 38 million synthetic images, then with contrastive learning on 475,845 real image‑text pairs from diverse ethnicities, devices, and regions. GlobeReady was evaluated on 488,448 images from multiple international centers and tested prospectively with 31 ophthalmologists, also exploring domain generalisability, uncertainty quantification, OOD detection, and feature-based case retrieval.

By Meng Wang, Tian Lin, Qingshan Hou, Aidi Lin, Lianyu Wang, Jingcheng Wang, Qingsheng Peng, Truong X. Nguyen, Zhi Da Soh, Xiayin Zhang, Jingyan Yang, Danqi Fang, Ke Zou, Ting Xu, Can Can Xue, Ten Cheer Quek, Qinkai Yu, Minxin Liu, Hui Zhou, Zixuan Xiao, Guiqin He, Huiyu Liang, Tingkun Shi, Man Chen, Zhuangling Lin, Linna Liu, Yuanyuan Peng, Li Jia Chen, Chi Ming Chan, Xiaohong Li, Junren He, Zhirong Xu, Tingbing Fang, Yanli Wang, Qingzhi Wang, Wenyi Hu, Yujie Wang, Li Li, Jiaying Ye, Tonghui Ye, Liang Lyu, Yongjian Lu, Ruoshi Cai, Yiwen Tang, Qiuming Hu, Junhong Chen, Zhenhua Zhang, Cheng Chen, Yitian Zhao, Dianbo Liu, Jianhua Wu, Xinjian Chen, Changqing Zhang, Xiaojun Wu, Triet Thanh Nguyen, Yanda Meng, Yalin Zheng, Daoqiang Zhang, Xiaochun Cao, Yih Chung Tham, Ye Zhang, Ying Han, Alvin L Young, Mary Ho, Carmen K M Chan, Clement C Tham, Zhuoting Zhu, Carol Y. Cheung, Tien Yin Wong, Huazhu Fu, Haoyu Chen, Ching-Yu Cheng
arXiv AI
Jul 17

Parameter-efficient Prompt Tuning of Vision Foundation Model With Adaptive Focal Loss for Interpretable MCI Screening

arXiv:2607. 15047v1 Announce Type: cross Abstract: Mild Cognitive Impairment is a critical early stage of cognitive decline that frequently precedes Alzheimer's disease, yet its automated detection from neuropsychological drawing tests remains fundamentally constrained by data scarcity, class imbalance, and diagnostic ambiguity near clinical boundaries.

By Javad Khoramdel, Farhad Hoseyni, Amirhossein Nikoofard
arXiv Machine Learning
Aug 11

Disentangling Co-Occurring Retinal Pathologies with Saliency-Guided Sparse Expert Routing

arXiv:2608. 09752v1 Announce Type: cross Abstract: Retinal fundus images frequently exhibit multiple co-occurring pathologies, yet standard deep learning classifiers apply static, identical computation to every image regardless of the underlying disease distribution.

By Nagur Shareef Shaik, Jeongwoo Park, Yeong-Jin Kim, Jaeuk Jung, Hyunjung Oh, Dong Hye Ye