← Back to all news
arXiv Computer Vision September 17, 2026 By Guoliang You, Haifan Gong, Xiaomeng Chu

Semantically Calibrated Evidence Composition for CT Vision-Language Learning

Read the original on arXiv Computer Vision →

The Flow has not summarised this story yet — read it at arXiv Computer Vision.

  • multimodal
  • benchmarks
  • safety

One email a morning, machine-written

One email a day, machine-written, one click to leave. We never share your address.

Related stories

arXiv AI
Jul 31

Anatomy Contextualized Adaption of CT Foundation Models

arXiv:2607. 27154v1 Announce Type: cross Abstract: CT vision-language foundation models have demonstrated promising performance across downstream tasks, but are typically trained with whole-volume representations that dilute fine-grained anatomical signals.

By Roshan Kenia, Stephanie L McNamara, William Lotter
llmsragmultimodalsafety
More like this →
arXiv AI
Aug 21

Anatomy Contextualized Adaptation of CT Foundation Models

arXiv:2607. 27154v2 Announce Type: replace-cross Abstract: CT vision-language foundation models have demonstrated promising performance across downstream tasks, but are typically trained with whole-volume representations that dilute fine-grained anatomical signals.

By Roshan Kenia, Stephanie L McNamara, William Lotter
llmsragmultimodalsafety
More like this →
arXiv Computer Vision
Sep 1

CheXGround: Anatomical Region Tokens for Grounded Longitudinal Chest X-ray Interpretation

arXiv:2608.30758v1 Announce Type: new Abstract: Recent radiology multi-modal language models have made substantial progress in chest X-ray report generation, visual question answering, and temporal r...

By Adonay Demewez Gebremedhin, Wessam Shehieb, Sara Alansari, Mohamad Alansari, Muzammal Naseer, Sajid Javed, Naoufel Werghi
llmsnlpmultimodalsafety
More like this →
arXiv Computer Vision
1d ago

Learning How Much, Not Just What: Cross-Patient Burden Order for CT Vision-Language Pretraining

arXiv:2608.00231v2 Announce Type: replace Abstract: Volumetric CT vision-language pretraining learns 3D representations from scan-report pairs, but global and anatomy-aware objectives supervise only...

By Guoliang You, Haifan Gong, Xiaomeng Chu
multimodalsafety
More like this →
arXiv Computer Vision
Aug 31

ARC-CT: Anatomy-Routed Contrastive Vision-Language Learning for 3D Chest CT

arXiv:2608.28455v1 Announce Type: new Abstract: Contrastive vision-language learning uses paired chest CT volumes and radiology reports to learn abnormality classifiers without manually annotated lab...

By Huseyin Umut Isik, Mehmet Alp Ozaydin, Sila Kurugol, \c{S}eyda Ertekin
llmsragmultimodalsafety
More like this →
arXiv AI
Jul 29

OrganLens: Organ-Specific Representation Learning for CT Foundation Models

arXiv:2607. 25164v1 Announce Type: cross Abstract: A CT examination captures multiple organs, but many biomedical questions concern abnormalities, prognosis, or longitudinal change in a specific organ.

By Zhixuan Ge, Anqi Li, Sadeer Al-Kindi, Hanwen Xu, Wei Qiu
diffusioncomputer-visionefficiency
More like this →
About Pricing API Newsletter Sources Privacy Terms Refunds Accessibility Provider info Contact RSS

The Flow links to publishers and never republishes their articles. Summaries are machine-generated.

v1.1.0 · 5f852ea