arXiv AI

An integrated diffusion-weighted imaging processing and interpretation platform for MR-guided radiotherapy

An integrated, web‑based platform has been developed to process diffusion‑weighted imaging (DWI) from MR‑guided linear accelerators and provide structured, literature‑grounded clinical interpretations. The system combines a deep‑learning pipeline for distortion correction, denoising, and IVIM/ADC fitting with a retrieval‑augmented generation (RAG) agent that references a curated knowledge base and traces each statement to its source. Independent expert ratings of nine glioblastoma cases showed high scores for clinical reasoning, citation quality, and overall utility, with a mean rating of 4.65 out of 5.

arXiv Computation and Language
1d ago

Comparison of techniques for fine-tuning open-weight models for entity extraction from radiology reports

The study evaluates whether a fine‑tuned open‑weight model (Gemma‑3‑12B) can match the performance of GPT‑4o in extracting multi‑label intracranial hemorrhage acuity from non‑contrast head‑CT reports. Using a 2×2 design that varied adaptation strategy (classification head vs. instruction fine‑tuning) and training‑data source (distilled real GPT‑4o labels vs. synthetic GPT‑4o‑generated reports), the distilled instruction‑tuned model achieved macro‑F1 scores comparable to GPT‑4o and surpassed the untuned base model. The key finding is that the source of training data—distilled real reports—was more important than the fine‑tuning method, and that the entire fine‑tuning and inference process fits on a single 24 GB consumer GPU.

By Aawez Mansuri, Kush Mehta, Mohammadreza Chavoshi, Jahanzaib Malik, Theodorus Dapamede, Frank Li, Rohan Isaac, Beatrice Brown-Mulry, Chiratidzo Rudado Sanyika, YoungSeok Jeon, Judy W. Gichoya, Ali Emami, Hari Trivedi
arXiv Machine Learning
Sep 10

A Specialized Large Multimodal Model for Interpreting PET/CT in Head and Neck Cancer

A specialized large multimodal model, LLaVA‑NeXT, was fine‑tuned on a curated two‑level curriculum of PET/CT image‑conversation pairs to interpret head and neck cancer scans. In external validation across four institutions, the model achieved high ROUGE and similarity scores and outperformed generalist models such as ChatGPT, with primary tumor classification accuracy of 83.14% internally and 69.03% externally. The study demonstrates that domain‑specific LMMs can provide fast, accurate diagnostic support for PET/CT imaging.

By Haengbok Chung, SunGyu Kim, Joo hyun Lee, Sangjin Bae, Min Jeong Cho, Minseok Suh, Jae Sung Lee
arXiv AI
Jun 9

Automatic Extraction of Structured Information from Brain MRI Reports Using an Open-Weight Large Language Model

arXiv:2606. 07721v1 Announce Type: new Abstract: Objectives: Automatic data extraction from free-text radiology reports enables large-scale research, but few studies assessed the performance of large language models (LLMs) on Dutch neuroradiology reports.

By Kaouther Mouheb, Amos Pomp, Antoine Manenti, Romy de Haan, Farog Faghir, Joy Martens, Harro Seelaar, Francesco Mattace-Raso, Meike W. Vernooij, Frank J. Wolters, Stefan Klein, Esther E. Bron
arXiv Computer Vision
Sep 3

The Diagnosis a Reporter Leaves Unspoken: Surfacing Frozen Tumor Features for Brain-Tumor MRI Reporting

The paper introduces NeuroFusion, an assistive brain‑MRI report generator that surfaces latent tumor signals from a frozen Mistral‑7B backbone. By adding discriminative field‑classifier heads over per‑lesion features, NeuroFusion restores accurate diagnoses (meningioma 0.92, metastasis 0.75) and improves prose quality while reducing latency 5–6×. A controlled negative result shows that overriding the decoder with a learned diagnosis pin harms performance, and grammar‑constrained decoding yields high schema‑validity (92.3%).

By Khawaja Murad ul Hassan, Ruqiyya Adil, Adil Qayyum, Rida Hassan, Asad Mansoor Khan, Muhammad Usman Akram, Mehran Ebrahimi
arXiv Computer Vision
Sep 16

Anatomy-Change-Aware Bidirectional Selective State-Space Memory for Clinically Deployed Thoracic Radiotherapy Auto-Contouring

The paper introduces DAMM‑Net++, a 2.5D neural network for thoracic organ‑at‑risk and target volume segmentation that tackles inter‑slice surface incoherence, small low‑contrast target failure, and lack of per‑case reliability signals. Its core is an anatomy‑change‑aware bidirectional selective state‑space memory that propagates context across axial slices, complemented by a boundary‑aware decoder and an uncertainty head for calibrated per‑voxel confidence. Evaluations on 2,146 patients, an external cohort, and a reader study show high Dice scores (0.955), low HD95 (3.78 mm), significant time savings (75‑80 %) for clinicians, and improved junior‑reader performance, with the system fully integrated into a clinical workflow.

By Galib Ahmed, Istiak Ahmed, Aritra Islam Saswato, Asib Mostakim Fony, Kazi Shahriar Sanjid, Md. Tanzim Hossain, Md. Anwarul Islam, Md. Nishan Khan, Md. Misbah Khan, Labiba Faiza Karim, Jobaer Rahman, S M Hasibul Hoque, Rahnuma Shahrin Rista, Kamruzzaman Rumman, Md Arifur Rahman, Syed Md. Akram Hussain, Mohammad Ashrafuzzaman Khan, M. Monir Uddin
arXiv AI
Jun 9

RadOT-Eval: Auditable Structured-Evidence Transport for Radiology Report Evaluation

arXiv:2606. 08769v1 Announce Type: cross Abstract: Automatic evaluation is critical for high-stakes text generation, where errors often involve omitted findings, hallucinated content, polarity reversals, location changes, uncertainty mismatches, and temporal-comparison errors rather than low surface similarity alone.

By Weixin Liu, Juming Xiong, Yang Li, Qingyuan Song, Susannah Rose, Murat Kantarcioglu, Bradley Malin, Zhijun Yin
arXiv AI
Sep 16

A Vision-Language Foundation Model for Precise and Comprehensive Brain Tumor Diagnosis from Preoperative Multimodal Data

arXiv:2609.16597v1 Announce Type: cross Abstract: Background Non-invasive presurgical diagnosis of brain tumor types from Magnetic Resonance Imaging (MRI) is essential but challenging due to overlapp...

By Yinong Wang (Joyce), Jianwen Chen (Joyce), Zhou Chen (Joyce), Shuwen Kuang (Joyce), Haoning Jiang (Joyce), Yanzhao Shi (Joyce), Huichun Yuan (Joyce), Yan-ran (Joyce), Wang, Bing Wang, Lei Wu, Bin Tang, Li Meng, Baihua Luo, Bin Zhou, Wei Ding, Weiming Zhong, Wei Hou, Yuanbing Chen, Zhiping Wan, Wei Wang, Zhenkun Xiao, Wenwu Wan, Allen He, Yuyin Zhou, Longbo Zhang, Feifei Wang, Zhixiong Liu, Michael Iv, Xuan Gong, Liangqiong Qu
arXiv AI
Jul 15

BAT-RM: A Boundary-Aware Transformer with Region-Aware Multi-Directional Mamba for Clinically Deployed Cervical Cancer Radiotherapy Auto-Contouring

arXiv:2607. 11949v1 Announce Type: cross Abstract: We present a clinically deployed end-to-end auto-contouring system for cervical cancer radiotherapy planning, anchored by the Boundary-Aware Transformer with Region-Aware Mamba (BAT-RM), a hybrid architecture that integrates Sobel-gated boundary attention, a linear-time, multi-directional Mamba module for long-range context, and a boundary-skeleton-guided fusion gate.

By Istiak Ahmed, Kazi Shahriar Sanjid, Galib Ahmed, Md. Tanzim Hossain, Md. Anwarul Islam, Shahrukh Khan, Md. Ashrif Rahman Arian, Md. Nishan Khan, Md. Misbah Khan, S M Hasibul Hoque, Rahnuma Shahrin Rista, Md. Jobairul Islam, Sheikh Anisul Haque, Md Arifur Rahman, Syed Md. Akram Hussain, Syeda Nashra, Sayeed Shafayet Chowdhury, Md. Mostafa Kamal Sarker, M. Monir Uddin
arXiv Computation and Language
Sep 23

MultiViewDx: Evidence-Linked Multi-View Clinical Diagnosis

MultiViewDx is a physician‑validated multimodal instruction dataset that links medical imaging studies with patient context and normalizes heterogeneous reports into an evidence‑linked workflow (evidence → findings → differential discussion → diagnosis). The dataset covers a wide range of imaging modalities and uses a unified image‑text retriever to ensure that instruction synthesis is grounded in source‑supported evidence. Fine‑tuned models on MultiViewDx achieve the highest average accuracy on four MedVQA benchmarks and receive the strongest overall rating on JAMA Clinical Challenge cases, with ablations confirming the importance of case‑level multi‑view organization and evidence‑linked reasoning.

By Junda Wang, Zonghai Yao, Yujan Ting, Eric Z. Chen, Hieu Tran, Hong Yu, Weijing Huang, Terrence Chen
arXiv AI
Jul 8

Harrison.Rad 1.5 Technical Report: A radiology foundation model that can draft reports from images, priors and clinical context

arXiv:2607. 05880v1 Announce Type: cross Abstract: Imaging demand is growing faster than the radiology workforce can expand, and reporting backlogs cannot be resolved through training and recruitment alone.

By Suneeta Mall, Vladimir Nekrasov, Ashnil Kumar, Sajith Karunasena, Aiden Nibali, Alix Bird, Mateo Diaz Shine, Jarrel Seah