arXiv Computer Vision

Evaluating the Safety of Deep Learning-Based Brain MRI Reconstruction

arXiv AI
3d ago

Federated Learning for MRI-based BrainAGE: a multicenter study on post-stroke functional outcome prediction

arXiv:2506.15626v3 Announce Type: replace-cross Abstract: $\textbf{Objective:}$ Brain-predicted age difference (BrainAGE) is a neuroimaging biomarker reflecting brain health. However, training robust...

By Vincent Roca, Marc Tommasi, Paul Andrey, Aur\'elien Bellet, Markus D. Schirmer, Hilde Henon, Laurent Puy, Julien Ramon, Gr\'egory Kuchcinski, Martin Bretzner, Renaud Lopes
arXiv Machine Learning
Aug 10

Recovering Lesion Parameters from Aphasic Picture Naming Error Profiles in Large Language Models

arXiv:2608. 06429v1 Announce Type: cross Abstract: Interpretability methods for large language models (LLMs) describe internal state but do not directly test whether that state is causally sufficient to produce the observed behavior.

By Yong Yang, Roger Newman-Norlund, Xiang Guan, Saeed Ahmadi, Regan Willis, Nadra Salman, Kalil Warren, Sophie Arheix-Parras, Srihari Nelakuditi, Leonardo Bonilha, Christopher Rorden, Rutvik H. Desai, Julius Fridriksson
arXiv Computer Vision
Aug 27

Deep Learning Segmentation of Diffusion-Weighted MRI Acute Ischaemic Stroke: A Pragmatic Evaluation Across Three Datasets

The study evaluated a pragmatic deep‑learning approach for segmenting acute ischemic stroke lesions on diffusion‑weighted MRI. Using a self‑configured nnU‑Net trained on 1,744 cases and tested on 436, the baseline model achieved a median Dice similarity coefficient of 0.84, outperforming the DeepISLES ensemble, especially for smaller infarcts. The approach required minimal preprocessing and fast inference, suggesting it could streamline clinical stroke imaging workflows.

By Atle Bj{\o}rnerud, Till Schellhorn, Thor H. Skatt{\o}r, Terje Nome, Jon Andr\'e Ottesen, Anne Hege Aamodt, Bradley J MacIntosh
arXiv Computer Vision
1d ago

The Diagnosis a Reporter Leaves Unspoken: Surfacing Frozen Tumor Features for Brain-Tumor MRI Reporting

The paper introduces NeuroFusion, an assistive brain‑MRI report generator that surfaces latent tumor signals from a frozen Mistral‑7B backbone. By adding discriminative field‑classifier heads over per‑lesion features, NeuroFusion restores accurate diagnoses (meningioma 0.92, metastasis 0.75) and improves prose quality while reducing latency 5–6×. A controlled negative result shows that overriding the decoder with a learned diagnosis pin harms performance, and grammar‑constrained decoding yields high schema‑validity (92.3%).

By Khawaja Murad ul Hassan, Ruqiyya Adil, Adil Qayyum, Rida Hassan, Asad Mansoor Khan, Muhammad Usman Akram, Mehran Ebrahimi
arXiv AI
Jun 9

Automatic Extraction of Structured Information from Brain MRI Reports Using an Open-Weight Large Language Model

arXiv:2606. 07721v1 Announce Type: new Abstract: Objectives: Automatic data extraction from free-text radiology reports enables large-scale research, but few studies assessed the performance of large language models (LLMs) on Dutch neuroradiology reports.

By Kaouther Mouheb, Amos Pomp, Antoine Manenti, Romy de Haan, Farog Faghir, Joy Martens, Harro Seelaar, Francesco Mattace-Raso, Meike W. Vernooij, Frank J. Wolters, Stefan Klein, Esther E. Bron
arXiv Machine Learning
Jun 25

Improving Factuality of 3D Brain MRI Report Generation with Paired Image-domain Retrieval and Text-domain Augmentation

arXiv:2411. 15490v2 Announce Type: replace-cross Abstract: Acute ischemic stroke (AIS) requires time-critical decision-making, where inaccurate interpretation of neuroimaging findings can lead to irreversible disability.

By Junhyeok Lee, Yujin Oh, Dahyoun Lee, Hyon Keun Joh, Chul-Ho Sohn, Sung Hyun Baik, Cheol Kyu Jung, Jung Hyun Park, Kyu Sung Choi, Byung-Hoon Kim, Jong Chul Ye
arXiv AI
Jun 2

CardioLens: Revealing the Clinical Reality Gap of MLLMs via Multi-Sequence Cardiac MRI Evaluations

arXiv:2606. 00123v1 Announce Type: cross Abstract: Multimodal Large Language Models (MLLMs) have shown strong performance on public medical benchmarks, yet existing evaluations often remain weak proxies for clinical use, relying on isolated inputs and simplified recognition-style tasks.

By Zixian Su, Hongkai Zhang, Fan Gao, Encheng Su, Taiping Qu, Jingwei Guo, Nan Zhang, Hui Wang, Zhen Zhou, Kairui Bo, Yan Chen, Yue Ren, Shuai Li, Lei Xu, Henggui Zhang