arXiv:2608. 12717v1 Announce Type: new Abstract: Mechanistic interpretability of large language models lacks spatially resolved, falsifiable tools for testing whether internal components are specialized for distinct cognitive operations.
By Xiang Guan, Roger D. Newman-Norlund, Yong Yang, Saeed Ahmadi, Regan Willis, Nadra Salman, Kalil Warren, Srihari Nelakuditi, Chris Rorden, Leonardo Bonilha, Julius Fridriksson
arXiv:2607. 11621v1 Announce Type: new Abstract: Aphasia following stroke commonly produces systematic naming errors with characteristic profiles, but whether general-purpose language models not designed for clinical simulation can reproduce these patterns remains untested.
By Yong Yang, Xiang Guan, Sophie Arheix-Parras, Saeed Ahmadi, Roger Newman-Norlund, Leonardo Bonilha, Christopher Rorden, Julius Fridriksson, Rutvik H. Desai, Srihari Nelakuditi
arXiv:2608.28714v1 Announce Type: cross
Abstract: Objective: Deep learning accelerates brain MRI four- to tenfold, but models can erase lesions or synthesize false tissue - failures pixel-averaged me...
By Dat Tat Mai, Thai Viet Pham, Thu Nguyen Thi Dang, James Jin Kang
arXiv:2603.28387v3 Announce Type: replace-cross
Abstract: Trustworthy clinical AI must use real evidence and avoid relying on surface-level artifacts. We evaluate 12 open-weight vision-language model...
By Doan Nam Long Vu, Simone Balloccu
The study evaluates whether disease can be identified from reactive, non‑lesional brain tissue in intracranial biopsies. Using four foundation‑model encoders within an attention‑based multiple‑instance learning framework on 245 whole‑slide images, the authors find that disease labels remain predictive even after controlling for slide size and sampling bias, and that performance is similar across all encoders. Signed instance‑contribution maps and expert review confirm that predictive signals localize to reactive parenchyma rather than artifacts such as blood.
"whyItMatters":"The findings demonstrate that weakly supervised models can recover disease signals from tissue traditionally considered non‑diagnostic, highlighting the need for provenance‑only baselines in computational pathology benchmarks."
By Jan Schnorrenberg, Jan Ernsting, Enrico K\"ullenberg, Tim Hahn, Benjamin Risse, Christian Thomas
arXiv:2609.37387v1 Announce Type: cross
Abstract: Enlarged perivascular spaces (PVS) visible in brain magnetic resonance imaging (MRI) are increasingly thought to be linked to poor brain health. PVS...
By Jesse Phitidis, William N. Whiteley, Joanna M. Wardlaw, Miguel O. Bernabeu, Yajun Cheng, Xiaodi Liu, Junfang Zhang, Una Clancy, Stephen Makin, Roberto Duarte Coello, Susana Mu\~noz Maniega, Mark E. Bastin, Simon R. Cox, Maria del C. Vald\'es Hern\'andez
The paper investigates how medical vision‑language models (VLMs) behave when faced with distribution shifts such as changes in acquisition domain, supervision, or evaluation protocol. Using datasets like NIH ChestXray14, CheXpert, PadChest, and OpenI, the authors isolate cross‑dataset visual transfer, evaluate multimodal alignment, and quantify source‑proxy leakage in frozen embeddings. They find that self‑supervised visual initialization improves transfer, adversarial adaptation is only marginally helpful, and that multimodal retrieval performance drops under external stress tests while source‑proxy information remains recoverable, highlighting hidden failure modes in medical VLMs.
By Ayoub Louaye Bouaziz, Lokmane Chebouba, Yassine Himeur
arXiv:2504. 06299v2 Announce Type: replace-cross Abstract: Multimodal prediction models based on imaging and clinical data are increasingly used for clinical decision support, yet their interpretability remains limited.
By Lisa Herzog, Jonas Br\"andli, Maurice Schneeberger, Loran Avci, Nordin Dari, Martin H\"ansel, Hakim Baazaoui, Pascal B\"uhler, Susanne Wegener, Beate Sick
arXiv:2603. 28387v2 Announce Type: replace Abstract: Trustworthy clinical AI requires that performance gains reflect genuine evidence integration rather than surface-level artifacts.
By Doan Nam Long Vu, Simone Balloccu
arXiv:2609.22271v1 Announce Type: new
Abstract: Multimodal stroke recurrence prediction requires effective integration of heterogeneous clinical and imaging data, yet modality imbalance often causes...
By Christian Gapp, Elias Tappeiner, Martin Welk, Karl Fritscher, Stephanie Mangesius, Constantin Eisenschink, Philipp Deisl, Michael Knoflach, Astrid E. Grams, Elke R. Gizewski, Rainer Schubert
arXiv:2606.28798v2 Announce Type: replace
Abstract: The main objective of this paper is to propose a general framework for prediction based on different sources of multimodal data in the healthcare d...
By Chengyuan Liu, Xinyue Zhang, Yao Li, Guanting Chen
arXiv:2608. 16507v1 Announce Type: new Abstract: Due to the limited amount of information, modeling longitudinal rare-disease data can benefit from integrating clinical knowledge.
By Clemens Sch\"achter, Astrid Pechmann, Janbernd Kirschner, Jan Hasenauer, Harald Binder