arXiv AI

Knee-xRAI: An Explainable AI Framework for Automatic Kellgren-Lawrence Grading of Knee Osteoarthritis

arXiv:2604. 23435v2 Announce Type: replace-cross Abstract: Grading knee osteoarthritis (KOA) on plain radiographs is poorly reproducible across readers.

arXiv Machine Learning
Sep 16

RAM-H1200: A Unified Evaluation and Dataset on Hand Radiographs for Rheumatoid Arthritis

RAM‑H1200 is a new public benchmark comprising 1,200 hand radiographs from six medical centers, each annotated with whole‑hand bone instance segmentation, pixel‑level bone erosion masks, joint regions of interest, and joint‑level Sharp‑van der Heijde (SvdH) scores for bone erosion and joint space narrowing. The dataset enables unified multi‑level analysis of anatomical structure, localized erosive pathology, and clinically standardized RA severity, filling a gap in existing resources that lack full‑hand coverage and fine‑grained annotations. Initial benchmark results show strong performance in bone segmentation but highlight that bone erosion segmentation remains a significant challenge, underscoring the dataset’s potential to advance quantitative RA analysis.

By Songxiao Yang, Haolin Wang, Yao Fu, Junmu Peng, Lin Fan, Hongruixuan Chen, Jian Song, Masayuki Ikebe, Shinya Takamaeda-Yamazaki, Masatoshi Okutomi, Tamotsu Kamishima, Yafei Ou
arXiv Computer Vision
Sep 25

Integrating Local Detail and Global Context: A Dual-Input Multi-Task Learning Framework for Bone Tumor Diagnosis

The paper introduces a dual‑input, multi‑task learning framework that jointly segments and classifies bone tumors by applying bidirectional cross‑modal attention between a lesion crop and the full radiograph. Using a YOLO‑based detector and a dual‑stream DenseNet121 architecture, the model fuses fine‑grained lesion detail with global anatomical context through a novel cross‑modal attention fusion strategy and hierarchical multi‑scale feature fusion. On the multi‑institutional Bone Tumor X‑ray Radiograph Dataset, the approach outperforms single‑input baselines, achieving a Dice coefficient of 0.896 and a macro‑averaged F1‑score of 0.928, with an AUC of 0.999 for malignant osteosarcoma.

By S. M. Nasif Uddin, Rusab Sarmun, Muhammad E. H. Chowdhury, Adam Mushtak, Israa Al-Hashimi, Sohaib Bassam Zoghoul
arXiv AI
Jun 6

An interpretable and trustworthy AI framework for large-scale longitudinal structure-pain association studies using data from the Osteoarthritis Initiative (OAI)

arXiv:2606. 05357v1 Announce Type: new Abstract: Purpose: To develop an interpretable and trustworthy AI framework that combines deep learning based MRI Osteoarthritis Knee Score (MOAKS) prediction with interpretable statistical modeling to study structure-pain relationships at scale using data from the Osteoarthritis Initiative (OAI).

By Jincheng Yu, Haoyang Li, Yiwen Liu, Shen Liu, Rachel Yuanbao Chen, C. Kent Kwoh, Hongxu Ding, Xiaoxiao Sun
arXiv Computer Vision
Sep 23

Radiomics--Foundation Fusion for Interpretable RCC Classification: Internal Benchmarking and Exploratory External Transfer

arXiv:2609.26578v1 Announce Type: new Abstract: Accurate preoperative subtype classification of renal cell carcinoma (RCC) from contrast-enhanced CT remains clinically challenging because clear cell...

By Yuan Liang, Fangyijie Wang, Kathleen M. Curran, Gu\'enol\'e Silvestre, Sourav Bhattacharjee, Abraham Campbell
arXiv AI
Sep 7

Cross-modal triage network: a multimodal deep learning framework for severity-based triage and visual explainability in chest radiographs

The paper introduces the Cross‑Modal Triage Network (CMTN), a multimodal deep‑learning model that fuses a Swin Transformer V2 visual encoder with a PubMedBERT text encoder to perform severity‑based triage, pathology detection, and generate visual explanations for chest radiographs. Trained on 34,639 image‑text pairs from MIMIC‑CXR‑JPG, the CMTN achieves high ordinal agreement with reference labels (QWK = 0.9341) and excellent pathology detection (macro‑AUROC = 0.9970) while operating with 34 ms latency. However, a blinded clinical audit revealed low agreement with expert radiologists (QWK = 0.1399) and only modest spatial‑semantic concordance in heatmaps, underscoring the gap between algorithmic performance and clinical judgment.

By Zinah Ghulam, Richa Mittal, Eranga Ukwatta
arXiv AI
Sep 1

Towards Accurate and Lightweight Peripheral Neuroblastic Tumor Diagnosis via Contrastive Multi-scale Pathological Image Analysis

The paper introduces CoPath, a lightweight framework for diagnosing peripheral neuroblastic tumors (pNTs) from whole-slide images. CoPath combines CoHisNet, a multi‑scale feature‑fusion network that replaces traditional MLPs with Kolmogorov‑Arnold Network layers for efficient nonlinear modeling, and PathVote, which aggregates patch‑level predictions using pathology‑informed priors. Experiments on a private pNT cohort and the public BreakHis dataset show that CoPath matches or surpasses existing classifiers while reducing computational complexity.

By Zhu Zhu, Shuo Jiang, Jingyuan Zheng, Yawen Li, Yifei Chen, Manli Zhao, Weizhong Gu, Feiwei Qin, Jinhu Wang, Gang Yu
arXiv AI
Aug 24

Volumetric Radiology AI in the Era of Multimodal Large Language Models

The article reviews how multimodal large language models (MLLMs) are expanding radiology AI beyond image‑specific tasks to multimodal reasoning, yet volumetric radiology poses a representational challenge because clinical interpretation needs full 3‑D spatial context and quantitative data. It surveys over 200 studies, categorizing advances in volumetric representation, multimodal understanding, and agentic orchestration, and introduces a Claim‑Design‑Validation framework to align technical, workflow, and clinical claims. The review emphasizes that native volumetric modeling and agentic capabilities must match spatial, quantitative, contextual, and workflow demands, and that clinical credibility hinges on faithful 3‑D representation, traceable behavior, proper validation, and defined human oversight.

By Zanting Ye, Shengyuan Liu, Xin Liu, Chenhui Wang, Zhisong Wang, Jiashuai Liu, Zipei Wang, Cheng Wang, Wentao Pan, Mengjie Fang, Di Dong, Mohammad Salmanpour, Arman Rahmim, Yu Gu, Yong Xia, Hongming Shan, Yixuan Yuan, Yefeng Zheng, Lijun Lu