Hugging Face Trending Papers

FCA-Guided Counterfactual Explanations for Multi-Modal Breast Cancer Diagnosis: A Framework Achieving Perfect Validity with Emergent Sparsity

The paper introduces FCA‑Guided Counterfactual (FCA‑CF) explanations for multi‑modal breast cancer diagnosis, leveraging a Formal Concept Analysis lattice as a hard structural constraint to generate counterfactuals. On the TCGA‑BRCA dataset, FCA‑CF achieves perfect validity (100% prediction flips), the lowest average feature changes (2.37), and competitive proximity (0.900), outperforming four established counterfactual methods. Ablation studies show the lattice constraint and a greedy refinement phase are key to its sparsity and validity.

arXiv AI
Sep 18

FCA-Guided Counterfactual Explanations for Multi-Modal Breast Cancer Diagnosis: A Framework Achieving Perfect Validity with Emergent Sparsity

The paper introduces FCA‑Guided Counterfactual (FCA‑CF) explanations for multi‑modal breast cancer diagnosis, leveraging a Formal Concept Analysis lattice as a hard structural constraint to search for counterfactuals. On the TCGA‑BRCA dataset, FCA‑CF achieves perfect validity (100% prediction flips), the lowest average feature changes (2.37), and competitive proximity (0.900) compared to four other methods. Ablation studies show the lattice constraint drives sparsity, while a greedy refinement phase further improves results.

By Abdullahi Isa, Souley Boukari, Muhammad Aliyu
arXiv Machine Learning
Jul 9

Counterfactual Modeling with Fine-Tuned LLMs for Health Intervention Design and Sensor Data Augmentation

arXiv:2601. 14590v3 Announce Type: replace Abstract: Counterfactual explanations (CFEs) provide human-centric interpretability by identifying the minimal, actionable changes required to alter a machine learning model's prediction.

By Shovito Barua Soumma, Asiful Arefeen, Stephanie M. Carpenter, Melanie Hingle, Hassan Ghasemzadeh
arXiv AI
Aug 5

CorePath: A Breast-Specialized Pathology Foundation Model for Core Needle Biopsy Diagnosis and Risk-Controlled Report Generation

arXiv:2608. 03079v1 Announce Type: cross Abstract: Breast core needle biopsy (CNB) is central to breast cancer diagnosis yet remains challenging because limited tissue sampling, lesion heterogeneity, and subtle morphologic overlap can obscure subtype distinctions.

By Ting Yin, Danning Li, Chen Shu, Xiaoxia Yao, Boyu Fu, Yujing Chang, Tianyu Shi, Mengna Feng, Jie Chen, Jing Fu, Xiuli Xiao, Tianlin Li, Mumin Shao, Jiaxin Bi, Wenchuan Zhang, Xiaoyan Wu, Xiao Han, Zhang Zhang, Yuhao Yi, Hong Bu
arXiv AI
Sep 15

Causal multi-modal AI for personalized chemosensitivity prediction

A causal multi-modal AI model was developed to predict personalized chemosensitivity in breast cancer patients using routine pathology and clinical data. Trained on 9,141 patients from nine countries and validated on 1,994 patients from three countries, the model produced treatment-specific recurrence probabilities with near-perfect calibration and strong prognostic discrimination over 5- and 10-year horizons. It outperformed existing recurrence-score tests and could reduce chemotherapy prescriptions by 30% while maintaining recurrence-free rates, with predictive performance also transferring to non-breast cancers.

By Dhruva Biswas, Jeroen Berrevoets, Alec McClean, Linus Bao, Jungkyu Park, Ken G. Zeng, Joseph Cappadona, Cerise Tang, Chuwen Liu, Bartosz Machura, Yin Wu, Valerie Speirs, Hatem Soliman, Rohit Bhargava, Sheheryar Kabraji, Thaer Khoury, David Page, Brian Piening, Carlo Bifulco, Claudia Meurs, Pieter Westenend, Sylvie Chabaud, Jerome Lemonnier, Paul H. Cottu, Florence Dalenc, Fabrice Andre, Frederique Madeleine Penault-Llorca, Thomas Bachelot, Frederick Howard, Francisco J. Esteva, Kevin Kalinsky, Lajos Pusztai, Jan Witowski, Krzysztof J. Geras
arXiv Machine Learning
1d ago

Fixing a Model That Learned Worse Cancer Means Lower Risk: Monotonic Constraints in Bladder Cancer Recurrence Prediction

In a UK multicentre trial, an unconstrained XGBoost model incorrectly learned that higher tumour stage and carcinoma in situ predicted lower bladder cancer recurrence risk, a finding that conventional metrics such as discrimination, calibration, and SHAP failed to detect. The authors introduced a counterfactual direction test and a monotonic‑constraint framework, which removed the inversion without harming model performance and even outperformed established risk systems. The study demonstrates that such tests should be routine before deploying predictive models in clinical settings.

By Saram Abbas, David Thomas, Naeem Soomro, Rishad Shafik, Rakesh Heer, Kabita Adhikari
arXiv AI
Jul 7

CaresAI at SMM4H-HeaRD 2026: Predicting TNM Staging

arXiv:2607. 03466v1 Announce Type: cross Abstract: This study aims to predict Tumor, Node, and Metastasis (TNM) stage labels independently, with the Cancer Genome Atlas (TCGA) pathology report as the sixth shared task of SMM4H-HeaRD 2026.

By Joseph Itopa Abubakar, Jorge Jarme, Favour Igwezeke, Mary Adewunmi
arXiv Machine Learning
Aug 19

Pathology Transport: Optimal-Transport Explanations for Clinical Data, and When Their Heatmaps (Fail to) Localize Disease

The paper presents an optimal‑transport based generative model that learns the distributional differences between healthy and diseased patients, producing per‑patient counterfactuals and label‑free attribution heatmaps. On tabular breast cancer data the model achieves high malignancy scoring (AUROC ≈ 0.91) and its attributions correlate moderately with a supervised classifier, yet it does not surpass logistic regression. In chest X‑ray experiments the transport heatmaps capture population‑level signals but fail to localize real lesions, revealing a synthetic‑to‑real gap that challenges the reliability of label‑free explanations.

By Lalit Kumar
arXiv Machine Learning
Sep 21

Purification and Regulation: Comorbidity-Aware Multi-Label Few-Shot Learning for Medical Image Classification

The paper introduces Prototype Purification and Regulation (PPR), a multi‑label few‑shot learning framework for medical image classification that addresses two key limitations of existing metric‑based meta‑learning methods. PPR first purifies prototypes by using sample‑level comorbidity scores to highlight disease‑specific features, then regulates inter‑class prototype distances with disease‑level comorbidity statistics to create a comorbidity‑aware embedding space. Experiments on four chest X‑ray datasets, including cross‑domain tests, show that PPR outperforms state‑of‑the‑art methods, improving disease detection and demonstrating robust generalization and clinical applicability.

By Ying-Chih Lin, Po-Chih Kuo, Yong-Sheng Chen
arXiv AI
2d ago

OpenMTB-Audit: Exposing Over-Refusal and Clinical Expert Perspectives in LLM-Based Molecular Tumor Board Safety Evaluation

OpenMTB‑Audit is an open‑source benchmark that tests large language models on 500 synthetic non‑small cell lung cancer cases, covering five adversarial error categories and four safety labels: Supported, Partially Supported, Unsupported, and Insufficient Information. The study found that all eight tested LLMs over‑refused Partially Supported recommendations, collapsing labels to achieve high safety scores. A deterministic seven‑module framework, MTB‑AuditAgent, was introduced to reduce over‑refusal to 6.7% and reach 91.2% accuracy, while an oncologist annotation study highlighted disagreement around the boundary between information sufficiency and treatment optimization.

By Negin Ashrafi, Jia Luo, Stacey M. Frumm, Roxana Daneshjou