arXiv AI

Benchmarking Machine Learning Models for Multi-Omics-Based Breast Cancer Prediction

arXiv:2607. 16250v1 Announce Type: cross Abstract: Estrogen Receptor (ER) status is a critical biomarker in breast cancer diagnosis, prognosis, and treatment selection.

arXiv Machine Learning
Sep 24

Benchmarking Active Spot Selection for Cost-Efficient Spatial Transcriptomics

The study benchmarks active spot selection methods against random sampling for spatial transcriptomics, focusing on cost‑efficient data acquisition. Using two public cohorts, the authors simulate multi‑round selection with uncertainty‑based (MC‑dropout, TOD) and diversity‑based (CoreSet, TypiClust) strategies, evaluating performance at 5%, 10%, 30%, and 50% of the spot pool. Results show that none of the active strategies consistently outperforms random sampling across all budgets or evaluation metrics, with performance varying by dataset and metric.

By Zheyu Zhu, Junchao Zhu, Fengbei Liu, Tianyuan Yao, Gelei Xu, John Cannon, Haichun Yang, Yuankai Huo, Mert R. Sabuncu, Ruining Deng
arXiv Machine Learning
Aug 11

TRAPS: Treatment-Assignment Prediction via Pathway-informed Stratification

arXiv:2606. 09898v2 Announce Type: replace Abstract: Cancer treatment involves decisions across multiple clinical outcomes, yet pathway-informed deep learning models are typically evaluated in isolation, making their relative benefits unclear.

By Sujoy Banik, Sayantan Chakraborty, Boishakhi Das Toma, Zainab Ghafoor, Ushashi Bhattacharjee, Koushik Howlader, Tirtho Roy
arXiv AI
Sep 10

The Accuracy Paradox: Empirical Diagnostic of Default Decision Thresholds in Multi-Label Enzyme Commission Prediction [With Code]

The study evaluates the use of default decision thresholds (t=0.50) in multi‑label enzyme commission (EC) number prediction across 14,096 compounds and six EC classes. It finds a high mean accuracy of 77.16% but low macro F1 (0.3976) and macro recall (0.3872), indicating severe class‑imbalance issues: majority classes are over‑predicted while minority classes, especially EC6, have zero recall despite reasonable ROC‑AUC. The authors recommend target‑specific threshold tuning and conformal calibration as post‑processing safeguards to expose and correct these hidden errors.

By Bilal Ahmad, Rajed Mehmood