arXiv Machine Learning By Ariel Fargion, Lahav Dabah, Tom Tirer

Enhancing Conformal Prediction via Class Similarity

Read the original on arXiv Machine Learning →

arXiv:2511. 19359v2 Announce Type: replace Abstract: Conformal Prediction (CP) has emerged as a powerful statistical framework for high-stakes classification applications.

Machine-generated by The Flow from the publisher's headline and feed description — not written or checked by a human. The full article lives at arXiv Machine Learning.

arXiv Machine Learning
Sep 11

Benchmarking non-conformity score functions in conformal prediction

Conformal prediction replaces single-class predictions with prediction sets that guarantee a pre-specified coverage probability. The paper reviews properties of non‑conformity score functions, presents examples from the literature, and proposes new modifications. It introduces a method to evaluate prediction set sizes and compares different score functions, including their effectiveness for class‑conditional conformal prediction with imbalanced classes.

By Sol Erika Boman
arXiv Machine Learning
Sep 18

Alliance Beats Isolation: Unifying Heterogeneous Allied Datasets Improves Classifier Performance

The paper introduces a method for combining heterogeneous, allied datasets—datasets that share the same class labels but have disjoint objects and largely distinct feature spaces—into a single unified feature space. By applying matrix completion to this merged space, the authors create a unified dataset that enables knowledge transfer between the original datasets. Experiments across multiple dataset pairs and classifiers show that models trained on the unified representation consistently outperform those trained separately on each dataset.

By Girish Keshav Palshikar
arXiv AI
Sep 1

CoLa-ICD: A Knowledge-Enhanced Framework for Long-Tail Automated Medical Coding

CoLa-ICD is a knowledge‑enhanced framework designed to improve automatic medical coding of ICD codes in long, imbalanced clinical documents. It enriches ICD labels with external terms, models dependencies among related codes, and strengthens the alignment between label semantics and clinical evidence, particularly for rare codes. Experiments demonstrate that CoLa-ICD achieves state‑of‑the‑art performance in AUC, F1, and P@k, with larger gains in larger and sparser label spaces.

By Yihang Cheng, Veronica Liesaputra, Andrew Trotman