arXiv Machine Learning

Does the Data Processing Inequality Reflect Practice? On the Utility of Low-Level Tasks

arXiv:2512. 21315v2 Announce Type: replace Abstract: The data processing inequality is an information-theoretic principle stating that the information content of a signal cannot be increased by processing the observations.

arXiv Machine Learning
Aug 27

Comparing Corrupted Constrained Learning Problems

The paper discusses the data processing inequality (DPI) in statistics, which states that a stochastically modified experiment cannot have a lower Bayes risk than the original. It shows that this classical DPI does not hold for constrained learning problems common in machine learning, where the model class is limited. The authors propose a generalized DPI that applies to constrained Bayes risks, linking it to a set containment condition on a superprediction set, and provide sufficient conditions for this containment.

By Laura Iacovissi, Rabanus Derr, Robert C. Williamson
arXiv Machine Learning
Jul 30

The Advantage of Fine-Grained Training

arXiv:2509. 05130v2 Announce Type: replace Abstract: In classification problems, models are trained to predict a class label based on the input data features.

By Davide Pirovano, Federico Milanesio, Michele Caselle, Piero Fariselli, Matteo Osella
arXiv Machine Learning
Jun 16

Imbalanced Classification under Capacity Constraints

arXiv:2605. 03289v2 Announce Type: replace-cross Abstract: Detecting observations from a minority class under severe class imbalance is a central challenge in applications such as fraud detection, medical screening, and industrial quality control.

By Daniel Fraiman, Ricardo Fraiman