arXiv Machine Learning By Hanti Lin

Pessimistic Meta-Induction and Its Limits: Lessons from Frequentist Statistics and Machine Learning Theory

Read the original on arXiv Machine Learning →

The paper titled "Pessimistic Meta-Induction and Its Limits: Lessons from Frequentist Statistics and Machine Learning Theory" critiques the pessimistic meta-inductive argument against scientific realism by attacking its inductive step rather than its historical premise. It introduces a new challenge, drawing on frequentist statistics, machine learning, and formal epistemology to assess induction through convergence to truth. The authors argue that ordinary enumerative induction can achieve convergence everywhere, whereas meta-induction fails to achieve even almost everywhere convergence, and in contexts where meta-induction applies, no inference method can achieve almost everywhere convergence.

Machine-generated by The Flow from the publisher's headline and feed description — not written or checked by a human. The full article lives at arXiv Machine Learning.

arXiv Machine Learning
Sep 25

An Order-Theoretic Characterization of Consistent Inductive Inference

The paper presents an order-theoretic characterization of consistent inductive inference for arbitrary binary hypothesis classes. It shows that consistency—making only finitely many prediction errors on any infinite sequence labeled by an unknown hypothesis—can be captured by a single linear order on finite realizable traces. This order must satisfy two conditions: conflicting traces select different least subtraces, and the order is well‑founded on traces of each fixed target, enabling a learner whose evidence decreases with each mistake. Conversely, any consistent learner induces such an order via canonical mistake transcripts and the Kleene–Brouwer ordering.

By Zhou Lu
arXiv Statistics ML
2d ago

The Benchmarking Epistemology: Validity Theory for Evaluating Machine Learning Models

The article discusses how predictive benchmarking—evaluating machine learning models by their performance and ranking—serves as a core method in machine learning research. It argues that benchmark scores only reflect performance on specific datasets and learning problems, and that drawing broader scientific conclusions requires explicit assumptions. By adapting concepts from psychological validity theory, the authors propose validity conditions to make these assumptions clear, and demonstrate their application in two case studies (ImageNet and the Fragile Families Challenge) to illustrate how benchmark results can inform inferences about research progress and limits of predictability.

By Timo Freiesleben, Sebastian Zezulka
arXiv Machine Learning
Aug 5

Benign interpolation and Occam's razor

arXiv:2608. 03386v1 Announce Type: new Abstract: Contemporary deep learning methods generalize well even when they fit their training data perfectly, a phenomenon known as benign interpolation.

By Tom F. Sterkenburg, Daniel A. Herrmann, Jan-Willem Romeijn