arXiv AI

One Knob to Rule Them All: A Unified Optimal Transport View of Cold-Start Active Learning

arXiv:2608. 03249v1 Announce Type: new Abstract: Cold-Start Active Learning (CSAL) aims to select a valuable subset from an unlabeled pool without any prior knowledge or human assistance.

arXiv Machine Learning
1d ago

Low-Budget Active Learning through Entropic Optimal Transport

The paper introduces a low-budget active learning approach that selects a small coreset of data points for training high-accuracy models, particularly useful when labeling is expensive, such as in medical contexts. It uses features from a pretrained self-supervised model and applies entropic optimal transport—specifically the Sinkhorn divergence—as the selection criterion, enabling dimension-free sample complexity and efficient gradient-based optimization. The method combines gradient-based candidate generation with a swap-based local search, achieving superior performance over existing heuristics on image and medical datasets.

By Rim Hajal, Mathieu Besan\c{c}on, J\'er\^ome Malick
arXiv AI
Jun 3

ASAP: Exploiting the Satisficing Generalization Edge in Neural Combinatorial Optimization

arXiv:2501. 17377v4 Announce Type: replace-cross Abstract: Deep Reinforcement Learning (DRL) has emerged as a promising approach for solving Combinatorial Optimization (CO) problems, such as the 3D Bin Packing Problem (3D-BPP), Traveling Salesman Problem (TSP), or Vehicle Routing Problem (VRP), but these neural solvers often exhibit brittleness when facing distribution shifts.

By Han Fang, Paul Weng, Yutong Ban
arXiv Machine Learning
Jun 9

LARP: Learner-Agnostic Robust Data Prefiltering

arXiv:2506. 20573v4 Announce Type: replace-cross Abstract: Public datasets, crucial for modern machine learning and statistical inference, often contain low-quality or contaminated samples that can harm model performance.

By Kristian Minchev, Dimitar I. Dimitrov, Nikola Konstantinov
arXiv AI
2d ago

Geometry-Aware Adaptation for Pretrained Models

arXiv:2307.12226v3 Announce Type: replace-cross Abstract: Machine learning models -- including prominent zero-shot models -- are often trained on datasets whose labels are only a small proportion of...

By Nicholas Roberts, Xintong Li, Dyah Adila, Sonia Cromp, Tzu-Heng Huang, Jitian Zhao, Frederic Sala