arXiv Machine Learning

EMS Coreset: An Efficient Expectation-Maximization Algorithm for Sinkhorn Coreset

arXiv:2608. 16101v1 Announce Type: cross Abstract: Coresets distill large datasets into small, representative subsets for efficient downstream learning.

arXiv Machine Learning
Jun 25

Sample complexity of unbalanced entropic OT

arXiv:2606. 24987v1 Announce Type: cross Abstract: Optimal transport (OT) has become a central language for comparing probability measures, but exact balanced OT is often both too rigid for data with missing, created, or destroyed mass and subject to unfavorable high-dimensional sample complexity.

By Francisco Andrade, Gabriel Peyr\'e, Clarice Poon
arXiv Machine Learning
Jul 23

Streaming Sliced Optimal Transport

arXiv:2505. 06835v5 Announce Type: replace Abstract: Sliced optimal transport (SOT), or sliced Wasserstein (SW) distance, is widely recognized for its statistical and computational scalability.

By Khai Nguyen
Hugging Face Trending Papers
Jun 29

The Fundamental Limits of Valid Transport Map Estimation

Many modern generative modeling methods, including diffusion models, normalizing flows, and flow matching, estimate transport maps or plans between distributions without explicitly targeting an optimal transport (OT) map. In applications like generative modeling, the transport cost itself is irrelevant, and this makes it natural to target maps which are more tractable from either a statistical or computational standpoint.