arXiv Machine Learning By Behrooz Tahmasebi, Melanie Weber, Stefanie Jegelka

Data Augmentation: A Fourier Analysis Perspective

Read the original on arXiv Machine Learning →

arXiv:2606. 24418v1 Announce Type: new Abstract: Data augmentation is a simple and model-agnostic approach for exploiting known invariances in learning problems.

Machine-generated by The Flow from the publisher's headline and feed description — not written or checked by a human. The full article lives at arXiv Machine Learning.

Hugging Face Trending Papers
Sep 8

Sparse Data Augmentation for Optimization with Provable Guarantees

The paper investigates sparse data augmentation for nonconvex optimization in geometric machine learning. It shows that using a small, fixed sample of transformations—obtained before optimization—allows gradient descent to achieve an ε‑stationary point of the fully augmented objective with ≤ O((log|G|+log(1/δ))/ε²) transformation queries. This is more efficient than both full augmentation and standard group‑SGD, which require O(1/ε⁴) queries.

arXiv Machine Learning
Sep 7

An Analysis of Self-supervised Pre-training with Dependent Samples

The paper investigates self‑supervised pre‑training that uses multiple data augmentations of the same unlabeled sample. It shows that pooling these dependent augmentations together yields statistical estimation error bounds that are never worse than, and sometimes better than, partitioning the data into independent subsets. The analysis explains why using many augmentations is practically advantageous, especially when their correlations have mild effects or reduce estimation variance.

By Maximilian Fleissner, Debarghya Ghoshdastidar, Samory Kpotufe