← Back to all news
arXiv Machine Learning September 15, 2026 By Seunghoon Paik, Kangjie Zhou, Matus Telgarsky, Ryan J. Tibshirani

Basic Inequalities for First-Order Optimization with Applications to Statistical Risk Analysis

Read the original on arXiv Machine Learning →

The Flow has not summarised this story yet — read it at arXiv Machine Learning.

One email a morning, machine-written

One email a day, machine-written, one click to leave. We never share your address.

Related stories

arXiv Machine Learning
Jul 28

Risk reversal for least squares estimators under nested convex constraints

arXiv:2601. 16041v2 Announce Type: replace-cross Abstract: In constrained stochastic optimization, one expects that restricting the feasible set, provided it still contains the true parameter, should not increase the statistical risk of the corresponding projection estimator.

By Omar Al-Ghattas
rag
More like this →
arXiv Machine Learning
Jun 30

Replica Symmetry Breaking and Algorithmic Thresholds in Empirical Risk Minimization under Multi-Index Model

arXiv:2606. 28573v1 Announce Type: new Abstract: Modern machine learning models are trained by optimizing high-dimensional non-convex empirical risk functions.

By Andrea Montanari, Kangjie Zhou
More like this →
arXiv Machine Learning
4d ago

High-Probability Convergence of Clipped SGD under Heavy-Tailed Noise and $(L_0,L_1)$-Smoothness

arXiv:2505.20817v3 Announce Type: replace-cross Abstract: Gradient clipping is widely used in language-model training to control heavy-tailed gradient noise and can improve convergence guarantees ove...

By Taha El Bakkali El Kadi, Savelii Chezhegov, Aleksandr Beznosikov, Samuel Horv\'ath, Eduard Gorbunov
safety
More like this →
arXiv Machine Learning
Jun 25

A Geometry-Aware Efficient Algorithm for Compositional Entropic Risk Minimization

arXiv:2602. 02877v2 Announce Type: replace Abstract: This paper studies optimization for a family of problems termed $\textbf{compositional entropic risk minimization}$, in which each data's loss is formulated as a Log-Expectation-Exponential (Log-E-Exp) function.

By Xiyuan Wei, Linli Zhou, Bokun Wang, Chih-Jen Lin, Tianbao Yang
More like this →
arXiv Machine Learning
Aug 27

Beyond Optimal Rates in Stochastic Optimization: Trajectory-Adaptive Stopping Rules

arXiv:2608. 25551v1 Announce Type: new Abstract: Stochastic gradient descent (SGD) is typically analyzed at a deterministic horizon chosen before the algorithm is run, even though practical stopping decisions are made adaptively by inspecting the evolving trajectory.

By Liviu Aolaritei, Lucas L\'evy, Francis Bach, Michael I. Jordan
More like this →
arXiv Machine Learning
Jul 13

Accelerated Fully First-Order Methods for Bilevel and Minimax Optimization

arXiv:2405. 00914v4 Announce Type: replace-cross Abstract: We present in this paper novel accelerated fully first-order methods in \emph{Bilevel Optimization} (BLO).

By Chris Junchi Li
benchmarks
More like this →
About Pricing API Newsletter Sources Privacy Terms Refunds Accessibility Provider info Contact RSS

The Flow links to publishers and never republishes their articles. Summaries are machine-generated.

v1.1.0 · 5f852ea