arXiv:2606. 04757v1 Announce Type: cross Abstract: We study decentralized stochastic smooth convex optimization, where $M$ workers minimize an average objective using local stochastic gradients and neighbor-only communication over a fixed gossip network.
By Nitai Kluger, Amit Attia, Tomer Koren
arXiv:2609.14953v1 Announce Type: cross
Abstract: This paper aims to develop new and efficient distributed algorithms for solving a class of monotone inclusions, $0 \in \sum_{i=1}^n (G_ix + T_ix)$, o...
By Nghia Nguyen-Trung, Ion Necoara, Quoc Tran-Dinh
arXiv:2606. 07496v1 Announce Type: new Abstract: Decentralized stochastic optimization is a fundamental paradigm for large-scale learning over networks, where agents communicate only with their neighbors and no central coordinator is required.
By Ming Sun, Kun Yuan
arXiv:2407.02765v4 Announce Type: replace-cross
Abstract: We study the distributed optimization problem over a graphon with a continuum of nodes, which is regarded as the limit of the distributed net...
By Yan Chen, Tao Li, Xiaofeng Zong
arXiv:2606. 09154v1 Announce Type: new Abstract: Decentralized SGD is a fundamental algorithm in decentralized learning, although the influence of an underlying network topology on its convergence behavior is not yet fully understood.
By Yuki Takezawa, Anastasia Koloskova, Sebastian U. Stich
In this paper, we consider the nonsmooth nonconvex decentralized optimization problem, where inter-agent communication is compressed. We propose a general framework that unifies various decentralized stochastic subgradient-type methods with unbiased compression and contractive compression with error compensation.
The paper investigates two strategies for incorporating heterogeneous node weights in decentralized learning: embedding the weights into local losses to use a doubly stochastic matrix, and keeping the original losses while using a λ‑induced row‑stochastic matrix. By developing a weighted Hilbert‑space framework, the authors derive tighter convergence rates and show that the row‑stochastic matrix becomes self‑adjoint, reducing penalty terms that otherwise amplify consensus error. They provide conditions under which the row‑stochastic design converges faster, even with a smaller spectral gap, and offer topology‑design guidelines based on eigenvalue comparisons.
By Bing Liu, Boao Kong, Limin Lu, Kun Yuan, Chengcheng Zhao
arXiv:2602. 03682v2 Announce Type: replace-cross Abstract: We analyze the Accelerated Noisy Power Method, an algorithm for Principal Component Analysis in the setting where only inexact matrix-vector products are available, which can arise for instance in decentralized PCA.
By Pierre Agui\'e, Mathieu Even, Laurent Massouli\'e
arXiv:2607. 01755v1 Announce Type: cross Abstract: In this paper, we consider the nonsmooth nonconvex decentralized optimization problem, where inter-agent communication is compressed.
By Siyuan Zhang, Nachuan Xiao, Xin Liu
arXiv:2608. 09565v1 Announce Type: cross Abstract: Optimization theory is a widely used tool for intelligent decision-making.
By Muhammad Faraz Ul Abrar, Nicol\`o Michelusi, Erik G. Larsson
We study decentralized online optimization of upper-linearizable payoffs over an action set under efficient separation access, with applications to online continuous diminishing-return (DR) submodular...
arXiv:2606. 27216v1 Announce Type: cross Abstract: Muon-type optimizers construct update directions for dense neural-network weights by applying a finite Newton-Schulz map to momentum-gradient matrices.
By Ziyuan Tang, Tianshi Xu, Yousef Saad, Yuanzhe Xi