arXiv AI

Decentralized Autoregressive Generation

arXiv:2601. 03184v3 Announce Type: replace-cross Abstract: The decentralization of autoregressive generation has attracted considerable attention in recent years as a solution to scaling bottlenecks.

arXiv AI
Jun 2

Heterogeneous Decentralized Diffusion Models

arXiv:2603. 06741v2 Announce Type: replace-cross Abstract: Training frontier-scale diffusion models often requires substantial computational resources concentrated in tightly-coupled clusters, limiting participation to well-resourced institutions.

By Zhiying Jiang, Raihan Seraj, Marcos Villagra, Bidhan Roy
arXiv Machine Learning
Sep 11

Partial GFlowNet: Accelerating Convergence in Large State Spaces via Strategic Partitioning

The paper introduces Partial GFlowNet, a method that partitions a large state space into overlapping partial state spaces to accelerate convergence of Generative Flow Networks. By restricting the actor’s exploration to these smaller regions and using a heuristic to switch between them, the approach enables efficient identification of high‑reward subregions. Experiments on popular datasets show that Partial GFlowNet converges faster, produces higher‑reward candidates, and improves diversity compared to existing methods.

By Xuan Yu, Xu Wang, Rui Zhu, Yudong Zhang, Yang Wang
arXiv Machine Learning
Aug 6

Stable GFlowNets with TV Monitoring and Probabilistic Guarantees

arXiv:2605. 01729v2 Announce Type: replace Abstract: Generative Flow Networks (GFlowNets) learn to sample states proportional to an unnormalized reward.

By Zengxiang Lei, Ananth Shreekumar, Jonathan Rosenthal, Ruoyu Song, Alvaro A. Cardenas, Daniel J. Fremont, Dongyan Xu, Satish Ukkusuri, Z. Berkay Celik
arXiv Machine Learning
Aug 31

Beyond Non-IID: Learner--Client Distribution Mismatch in Federated Learning

The paper addresses the mismatch between learner and client data distributions in federated learning, noting that traditional client selection methods often ignore this misalignment. It introduces a dynamic, influence-aware client selection framework that uses a small proxy dataset to estimate each client's utility for the learner’s objective, prioritizing informative sources while mitigating noise and heterogeneity. Experiments on CIFAR-10 with heterogeneous partitions show the proposed method outperforms static and dynamic baselines, achieving faster convergence and higher accuracy.

By Yiming Xie, Lili Su, Ningfang Mi
arXiv AI
Jul 7

Decentralised Federated Learning over Temporal Networks: The Role of Heterogeneities

arXiv:2607. 03171v1 Announce Type: cross Abstract: Decentralised federated learning, based on peer-to-peer communication, is increasingly proposed for on-device training of machine learning models, promising a privacy-preserving, communication-efficient training process with no risk of single-point failure.

By Arash Badie-Modiri, Chiara Boldrini, Lorenzo Valerio, J\'anos Kert\'esz, M\'arton Karsai
arXiv Machine Learning
Aug 27

Cooperative Multi-Agent Reinforcement Learning for Adaptive Aggregation in Semi-Supervised Federated Learning with non-IID Data

The paper introduces pFedMARL, a federated learning framework that uses multi‑agent reinforcement learning with TD3 to dynamically adjust client contributions and personalize models. It applies a server‑side agent to optimize global aggregation and client‑side agents to balance global and local updates, eliminating the need for pre‑training. Experiments on a semi‑supervised audio spectrogram transformer show that pFedMARL outperforms or matches FedAvg, Ditto, and local training across various non‑IID settings and against adversarial clients, improving accuracy, robustness, and fairness.

By Rene Glitza, Luca Becker, Rainer Martin