arXiv AI
2d ago

FractalNet-Based Heterogeneous Federated Learning for Orbital Edge Intelligence in Satellite Mega-Constellations: A Wildfire Case Study

The paper introduces a heterogeneous federated learning approach using the FractalNet architecture tailored for satellite mega‑constellations. It formalizes contact‑window‑constrained, depth‑heterogeneous optimization and proposes a distributed path scheduler that assigns model depth based on satellite SWAP‑C constraints, predicted contacts, and training statistics. The framework includes periodic update pooling and a three‑tier agentic control plane, and is validated through a wildfire detection case study across LEO, MEO, and GEO/HEO shells, demonstrating improvements in convergence, communication efficiency, energy adaptation, and robustness.

By Sai Puppala, Koushik Sinha
arXiv Machine Learning
Aug 5

When RL Meets Adaptive Speculative Training: A Unified Training-Serving System

arXiv:2602. 06932v5 Announce Type: replace Abstract: Speculative decoding can significantly accelerate LLM serving, yet most deployments today disentangle speculator training from serving, treating speculator training as a standalone offline modeling problem.

By Junxiong Wang, Fengxiang Bie, Jisen Li, Zhongzhu Zhou, Zelei Shao, Yubo Wang, Yinghui Liu, Qingyang Wu, Avner May, Sri Yanamandra, Ce Zhang, Tri Dao, Percy Liang, Ben Athiwaratkun, Shuaiwen Leon Song, Chenfeng Xu, Xiaoxia Wu
arXiv Machine Learning
4d ago

Ampere: Communication-Efficient and High-Accuracy Split Federated Learning

Ampere is a new split federated learning system that reduces both on‑device computation and device‑server communication while improving accuracy. It trains device and server blocks sequentially with local losses, eliminating gradient transfers, and uses a lightweight auxiliary network to consolidate activations into a single transfer. Experiments on CNNs and Transformers show up to 11.70 pp accuracy gains, 18.6× faster training, 911× less communication, and 14.5× less computation compared to state‑of‑the‑art SFL baselines.

By Zihan Zhang, Leon Wong, Blesson Varghese