Hugging Face Trending Papers

Transfer learning-based method for automated ewaste recycling in smart cities

Sorting a huge stream of waste accurately within a short period can be done with the support of digitalization, particularly Artificial Intelligence, instead of traditional methods. The overlap of Artificial Intelligence and Circular Economy can flourish many services in the environmental technology domain, in particular smart ewaste recycling, resulting in enabling circular smart cities.

arXiv Machine Learning
Sep 10

Constrained Bayesian Optimization for Hierarchical Federated Learning in IoT Networks for Plant Disease Classification

The paper introduces a constrained Bayesian Optimization framework to efficiently configure Hierarchical Federated Learning (HFL) for plant disease classification in IoT networks. It jointly optimizes the deep learning backbone, aggregation strategy, and communication rounds while respecting energy, execution time, and accuracy constraints. Experiments on an IoT-based plant disease task show the method explores only 11.11% of the search space yet finds solutions within 1% of exhaustive search, achieving a mean optimality gap of 0.056%.

By Athanasios Papanikolaou, Athanasios Tziouvaras, Apostolos Xenakis, Periklis Chatzimisios, Shameem A. Puthiya Parambath, George Floros, Enrica Zereik, Ivan Petrovic, Fabio Bonsignorio
Hugging Face Trending Papers
Aug 19

An Empirical Benchmark of Deep Time-Series Models for Smart Meter Energy Forecasting

The paper presents an empirical benchmark of nine deep learning models for smart meter energy forecasting, evaluating them on two public datasets. It examines how historical input length, prediction horizon, and model architecture affect accuracy, finding that longer historical context improves performance up to a saturation point while accuracy declines with longer horizons. The study also compares computational cost, showing lightweight models achieve similar accuracy to heavier ones, and notes that model choice matters less across most population segments.

arXiv AI
Sep 10

TASTE: Throughput-Aware Batch Size Tuning for On-Device Edge Learning

The paper presents TASTE, a method that uses Bayesian optimization to tune batch size for on‑device edge learning, aiming to maximize hardware throughput while preserving accuracy. Experiments on devices like the Raspberry Pi 4 show that the tuned batch size, combined with gradient accumulation and linear learning‑rate scaling, can double training throughput compared to using the maximum batch size. In online continual learning, the optimal batch size also helps balance stability and plasticity, reducing catastrophic forgetting without sacrificing efficiency.

By Avik Bhatnagar, Federico Nicolas Peccia, Oliver Bringmann
arXiv Machine Learning
Aug 24

BIPPO: Budget-Aware Independent PPO for Energy-Efficient Federated Learning Services

BIPPO (Budget-aware Independent Proximal Policy Optimization) is a multi‑agent reinforcement learning framework designed for energy‑efficient client selection in federated learning (FL) over IoT systems. It addresses infrastructure constraints such as limited resources and device churn, which traditional FL and RL approaches overlook. Evaluated on two image‑classification tasks with non‑IID data, BIPPO improves mean accuracy over non‑RL methods, standard PPO, and IPPO while consuming only a negligible portion of the budget, even as client numbers grow.

By Anna Lackinger, Andrea Morichetta, Pantelis A. Frangoudis, Schahram Dustdar
arXiv Machine Learning
Sep 18

Fast Training of Mixture-of-Experts for Time Series Forecasting via Expert Loss Integration

The paper introduces an adaptive Mixture-of-Experts (MoE) framework for time series forecasting that incorporates expert-specific losses to give each expert a direct learning signal independent of gating weights. The overall objective combines base forecasting loss with these expert losses, encouraging experts to specialize on different temporal segments. A partial online learning strategy is added for efficient incremental updates, and experiments on economic, tourism, and energy datasets show the method outperforms state‑of‑the‑art neural models and foundation models, with ablation studies confirming the benefit of expert loss integration.

By Btissame El Mahtout, Florian Ziel