arXiv:2606. 15004v1 Announce Type: cross Abstract: Deploying neural networks on low-power microcontrollers (MCUs) requires selecting model architectures under tight memory, latency, and energy constraints.
By Joseph Q. Zales, Pragya Sharma, Mani Srivastava
The paper presents a rapid pipeline for training and deploying machine‑learning models on the WeBe Band, a wrist‑worn wearable device. It automates the creation of hardware‑efficient models, integrates with the Piccolo AI ecosystem, and supports OTA deployment while profiling latency and memory usage. Experimental results show trade‑offs between classical models and lightweight neural networks for real‑time performance on a microcontroller.
By Ehsan Kourkchi, Asmita Asmita, Houman Homayoun, Mahdi Eslamimehr
arXiv:2607. 18171v1 Announce Type: new Abstract: Real-time multimodal applications, including voice agents and interactive video generation, compose heterogeneous models into pipelines whose efficient deployment requires application-specific decisions about placement, streaming, and intra-model parallelism.
By Krish Agarwal, Zhuoming Chen, Yanyuan Qin, Zhenyu Gu, Atri Rudra, Beidi Chen
The paper explores how the number of bird species (target classes) affects the compressibility of neural networks for passive acoustic monitoring on microcontroller units (MCUs). By training and compressing models with varying class counts, the authors show that significant compression can be achieved with minimal performance loss. They also benchmark different hardware platforms and assess the feasibility of deploying energy‑autonomous monitoring devices.
By Nina Brolich, Simon Geis, Maximilian Kasper, Alexander Barnhill, Axel Plinge, Dominik Seu{\ss}
arXiv:2608.28652v1 Announce Type: new
Abstract: Artificial intelligence (AI) models have demonstrated remarkable capabilities across various domains, yet their widespread deployment is impeded by sig...
By Venkat R. Dasari, Jakob A. Adams, Vinod K. Mishra, Brian Jalaian
arXiv:2608. 26418v1 Announce Type: cross Abstract: Modern AI workloads and the hardware that runs them evolve on different timescales: architectural definition precedes volume silicon by years, while target workloads shift in months.
By Architect Labs