arXiv AI By Varun Manjunath, Ruokai Yin, Donghyun Lee, Arkapravo Ghosh, Priyadarshini Panda

BRIM: Workload-Balanced Dual-Sided Bit-Serial Sparse Inference Accelerator

Read the original on arXiv AI →

arXiv:2607. 19431v1 Announce Type: cross Abstract: Bit-serial accelerators exploit bit-level sparsity to reduce DNN inference cost, but existing designs exploit sparsity on only one operand, bounding the speedup.

Summary generated by The Flow from the publisher's feed. The full article lives at arXiv AI.