arXiv:2606.03026v2 Announce Type: replace-cross
Abstract: Binary spike activations allow a language-model runtime to read only active weight columns and replace multiplications by weight sums. We imp...
By Ting Liu
arXiv:2609.09772v1 Announce Type: new
Abstract: SymbolicLight V2 combines sparse event computation with continuous-state processing in a hybrid neuromorphic language architecture. Extending V1's spik...
By Ting Liu
arXiv:2605.21333v3 Announce Type: replace-cross
Abstract: Natively trained spiking language models must preserve information across time while operating through sparse binary activations, a combinati...
By Ting Liu
arXiv:2605. 21333v2 Announce Type: replace-cross Abstract: Natively trained spiking language models must preserve information across time while operating through sparse binary activations, a combination that has produced a persistent quality gap relative to dense Transformers.
By Ting Liu
arXiv:2608. 01536v1 Announce Type: cross Abstract: Large Language Models (LLMs) increasingly rely on sparsity to reduce inference cost, but most prior work targets a single sparsity source-either weight or activation-and optimizes for batched multi-user inference.
By Ruokai Yin, Priyadarshini Panda
arXiv:2608.30439v1 Announce Type: cross
Abstract: Inference with transformer-based large language models (LLMs) is often limited by the memory-bound KV cache and quadratic attention cost. State-space...
By Simon Richter, Ruhai Lin, Jason Yik, Taylor Kergan, Rui-Jie Zhu, Farshad Moradi, Jason Eshraghian