arXiv Machine Learning By Jiamu Zhang, Liang Wu, Kelly Wan, Hanjie Chen, Liangjie Hong

SPECTRA: Pushing the KV Cache Beyond the 2-Bit Cliff via Spectral Transform Coding

Read the original on arXiv Machine Learning →

arXiv:2608. 07915v1 Announce Type: new Abstract: Large language models (LLMs) increasingly read long inputs in the agentic era, from whole documents and codebases to conversations across many turns.

Summary generated by The Flow from the publisher's feed. The full article lives at arXiv Machine Learning.