arXiv AI By Yongwan Jo, Jinyoung Park, Euihyun Lee, Dokyung Song

SparSEEty: Extracting Tokens from Sparsity-Exploiting LLM Serving Systems via Deterministic Side Channels

Read the original on arXiv AI →

arXiv:2608. 02995v1 Announce Type: cross Abstract: Modern large language models (LLMs) exhibit activation sparsity, wherein only a subset of their neurons is activated for given input tokens.

Summary generated by The Flow from the publisher's feed. The full article lives at arXiv AI.