← Back to all news
arXiv Machine Learning August 13, 2026 By Junyi Zou, Avrova Donz

MMLA: How Memory Lets the Past Shape the Future

Read the original on arXiv Machine Learning →

arXiv:2606. 28876v3 Announce Type: replace-cross Abstract: Proposal.

Summary generated by The Flow from the publisher's feed. The full article lives at arXiv Machine Learning.

  • llms

Related stories

arXiv Machine Learning
Aug 11

When Does Trace-Driven Evaluation Mislead MoE Expert Caching? Replay Semantics, Workload Contamination, and Operating Regimes

arXiv:2608. 07911v1 Announce Type: new Abstract: Mixture-of-Experts (MoE) models have outgrown accelerator memory, and offloading expert weights to host memory is now standard.

By Yu Zhang
More like this →
arXiv Machine Learning
5d ago

Subtract, Transport, or Replay? Auditable Deletion from Language-Model Memory

arXiv:2607. 27539v2 Announce Type: replace Abstract: Exact deletion from persistent language-model memory depends on whether a record's effect remains addressable after later computation.

By Vishwajith Ramesh
efficiency
More like this →
arXiv AI
Jun 10

Less Context, More Accuracy: A Bi-Temporal Memory Engine for LLM Agents Where a Lean Retrieved Context Beats the Full History

arXiv:2606. 09900v1 Announce Type: cross Abstract: Long-term memory is the missing layer for LLM agents: across sessions they forget, and the common workaround -- replaying the whole history into the prompt -- is expensive, slow, and, as distractors accumulate, less accurate.

By Liuyin Wang
llmsagentsbenchmarks
More like this →
arXiv Machine Learning
Jun 25

Reclaim Evaluation: A Lossy Memory Is Worse Than an Empty One

arXiv:2606. 25449v1 Announce Type: cross Abstract: A language model's memory can be worse than having no memory at all.

By Alex Kwon
llmsbenchmarks
More like this →
arXiv AI
Jun 4

memorywire: A Vendor-Neutral Wire Format for Agent Memory Operations

arXiv:2606. 01138v2 Announce Type: replace-cross Abstract: Agent-memory frameworks -- mem0, Letta/MemGPT, Cognee, Zep/Graphiti, MemoryOS, MemTensor -- each ship their own SDK, storage layout, and operational vocabulary.

By Thamilvendhan Munirathinam
agentssafety
More like this →
arXiv AI
Aug 7

Runtime Observability for Heterogeneous Attention Memory

arXiv:2608. 05863v1 Announce Type: new Abstract: Modern models no longer keep a plain KV cache: latent caches, learned sparse selectors and recurrent states each carry the model's memory in a different form, and each fails differently under compression.

By Fanzhe Wei, Li Liu, Ziyang Wang, Chenyu Wang
efficiency
More like this →