arXiv AI
Sep 11

What Should an Agent Forget? Separating What Is Stored from What Is Used

The paper introduces RD-Forget, a training‑free framework that separates what a persistent language agent stores from what it uses at answer time. It keeps a source archive of all observations while a query‑conditioned memory view filters evidence relevant to the current question, using a frozen language‑model curator to group facts into semantic slots and preserve multi‑hop relations. The approach employs rate‑distortion principles to stay within a memory budget and demonstrates improvements across conversational memory, knowledge updating, fact consolidation, long‑context reasoning, and personalization tasks.

By Yuhang Li, Yuchen Li
arXiv AI
Oct 2

What Should an Agent Remember? Disentangling Retention from Retrieval in Bounded-Memory Evaluation

The paper introduces a streaming-recall benchmark that separates retention and selection decisions for persistent agents. It shows that query‑aware selection boosts recall by 15.5 points when access is fixed, while mixed comparisons inflate gains due to changes in history access. The study finds that under bounded retention, failures stem from eviction rather than ranking errors, and that dense retrieval can outperform lexical retrieval on natural text.

By Juli Huang
arXiv Machine Learning
Jun 15

Efficient Rationale-based Retrieval: On-policy Distillation from Generative Rerankers based on JEPA

arXiv:2604. 23336v3 Announce Type: replace-cross Abstract: Unlike traditional fact-based retrieval, rationale-based retrieval typically necessitates cross-encoding of query-document pairs using large language models, incurring substantial computational costs.

By Teng Chen, Sheng Xu, Feixiang Guo, Xiaoyu Wang, Qingqing Gu, Hongyan Li, Luo Ji