arXiv:2601. 00821v3 Announce Type: replace Abstract: A growing class of conversational-memory systems compresses dialogue history into structured artifacts -- extracted facts, decisions, or events -- on the premise that distilled structure retrieves better than raw text.
By Tao An
arXiv:2601. 00821v4 Announce Type: replace Abstract: A growing class of conversational-memory systems compresses dialogue history into structured artifacts (extracted facts, decisions, or events) on the premise that distilled structure retrieves better than raw text.
By Tao An
arXiv:2608. 03463v1 Announce Type: new Abstract: Long-term memory is essential for LLM-based agents to sustain interactions and reliably leverage distant history.
By Yuxin Liao, Le Wu, Min Hou, Hao Liu, Han Wu, Zishu Wang
arXiv:2609.26780v1 Announce Type: cross
Abstract: Long-term conversational memory in multi-party settings requires more than retrieving relevant content from long-term conversations: it must distingu...
By Haobo Zheng, Tan Tang, Yan Chen, Weijie Wang, Yingcai Wu
The paper introduces Hard-Origin Adaptively Softened Memory (HasMem), a memory system for large language model agents that combines frozen hard‑prompt embeddings with a controller, writer, reader, and global module to adaptively resize and re‑encode memory entries. On a reconstruction probe of 535 questions, HasMem achieves a lexical F1 of 95.3, outperforming the hard reference by 4.4 percentage points while maintaining 93.6% of the reference’s memory positions. Across six configurations with similar per‑question budgets, the system surpasses rule‑based re‑encoding by 8.0–23.6 exact‑match points, and on LongMemEval‑S it improves local lexical F1 from 3.4 to 8.9 and reduces answer negative log‑likelihood from 12.257 to 5.274.
By Zihong He, Junxiao Shen, Chen Liang, Hai-Ning Liang
The paper introduces AWM, a framework that treats the terminal working memory of long‑document VQA agents as an answerable evidence artifact. It proposes a memory‑only answerability diagnostic and incorporates this signal into the GRPO reward, giving higher advantage to trajectories whose final memory can answer the question alone. Experiments on MMLongBench‑Doc and LongDocURL show that AWM‑GRPO boosts final‑answer accuracy by up to 11.9 points and reduces the rate of correct answers that cannot be supported by memory alone.
By Dongzhuoran Zhou, Yuqicheng Zhu, Yule Liu, Zhen Yang, Rui Lu, Yuxiao Dong, Jie Tang, Evgeny Kharlamov