arXiv AI By Nayoung Choi, Jonathan Zhang, Jinho D. Choi

DYCP: Dynamic Context Pruning for Long-Form Dialogue with LLMs

Read the original on arXiv AI →

arXiv:2601. 07994v5 Announce Type: replace-cross Abstract: Large Language Models (LLMs) increasingly operate over long-form dialogues with frequent topic shifts.

Machine-generated by The Flow from the publisher's headline and feed description — not written or checked by a human. The full article lives at arXiv AI.

Hugging Face Trending Papers
Jun 10

Context-Driven Incremental Compression for Multi-Turn Dialogue Generation

Modern conversational agents condition on an ever-growing dialogue history at each turn, incurring redundant attention and encoding costs that grow with conversation length. Naive truncation or summarization degrades fidelity, while existing context compressors lack cross-turn memory sharing or revision, causing information loss and compounding errors in long dialogues.

arXiv Computation and Language
Sep 14

Residual Vector-based Reconstruction as Long-Context Recall Regardless of Context Window Size

The paper introduces a long‑context recall technique that keeps GPU memory usage nearly constant regardless of context length, without requiring additional training. It reconstructs facts by leveraging residual vectors stored in the LLM’s feed‑forward layers, enabling deterministic retrieval of query‑relevant information without accessing the original document. Experiments demonstrate the method can answer single‑fact questions in two‑million‑token stories, outperforming prior approaches.

By MyungHoon Ryu, XinYu Piao, Jong-Kook Kim
arXiv AI
3d ago

StateTree: Enhancing Long-Term Dialogue Reasoning via Reinforcement Learning

StateTree is a reinforcement learning approach that improves long‑term dialogue reasoning by building a tree‑structured auxiliary task from limited dialogue data. The method embeds key‑value records across multiple sessions into a binary tree, requiring the model to traverse from root to leaf, retrieve records, compare timestamps, and identify a target question among distractors. Curriculum RL training increases tree depth, and a compositional variant trains the model to combine partial reasoning fragments, enabling cross‑session retrieval, temporal reasoning, knowledge updates, and multi‑hop reasoning while generalizing from 10K‑token to 128K‑token contexts.

By Naen Xu, Wanqing Cui, Yibo Hu, Shixin Hong, Hengyu An, Meiguang Jin, Junfeng Ma, Tianyu Du