arXiv AI By Harish Gaggar

Protocol-Preserving Context Trimming for Agentic Workflows: Benefits, Failure Regimes, and Budget Guardrails

Read the original on arXiv AI →

The paper evaluates five context‑trimming strategies for agentic large language model workflows, comparing them on metrics such as task success, protocol adherence, token savings, and latency. Conventional trimming methods save about 60% of tokens but achieve lower success rates, while protocol‑aware trimming raises success to 92.2% and adaptive guardrails further improve it to 96% success with 56% token savings. The study shows that preserving protocol‑critical state is more important than aggressive token removal, and that adaptive guardrails enhance efficiency, scalability, and reliability for long‑horizon agentic systems.

Machine-generated by The Flow from the publisher's headline and feed description — not written or checked by a human. The full article lives at arXiv AI.

arXiv AI
Jun 10

Less Context, Better Agents: Efficient Context Engineering for Long-Horizon Tool-Using LLM Agents

arXiv:2606. 10209v1 Announce Type: new Abstract: Large language models deployed as autonomous agents for enterprise workflows face a key challenge: verbose tool responses from enterprise systems can cause context overflow, stale-state errors, and high inference cost.

By Abhilasha Lodha, Mahsa Pahlavikhah Varnosfaderani, Abir Chakraborty, Abhinav Mithal
arXiv AI
Aug 26

When "Must" Becomes "Maybe": Constraint Weakening in LLM Agent Workflows

Large language model agents coordinate tasks via multi‑role, multi‑stage workflows that transform upstream state into intermediate artifacts such as summaries and plans. The study shows that when these artifacts are transformed—through compression, plan assimilation, or other handoff methods—the strict action‑binding constraints on upstream state can be weakened, turning mandatory requirements into optional information. In 1,296 synthetic episodes, direct handoff preserved all safety blockers, whereas transformed handoffs frequently deactivated or forbidden actions, but restoring full state fields or applying downstream verification can recover preservation.

By Yiheng Sun, Huifei Wang, Yancheng Zhu, Zhenyu Li, Zebin Zhao, Yifan Yuan
Hugging Face Trending Papers
Aug 2

Control Under Compression: Reliability Frontiers for Tool-Using Agents

Tool-using language-model agents are governed not only by task prompts but also by persistent system-side instructions that specify tools, arguments, policies, execution protocols, and recovery. Compressing these agent control contexts (ACCs) can reduce input cost and context use, yet existing prompt-compression evaluations do not reveal whether the resulting control remains operationally reliable.

arXiv AI
Aug 19

Token Optimization and Context Window Management in Multi-Agent AI Workflows

The paper "Token Optimization and Context Window Management in Multi‑Agent AI Workflows" introduces a practitioner framework that reduces token usage and latency in multi‑agent AI systems. It outlines six patterns—context stratification, fetch‑once/process‑locally architecture, schema‑contracted prompts, token‑aware fallback chains, semantic caching, and inter‑agent communication compression—and reports a 60‑70% token reduction and a 61‑116 second cold‑load latency improvement in production. A controlled study on relevance‑contrast context shows that mixing high‑ and low‑relevance items in prompts can improve relevance accuracy by up to +0.084. whyItMatters":"The work provides concrete, repeatable engineering patterns that bridge research and production, enabling faster, cheaper, and more reliable AI workflows."

By Dvir Shamay
arXiv AI
Aug 18

AstronOS: A Unified Execution Model and Runtime for Long-Horizon Agentic Systems

arXiv:2608. 16381v1 Announce Type: new Abstract: Agentic systems often organize execution and state around a single conversation, model invocation, or agent instance, even when real work spans many calls and stages.

By Zhenhang Nie (iFLYTEK Co., Ltd., Hefei, China), Gui Zheng (iFLYTEK Co., Ltd., Hefei, China), Xudong Sun (iFLYTEK Co., Ltd., Hefei, China), Tailong Zhu (iFLYTEK Co., Ltd., Hefei, China), Bin Zhang (iFLYTEK Co., Ltd., Hefei, China)