arXiv AI

Why Does CLAUDE.md Keep Growing? Catastrophic Remembering in Agentic Coding

arXiv:2608. 11095v1 Announce Type: new Abstract: Agentic coding READMEs like CLAUDE.

Hugging Face Trending Papers
Jul 7

Doomed from the Start: Early Abort of LLM Agent Episodes via a Recall-Controlled Probe Cascade

Large language model (LLM) agents solving multi-step tasks frequently commit to trajectories that are doomed to fail, yet continue to consume substantial inference compute before the failure becomes observable. We show that failure is predictable early from the agent's internal representations: lightweight per-round probes on hidden activations anticipate eventual episode failure as early as the first interaction round, where scorers reading only the agent's observable behavior are barely better than chance.