arXiv AI By Hongyao Tang, Yi Ma, Pengyi Li, Yifu Yuan

Generalized Agent Iteration: One Formal Framework for Iterative Policy Improvement and Recursive Self-Improvement

Read the original on arXiv AI →

The paper introduces Generalized Agent Iteration (GAI), a formal framework that unifies iterative policy improvement and recursive self‑improvement (RSI) under a single learning paradigm. GAI treats an agent as a configuration of modifiable components and models learning as a cycle of evaluation and improvement, with two key dials: whether the improving mechanism is part of the agent and whether the evaluation standard is external. These dials distinguish between generalized policy iteration (GPI) and RSI, and classify systems as anchored, goal‑drift, or fully self‑referential, allowing existing systems to be mapped and RSI defects to be analyzed systematically.

Machine-generated by The Flow from the publisher's headline and feed description — not written or checked by a human. The full article lives at arXiv AI.

Hugging Face Trending Papers
Jul 8

Recursive Self-Improvement in AI: From Bounded Self-Refinement to Autonomous Research Loops

AI systems increasingly participate in their own improvement: revising their outputs, adapting their own harnesses during deployment, training on data they generate, and, increasingly, conducting AI research itself. This literature is described under a vocabulary ("self-refine," "self-reward," "self-play," "self-evolve") that conflates fundamentally different ambitions.

arXiv AI
Jul 16

Self-Improvements in Modern Agentic Systems: A Survey

arXiv:2607. 13104v1 Announce Type: new Abstract: Self-improving autonomous agents are moving from research prototypes to deployed systems.

By Zhe Ren, Yimeng Chen, Dandan Guo, Guowei Rong, Tonghui Li, R. B. Xiong, Qingfeng Lan, Wenyi Wang, Li Nanbo, Yibo Yang, Mingchen Zhuge, J\"urgen Schmidhuber
arXiv AI
Jun 26

The Red Queen G\"odel Machine: Co-Evolving Agents and Their Evaluators

arXiv:2606. 26294v1 Announce Type: cross Abstract: Self-improving agents are state-of-the-art (SOTA) on agentic coding benchmarks and have recently been extended to general domains.

By Alex Iacob, Andrej Jovanovi\'c, William F. Shen, Daniel Burkhardt, Meghdad Kurmanji, Nurbek Tastan, Lorenzo Sani, Niccol\`o Alberto Elia Venanzi, Ambroise Odonnat, Zeyu Cao, Bill Marino, Xinchi Qiu, Nicholas D. Lane