Meta$^n$ is a recursive self‑improvement framework for large language models that keeps a fixed meta‑operation Ω and repeatedly applies it to its own outputs, creating deeper layers that reason from higher perspectives. By avoiding changes to the meta‑operation, the system remains stable while the input grows, allowing depth to emerge through convergence and evolutionary search. Experiments on two backbone models show Meta$^n$ surpasses prior self‑improving agents across eight benchmark families, notably achieving positive scores on the ARC‑AGI‑2 benchmark designed to resist skill memorization.
By Zae Myung Kim, Young-Jun Lee, Seungyeon Jwa, Dongyeop Kang
arXiv:2608. 02636v1 Announce Type: cross Abstract: Self-evolving skill systems promise to improve agents by turning execution feedback into persistent skill updates without changing the underlying model.
By Yuxuan Liu, Zhaochen Su, Yuhao Zhang, Jiahe Guo, Zhongwei Xie, Huihao Jing, Lingyun Xie, Qing Zong, Yauwai Yim, Zhixiong Zhang, Haoran Li, Yangqiu Song
arXiv:2608. 08466v1 Announce Type: new Abstract: Modern LLM agents are often improved by modifying prompts, tools, or workflows manually, while the executable scaffold surrounding the model---the \emph{harness}---is typically treated as a fixed artifact after deployment.
By Tailin Zhou
The paper introduces a self‑evolving harness framework where a frozen language‑model agent first solves tasks and then edits its own harness based on run records. Using a 49‑line seed harness, the evolved harness improves average scores on in‑distribution benchmarks by 4.48 points and on out‑of‑distribution benchmarks by 12.64 points, surpassing Codex on the former and matching it on the latter. Continued evolution on a specific out‑of‑distribution benchmark further raises performance, and the study analyzes emergent mechanisms such as output truncation and history compaction.
By Qiankai Xu
arXiv:2608. 09629v1 Announce Type: new Abstract: Self-evolving agents are usually built around prescribed optimization pipelines: the framework decides how to gather evidence, revise a persistent artifact, select candidates, and stop.
By Hui Xue, Fan Yang
arXiv:2609.36746v1 Announce Type: new
Abstract: Agent skills provide a lightweight mechanism for self-evolving agents to accumulate reusable procedural knowledge without updating model parameters. Ho...
By Zhen Xiong, Qiaoyu Tan