arXiv AI

Self-Reference in Large Language Models: The Introspection Threshold for Recursive Self-Improvement

arXiv:2607. 04277v1 Announce Type: cross Abstract: The pursuit of self-evolving AI raises a critical question: when is autonomous self-improvement sustainable rather than degenerative?

arXiv AI
Aug 26

Meta$^n$: Recursive Self-Improvement through Emergent Depth

Meta$^n$ is a recursive self‑improvement framework for large language models that keeps a fixed meta‑operation Ω and repeatedly applies it to its own outputs, creating deeper layers that reason from higher perspectives. By avoiding changes to the meta‑operation, the system remains stable while the input grows, allowing depth to emerge through convergence and evolutionary search. Experiments on two backbone models show Meta$^n$ surpasses prior self‑improving agents across eight benchmark families, notably achieving positive scores on the ARC‑AGI‑2 benchmark designed to resist skill memorization.

By Zae Myung Kim, Young-Jun Lee, Seungyeon Jwa, Dongyeop Kang
arXiv Machine Learning
Sep 10

MetaRSI / RSI2: A Meta-Recursive Self-Improving System for Recursive Self-Improving Systems Themselves

arXiv:2609.06396v2 Announce Type: new Abstract: Recursive self-improvement (RSI) lets a system improve the model-building machinery from its own failures, so every later model inherits the gain. Yet...

By Zihan Tan, Leixin Sun, Zitong Shi, Yitao Liu, Jiajun Wu, Nathaniel Brooks, Jiaru Qian, Xiaoran Shang, Suyuan Huang, Yi Ding, Yangxu Liao, Mukai Li, Qiushi Sun, Shudong Liu, Xuankun Rong, Xiaohang Yu, Zhuo Chen, Hejia Geng, Chenxin Li, Aozhou Wang, Zengji Tu, Robert Tang, Yuxin Zhan, Eric Jiang, Yuxin Wu, Jianqing Zhang, Xiao Liang, Fang Wu, Haochi Zhang, Alexander Marlow, Guancheng Wan
arXiv AI
Sep 16

Self-Emergence Agent Architecture:Behavior-Inertia HMM, Reflexive Metacognition,and Social-Contrastive Self-Modeling

The paper introduces the Self‑Emergence Agent Architecture (SEAA), a framework that combines a Hidden Markov Model for behavioral inertia, a reflexive metacognition loop that updates the HMM, and a social environment where agents compare behaviors. This closed loop enables agents to develop distinct, stable personalities and social structures without external prompts. Experiments with both a language‑model‑free prototype and hosted LLMs demonstrate spontaneous symmetry breaking and the emergence of consensus hubs and outliers.

By Xiaoyang Liu
arXiv Machine Learning
Aug 12

Scaling Self-Play with Self-Guidance

arXiv:2604. 20209v2 Announce Type: replace Abstract: LLM self-play algorithms are notable in that, in principle, nothing bounds their learning: a Conjecturer model creates problems for a Solver, and both improve together.

By Luke Bailey, Kaiyue Wen, Kefan Dong, Tatsunori Hashimoto, Tengyu Ma
arXiv Machine Learning
Aug 27

Emergent Abilities in Large Language Models: A Survey

Emergent Abilities in Large Language Models: A Survey reviews how scaling LLMs leads to previously unseen capabilities such as advanced reasoning, in-context learning, coding, and problem-solving. The paper critically examines definitions, inconsistencies, and the conditions that foster these abilities, including scaling laws, task complexity, pre‑training loss, quantization, and prompting strategies. It also discusses the extension to Large Reasoning Models and highlights safety concerns like deception, manipulation, and reward hacking, calling for improved evaluation and governance.

By Leonardo Berti, Flavio Giorgi, Gjergji Kasneci
arXiv AI
3d ago

Rep2Skill: Representation-Guided Skill Self-Evolution for LLM Agents

Rep2Skill introduces a representation-guided framework that enables large language model agents to self-evolve their textual skills by analyzing internal representation trajectories from agent rollouts. The method identifies execution turns that deviate from successful dynamics and uses these signals, together with execution contexts, as actionable feedback for targeted skill revision. Experiments with two open-source LLMs across two agent environments demonstrate that Rep2Skill consistently outperforms purely text-based approaches, showing that incorporating internal representations can enhance agent self-improvement.

By Kaixing Zhang, Changming Li, Yingdong Shi, Zheng Zhang, Kaitao Song, Wenjie Shi, Jingang Wang, Kan Ren