arXiv:2605. 27628v2 Announce Type: replace Abstract: As autonomous and agentic AI systems scale in robotic and human-machine environments, managing hallucination and persistent but unjustified action remains an open challenge.
By Srini Ramaswamy
The paper introduces the concept of Evolutionary Safety for recursive self-improving AI, focusing on how safety properties evolve as an AI system and its successors change. It identifies key risks such as intent drift, error accumulation, and safety-property erosion, and presents a taxonomy covering agent state, model state, evaluation, environment, and update mechanisms. The authors propose methods for discovering and evaluating evolutionary risks, and outline governance principles for modification, selection, authorization, provenance, and recovery, while highlighting open problems for maintaining safety in persistent, adaptive, and recursively self-improving systems.
By Chang Gong, Jingping Bi, Di Yao, Xinjian Liang, Chao Xiang, Ruijie Guo
arXiv:2602. 10429v2 Announce Type: replace-cross Abstract: AIvilization v0 is a publicly deployed large-scale artificial society that couples a resource-constrained sandbox with a unified LLM-agent architecture, aiming to sustain long-horizon autonomy while remaining executable under a rapidly changing environment.
By Wenkai Fan, Shurui Zhang, Xiaolong Wang, Haowei Yang, Tsz Wai Chan, Xingyan Chen, Junquan Bi, Zirui Zhou, Jia Liu, Kani Chen
arXiv:2512. 15044v2 Announce Type: replace Abstract: Integrated sensing and communication (ISAC) has emerged as a key development direction in the sixth-generation (6G) era, which provides essential support for the collaborative sensing and communication of future intelligent networks.
By Wenwen Xie, Geng Sun, Chuang Zhang, Xuejie Liu, Dong In Kim
The paper investigates whether large language model (LLM) agents can autonomously manage long‑horizon physical tasks without human intervention. It proposes a multi‑agent framework that combines planning, tool calling, observation, and verification, and tests it on agricultural tasks under varying weather conditions. Results show that zero‑shot LLM agents match reinforcement learning (RL) agents in the same environment and outperform RL when the environment shifts, suggesting a viable path for self‑adaptive physical AI.
By Varun Kaushik, Yayun Tan, Xiaofan Yu
The paper introduces epistemic memory, a validity-maintenance layer for intelligent systems that tracks when stored knowledge remains applicable. It formalizes a dynamic epistemic quotient and shows that fixed semantic representations inevitably incur error as epistemic boundaries shift. The authors propose Observable Belief Memory (OBM), which combines current epistemic quotients, belief over quotient classes, and within-class provenance, and demonstrate that explicit epistemic tracking improves robustness under changing observation conditions.
By Pin-Han Ho, Limei Peng, Yiming Miao, Yan Jiao
The paper introduces Env‑Rethink, a 27B post‑trained model system designed to help large language model agents better interact with complex, evolving environments. It builds Collection Maps and Event Logs to organize scattered information, uses offline trajectory learning to detect noise, and generates virtual event histories to evolve environments for more challenging tasks. Experiments show that Env‑Rethink improves downstream task performance by over 15.1% rubric pass rate across nine models on 30 tasks.
By Yukai Wu, Yuanjing Yang, Le Zhou, Shaokun Han, Haoyu Wang, Zirui Tang, Weihuang Zheng, Maxm Pan, Xuanhe Zhou, Fan Wu
arXiv:2607. 28629v1 Announce Type: new Abstract: The rapid transition from reactive large language models (LLMs) to persistent, action-capable systems has exposed critical gaps in the architectural understanding of Agentic AI, particularly in separating inference, orchestration, and execution layers for autonomous AI agents.
By Konstantinos I. Roumeliotis, Ranjan Sapkota
The paper proposes a developmental framework for autonomous artificial agents that emphasizes learning social norms and alignment through direct interaction with dynamic environments. It argues that intrinsic motivations such as curiosity and competence can guide exploration, but also complicate alignment with human goals. By drawing parallels to child development, the authors suggest that regulatory sandboxes serve as pedagogical spaces where agents gradually acquire moral agency and adapt their behaviors through experience and cooperation.
By Marica Notte, Ludovica Marinucci, Vieri Giuliano Santucci
arXiv:2608.30478v1 Announce Type: new
Abstract: Cognitive language agents have achieved substantial progress by equipping language models with memory, tools, and decision-making procedures, enabling...
By Shihan Dou, Haoxiang Jia, Shichun Liu, Feng Chen, Chenhao Huang, Yujiong Shen, Shaofan Liu, Jiayi Chen, Jiahang Lin, Honglin Guo, Qianyu He, Minghao Guo, Ziyi Ye, Pluto Zhou, Tao Gui, Qi Zhang, Xuanjing Huang
arXiv:2606. 26575v1 Announce Type: cross Abstract: Complex multi-agent control tasks remain challenging for traditional rule-based and model-based approaches, motivating the adoption of learning-based methods.
By Chenlong Liu, Zhuohui Zhang, Xinyan Chen, Zhipeng Wang, Bin Cheng, Bin He
arXiv:2609.07741v1 Announce Type: new
Abstract: Persistent AI assistants are intended to extend human attention, memory, and coordination across changing digital and physical environments. To be trul...
By Jo\~ao Dias Ferreira