arXiv:2606. 20858v2 Announce Type: replace Abstract: The temporal structure of reward composition in reinforcement learning (RL) is typically hand-designed and held fixed throughout training, leaving the progression of motivational priorities largely unexplored.
By Alan Nadelsticher Ruvalcaba
arXiv:2608. 10323v1 Announce Type: new Abstract: Competitive artificial-life systems can rank trained controllers differently under training and ecological evaluation.
By Yuxu Ge, Yifei Cheng
arXiv:2609.00129v1 Announce Type: cross
Abstract: The performance of artificial intelligence (AI) and machine learning (ML) models degrades when the problem they were trained on drifts. This is a nea...
By J. M. Diederik Kruijssen (Allora Foundation)
arXiv:2609.14418v1 Announce Type: cross
Abstract: Dynamic multi-mode resource-constrained project scheduling requires decisions to be made under precedence constraints, limited resources, multiple ex...
By Yuan Tian, Yi Mei, Mengjie Zhang
The paper proposes a single variational principle that explains how gating mechanisms, their dynamics, and neural implementations for behavioral composition can be unified. This principle yields softmax gating, an energy‑based dynamical system with guaranteed convergence, and a recurrent neural network model with context‑dependent, local interactions. Experiments across collective behavior, human decision‑making, and layered control show that the mechanism reproduces known behavioral patterns, offers interpretable accounts of behavior combination, and matches or outperforms existing methods.
By Francesca Rossi, Veronica Centorrino, Francesco Bullo, Giovanni Russo
arXiv:2607. 21971v1 Announce Type: new Abstract: Test-time scaling through iterative self-evolution with environment feedback, as demonstrated by AlphaEvolve, shows remarkable performance gains.
By Shujin Wu, Cheng Qian, Xiusi Chen, Heng Ji