arXiv:2608. 14565v1 Announce Type: new Abstract: AI safety research has mainly focused on two areas: technical alignment (ensuring AI systems produce human-aligned outputs) and the regulation of generative AI's societal impacts (including unemployment risk and labor market disruption).
By Jaeho Kim, Seokhyun Lee, Jieun Lee, Changhee Lee
arXiv:2609.05749v1 Announce Type: new
Abstract: Work on the risks of artificial intelligence has focused predominantly on capability risk: the danger that systems become too powerful, too autonomous,...
By Emilio Barkett, Alexander Kimpton, Daniel Graham, Yusuf Kundgol
arXiv:2606. 12442v2 Announce Type: replace-cross Abstract: At present, loss of control risks have gained much prominence in public discussion, particularly in relation to AI, with extensive discourse present among academics, frontier labs, and even governments.
By Ze Shen Chin, Maurice Chiodo, Dennis M\"uller, Coleman Snell
arXiv:2606. 12442v1 Announce Type: cross Abstract: At present, loss of control risks have gained much prominence in public discussion, particularly in relation to AI, with extensive discourse present among academics, frontier labs, and even governments.
By Ze Shen Chin, Maurice Chiodo, Dennis M\"uller, Coleman Snell
arXiv:1912. 08786v3 Announce Type: replace-cross Abstract: Three generations of software have transformed the role of artificial intelligence in society.
By Thomas Bartz-Beielstein
The paper argues that deploying generative AI agents requires more than isolated task success; they must remain useful across repeated interactions, changing conditions, and dependencies on people within shared workflows. The authors introduce two complementary evaluation aspects—operational resilience and considerate participation—to assess how agents recover from blocked work, communicate limits, and adapt to affected people and role boundaries. Using 120 simulated healthcare trajectories across two AI models and twelve stakeholder-derived tasks under varying challenge levels, the study finds that agents shift toward greater human dependence and increased workload as challenge accumulates, while also broadening from task-focused adaptation to task reframing and wider coordination.
By Yuanchen Bai, Zijian Ding, Angelique Taylor
arXiv:2607. 19292v1 Announce Type: cross Abstract: Current AI safety discourse still focuses disproportionately on visible failures, including obvious harms, dramatic misuse, and hypothetical catastrophic scenarios.
By Gjergji Kasneci, Enkelejda Kasneci
arXiv:2607. 03215v1 Announce Type: cross Abstract: Artificial intelligence has spread across the whole of the security lifecycle.
By Mohamed Chahine Ghanem
OpenAI is investing in stronger safeguards and defensive capabilities as AI models become more powerful in cybersecurity. We explain how we assess risk, limit misuse, and work with the security community to strengthen cyber resilience.
arXiv:2607. 08285v1 Announce Type: new Abstract: Current AI evaluation frameworks focus primarily on technical performance, including accuracy, robustness, reasoning ability, and policy compliance.
By Marcos Economides, Paul M. Sacher, Samuel Salzer, Alexis Michelle Abellar, Fendi Tsim, Antoine Ferr\`ere
arXiv:2607. 21268v1 Announce Type: cross Abstract: In many social-science research tasks, such as economics, LLM-based agents must produce outputs for which no cheap, task-complete, machine-readable correctness signal exists.
By Chen Zhu, Xiaolu Wang, Weilong Zhang
arXiv:2606. 18259v1 Announce Type: cross Abstract: AI agents that plan, retain memory across sessions, invoke external tools and act with partial autonomy are transforming human--AI collaboration.
By Junjie Xu, Xingjiao Wu, Zihao Zhang, Yujia Xu, Yuzhe Yang, Jin Zhu, Luwei Xiao, Wen Wu, Liang He