arXiv:2607. 00002v1 Announce Type: new Abstract: Moral cognition has traditionally been modeled as adherence to fixed ethical theories--deontology, consequentialism, virtue ethics--implemented as static rules or value functions.
By Max Kanwal, Caryn Tran, Patrick Mineault
The paper argues that AI alignment depends on a system’s ability to exhibit a coherent moral policy—stable, monotonic, decisive, and Pareto‑viable—rather than on any specific moral standard. The authors test nine large language models across varied moral scenarios and find that none maintain consistent verdicts, with surface‑form changes causing up to 99% shifts in outcomes. This indicates that current LLM agents lack the structural moral competence required for meaningful alignment.
By Arno Libert, Derck W. E. Prinzhorn, Daan R. Henselmans
arXiv:2606. 11232v1 Announce Type: cross Abstract: Existing LLM moral benchmarks usually ask which isolated moral act, value, or foundation a model prefers.
By Weijia Zhang, Ruiqi Chen, Yunze Xiao, Weihao Xuan
arXiv:2603.23114v2 Announce Type: replace
Abstract: A human's moral decision depends heavily on the context. Yet research on LLM morality has largely studied fixed scenarios. We address this gap by i...
By Adrian Sauter, Mona Schirmer
arXiv:2604. 24155v3 Announce Type: replace-cross Abstract: The project of aligning machine behavior with human values raises a basic problem: whose moral expectations should guide AI decision-making?
By Benjamin Minhao Chen, Xinyu Xie
arXiv:2608. 14522v1 Announce Type: new Abstract: As AI systems make more morally loaded decisions across society, one response has been moral preference elicitation.
By Taenyun Kim, Edyta Bogucka, Daniele Quercia
arXiv:2608. 15354v1 Announce Type: new Abstract: LLMs are increasingly used in morally sensitive contexts, yet it is unclear whether they apply ethical principles consistently across situations.
By Pegah Nokhiz, Aravinda Kanchana Ruwanpathirana, Helen Nissenbaum
arXiv:2608. 12368v1 Announce Type: new Abstract: Agreement with human judgments is a common proxy for evaluating the alignment of large language models (LLMs).
By Octavian M. Machidon, Alina L. Machidon, Vojko Strahovnik, Mateja Centa Strahovnik, Jonas Miklav\v{c}i\v{c}, Marko Robnik \v{S}ikonja
The paper reports the first empirical study comparing how humans and large language models (LLMs) evaluate perceived moral agency (PMA) in both human and autonomous artificial agents within smart city scenarios. Using a validated PMA scale, 190 human participants and various LLMs were assessed, revealing that humans are perceived to have higher moral agency than artificial agents. When confronted with moral dilemmas, LLMs focus on situational factors such as harm severity and urgency, mirroring the context‑sensitivity observed in human raters.
By Fernanda Mansilla, Aloysius Tok, Bahia Guella\"i, Farah Benamara, Nancy F. Chen
arXiv:2606. 31213v1 Announce Type: cross Abstract: As large language models (LLMs) are increasingly deployed as moral advisors and agents, they need to address dilemmas between two competing values.
By Jongchan Choi, Nari Yang, Sung Soo Park, Jaemin Cho, Han Seoyoung, Haerin Shin, Jun-Hyung Park
arXiv:2609.21992v1 Announce Type: new
Abstract: Most work in computational ethics treats annotator disagreement on moral content as noise to be voted away, collapsed into majority vote or the more pe...
By Maciej Skorski
The article argues that current evaluations of large language models’ moral competence focus mainly on whether outputs align with human moral values—the so‑called moral value problem—while neglecting the moral norm problem, which concerns the models’ ability to identify and apply context‑sensitive moral norms. It attributes this imbalance to the field’s reliance on descriptive ethics frameworks that emphasize value representation over normative application. The authors review existing benchmarks, highlight three gaps—lack of ground‑truth norm data, insufficient evaluation of intermediate reasoning, and limited focus on context‑relevant features—and propose a research agenda to develop formal normative representations, expert‑annotated datasets, and evaluation protocols that distinguish between value‑level and norm‑level competence.
By Aidan Kierans, Ritam Dutt, Kaley Rittichier, Shiri Dori-Hacohen, Avijit Ghosh