The paper introduces PAI‑Bench, a benchmark designed to evaluate persistent AI agents on how faithfully they adhere to a versioned identity contract. It separates several dimensions—recall, composition, behavioral enactment, resistance, persistence, lineage, and role‑conditioned updates—while keeping scoring oracles independent of the target process. Experiments on synthetic profiles show that explicit cues can significantly alter the presence of identity identifiers, revealing prompt‑dependent component selection and sensitivity to startup cues.
By Zhenyu Zhao, Roy Zhao
The study investigates how the visibility of speaker biographies to interlocutors during training, inference, and evaluation affects persona-based dialogue generation. It finds that training-time visibility is the primary factor determining whether models express persona traits or simply copy biographical text, and that providing interlocutor-biography visibility during training reduces target-biography copying. Additionally, asymmetric disclosure—where only the interlocutor sees the target biography—leads to more frequent leakage of target content into interlocutor turns, making such dialogues easier for a judge to identify.
By Daniela Occhipinti, Malvina Nissim, Marco Guerini
arXiv:2607. 15883v1 Announce Type: cross Abstract: Large language models are broadly capable, yet in sustained one-to-one conversation they still read as flat: competent, responsive, and somehow not quite the presence of a mind.
By Sebastian Cochinescu
Emergi-PersonaOS is a psychology‑grounded operating system designed to manage persona agents throughout their lifecycle. It structures personas into three layers—dispositional traits, characteristic adaptations, and narrative identity—allowing the system to infer current persona states from situational cues and generate appropriate actions. The OS records experiences, evaluates revision candidates, and controls belief updates through explicit review and traceable evidence, enabling controllable evolution of persona agents over long interactions.
By Haoluan Fu, Keni Chen, Xinyu Jia, Jinpeng Wang, Yuyu Yin
The study evaluates whether large language models (LLMs) used as synthetic personas can predict real audience responses to marketing copy. Using thousands of headline A/B tests from the Upworthy Research Archive, the authors compare a ten-persona panel grounded in real audience demographics to a no-persona zero‑shot baseline that asks the model for a typical reader’s click likelihood. Results show that the no‑persona baseline outperforms the persona‑based approach, with higher predictive validity and top‑1 accuracy, indicating that forcing the model to role‑play specific personas introduces bias and noise.
By Alexandre Cristov\~ao Maiorano
arXiv:2607. 28818v1 Announce Type: new Abstract: As AI companions increasingly mediate repeated social interaction, users may rely on a stable role and shared history, yet locally acceptable replies do not ensure that either persists.
By Pranav Narayanan Venkit, Akshara Prabhakar, Yu Li, Daniel Lee, Chien-Sheng Wu