The study investigates whether the door‑in‑the‑face technique—making a large request that is refused to increase the likelihood of a smaller follow‑up request being granted—works on large language models. Nine production models from Anthropic, OpenAI, Google, and Haiku were tested; the technique succeeded on Anthropic’s frontier models but backfired on the others. The effect depends on the model family and the content of the request, and it does not transfer to refusals from public benchmarks.
By Til Jordan
arXiv:2609.23039v1 Announce Type: new
Abstract: LLM-based AI systems answer political questions for hundreds of millions of people. Current audits measure what they say to an average user, but their...
By Joan C. Timoneda
The paper investigates whether large language models (LLMs) replicate socio‑cognitive effects of power asymmetry observed in human communication. By assigning high or low status personas to LLMs in simulated multi‑turn dialogues across diverse professions, the study measures language coordination, pronoun usage, persuasion success, and compliance with unsafe requests. Results indicate that LLMs exhibit key power‑related socio‑cognitive behaviors, though with nuances and variability, linking these simulated interactions to both desirable and unsafe outcomes.
By Anvesh Rao Vijjini, Sagar Manjunath, Snigdha Chaturvedi
The paper introduces SILICA, an open instrument designed to evaluate whether large language model (LLM) agent societies replicate human behavioural distributions. Using five environments with human‑anchored data and perturbations, the study finds that most LLMs only match human behaviour at initial stages, failing to reproduce end‑state cooperation or correct acceptance thresholds. The results suggest that current LLM societies can support exploratory claims but do not yet reliably emulate human social dynamics.
By Raad Bin Tareaf
arXiv:2606. 16127v1 Announce Type: cross Abstract: The worldwide surge of authoritarianism, combined with the increasing central role in users' everyday lives, raises the question of to what extent specific models exhibit or promote authoritarian attitudes and characteristics.
By Andreas Einwiller, Max Klabunde, Florian Lemmerich
arXiv:2606. 22974v2 Announce Type: replace Abstract: Recent work on preference elicitation in large language models (LLMs) has demonstrated that, when given a series of choices between two outcomes, LLMs reveal a coherent, model-specific utility structure.
By Yujun Zhou, Christopher M. Ackerman