arXiv AI By Andreas Einwiller, Max Klabunde, Florian Lemmerich

AuAu: A Benchmark for Auditing Authoritarian Alignment in Large Language Models

Read the original on arXiv AI →

arXiv:2606. 16127v1 Announce Type: cross Abstract: The worldwide surge of authoritarianism, combined with the increasing central role in users' everyday lives, raises the question of to what extent specific models exhibit or promote authoritarian attitudes and characteristics.

Summary generated by The Flow from the publisher's feed. The full article lives at arXiv AI.