arXiv AI By Avni Mittal

Triadic Werewolf: A Jester Role for Multi-Hop Theory of Mind in LLMs

Read the original on arXiv AI →

arXiv:2606. 27909v1 Announce Type: cross Abstract: Theory-of-mind evaluations of large language models typically use dyadic social-deduction games, where every observable cue points to a single hidden side, so a model with strong language priors can score well without ever simulating opponents' incentives.

Summary generated by The Flow from the publisher's feed. The full article lives at arXiv AI.