arXiv AI

Interpretation as Linear Transformation: A Cognitive-Geometric Model of Concepts and Meaning

arXiv:2512. 09831v2 Announce Type: replace Abstract: This paper develops a geometric framework for modeling concepts, motivation, and influence across cognitively heterogeneous agents.

arXiv AI
Jun 26

Radical AI Interpretability

arXiv:2606. 26523v1 Announce Type: new Abstract: We develop a framework for interpreting AI systems as agents, drawing on the philosophical tradition of radical interpretation and the tools of mechanistic interpretability.

By Daniel A. Herrmann, Benjamin A. Levinstein
arXiv AI
Jul 16

A Survey on Hypergame Theory: Modelling Misaligned Perceptions and Nested Beliefs for Multi-Agent Systems

arXiv:2507. 19593v3 Announce Type: replace Abstract: Classical game-theoretic models typically assume rational agents, complete information, and common knowledge of payoffs - assumptions that are often violated in real-world MAS characterized by uncertainty, misaligned perceptions, and nested beliefs.

By Vince Trencsenyi, Agnieszka Mensfelt, Kostas Stathis
arXiv AI
Aug 28

Toward a New Science of AI as Cognitive Infrastructure

The paper proposes a new interdisciplinary field called Cognitive Infrastructure Studies (CIS) to examine how AI systems act as invisible, foundational cognitive infrastructures that shape what people can know and do in digital societies. It argues that these infrastructures, through anticipatory personalization and adaptive invisibility, automate relevance judgments and shift epistemic agency to non‑human systems. CIS offers methodological tools, such as infrastructure breakdown experiments, to uncover the hidden cognitive dependencies created by AI preprocessing across individual, collective, and societal levels.

By Giuseppe Riva