arXiv AI By Sewoong Lee, Marc E. Canby, Ikhyun Cho, Julia Hockenmaier

A Survey on the Linear Representation Hypothesis

Read the original on arXiv AI →

The Flow has not summarised this story yet — read it at arXiv AI.

arXiv AI
Sep 24

The Linear Representation Hypothesis Needs a Group Action

The paper argues that the Linear Representation Hypothesis (LRH) should not be treated as a single claim but as a family of claims differentiated by how representations are considered equivalent. It highlights that different equivalence notions preserve different structures, leading to metrics, probes, and interventions that may actually test distinct hypotheses. By formalizing these ideas with group actions, the authors provide a framework that clarifies how assumptions vary across metrics, reading points, and analysis stages, and they apply it to audit common representation quantities and recent interpretability analyses.

By Louie Hong Yao, Yuhao Li, Shengchao Liu