The Linear Representation Hypothesis Needs a Group Action
Abstract
To make claims about representations that generalize beyond a particular trained model, we need to specify when two representations should count as equivalent. The Linear Representation Hypothesis is often discussed without making this equivalence explicit. Different notions of equivalence preserve different structures, so metrics, probes, and interventions that appear to study the same representation may in fact correspond to different hypotheses. We therefore argue that the Linear Representation Hypothesis is not one hypothesis but a family of claims distinguished by representation equivalence. We formalize this idea using group actions, specifying the representation object, the procedure that produces it, and the property ultimately asserted, while accounting for equivalences imposed by the model architecture. This framework clarifies how assumptions can change across metrics, reading points, and analysis stages, and we use it to audit common representation quantities and recent interpretability analyses.
Community
This paper argues that the Linear Representation Hypothesis needs an explicit notion of representation equivalence. Feedback welcome!
This is an automated message from the Librarian Bot. I found the following papers similar to this paper.
The following papers were recommended by the Semantic Scholar API
- A Survey on the Linear Representation Hypothesis (2026)
- The Answer-Basin Representation Hypothesis: We Are Not Probing or Steering Concepts (2026)
- InfluenceField: A Differentiable Field with Interventionally Identifiable Causal Structure for Multimodal World Modeling (2026)
- Map Users and Mapmakers: The Scope of Cognitive Attribution from Acquired Representations (2026)
- Beyond expressiveness in pairwise and higher-order models (2026)
- Xeno-Interpretability: Investigating the Alien Minds of LLMs (2026)
- A Unifying Perspective on Causal World Models: From Observations to Representations to Structure (2026)
Please give a thumbs up to this comment if you found it helpful!
If you want recommendations for any Paper on Hugging Face checkout this Space
You can directly ask Librarian Bot for paper recommendations by tagging it in a comment: @librarian-bot recommend
Models citing this paper 0
No model linking this paper
Datasets citing this paper 0
No dataset linking this paper
Spaces citing this paper 0
No Space linking this paper
Collections including this paper 0
No Collection including this paper