arXiv AI

A Geometric View for Understanding Concept Learning and Neuron Interpretation in Sparse Autoencoders

arXiv:2606. 07007v1 Announce Type: cross Abstract: We propose a unified mathematical framework for a geometric understanding of concept learning and neuron interpretation in sparse autoencoders (SAEs).