← Back to all news
arXiv Machine Learning October 7, 2026 By Nils Grandien, David Steinmann, Felix Friedrich, Kristian Kersting

Do Sparse Autoencoders Learn Meaningful Concept Hierarchies?

Read the original on arXiv Machine Learning →

The Flow has not summarised this story yet — read it at arXiv Machine Learning.

One email a morning, machine-written

One email a day, machine-written, one click to leave. We never share your address.

Related stories

arXiv AI
Jun 16

Cascaded Sparse Autoencoders Learn Multi-Level Visual Concepts in Multimodal LLMs

arXiv:2606. 16193v1 Announce Type: cross Abstract: Multimodal Large Language Models (MLLMs) have demonstrated strong performance on vision-language tasks, yet their internal visual representations remain difficult to interpret.

By Yusong Zhao, Hengyi Wang, Tanuja Ganu, Akshay Nambi, Hao Wang
llmsmultimodalbenchmarkssafety
More like this →
arXiv AI
1d ago

SAE++: Cascaded Sparse Autoencoders Learn Multi-Level Visual Concepts in Multimodal LLMs

arXiv:2606.16193v2 Announce Type: replace-cross Abstract: Multimodal Large Language Models (MLLMs) have demonstrated strong performance on vision-language tasks, yet their internal visual representat...

By Yusong Zhao, Hengyi Wang, Tanuja Ganu, Akshay Nambi, Hao Wang
llmsmultimodalbenchmarkssafety
More like this →
arXiv AI
Jun 9

A Geometric Unification of Concept Learning with Concept Cones

arXiv:2512. 07355v2 Announce Type: replace Abstract: Two traditions of interpretability have evolved side by side but seldom spoken to each other: Concept Bottleneck Models (CBMs), which prescribe what a concept should be, and Sparse Autoencoders (SAEs), which discover what concepts emerge.

By Alexandre Rocchi, Thomas Fel, Gianni Franchi
efficiencysafety
More like this →
arXiv Computer Vision
Oct 1

Towards Open-Ended Visual Scientific Discovery with Sparse Autoencoders

arXiv:2511.17735v2 Announce Type: replace Abstract: Foundation models in several scientific domains, including visual domains, learn representations that capture complex semantics from their respecti...

By Jacob Beattie, Samuel Stevens, Neil Rosser, Yu Su, Tanya Berger-Wolf
More like this →
arXiv Machine Learning
Jun 5

Concept-SAE: A Controllable and Invertible Concept Interface for Sparse Autoencoders

arXiv:2509. 22015v2 Announce Type: replace Abstract: Standard Sparse Autoencoders (SAEs) excel at discovering a dictionary of a model's learned features, providing a powerful lens for passive feature discovery.

By Jianrong Ding, Muxi Chen, Chenchen Zhao, Qiang Xu
safety
More like this →
arXiv AI
Jul 2

Identifying Latent Concepts and Structures for Generalized Category Discovery

arXiv:2607. 00620v1 Announce Type: cross Abstract: Generalized Category Discovery (GCD) aims to recognize known classes while autonomously discovering novel ones in open-world settings.

By Boyang Dai, Chaoqi Chen, Yizhou Yu
ragsafety
More like this →
About Pricing API Newsletter Sources Privacy Terms Refunds Accessibility Provider info Contact RSS

The Flow links to publishers and never republishes their articles. Summaries are machine-generated.

v1.1.0 · 5f852ea