← Back to all news
Hugging Face Trending Papers September 21, 2026

The Answer-Basin Representation Hypothesis: We Are Not Probing or Steering Concepts

Read the original on Hugging Face Trending Papers →

The Flow has not summarised this story yet — read it at Hugging Face Trending Papers.

  • llms
  • safety

One email a morning, machine-written

One email a day, machine-written, one click to leave. We never share your address.

Related stories

arXiv Computation and Language
Sep 22

The Answer-Basin Representation Hypothesis: We Are Not Probing or Steering Concepts

arXiv:2609.24821v1 Announce Type: new Abstract: The Linear Representation Hypothesis associates high-level concepts with directions in language models, but it remains unclear how these concept-relate...

By Manjiang Yu, Hongji Li, Zihan Wang, Junwei Chen, Xue Li, Priyanka Singh, Yang Cao, Lijie Hu
llmssafety
More like this →
arXiv Machine Learning
Aug 5

Inverted Detection and Control in Steering Vectors

arXiv:2608. 02957v1 Announce Type: new Abstract: Steering vectors (SVs) are widely used to influence the expression of concepts (e.

By Max Torop, Aria Masoomi, Jennifer Dy
llms
More like this →
Hugging Face Trending Papers
Jul 5

Language Models Represent and Transform Concepts with Shared Geometry

How concepts are represented in neural networks is a fundamental question in machine learning. The dominant view treats concept representations as stationary geometric objects.

llms
More like this →
arXiv AI
Jul 7

Language Models Represent and Transform Concepts with Shared Geometry

arXiv:2607. 04525v1 Announce Type: cross Abstract: How concepts are represented in neural networks is a fundamental question in machine learning.

By Zhimin Hu, Lanhao Niu, Sashank Varma
llms
More like this →
arXiv Computation and Language
Sep 16

A Data-free Universal Prior over Syntactic Structures

arXiv:2609.16854v1 Announce Type: new Abstract: Probability is fundamental to theories of language comprehension, production, acquisition, and evolution, as well as to large language models. Existing...

By Ferm\'{\i}n Moscoso del Prado Mart\'{\i}n
llmssafety
More like this →
Hugging Face Trending Papers
Aug 26

How Unlikely Is "Unlikely"? Assessing Verbal Probability Perception Across Large Language Models

Large language models increasingly produce and interpret verbal probability expressions, yet whether these expressions carry consistent meaning across models (or match human perceptions of uncertainty...

llmsbenchmarkssafety
More like this →
About Pricing API Newsletter Sources Privacy Terms Refunds Accessibility Provider info Contact RSS

The Flow links to publishers and never republishes their articles. Summaries are machine-generated.

v1.1.0 · 5f852ea