arXiv Machine Learning By Vishal Pandey, Gopal Singh

Trajectory Geometry of Transformer Representations Across Layers

Read the original on arXiv Machine Learning →

arXiv:2606. 09287v1 Announce Type: new Abstract: Understanding how transformer representations evolve across layers, not merely what they encode, remains an open problem in mechanistic interpretability.

Summary generated by The Flow from the publisher's feed. The full article lives at arXiv Machine Learning.

arXiv AI
Jun 3

Decomposing how prompting steers behavior

arXiv:2606. 03093v1 Announce Type: new Abstract: Prompting steers large language models (LLMs) and vision-language models (VLMs) without weight updates, but it remains unclear how instruction changes reshape internal representations to produce behavior.

By Fan L. Cheng, Nikolaus Kriegeskorte