arXiv AI By Georgii Aparin, Tatiana Gaintseva

A Geometric Account of Activation Steering through Angle-Norm Decomposition

Read the original on arXiv AI →

arXiv:2606. 06735v1 Announce Type: new Abstract: Linear activation steering has gained popularity as a simple and empirically effective way to control language model behavior.

Summary generated by The Flow from the publisher's feed. The full article lives at arXiv AI.