arXiv AI By Yilun Hua, Evan Wang, Yoav Artzi

Post-training for Efficient Communication via Convention Formation

Read the original on arXiv AI →

arXiv:2508. 06482v2 Announce Type: replace-cross Abstract: Humans communicate with increasing efficiency in multi-turn interactions, by adapting their language and forming ad-hoc conventions.

Machine-generated by The Flow from the publisher's headline and feed description — not written or checked by a human. The full article lives at arXiv AI.

arXiv AI
2d ago

Predicting Steering Vectors and Adapter Weights for Few-Shot Author-Style Transfer

The paper investigates how to adapt large language models to mimic an individual author's style using only a few example abstracts, a task made harder by the formal nature of scientific writing. It proposes three style‑conditioning methods—contrastive activation steering, a network predicting steering vectors, and a hypernetwork predicting LoRA adapters—and finds that while fine‑tuning captures the strongest style signal, it harms fluency; the hypernetwork offers the best balance between style imitation and output quality for both seen and unseen authors. The authors also show that steering can be performed at the author level by contrasting author abstracts against style‑neutral generations for the same content, eliminating the need for a predefined style inventory and outperforming inventory‑based approaches. whyItMatters":"The study provides practical techniques for author‑style transfer in scientific writing, revealing a trade‑off between style fidelity and fluency and demonstrating that hypernetworks can effectively balance these aspects."

By Leonard Popp, Danni Liu, Supriti Sinhamahapatra, Jan Niehues
arXiv Machine Learning
Jul 9

Towards Understanding Steering Strength

arXiv:2602. 02712v2 Announce Type: replace Abstract: A popular approach to post-training control of large language models (LLMs) is the steering of intermediate latent representations.

By Magamed Taimeskhanov, Samuel Vaiter, Damien Garreau