arXiv Machine Learning By Beatrix M. G. Nielsen, Emanuele Marconato, Luigi Gresele, Andrea Dittadi, Simon Buchholz

Logit Distance Bounds Representational Similarity

Read the original on arXiv Machine Learning →

arXiv:2602. 15438v3 Announce Type: replace Abstract: For a broad family of discriminative models that includes autoregressive language models, identifiability results imply that if two models induce the same conditional distributions, then their internal representations are equal up to an invertible linear transformation.

Summary generated by The Flow from the publisher's feed. The full article lives at arXiv Machine Learning.