arXiv Machine Learning By Chester Tan, Moritz Lampert, Courtney Maynard, Ankit Ramakrishnan, Tina Eliassi-Rad, Ingo Scholtes

Can Graph Learning Learn Circuits?

Read the original on arXiv Machine Learning →

arXiv:2608. 08536v1 Announce Type: new Abstract: Circuit localization is a mechanistic interpretability task whose goal is to identify a sparse subgraph of a transformer's computation graph sufficient to reproduce a particular behavior.

Summary generated by The Flow from the publisher's feed. The full article lives at arXiv Machine Learning.