arXiv AI By Zhiren Gong, Zihao Zeng, Chau Yuen, Wei Yang Bryan Lim

Conditional Co-Ablation: Recovering Self-Repair Backups in Transformer Circuits

Read the original on arXiv AI →

arXiv:2607. 01940v1 Announce Type: cross Abstract: Mechanistic interpretability often relies on component-level interventions to discover how a model produces a behavior.

Summary generated by The Flow from the publisher's feed. The full article lives at arXiv AI.