arXiv AI By Joshua Penman

Epistemic Goggles: A Pretrained Module that Induces an Epistemic Frame via Gradient Editing

Read the original on arXiv AI →

arXiv:2607. 01690v1 Announce Type: new Abstract: Finetuning a language model on documents that are explicitly annotated as fictional results in a model that still actually believes the documents' core claims, an effect known as Negation Neglect.

Summary generated by The Flow from the publisher's feed. The full article lives at arXiv AI.