arXiv Machine Learning By Kotaro Yoshida, Laura Gomezjurado Gonzalez, Yukinori Yamamoto, Yuji Naraki, Ryotaro Shimizu, Wenya Wang

Looking in the Mirror: Introspecting Side-Effect Misalignments Induced by Fine-Tuning

Read the original on arXiv Machine Learning →

arXiv:2608. 04347v1 Announce Type: new Abstract: Fine-tuning enables a source model to acquire desired capabilities and behaviors in a target domain while retaining much of its general-purpose competence.

Summary generated by The Flow from the publisher's feed. The full article lives at arXiv Machine Learning.