arXiv Machine Learning By Omer Tariq, Syed Muhammad Raza, Jeongbae Son

SAD-LoRA: Spectral Alignment for Low-Rank Knowledge Distillation

Read the original on arXiv Machine Learning →

arXiv:2607. 04306v1 Announce Type: new Abstract: Distilling a fine-tuned teacher into a LoRA-adapted student is a standard recipe for parameter-efficient compression, but output-level KD does not explicitly control which rank-$r$ weight subspace the adapter occupies.

Summary generated by The Flow from the publisher's feed. The full article lives at arXiv Machine Learning.