arXiv Machine Learning By Shuzhong Lai, Junhong Lai, Chenxi Li, Qing Zhou, Haifeng Li, Gang Pan, Lin Yao, Yueming Wang

TD-DPO: Difference-Aware Preference Optimization for Mitigating Sycophancy in Clinical Autism Intervention Dialogue

Read the original on arXiv Machine Learning →

arXiv:2607. 18304v1 Announce Type: new Abstract: The sycophancy of large language models can increase the safety risk in intervention dialogue for autistic children.

Summary generated by The Flow from the publisher's feed. The full article lives at arXiv Machine Learning.