arXiv Machine Learning By Xiang Yuan, Kaiqing Lei, Zhenyu Jin, Jun Shu, Deyu Meng, Zongben Xu

Harnessing the Potential of Optimizing Data Mixtures via Bayesian Domain Reweighting

Read the original on arXiv Machine Learning →

arXiv:2607. 27928v1 Announce Type: new Abstract: The performance of Large Language Models (LLMs) is fundamentally influenced by the distributional composition of multi-domain pre-training data.

Summary generated by The Flow from the publisher's feed. The full article lives at arXiv Machine Learning.