arXiv AI By Ziyun Qiao, Yue Min, Ruining Chen, Yujun Li

HERMES: A Multi-Granularity Labeling Substrate for Pre-training Data Mixtures

Read the original on arXiv AI →

arXiv:2607. 02266v1 Announce Type: cross Abstract: Most data-mixing methods assume the corpus has already been partitioned into groups, and the choice of those groups determines what a mixer can express.

Summary generated by The Flow from the publisher's feed. The full article lives at arXiv AI.