arXiv AI By Zhixuan Li, Jiangan Yuan, Han Xu

Data and Evaluation Closed-Loop for Model Capability Enhancement

Read the original on arXiv AI →

arXiv:2606. 28471v1 Announce Type: new Abstract: Model capability is the central variable in LLM pre-training, yet is never observed directly: data shapes it prospectively, while evaluation reveals it only retrospectively, compressing samples, prompts, decoding, and scoring rules into one noisy score.

Summary generated by The Flow from the publisher's feed. The full article lives at arXiv AI.