arXiv AI By Xuefeng Liu, Mingxuan Cao, Qinan Huang, Thomas Brettin, Rick Stevens, Le Cong

Active-GRPO: Adaptive Imitation and Self-Improving Reasoning for Molecular Optimization

Read the original on arXiv AI →

arXiv:2607. 00531v1 Announce Type: cross Abstract: Scientific reasoning is an increasingly important capability of large language models, yet improving the robustness and efficiency of training such reasoning remains a key open challenge.

Summary generated by The Flow from the publisher's feed. The full article lives at arXiv AI.