arXiv AI By Jongchan Park, Seungjun Oh, Seungho Baek, Yusung Kim

Learning Generalizable Skill Policy with Data-Efficient Unsupervised RL

Read the original on arXiv AI →

arXiv:2607. 00392v1 Announce Type: cross Abstract: Unsupervised Reinforcement Learning (URL) aims to pre-train scalable, skill-conditioned policies without extrinsic rewards, serving as a foundation for downstream control tasks.

Summary generated by The Flow from the publisher's feed. The full article lives at arXiv AI.