arXiv AI By Haolin Liu, Braham Snyder, Chen-Yu Wei

On the Complexity of Offline Reinforcement Learning with $Q^\star$-Approximation and Partial Coverage

Read the original on arXiv AI →

arXiv:2602. 12107v2 Announce Type: replace-cross Abstract: We study offline reinforcement learning under $Q^\star$-approximation and partial coverage, a setting that motivates practical algorithms such as Conservative $Q$-Learning (CQL; Kumar et al.

Summary generated by The Flow from the publisher's feed. The full article lives at arXiv AI.