Hugging Face Trending Papers

Heavy-Ball Q-Learning with Residual Weighting Correction

Read the original on Hugging Face Trending Papers →

This paper proposes a corrected heavy-ball Q-learning method for reinforcement learning (RL) and establishes its convergence. It also identifies conditions under which the method is theoretically guaranteed to converge faster than standard Q-learning.

Summary generated by The Flow from the publisher's feed. The full article lives at Hugging Face Trending Papers.