arXiv Machine Learning By Seyeon Kim, Joonhun Lee, Namhoon Cho, Sungjun Han, Wooseop Hwang

Generalized Gaussian Temporal Difference Error for Uncertainty-aware Reinforcement Learning

Read the original on arXiv Machine Learning →

arXiv:2408. 02295v4 Announce Type: replace Abstract: Conventional uncertainty-aware temporal difference (TD) learning often models TD errors as zero-mean Gaussian.

Summary generated by The Flow from the publisher's feed. The full article lives at arXiv Machine Learning.