arXiv Machine Learning By Yuanlong Chen

On the convergence of optimistic policy iteration for stochastic shortest path problem

Read the original on arXiv Machine Learning →

arXiv:1808. 08763v3 Announce Type: replace Abstract: In this paper, we prove some convergence results of a special case of optimistic policy iteration algorithm for stochastic shortest path problem.

Machine-generated by The Flow from the publisher's headline and feed description — not written or checked by a human. The full article lives at arXiv Machine Learning.