arXiv AI By Ali Asadi, Krishnendu Chatterjee, Ehsan Goharshady, Mehrdad Karrabi, Alipasha Montaseri, Carlo Pagano

Strongly Polynomial Time Complexity of Policy Iteration for $L_\infty$ Robust MDPs

Read the original on arXiv AI →

arXiv:2601. 23229v2 Announce Type: replace Abstract: Markov decision processes (MDPs) are a fundamental model in sequential decision making.

Summary generated by The Flow from the publisher's feed. The full article lives at arXiv AI.