arXiv Machine Learning By Fredy Pokou (CRIStAL)

Low-Complexity Policy Tessellations in Structured Markov Decision Processes

Read the original on arXiv Machine Learning →

arXiv:2606. 25593v1 Announce Type: new Abstract: We study optimal-policy geometry in structured Markov decision processes.

Machine-generated by The Flow from the publisher's headline and feed description — not written or checked by a human. The full article lives at arXiv Machine Learning.

arXiv Machine Learning
1d ago

Towards Optimal Policy Improvement

The paper introduces a framework for optimal policy improvement in reinforcement learning, defining it as the best single update under given constraints. It shows that restricting improvement to a subset of states is equivalent to solving an induced Markov Decision Process, linking planning with explicit or implicit models to optimal policy improvement. The authors develop a novel operator for greedification under approximate evaluation, demonstrating empirical gains across several RL algorithms and settings.

By Yaniv Oren, Viliam Vadocz, Wiktor Zabka, Thomas Evers, Jan Robine, Wendelin B\"ohmer, Matthijs T. J. Spaan, Martha White, Hendrik Baier, Fenghui Yu
arXiv Machine Learning
Sep 17

A Geometric Theory of Decision Boundaries in Structured Markov Decision Processes

The paper develops a geometric theory of decision boundaries for structured Markov Decision Processes, treating the geometry induced by optimal policies as the key analytical object. It shows that, under structural regularity, this geometry yields the minimal representation needed for policy reconstruction and dictates the statistical and computational complexity of the reconstruction problem. The authors introduce intrinsic notions of boundary and decision complexity, derive information-theoretic measures of decision compression, and provide statistical guarantees for boundary estimation and policy reconstruction from black-box queries, supported by controlled numerical experiments.

By Fredy Pokou (MRE, INOCS)