arXiv AI By Max Weltevrede, Caroline Horsch, Matthijs T. J. Spaan, Wendelin B\"ohmer

Training on Irrelevant States Implies Data Augmentation: Generalization in Contextual MDPs

Read the original on arXiv AI →

arXiv:2410. 03565v4 Announce Type: replace-cross Abstract: In the zero-shot policy transfer (ZSPT) setting for contextual Markov decision processes (CMDP), agents train on a fixed, finite set of contexts and must generalize to new ones.

Summary generated by The Flow from the publisher's feed. The full article lives at arXiv AI.