arXiv Machine Learning By Hikaru Shindo, Yu Deng, Teng Cao, Quentin Delfosse, Christopher Tauchmann, Jannis Bl\"uml, Gopika Sudhakaran, Kristian Kersting

Learning Explicit Behavioral Models with Adaptive Questions and World-Model Probes

Read the original on arXiv Machine Learning →

arXiv:2606. 07127v1 Announce Type: new Abstract: Interactive agents trained only against task return can achieve high scores while failing to represent the mechanisms that make their actions succeed.

Summary generated by The Flow from the publisher's feed. The full article lives at arXiv Machine Learning.