arXiv AI By Andreas Chouliaras, Luke Connolly, Dimitris Chatzpoulos

Themis: An explainable AI-enabled framework for Reinforcement Learning with Human Feedback

Read the original on arXiv AI →

arXiv:2606. 24622v1 Announce Type: new Abstract: Training safe Reinforcement Learning (RL) systems is inherently challenging, with no guarantee of avoiding unwanted behaviors.

Summary generated by The Flow from the publisher's feed. The full article lives at arXiv AI.

OpenAI Blog
Aug 3, 2017

Gathering human feedback

RL-Teacher is an open-source implementation of our interface to train AIs via occasional human feedback rather than hand-crafted reward functions. The underlying technique was developed as a step towards safe AI systems, but also applies to reinforcement learning problems with rewards that are hard to specify.