arXiv AI By Andreas Chouliaras, Luke Connolly, Dimitris Chatzpoulos

Themis: An explainable AI-enabled framework for Reinforcement Learning with Human Feedback

Read the original on arXiv AI →

arXiv:2606. 24622v1 Announce Type: new Abstract: Training safe Reinforcement Learning (RL) systems is inherently challenging, with no guarantee of avoiding unwanted behaviors.

Machine-generated by The Flow from the publisher's headline and feed description — not written or checked by a human. The full article lives at arXiv AI.

OpenAI Blog
Aug 3, 2017

Gathering human feedback

RL-Teacher is an open-source implementation of our interface to train AIs via occasional human feedback rather than hand-crafted reward functions. The underlying technique was developed as a step towards safe AI systems, but also applies to reinforcement learning problems with rewards that are hard to specify.