arXiv Machine Learning By Di Yang Shi, W. Bradley Knox

A Framework for Designing Reward Functions: From Objectives to Features to Human-Aligned Reward Functions

Read the original on arXiv Machine Learning →

arXiv:2608. 12302v1 Announce Type: new Abstract: We present a formal process to enable non-experts to instantiate and iterate on human-aligned reward functions, i.

Summary generated by The Flow from the publisher's feed. The full article lives at arXiv Machine Learning.