arXiv AI By Debmalya Mandal, Andi Nika, Parameswaran Kamalaruban, Adish Singla, Goran Radanovi\'c

Corruption Robust Offline Reinforcement Learning with Human Feedback

Read the original on arXiv AI →

arXiv:2402. 06734v2 Announce Type: replace-cross Abstract: We study data corruption robustness for reinforcement learning with human feedback (RLHF) in an offline setting.

Summary generated by The Flow from the publisher's feed. The full article lives at arXiv AI.