arXiv AI By Majid Ghasemi, Mark Crowley

Learning When to Trust in Contextual Social Bandits

Read the original on arXiv AI →

arXiv:2603. 13356v2 Announce Type: replace Abstract: Robust reinforcement learning typically assumes that feedback sources are either globally trustworthy or corrupted within a fixed global budget.

Summary generated by The Flow from the publisher's feed. The full article lives at arXiv AI.