arXiv Machine Learning By Deqi Zheng, Xiaoyang Xu, Yuhong Yang

Multi-Armed Bandits with Arriving Arms: Sequential Screening, Dynamic Regret, and Sublinear Guarantees

Read the original on arXiv Machine Learning →

arXiv:2606. 09002v1 Announce Type: cross Abstract: We study a stochastic multi-armed bandit problem in which the set of available arms expands over time.

Summary generated by The Flow from the publisher's feed. The full article lives at arXiv Machine Learning.