arXiv Machine Learning By Junpei Komiyama, Shinji Ito, Yuichi Yoshida, Souta Koshino

Replicability is Asymptotically Free in Multi-armed Bandits

Read the original on arXiv Machine Learning →

arXiv:2402. 07391v3 Announce Type: replace-cross Abstract: We consider a replicable stochastic multi-armed bandit algorithm that ensures, with high probability, that the algorithm's sequence of actions is not affected by the randomness inherent in the dataset.

Machine-generated by The Flow from the publisher's headline and feed description — not written or checked by a human. The full article lives at arXiv Machine Learning.