arXiv Machine Learning By Mengxiao Zhang

Toward Optimal Second-Order Path-Length Guarantee for Adversarial Multi-Armed Bandits

Read the original on arXiv Machine Learning →

arXiv:2608. 15996v1 Announce Type: new Abstract: We study second-order path-length regret in adversarial $K$-armed bandits against oblivious loss sequences.

Machine-generated by The Flow from the publisher's headline and feed description — not written or checked by a human. The full article lives at arXiv Machine Learning.