arXiv Machine Learning By Sourav Chakraborty, Amit Kiran Rege, Claire Monteleoni, Lijun Chen

Multi-Agent Lipschitz Bandits

Read the original on arXiv Machine Learning →

arXiv:2602. 16965v2 Announce Type: replace Abstract: We study the decentralized multi-player stochastic bandit problem over a continuous, Lipschitz-structured action space where hard collisions yield zero reward.

Summary generated by The Flow from the publisher's feed. The full article lives at arXiv Machine Learning.