arXiv Machine Learning By Yaohui Guo, X. Jessie Yang, Cong Shi

A Linear Matching Bandit Approach to Online Multi-Human Multi-Robot Teaming

Read the original on arXiv Machine Learning →

arXiv:2606. 29221v1 Announce Type: new Abstract: We address the problem of online multi-human multi-robot teaming through the lens of a linear matching bandit framework, where a learner assigns robots with unknown features from a fixed pool to distinct sets of human agents over multiple rounds.

Summary generated by The Flow from the publisher's feed. The full article lives at arXiv Machine Learning.