arxivcs.LG2026-06-28
A Linear Matching Bandit Approach to Online Multi-Human Multi-Robot Teaming
Yaohui Guo, X. Jessie Yang, Cong Shi
We address the problem of online multi-human multi-robot teaming through the lens of a linear matching bandit framework, where a learner assigns robots with unknown features from a fixed pool to distinct sets of human agents over multiple rounds. To solve this problem, we propose…