arXiv:2606. 19883v1 Announce Type: new Abstract: We study a multi-agent multi-armed bandit problem in the competitive setup with two-sided matching markets under a human centric decision making model.
By Ananya Kunisetty, Avishek Ghosh
arXiv:2607. 04824v1 Announce Type: new Abstract: We study a sequential learning problem for stable matchings in two-sided markets where preferences on both sides are initially unknown.
By Andreas Athanasopoulos, Anne-Marie George, Christos Dimitrakakis
arXiv:2506. 03802v2 Announce Type: replace Abstract: We introduce a learning problem in a generalized two-sided matching market, where agents select actions to interact with their match.
By Andreas Athanasopoulos, Christos Dimitrakakis
arXiv:2607. 08979v1 Announce Type: new Abstract: We study the active learning problem of fixed-confidence top-$k$ identification from noisy pairwise comparisons.
By Motti Goldberger, Nils Rudi
arXiv:2604. 00531v2 Announce Type: replace Abstract: Multi-task representation learning exploits the shared structure among related tasks by learning a common latent representation, thereby improving sample efficiency.
By Jiabin Lin, Shana Moothedath
arXiv:2606. 04305v1 Announce Type: new Abstract: We study online learning with an additional offline dataset in the stochastic linear bandit setting.
By Kushagra Chandak, Toshinori Kitamura, Xiaoqi Tan