← Back to all news
arXiv Machine Learning September 15, 2026 By Mengxiao Zhang

Toward Optimal Switching Regret for Multi-Armed Bandits with Oblivious Adversary

Read the original on arXiv Machine Learning →

The Flow has not summarised this story yet — read it at arXiv Machine Learning.

  • safety

One email a morning, machine-written

One email a day, machine-written, one click to leave. We never share your address.

Related stories

arXiv Machine Learning
Aug 26

An Efficient Minimax-Optimal Algorithm for Adversarial $m$-Set Bandits

arXiv:2608. 12231v2 Announce Type: replace Abstract: We study adversarial combinatorial bandits with $m$-set actions, where at each round the learner selects $m$ out of $d$ items and observes only the aggregate loss of the selected items.

By Francesco Bacchiocchi, Tommaso Cesari, Roberto Colomboni
safety
More like this →
arXiv Machine Learning
Jun 29

Learning in Markovian bandits with non-observable states and constrained decision epochs

arXiv:2606. 27448v1 Announce Type: new Abstract: This paper studies the problem of regret minimization in Markovian bandits with \emph{non-observable states} and possibly \emph{constrained} decision epochs.

By Thomas Hira, Victor Boone, Urtzi Ayesta, Ina Maria Verloop
reinforcement-learningbenchmarkssafety
More like this →
arXiv Machine Learning
Aug 11

Tracking the Best Strategy in an Extensive-Form Game

arXiv:2608. 09501v1 Announce Type: new Abstract: We consider the extensive-form bandit problem where on each trial the learner plays an extensive-form game against an oblivious adversary.

By Stephen Pasteris, Rahul Savani, Theodore Turocy
reinforcement-learning
More like this →
arXiv Machine Learning
Aug 4

A Perturbation Approach to Unconstrained Linear Bandits

arXiv:2603. 28201v3 Announce Type: replace Abstract: We revisit the standard perturbation-based approach of Abernethy et al.

By Andrew Jacobsen, Dorian Baudry, Shinji Ito, Nicol\`o Cesa-Bianchi
reinforcement-learningsafety
More like this →
arXiv Machine Learning
Jul 1

A Complete Characterization of Learnability for Adversarial Noisy Bandits

arXiv:2605. 09200v2 Announce Type: replace Abstract: We study adversarial noisy bandits given a known function class $\mathcal{F}$.

By Steve Hanneke, Kun Wang
reinforcement-learningsafety
More like this →
arXiv Machine Learning
Aug 13

An Efficient Near-Optimal Algorithm for Adversarial $m$-Set Bandits

arXiv:2608. 12231v1 Announce Type: new Abstract: We study adversarial combinatorial bandits with $m$-set actions, where at each round the learner selects $m$ out of $d$ items and observes only the aggregate loss of the selected items.

By Francesco Bacchiocchi, Tommaso Cesari, Roberto Colomboni
safety
More like this →
About Pricing API Newsletter Sources Privacy Terms Refunds Accessibility Provider info Contact RSS

The Flow links to publishers and never republishes their articles. Summaries are machine-generated.

v1.1.0 · 5f852ea