arxiv
PublishedOctober 1, 2026 at 4:00 AM
Bandits with Multiple Optimal Arms: Minimax Regret and Non-Adaptivit
Publisher summary· verbatim
arXiv:2609.38659v1 Announce Type: cross Abstract: We study multi-armed bandits (MAB) with multiple optimal arms, motivated by the fact that many practical decision making problems admit multiple correct answers. For $K$-armed bandits with $A$ optimal arms, we first provide a sharper analysis of prev
Stay posted· Newsletter
A 5-min weekly brief — top movers, price watch, story of the week.
Discussion
No replies yet. Be first.
Related coverage
More from ARXIV
arxivSequential Capacity of Quantum Processes with Finite Memory1darxivGraph Representation via Elements of Discrete Morse and Cobordism Theories1darxivAF-Muon: An AdamW-Free Muon Optimizer for Tied-Embedding Models1darxivDo Your Own Research: Learning to Forecast by Learning to Search1dThe Bubble Brief
WEEKLYRead AI insights every Tuesday — top movers, new releases, story of the week.
Originally published on arxiv ↗