arxiv
PublishedMay 21, 2026 at 4:00 AM
—neutral
Batched Single-Index Global Multi-Armed Bandits with Covariates
Publisher summary· verbatim
arXiv:2503.00565v3 Announce Type: replace-cross Abstract: The multi-armed bandits (MAB) framework is a widely used approach for sequential decision-making, where a decision-maker selects an arm in each round with the goal of maximizing long-term rewards. In many practical applications, such as perso
Stay posted· Newsletter
A 5-min weekly brief — top movers, price watch, story of the week.
Discussion
No replies yet. Be first.
Related coverage
More from ARXIV
arxivFederatedSkill: Federated Learning for Agentic Skill Evolution8harxivToward a Modular Architecture for Embedded AI Agent Systems at the Edge8harxivA Graph Foundation Model with Spectral Parsing and Prototype-Guided Spatial Propagation8harxivAnomalies in Multivariate Time Series Benchmarks Are Mostly Univariate8hThe Bubble Brief
WEEKLYRead machine-learning insights every Tuesday — top movers, new releases, story of the week.
Originally published on arxiv ↗