Вход на сайт

Просмотр новости

Найдите то, что Вас интересует

The Role of Contextual Information in Best Arm Identification

Дата публикации: 17-08-2026 20:26:00


We study the best-arm identification problem with fixed confidence when contextual (covariate) information is available in stochastic bandits. In each round, we observe contextual information before selecting an arm. The distribution of the reward associated with the selected arm depends on the observed contextual information. We are interested in finding the arm with the maximum mean reward marginalized over the contextual distribution and not the mean reward conditioned on contexts. Our goal is to identify the best arm with a minimal number of samples under a given error probability. First, we derive the instance-specific sample-complexity lower bounds under the contextual information. Then, we propose a context-aware version of the Track-and-Stop strategy, wherein the proportions of arm draws track the set of optimal allocations, and prove that the expected number of arm draws asymptotically matches the lower bound. We demonstrate that the contextual information can be used to improve the efficiency of the identification of the best marginalized mean reward when compared with the results of Garivier and Kaufmann
(2016). Furthermore, we experimentally confirm that contextual information contributes to faster best-arm identification.

Схожие новости

#Наименование новостиТональностьИнформативностьДата публикации
1 Best Arm Identification with Minimal Regret 08.417-08-2026
2 Bayesian Inference of Contextual Bandit Policies via Empirical Likelihood 03.9717-08-2026
3 An Anytime Algorithm for Good Arm Identification 07.1717-08-2026
4 A Convex Framework for Confounding Robust Inference 05.4517-08-2026
5 Differentially Private Best-Arm Identification 012.1417-08-2026
6 Neural Exploitation and Exploration of Contextual Bandits 06.3417-08-2026
7 Cheap Bootstrap for Fast Uncertainty Quantification of Stochastic Gradient Descent 06.3817-08-2026
8 The Sample Complexity of Parameter-Free Stochastic Convex Optimization 05.717-08-2026
9 A Reinforcement Learning Approach in Multi-Phase Second-Price Auction Design 07.3417-08-2026
10 A Two-Timescale Primal-Dual Framework for Reinforcement Learning via Online Dual Variable Guidance 013.1217-08-2026

Классификация: Пресс-релизы. Схожих патентов: 0. Схожих новостей: 10. Тональность: 0. Информативность: 4.07. Источник: jmlr.org.