ResearchTrend.AI
  • Papers
  • Communities
  • Events
  • Blog
  • Pricing
Papers
Communities
Social Events
Terms and Conditions
Pricing
Parameter LabParameter LabTwitterGitHubLinkedInBlueskyYoutube

© 2025 ResearchTrend.AI, All rights reserved.

  1. Home
  2. Papers
  3. 2502.03061
59
2

Optimal Best Arm Identification with Post-Action Context

5 February 2025
Mohammad Shahverdikondori
Amir Mohammad Abouei
Alireza Rezaeimoghadam
Negar Kiyavash
ArXiv (abs)PDFHTML
Abstract

We introduce the problem of best arm identification (BAI) with post-action context, a new BAI problem in a stochastic multi-armed bandit environment and the fixed-confidence setting. The problem addresses the scenarios in which the learner receives a post-action context\textit{post-action context}post-action context in addition to the reward after playing each action. This post-action context provides additional information that can significantly facilitate the decision process. We analyze two different types of the post-action context: (i) non-separator\textit{non-separator}non-separator, where the reward depends on both the action and the context, and (ii) separator\textit{separator}separator, where the reward depends solely on the context. For both cases, we derive instance-dependent lower bounds on the sample complexity and propose algorithms that asymptotically achieve the optimal sample complexity. For the non-separator setting, we do so by demonstrating that the Track-and-Stop algorithm can be extended to this setting. For the separator setting, we propose a novel sampling rule called G-tracking\textit{G-tracking}G-tracking, which uses the geometry of the context space to directly track the contexts rather than the actions. Finally, our empirical results showcase the advantage of our approaches compared to the state of the art.

View on arXiv
@article{shahverdikondori2025_2502.03061,
  title={ Optimal Best Arm Identification with Post-Action Context },
  author={ Mohammad Shahverdikondori and Amir Mohammad Abouei and Alireza Rezaeimoghadam and Negar Kiyavash },
  journal={arXiv preprint arXiv:2502.03061},
  year={ 2025 }
}
Comments on this paper