SPRT-Based Efficient Best Arm Identification in Stochastic Bandits

Arpan Mukherjee,Ali Tajer

doi:10.1109/jsait.2023.3288988

Abstract

This paper investigates the best arm identification (BAI) problem in stochastic multi-armed bandits in the fixed confidence setting. The general class of the exponential family of bandits is considered. The existing algorithms for the exponential family of bandits face computational challenges. To mitigate these challenges, the BAI problem is viewed and analyzed as a sequential composite hypothesis testing task, and a framework is proposed that adopts the likelihood ratio-based tests known to be effective for sequential testing. Based on this test statistic, a BAI algorithm is designed that leverages the canonical sequential probability ratio tests for arm selection and is amenable to tractable analysis for the exponential family of bandits. This algorithm has two key features: (1) its sample complexity is asymptotically optimal, and (2) it is guaranteed to be δ-PAC. Existing efficient approaches focus on the Gaussian setting and require Thompson sampling for the arm deemed the best and the challenger arm. Additionally, this paper analytically quantifies the computational expense of identifying the challenger in an existing approach. Finally, numerical experiments are provided to support the analysis.

Talk to us

Join us for a 30 min session where you can share your feedback and ask us any queries you have

Schedule a call

R Discovery Prime

R Discovery Prime

SPRT-Based Efficient Best Arm Identification in Stochastic Bandits

Abstract

Talk to us

Similar Papers

More From: IEEE Journal on Selected Areas in Information Theory

Lead the way for us

Journal: IEEE Journal on Selected Areas in Information Theory	Publication Date: Jan 1, 2023
Citations: 1

Similar Papers

SPRT-based Best Arm Identification in Stochastic Bandits
Arpan Mukherjee ... Ali Tajer
-
Arpan Mukherjee, et. al.Arpan Mukherjee ... Ali Tajer
26 Jun 2022
26 Jun 2022

Secure Best Arm Identification in Multi-armed Bandits
Radu Ciucanu ... Marta Soare
-
Radu Ciucanu, et. al.Radu Ciucanu ... Marta Soare
01 Jan 2019
01 Jan 2019

Best Arm Identification in Spectral Bandits
Tomáš Kocák ... Aurélien Garivier
-
Tomáš Kocák, et. al.Tomáš Kocák ... Aurélien Garivier
01 Jul 2020
01 Jul 2020

Best Arm Identification under Additive Transfer Bandits
Ojash Neopane ... Aaditya Ramdas
-
Ojash Neopane, et. al.Ojash Neopane ... Aaditya Ramdas
31 Oct 2021
31 Oct 2021

Editage

Paperpal

R Discovery

Mind the Graph

R Discovery Prime

R Discovery Prime

SPRT-Based Efficient Best Arm Identification in Stochastic Bandits

Abstract

Talk to us

Similar Papers

More From: IEEE Journal on Selected Areas in Information Theory