2604.02255
2026-04-03
cs.IT
cs.LG
math.IT
Best-Arm Identification with Noisy Actuation
Merve Karakas, Osama Hanna, Lin F. Yang, Christina Fragouli
详情
英文摘要
In this paper, we consider a multi-armed bandit (MAB) instance and study how to identify the best arm when arm commands are conveyed from a central learner to a distributed agent over a discrete memoryless channel (DMC). Depending on the agent capabilities, we provide communication schemes along with their analysis, which interestingly relate to the zero-error capacity of the underlying DMC.