Best-Arm Identification with Noisy Actuation

By Merve Karakas, Osama Hanna, Lin F. Yang, Christina Fragouli / April 3, 2026

arXiv:2604.02255v1 Announce Type: cross Abstract: In this paper, we consider a multi-armed bandit (MAB) instance and study how to identify the best arm when arm commands are conveyed from a central learner to a distributed agent over a discrete memoryless channel (DMC). Depending on the agent capabilities, we provide communication schemes along with their analysis, which interestingly relate to the zero-error capacity of the underlying DMC.

Leave a Comment