Preprint
Aug 2026
Projection-Free Bandit Online Optimization for Multi-Agent Systems with Dynamic Regret
This paper proposes a distributed bandit online feedback optimization algorithm that relies solely on real-time input-output data and establishes a sublinear dynamic regret bound that depends on a temporal variation measure of system non-stationarity.
Xia Jiang, Lu Liu, Gang Feng
· 0 citations