Online Learning with Switching Costs and Other Adaptive Adversaries
Establishes Ω(T^{2/3}) regret lower bound for online learning with switching costs under bandit feedback, highlighting the difficulty compared to full-information settings.
Nicolo Cesa-Bianchi, Ofer Dekel, Ohad Shamir