Unifying Clustered and Non-stationary Bandits

Chuanhao Li, Qingyun Wu

Citation Details

Non-stationary bandits and clustered bandits lift the restrictive assumptions in contextual bandits and provide solutions to many important real-world scenarios. Though they have been studied independently so far, we point out the essence in solving these two problems overlaps considerably. In this work, we connect these two strands of bandit research under the notion of test of homogeneity, which seamlessly addresses change detection for non-stationary bandit and cluster identification for clustered bandit in a unified solution framework. Rigorous regret analysis and extensive empirical evaluations demonstrate the value of our proposed solution, especially its flexibility in handling various environment assumptions, e.g., a clustered non-stationary environment. more »

Award ID(s):: 1553568 1838615 1618948

PAR ID:: 10300520

Author(s) / Creator(s):: Chuanhao Li, Qingyun Wu

Date Published:: 2021-04-13

Journal Name:: Proceedings of The 24th International Conference on Artificial Intelligence and Statistics

Volume:: 130

Format(s):: Medium: X

Sponsoring Org:: National Science Foundation

Free Publicly Accessible Full Text
Accepted Manuscript1.0
Conference Paper:
The DOI is not currently available.

More Like this