Concurrent bandits and cognitive radio networks

Orly Avner; Shie Mannor

Conference ProceedingsOPEN ACCESS

Concurrent bandits and cognitive radio networks

Lecture Notes in Computer Science (including subseries Lecture Notes in Artificial Intelligence and Lecture Notes in Bioinformatics) (2014) 8724 LNAI(PART 1) 66-81

DOI: 10.1007/978-3-662-44848-9_5

56Citations

24Readers

Abstract

We consider the problem of multiple users targeting the arms of a single multi-armed stochastic bandit. The motivation for this problem comes from cognitive radio networks, where selfish users need to coexist without any side communication between them, implicit cooperation or common control. Even the number of users may be unknown and can vary as users join or leave the network. We propose an algorithm that combines an ε-greedy learning rule with a collision avoidance mechanism. We analyze its regret with respect to the system-wide optimum and show that sub-linear regret can be obtained in this setting. Experiments show dramatic improvement compared to other algorithms for this setting. © 2014 Springer-Verlag.

Author supplied keywords

Cite

CITATION STYLE

APA

Avner, O., & Mannor, S. (2014). Concurrent bandits and cognitive radio networks. In Lecture Notes in Computer Science (including subseries Lecture Notes in Artificial Intelligence and Lecture Notes in Bioinformatics) (Vol. 8724 LNAI, pp. 66–81). Springer Verlag. https://doi.org/10.1007/978-3-662-44848-9_5

Concurrent bandits and cognitive radio networks

Abstract

Author supplied keywords

Cite

Register to see more suggestions