• Home
  • Search
  • Decentralized Learning for Multiplayer Multiarmed Bandits
  • Open Access IconOpen Access
  • Cite Icon205
  • https://doi.org/10.1109/tit.2014.2302471Copy DOI Icon

Decentralized Learning for Multiplayer Multiarmed Bandits

Show More
  • Abstract
  • Literature Map
  • References
  • Citations
  • Similar Papers
Abstract

We consider the problem of distributed online learning with multiple players in multiarmed bandit (MAB) models. Each player can pick among multiple arms. When a player picks an arm, it gets a reward. We consider both independent identically distributed (i.i.d.) reward model and Markovian reward model. In the i.i.d. model, each arm is modeled as an i.i.d. process with an unknown distribution with an unknown mean. In the Markovian model, each arm is modeled as a finite, irreducible, aperiodic and reversible Markov chain with an unknown probability transition matrix and stationary distribution. The arms give different rewards to different players. If two players pick the same arm, there is a collision, and neither of them get any reward. There is no dedicated control channel for coordination or communication among the players. Any other communication between the users is costly and will add to the regret. We propose an online index-based distributed learning policy called dUCB4 algorithm that trades off exploration versus exploitation in the right way, and achieves expected regret that grows at most as near- O(log <sup xmlns:mml="http://www.w3.org/1998/Math/MathML" xmlns:xlink="http://www.w3.org/1999/xlink">2</sup> T). The motivation comes from opportunistic spectrum access by multiple secondary users in cognitive radio networks wherein they must pick among various wireless channels that look different to different users. This is the first distributed learning algorithm for multiplayer MABs with heterogeneous players (that have player-dependent rewards) to the best of our knowledge.

Similar Papers
  • PDF
  • Research Article
  • Citations6

Multiple-level power allocation strategy for secondary users in cognitive radio networks

  • Apr 14, 2014
  • EURASIP Journal on Advances in Signal Processing
  • Zhong Chen +4
  • Conference Article
  • Citations5

An Energy-Efficient Cooperative Strategy for Secondary Users in Cognitive Radio Networks

  • Dec 01, 2015
  • Jianqing Liu +4
  • Research Article
  • Citations21

Distributed Approach for Power and Rate Allocation to Secondary Users in Cognitive Radio Networks

  • May 01, 2011
  • IEEE Transactions on Vehicular Technology
  • L Akter +1
  • Conference Article
  • Citations27

Auction-based spectrum sharing for multiple primary and secondary users in cognitive radio networks

  • Apr 01, 2010
  • Hoda Shah Mohammadian +1
  • Research Article
  • Citations16

Decentralized Opportunistic Channel Access in CRNs Using Big-Data Driven Learning Algorithm

  • Jan 22, 2021
  • IEEE Transactions on Emerging Topics in Computational Intelligence
  • Zuohong Xu +6
  • Conference Article
  • Citations1

Spectrum Pricing Research Based on Game Theory in Cognitive Radio Networks

  • Sep 01, 2013
  • Xiaolong Yang +2
  • PDF
  • Research Article
  • Citations18

Cooperative DoA-only localization of primary users in cognitive radio networks

  • Apr 19, 2013
  • EURASIP Journal on Wireless Communications and Networking
  • Federico Penna +1
  • Conference Article
  • Citations19

MAC protocol for energy-harvesting users in cognitive radio networks

  • Jan 09, 2014
  • Yunmin Kim +2
  • Conference Article
  • Citations11

Adaptive modulation in multi-user cognitive radio networks over fading channels

  • Jul 01, 2013
  • F Foukalas +2
  • Conference Article
  • Citations1

Link-Layer Resource Allocation for Voice Users in Cognitive Radio Networks

  • Jun 01, 2011
  • Khaled Ali +1
  • PDF
  • Research Article

Performance Analysis of a Communication Failure and Repair Mechanism with Classified Primary Users in CRNs

  • Aug 08, 2024
  • Applied Sciences
  • Yuan Zhao +3
  • Research Article

Signal technique for friend or foe detection of intelligent malicious user in cognitive radio network

  • Jan 01, 2019
  • International Journal of Ad Hoc and Ubiquitous Computing
  • Masanori Hamamura +1
  • Research Article
  • Citations6

A Review of Cross-layer Design in Dynamic Spectrum Access for Cognitive Radio Networks

  • Jan 01, 2014
  • Journal of Computing and Information Technology
  • G Shine Let +1
  • Research Article
  • Citations9

Design, Implementation and Characterization of Practical Distributed Cognitive Radio Networks

  • Oct 01, 2013
  • IEEE Transactions on Communications
  • Ahmed Khattab +2
  • Conference Article
  • Citations3

Channel allocation based on Game theory in CR networks over Rayleigh fading channel

  • Jun 01, 2016
  • Tanapong Khumyat +1
Cactus Communications logo

Copyright 2026 Cactus Communications. All rights reserved.