• Home
  • Search
  • Approximation Algorithms for Restless Bandit Problems
  • Open Access IconOpen Access
  • Cite Icon11
  • https://doi.org/10.1137/1.9781611973068.4Copy DOI Icon

Approximation Algorithms for Restless Bandit Problems

  • Jan 4, 2009
  • Sudipto Guha +2 more
Show More
  • Abstract
  • Literature Map
  • References
  • Citations
  • Similar Papers
Abstract

In this paper, we consider the restless bandit problem, which is one of the most well-studied generalizations of the celebrated stochastic multi-armed bandit problem in decision theory. In its ultimate generality, the restless bandit problem is known to be PSPACE-Hard to approximate to any non-trivial factor, and little progress has been made on this problem despite its significance in modeling activity allocation under uncertainty. We make progress on this problem by showing that for an interesting and general subclass that we term Monotone bandits, a surprisingly simple and intuitive greedy policy yields a factor 2 approximation. Such greedy policies are termed index policies, and are popular due to their simplicity and their optimality for the stochastic multi-armed bandit problem. The Monotone bandit problem strictly generalizes the stochastic multi-armed bandit problem, and naturally models multi-project scheduling where the state of a project becomes increasingly uncertain when the project is not scheduled. We develop several novel techniques in the design and analysis of the index policy. Our algorithm proceeds by introducing a novel “balance” constraint to the dual of a well-known LP relaxation to the restless bandit problem. This is followed by a structural characterization of the optimal solution by using both the exact primal as well as dual complementary slackness conditions. This yields an interpretation of the dual variables as potential functions from which we derive the index policy and the associated analysis.

Similar Papers
  • Book Chapter
  • Citations40

Deviations of Stochastic Bandit Regret

  • Jan 01, 2011
  • Antoine Salomon +1
  • Research Article
  • Citations24

Index policies for discounted bandit problems with availability constraints

  • Jun 01, 2008
  • Advances in Applied Probability
  • Savas Dayanik +2
  • Research Article
  • Citations4

Asymptotic Optimal Control of Markov-Modulated Restless Bandits

  • Apr 03, 2018
  • Proceedings of the ACM on Measurement and Analysis of Computing Systems
  • Santiago Duran +1
  • Research Article
  • Citations33

Modeling Human Performance in Restless Bandits with Particle Filters

  • Dec 16, 2009
  • The Journal of Problem Solving
  • Michael S.K Yi +2
  • Conference Article

Multi-armed bandit based approach for performance enhancement of window intensity test(WIT) detector

  • Dec 01, 2017
  • Kusal B Tennakoon +3
  • Conference Article
  • Citations3

Opportunistic Multichannel Access with Imperfect Observation: A Fixed Point Analysis on Indexability and Index-based Policy

  • Apr 01, 2018
  • Kehao Wang +3
  • Research Article
  • Citations70

Achieving Fairness in the Stochastic Multi-Armed Bandit Problem

  • Apr 03, 2020
  • Proceedings of the AAAI Conference on Artificial Intelligence
  • Vishakha Patil +3
  • Research Article
  • Citations1

Pattern Recognition as a Problem in Decision Theory and an Application to Speech Recognition

  • Apr 01, 1963
  • IEEE Transactions on Military Electronics
  • V E Sackschewsky +1
  • Research Article
  • Citations570

The Complexity of Optimal Queuing Network Control

  • May 01, 1999
  • Mathematics of Operations Research
  • Christos H Papadimitriou +1
  • Research Article
  • Citations14

Almost optimal policies for stochastic systemswhich almost satisfy conservation laws

  • Jan 01, 1999
  • Annals of Operations Research
  • K.D Glazebrook +1
  • Conference Article
  • Citations1

Computing an Index Policy for Multiarmed Bandits with Deadlines

  • Jan 01, 2008
  • José Nino-Mora
  • Research Article
  • Citations25

Gap-free Bounds for Stochastic Multi-Armed Bandit

  • Jan 01, 2008
  • IFAC Proceedings Volumes
  • A Juditsky +3
  • PDF
  • Research Article
  • Citations72

Risk-aware multi-armed bandit problem with application to portfolio selection

  • Nov 01, 2017
  • Royal Society Open Science
  • Xiaoguang Huo +1
  • Research Article
  • Citations15

A General Framework for Bandit Problems Beyond Cumulative Objectives

  • Jan 06, 2023
  • Mathematics of Operations Research
  • Asaf Cassel +2
  • Research Article
  • Citations674

Sample mean based index policies byO(logn) regret for the multi-armed bandit problem

  • Dec 01, 1995
  • Advances in Applied Probability
  • Rajeev Agrawal
Cactus Communications logo

Copyright 2026 Cactus Communications. All rights reserved.