• Home
  • Search
  • Planning in Markov Decision Processes with Gap-Dependent Sample Complexity
  • Cite Icon2
  • https://doi.org/10.48550/arxiv.2006.05879Copy DOI Icon

Planning in Markov Decision Processes with Gap-Dependent Sample Complexity

Show More
  • Abstract
  • Literature Map
  • Citations
  • Similar Papers
Abstract

We propose MDP-GapE, a new trajectory-based Monte-Carlo Tree Search algorithm for planning in a Markov Decision Process in which transitions have a finite support. We prove an upper bound on the number of calls to the generative models needed for MDP-GapE to identify a near-optimal action with high probability. This problem-dependent sample complexity result is expressed in terms of the sub-optimality gaps of the state-action pairs that are visited during exploration. Our experiments reveal that MDP-GapE is also effective in practice, in contrast with other algorithms with sample complexity guarantees in the fixed-confidence setting, that are mostly theoretical.

Similar Papers
  • Conference Article
  • Citations1

Towards scalable mdp algorithms

  • Dec 24, 2016
  • Andrey Kolobov
  • Research Article
  • Citations5

Suboptimal policy determination for large-scale Markov decision processes, Part 1: Description and bounds

  • Jul 01, 1985
  • Journal of Optimization Theory and Applications
  • C C White +1
  • Book Chapter
  • Citations17

Solving Markov Decision Processes via Simulation

  • Sep 18, 2014
  • Abhijit Gosavi
  • Book Chapter

GEMBench: A Platform for Collaborative Development of GPU Accelerated Embedded Markov Decision Systems

  • Jan 01, 2019
  • Adrian E Sapio +3
  • Research Article
  • Citations26

Probably Approximately Correct (PAC) exploration in reinforcement learning

  • Jan 01, 2007
  • Rutgers University Community Repository (Rutgers University)
  • Alexander L Strehl
  • Conference Article
  • Citations51

From Greedy Selection to Exploratory Decision-Making

  • Jun 27, 2018
  • Yue Feng +5
  • Research Article

Application of Reinforcement Learning to Solve Rubrik’s Cube with Markov Decision Process

  • Oct 08, 2025
  • Jurnal RESTI (Rekayasa Sistem dan Teknologi Informasi)
  • Defni +6
  • Book Chapter

Task Coordination for Service Robots Based on Multiple Markov Decision Processes

  • Jan 01, 2012
  • Elva Corona +1
  • Conference Article

Stochastic control with graphical models: the influence view approach

  • Oct 31, 1996
  • P Magni +1
  • Conference Article
  • Citations75

State aggregation in Markov decision processes

  • Dec 10, 2002
  • Zhiyuan Ren +1
  • Single Book
  • Citations226

Simulation-based Algorithms for Markov Decision Processes

  • Jan 01, 2007
  • Hyeong Soo Chang +3
  • Conference Article
  • Citations6

Wavefront-MCTS

  • Nov 05, 2018
  • Yong Hu +2
  • Research Article
  • Citations5

Online Dynamic Pricing for Electric Vehicle Charging Stations With Reservations

  • Sep 01, 2025
  • IEEE Transactions on Intelligent Transportation Systems
  • Jan Mrkos +3
  • Supplementary Content
  • Citations11

A Promising Approach to Optimizing Sequential Treatment Decisions for Depression: Markov Decision Process

  • Jan 01, 2022
  • Pharmacoeconomics
  • Fang Li +3
  • Book Chapter
  • Citations7

Automata Learning Meets Shielding

  • Jan 01, 2022
  • Martin Tappler +5
Cactus Communications logo

Copyright 2026 Cactus Communications. All rights reserved.