• Home
  • Search
  • Simulation-based Algorithms for Markov Decision Processes
  • Open Access IconOpen Access
  • Cite Icon226
  • https://doi.org/10.1007/978-1-84628-690-2Copy DOI Icon

Simulation-based Algorithms for Markov Decision Processes

  • Jan 1, 2007
  • Hyeong Soo Chang +3 more
Show More
  • Abstract
  • Literature Map
  • References
  • Citations
  • Similar Papers
Abstract

Markov decision process (MDP) models are widely used for modeling sequential decision-making problems that arise in engineering, economics, computer science, and the social sciences. Many real-world problems modeled by MDPs have huge state and/or action spaces, giving an opening to the curse of dimensionality and so making practical solution of the resulting models intractable. In other cases, the system of interest is too complex to allow explicit specification of some of the MDP model parameters, but simulation samples are readily available (e.g., for random transitions and costs). For these settings, various sampling and population-based algorithms have been developed to overcome the difficulties of computing an optimal solution in terms of a policy and/or value function. Specific approaches include adaptive sampling, evolutionary policy iteration, evolutionary random policy search, and model reference adaptive search. This substantially enlarged new edition reflects the latest developments in novel algorithms and their underpinning theories, and presents an updated account of the topics that have emerged since the publication of the first edition. Includes: innovative material on MDPs, both in constrained settings and with uncertain transition properties; game-theoretic method for solving MDPs; theories for developing roll-out based algorithms; and details of approximation stochastic annealing, a population-based on-line simulation-based algorithm. The self-contained approach of this book will appeal not only to researchers in MDPs, stochastic modeling, and control, and simulation but will be a valuable source of tuition and reference for students of control and operations research.

Similar Papers
  • Research Article
  • Citations20

Simulation-Based Algorithms for Markov Decision Processes: Monte Carlo Tree Search from AlphaGo to AlphaZero

  • Dec 01, 2019
  • Asia-Pacific Journal of Operational Research
  • Michael C Fu
  • Research Article
  • Citations50

Generalization of Faustmann's Formula for Stochastic Forest Growth and Prices with Markov Decision Process Models

  • Nov 01, 2001
  • Forest Science
  • Joseph Buongiorno
  • Research Article
  • Citations13

Towards enhanced threat modelling and analysis using a Markov Decision Process

  • Jul 30, 2022
  • Computer Communications
  • Saif U.R Malik +3
  • Research Article
  • Citations59

Tree Diversity, Landscape Diversity, and Economics of Maple-Birch Forests: Implications of Markovian Models

  • Oct 01, 1998
  • Management Science
  • Ching-Rong Lin +1
  • Conference Article

A Markov decision process model for inventory control under invisible stock loss and inaccurate record

  • Dec 01, 2017
  • Yajun Zhang +1
  • Research Article
  • Citations11

Linear programming formulation for non-stationary, finite-horizon Markov decision process models

  • Sep 17, 2017
  • Operations Research Letters
  • Arnab Bhattacharya +1
  • Book Chapter

Dynamic Adjustment Policy of Search Driver Matching Distance via Markov Decision Process

  • Jan 01, 2022
  • Suiming Guo +2
  • Research Article
  • Citations11

Dynamic selling of quality-graded products under demand uncertainties

  • Mar 15, 2011
  • Computers & Industrial Engineering
  • Hua-Hsuan Wu +2
  • Research Article
  • Citations2

Average case analysis of the classical algorithm for Markov decision processes with Büchi objectives

  • Feb 07, 2015
  • Theoretical Computer Science
  • Krishnendu Chatterjee +2
  • Research Article
  • Citations21

A MDP model for breast and ovarian cancer intervention strategies for BRCA1/2 mutation carriers.

  • Apr 22, 2014
  • IEEE Journal of Biomedical and Health Informatics
  • Mehrnaz Abdollahian +1
  • Research Article
  • Citations114

UAV-Assisted Wireless Energy and Data Transfer With Deep Reinforcement Learning

  • Sep 30, 2020
  • IEEE Transactions on Cognitive Communications and Networking
  • Zehui Xiong +6
  • Research Article
  • Citations67

Least Squares Temporal Difference Methods: An Analysis under General Conditions

  • Jan 01, 2012
  • SIAM Journal on Control and Optimization
  • Huizhen Yu
  • Book Chapter
  • Citations1

Optimal Overbooking Appointment Scheduling in Hospitals Using Evolutionary Markov Decision Process

  • Jan 01, 2022
  • Wenlong Ni +7
  • PDF
  • Book Chapter
  • Citations8

Symbolic Algorithms for Graphs and Markov Decision Processes with Fairness Objectives

  • Jan 01, 2018
  • Krishnendu Chatterjee +4
  • Research Article
  • Citations5

Suboptimal policy determination for large-scale Markov decision processes, Part 1: Description and bounds

  • Jul 01, 1985
  • Journal of Optimization Theory and Applications
  • C C White +1
Cactus Communications logo

Copyright 2026 Cactus Communications. All rights reserved.