• Cite Icon11
  • https://doi.org/10.1002/9781118557426.ch3Copy DOI Icon

Approximate Dynamic Programming

  • Feb 28, 2013
  • Rémi Munos
Show More
  • Abstract
  • Literature Map
  • References
  • Citations
  • Similar Papers
Abstract

In any complex or large scale sequential decision making problem, there is a crucial need to use function approximation to represent the relevant functions such as the value function or the policy. The Dynamic Programming (DP) and Reinforcement Learning (RL) methods introduced in previous chapters make the implicit assumption that the value function can be perfectly represented (i.e. kept in memory), for example by using a look-up table (with a finite number of entries) assigning a value to all possible states (assumed to be finite) of the system. Those methods are called exact because they provide an exact computation of the optimal solution of the considered problem (or at least, enable the computations to converge to this optimal solution). However, such methods often apply to toy problems only, since in most interesting applications, the number of possible states is so large (and possibly infinite if we consider continuous spaces) that a perfect representation of the function at all states is impossible. It becomes necessary to approximate the function by using a moderate number of coefficients (which can be stored in a computer), and therefore extend the range of DP and RL to methods using such approximate representations. These approximate methods combine DP and RL methods with function approximation tools.

Similar Papers
  • Research Article
  • Citations13

Reinforcement Learning for Clinical Applications.

  • Feb 08, 2023
  • Clinical Journal of the American Society of Nephrology
  • Kia Khezeli +5
  • Conference Article
  • Citations1

Comparison of Reinforcement Learning Methods for Production Control in Discrete Manufacturing Systems

  • Jun 12, 2023
  • Lingxiang Yun +3
  • Conference Article
  • Citations15

Cola-HRL: Continuous-Lattice Hierarchical Reinforcement Learning for Autonomous Driving

  • Oct 23, 2022
  • Lingping Gao +7
  • Conference Article

QMDP: DASH Adaptation using Queueing Theory within a Markov Decision Process

  • Jan 09, 2021
  • Kevin Gatimu +1
  • Research Article
  • Citations5

5G Network Slicing: Methods to Support Blockchain and Reinforcement Learning

  • Mar 24, 2022
  • Computational Intelligence and Neuroscience
  • Juan Hu +1
  • Research Article
  • Citations31

The reinforcement learning method for occupant behavior in building control: A review

  • Sep 02, 2020
  • Energy and Built Environment
  • Mengjie Han +4
  • PDF
  • Research Article
  • Citations27

A new solution to distributed permutation flow shop scheduling problem based on NASH Q-Learning

  • Sep 30, 2021
  • Advances in Production Engineering & Management
  • J.F Ren +2
  • Research Article

LPPG-RL: Lexicographically Projected Policy Gradient Reinforcement Learning with Subproblem Exploration

  • Mar 14, 2026
  • Ruiyu Qiu +4
  • Conference Article
  • Citations5

Cohesion-driven Online Actor-Critic Reinforcement Learning for mHealth Intervention

  • Aug 15, 2018
  • Feiyun Zhu +4
  • Research Article
  • Citations194

Deep reinforcement learning based energy management for a hybrid electric vehicle

  • Apr 14, 2020
  • Energy
  • Guodong Du +5
  • Conference Article
  • Citations23

Meta Reinforcement Learning-Based Lane Change Strategy for Autonomous Vehicles

  • Jul 11, 2021
  • Fei Ye +3
  • Dissertation

The effectiveness of numerical approximation for dynamic programming problems

  • Nov 07, 2019
  • Wyatt Jones
  • Conference Article
  • Citations9

A Concave Value Function Extension for the Dynamic Programming Approach to Revenue Management in Attended Home Delivery

  • Jun 01, 2019
  • Denis Lebedev +2
  • Research Article
  • Citations5

A Reinforcement Learning Method Using Reward Acquisition Efficiency for POMDP Environments

  • Jan 01, 2008
  • Transactions of the Japanese Society for Artificial Intelligence
  • Hirokazu Kawai +2
  • Research Article
  • Citations7

Learning to Acquire Whole-Body Humanoid Center of Mass Movements to Achieve Dynamic Tasks

  • Jan 01, 2008
  • Advanced Robotics
  • Takamitsu Matsubara +5
Cactus Communications logo

Copyright 2026 Cactus Communications. All rights reserved.