• Home
  • Search
  • Prioritized Sweeping Converges to the Optimal Value Function
  • Cite Icon13
  • https://doi.org/10.7282/t3tx3jsxCopy DOI Icon

Prioritized Sweeping Converges to the Optimal Value Function

  • Apr 1, 2008
  • View
  • Lihong Li +1 more
Show More
  • Abstract
  • Literature Map
  • References
  • Citations
  • Similar Papers
Abstract

Prioritized sweeping (PS) and its variants are model-based reinforcement-learning algorithms that have demonstrated superior performance in terms of computational and experience e‐ciency in practice. This note establishes the flrst|to the best of our knowledge|formal proof of convergence to the optimal value function when they are used as planning algorithms. We also describe applications of this result to provably e‐cient model-based reinforcement learning in the PAC-MDP framework. We do not address the issue of convergence rate in the present paper.

Similar Papers
  • PDF
  • Research Article
  • Citations7

Improving Model-Based Deep Reinforcement Learning with Learning Degree Networks and Its Application in Robot Control

  • Mar 04, 2022
  • Journal of Robotics
  • Guoqing Ma +3
  • Conference Article

Learning Control for Robotic Manipulator with Free Energy

  • Jul 01, 2020
  • Yazhou Hu +4
  • Supplementary Content
  • Citations6

Strategic Exploration in Reinforcement Learning - New Algorithms and Learning Guarantees

  • Feb 24, 2020
  • Figshare
  • Christoph Dann
  • Conference Article
  • Citations2

Context-dependent meta-control for reinforcement learning using a Dirichlet process Gaussian mixture model

  • Jan 01, 2018
  • Dongjae Kim +1
  • Research Article

Online learning algorithms : For passivity-based and distributed control

  • May 03, 2016
  • Research Repository (Delft University of Technology)
  • Subramanya Nageshrao
  • PDF
  • Research Article
  • Citations32

Parallel model-based and model-free reinforcement learning for card sorting performance

  • Sep 22, 2020
  • Scientific Reports
  • Alexander Steinke +2
  • Research Article
  • Citations39

Navigating complex decision spaces: Problems and paradigms in sequential choice.

  • Jan 01, 2014
  • Psychological Bulletin
  • Matthew M Walsh +1
  • Research Article
  • Citations91

Energy efficient speed planning of electric vehicles for car-following scenario using model-based reinforcement learning

  • Mar 15, 2022
  • Applied Energy
  • Heeyun Lee +3
  • Conference Article
  • Citations13

Reinforcement learning for model building and variance-penalized control

  • Dec 01, 2009
  • Abhijit Gosavi
  • Research Article
  • Citations20

On Value Function Representation of Long Horizon Problems

  • Apr 29, 2018
  • Proceedings of the AAAI Conference on Artificial Intelligence
  • Lucas Lehnert +2
  • Research Article
  • Citations108

Safety Augmented Value Estimation From Demonstrations (SAVED): Safe Deep Model-Based RL for Sparse Cost Robotic Tasks

  • Mar 06, 2020
  • IEEE Robotics and Automation Letters
  • Brijen Thananjeyan +8
  • Supplementary Content

Bidirectional Human-Robot Learning: Imitation and Skill Improvement

  • Jun 23, 2020
  • TUbilio (Technical University of Darmstadt)
  • Sousa Ewerton +1
  • Supplementary Content
  • Citations1

Solution of Large-scale Structured Optimization Problems with Schur-complement and Augmented Lagrangian Decomposition Methods

  • Aug 02, 2019
  • Figshare
  • Jose S Rodriguez
  • Conference Article
  • Citations14

Using training regimens to teach expanding function approximators

  • May 10, 2010
  • Ping Zang +4
  • Conference Article
  • Citations1

Hierarchical Control Architecture Regulating Competition between Model-Based and Context-Dependent Model-Free Reinforcement Learning Strategies

  • Oct 01, 2018
  • Dongjae Kim +2
Cactus Communications logo

Copyright 2026 Cactus Communications. All rights reserved.