• Home
  • Search
  • Finding Approximate POMDP solutions Through Belief Compression
  • Cite Icon268
  • https://doi.org/10.1613/jair.1496Copy DOI Icon

Finding Approximate POMDP solutions Through Belief Compression

Show More
  • Abstract
  • Highlights & Summary
  • PDF
  • Literature Map
  • References
  • Citations
  • Similar Papers
Abstract

Standard value function approaches to finding policies for Partially Observable Markov Decision Processes (POMDPs) are generally considered to be intractable for large models. The intractability of these algorithms is to a large extent a consequence of computing an exact, optimal policy over the entire belief space. However, in real-world POMDP problems, computing the optimal policy for the full belief space is often unnecessary for good control even for problems with complicated policy classes. The beliefs experienced by the controller often lie near a structured, low-dimensional subspace embedded in the high-dimensional belief space. Finding a good approximation to the optimal value function for only this subspace can be much easier than computing the full value function. We introduce a new method for solving large-scale POMDPs by reducing the dimensionality of the belief space. We use Exponential family Principal Components Analysis (Collins, Dasgupta & Schapire, 2002) to represent sparse, high-dimensional belief spaces using small sets of learned features of the belief state. We then plan only in terms of the low-dimensional belief features. By planning in this low-dimensional space, we can find policies for POMDP models that are orders of magnitude larger than models that can be handled by conventional techniques. We demonstrate the use of this algorithm on a synthetic problem and on mobile robot navigation tasks.

Loading PDF

Similar Papers
  • Research Article

A Novel Point-Based Incremental Pruning Algorithm for POMDP

  • Feb 06, 2014
  • Applied Mechanics and Materials
  • Bo Wu +2
  • PDF
  • Research Article
  • Citations88

Monte Carlo Sampling Methods for Approximating Interactive POMDPs

  • Mar 24, 2009
  • Journal of Artificial Intelligence Research
  • P Doshi +1
  • Research Article
  • Citations151

Motion planning under uncertainty for robotic tasks with long time horizons

  • Dec 07, 2010
  • The International Journal of Robotics Research
  • Hanna Kurniawati +3
  • Research Article
  • Citations35

Efficient Approximate Value Iteration for Continuous Gaussian POMDPs

  • Sep 20, 2021
  • Proceedings of the AAAI Conference on Artificial Intelligence
  • Jur Van Den Berg +2
  • Research Article
  • Citations271

Planning under Uncertainty for Robotic Tasks with Mixed Observability

  • May 04, 2010
  • The International Journal of Robotics Research
  • Sylvie C W Ong +3
  • Conference Article
  • Citations19

Motion planning under uncertainty for medical needle steering using optimization in belief space

  • Sep 01, 2014
  • Wen Sun +1
  • Research Article
  • Citations14

Smoother Entropy for Active State Trajectory Estimation and Obfuscation in POMDPs

  • Jun 01, 2023
  • IEEE Transactions on Automatic Control
  • Timothy L Molloy +1
  • Conference Article
  • Citations1

A POMDP Based Routing Model to Enhance Directed Diffusion in Wireless Sensor Networks

  • Dec 07, 2013
  • Yu Pang +3
  • PDF
  • Research Article
  • Citations2

Future memories are not needed for large classes of POMDPs

  • Mar 03, 2023
  • Operations Research Letters
  • Victor Cohen +1
  • Book Chapter
  • Citations110

Monte Carlo Value Iteration for Continuous-State POMDPs

  • Jan 01, 2010
  • Haoyu Bai +3
  • Research Article
  • Citations9

Observation-Based Optimization for POMDPs With Continuous State, Observation, and Action Spaces

  • May 01, 2019
  • IEEE Transactions on Automatic Control
  • Xiaofeng Jiang +3
  • Conference Article
  • Citations54

FIRM: Feedback controller-based information-state roadmap - A framework for motion planning under uncertainty

  • Sep 01, 2011
  • Ali-Akbar Agha-Mohammadi +2
  • Research Article
  • Citations23

A Machine Learning–Enabled Partially Observable Markov Decision Process Framework for Early Sepsis Prediction

  • Mar 22, 2022
  • INFORMS Journal on Computing
  • Zeyu Liu +5
  • Research Article
  • Citations8

Information Gathering and Reward Exploitation of Subgoals for POMDPs

  • Mar 04, 2015
  • Proceedings of the AAAI Conference on Artificial Intelligence
  • Hang Ma +1
  • Research Article
  • Citations17

Point-Based POMDP Solving with Factored Value Function Approximation

  • Jun 21, 2014
  • Proceedings of the AAAI Conference on Artificial Intelligence
  • Tiago Veiga +2
Cactus Communications logo

Copyright 2026 Cactus Communications. All rights reserved.