• Home
  • Search
  • Hierarchical Policy Learning for Mechanical Search
  • Open Access IconOpen Access
  • Cite Icon3
  • https://doi.org/10.1109/icra46639.2022.9811572Copy DOI Icon

Hierarchical Policy Learning for Mechanical Search

  • May 23, 2022
  • Oussama Zenkri +2 more
Show More
  • Abstract
  • Literature Map
  • References
  • Citations
  • Similar Papers
Abstract

Retrieving objects from clutters is a complex task, which requires multiple interactions with the environment until the target object can be extracted. These interactions involve executing action primitives like grasping or pushing as well as setting priorities for the objects to manipulate and the actions to execute. Mechanical Search (MS) [1] is a framework for object retrieval, which uses a heuristic algorithm for pushing and rule-based algorithms for high-level planning. While rule-based policies profit from human intuition in how they work, they usually perform sub-optimally in many cases. Deep reinforcement learning (RL) has shown great performance in complex tasks such as taking decisions through evaluating pixels, which makes it suitable for training policies in the context of object-retrieval. In this work, we first formulate the MS problem in a principled formulation as a hierarchical POMDP. Based on this formulation, we propose a hierarchical policy learning approach for the MS problem. For demonstration, we present two main parameterized sub-policies: a push policy and an action selection policy. When integrated into the hierarchical POMDP's policy, our proposed sub-policies increase the success rate of retrieving the target object from less than 32% to nearly 80%, while reducing the computation time for push actions from multiple seconds to less than 10 milliseconds.

Similar Papers
  • Research Article
  • Citations6

Break through the limits of learning by machines

  • Sep 20, 2016
  • Chinese Science Bulletin
  • Zhongzhi Shi
  • Book Chapter
  • Citations4

Nitty-Gritty of Deep Reinforcement Learning for the Healthcare Sector

  • Oct 18, 2023
  • Vaishnavi Kumari +5
  • PDF
  • Research Article
  • Citations22

A Review of Mobile Robot Path Planning Based on Deep Reinforcement Learning Algorithm

  • Dec 01, 2021
  • Journal of Physics: Conference Series
  • Yanwei Zhao +2
  • Research Article
  • Citations4

Leveraging deep reinforcement learning for design space exploration with multi-fidelity surrogate model

  • Jun 25, 2024
  • Journal of Engineering Design
  • Haokun Li +5
  • Research Article
  • Citations1

DDPG Agent to Swing Up and Balance Cart- Pole System

  • Apr 09, 2021
  • International Journal of Advanced Research in Science, Communication and Technology
  • Buvanesh Pandian V
  • PDF
  • Research Article
  • Citations5

Deep imitation reinforcement learning with expert demonstration data

  • Oct 31, 2018
  • The Journal of Engineering
  • Menglong Yi +3
  • Research Article
  • Citations57

A new ensemble deep graph reinforcement learning network for spatio-temporal traffic volume forecasting in a freeway network

  • Jan 29, 2022
  • Digital Signal Processing
  • Pan Shang +5
  • Research Article

Autonomous Driving Control Strategy Based on Deep Reinforcement Learning

  • Jan 13, 2025
  • Applied and Computational Engineering
  • Liangjun Yang
  • Research Article
  • Citations31

Deep Reinforcement Learning for Online Resource Allocation in IoT Networks: Technology, Development, and Future Challenges

  • Jun 01, 2023
  • IEEE Communications Magazine
  • Peng Cheng +5
  • PDF
  • Research Article
  • Citations38

Deep Reinforcement Learning Unleashing the Power of AI in Decision-Making

  • Feb 02, 2024
  • Journal of Artificial Intelligence General science (JAIGS) ISSN:3006-4023
  • Jeff Shuford
  • Research Article
  • Citations14

MASAC-based confrontation game method of UAV clusters

  • Dec 01, 2022
  • SCIENTIA SINICA Informationis
  • 健 薛 +6
  • Conference Article
  • Citations1

Research on AI-driven personalized learning path planning and effectiveness under dual-system teaching mode

  • Apr 18, 2025
  • Ling Chen
  • Conference Article

Safe Deep Reinforcement Learning Based on Sample Value Evaluation

  • Dec 02, 2022
  • Rongjun Ye +2
  • Conference Article
  • Citations16

An Adaptive Control Method for Arterial Signal Coordination Based on Deep Reinforcement Learning

  • Oct 01, 2019
  • Peng Chen +2
  • Research Article
  • Citations1

Deep Reinforcement Learning for irrigation optimization: Advantages, opportunities, and challenges

  • Dec 01, 2025
  • Agricultural Water Management
  • Jiamei Liu +8
Cactus Communications logo

Copyright 2026 Cactus Communications. All rights reserved.