• Home
  • Search
  • Omnidirectional-Wheel Conveyor Path Planning and Sorting Using Reinforcement Learning Algorithms
  • Cite Icon23
  • https://doi.org/10.1109/access.2022.3156924Copy DOI Icon

Omnidirectional-Wheel Conveyor Path Planning and Sorting Using Reinforcement Learning Algorithms

Show More
  • Abstract
  • Literature Map
  • References
  • Citations
  • Similar Papers
Abstract

In this paper, path planning and sorting of packages for Omnidirectional-Wheel conveyor are presented using Reinforcement Learning (RL). Q-learning, Double Q-learning, Deep Q-learning, and the Double Deep Q-learning algorithms are investigated. The RL algorithms enable the conveyor to self-learn the packages path and sort them without using conventional control or path planning theories. The RL algorithms are used for two different case studies on conveyors structures with different numbers of cells to compare and evaluate their performances in large- and small-scale sized structures. To explore the proposed methods response to external environment effects, two types of collisions between multiple packages were considered, the proposed RL algorithms showed their ability to resolve both types successfully. Comparative study between multiple RL algorithms for path planning showed that the Q-learning and Double Q-learning algorithms had outperformed their Deep learning versions for path planning in the two case studies. Furthermore, the proposed RL methods are compared experimentally to classic control and path planning theories using a hardware prototype for one of the presented case studies. The hardware experimental results showed that the proposed RL methods were as successful as the conventional methods in path planning and sorting in much less processing time. Two types of sorting scenarios (Type I and II) were tested for same package type and for multiple ones. For Type I sorting the Q-learning algorithm performed better than the Q-learning with weights approach, achieving better mean and minimum rewards while maximum rewards remain the same for both techniques. As for Type II sorting, only the Q-learning with weights approach was able to achieve it and converge in a reasonable time.

Similar Papers
  • Book Chapter
  • Citations1

Grid Path Planning for Mobile Robots with Improved Q-learning Algorithm

  • Jan 01, 2020
  • Lingling Peng +1
  • Conference Article
  • Citations1

Recent Trends in Computational Guidance and Control for Space Applications

  • Jul 31, 2023
  • Marilena Di Carlo +5
  • Conference Article
  • Citations4

Agent Maze Path Planning Based on Simulated Annealing Q-Learning Algorithm

  • Jul 25, 2022
  • Zhongtian Mao +4
  • PDF
  • Research Article
  • Citations14

Optimized-Weighted-Speedy Q-Learning Algorithm for Multi-UGV in Static Environment Path Planning under Anti-Collision Cooperation Mechanism

  • May 27, 2023
  • Mathematics
  • Yuanying Cao +1
  • Research Article
  • Citations12

State-chain sequential feedback reinforcement learning for path planning of autonomous mobile robots

  • Mar 01, 2013
  • Journal of Zhejiang University SCIENCE C
  • Xin Ma +4
  • Research Article

Three Dimensional Path Planning System for Unmanned Aerial Vehicles Based on Reinforcement Learning Algorithm

  • Jan 01, 2024
  • Academic Journal of Engineering and Technology Science
  • Zhengxiang Huang +1
  • Conference Article

On the Design of Safe Continual RL Methods for Control of Nonlinear Systems

  • Jun 24, 2025
  • Austin Coursey +2
  • Research Article
  • Citations5

Multimachine Collaborative Path Planning Method Based on A* Mechanism Connection Depth Neural Network Model

  • Jan 01, 2022
  • IEEE Access
  • Yipeng Zheng
  • Conference Article
  • Citations1

Double Learning for Suppliers' Bidding Strategy in the Electricity Market

  • May 11, 2022
  • Amir Bayati +1
  • Research Article
  • Citations5

P2P power trading based on reinforcement learning for nanogrid clusters

  • Jul 19, 2024
  • Expert Systems With Applications
  • Hojun Jin +4
  • Research Article

A welding manipulator path planning method combining reinforcement learning and intelligent optimisation algorithm

  • Jan 01, 2019
  • International Journal of Modelling, Identification and Control
  • Xianyun Duan +6
  • Research Article
  • Citations4

Path planning based on reinforcement learning

  • May 31, 2023
  • Applied and Computational Engineering
  • Jin Lin
  • Research Article
  • Citations1

A Two-Stage Reinforcement Learning Algorithm for AUV Path Planning Based on Trajectory Exploration and Sequence Modeling

  • Jan 01, 2025
  • IEEE Transactions on Automation Science and Engineering
  • Yue Liu +4
  • PDF
  • Research Article
  • Citations14

Secure State Estimation of Cyber-Physical System under Cyber Attacks: Q-Learning vs. SARSA

  • Oct 01, 2022
  • Electronics
  • Zengwang Jin +5
  • PDF
  • Research Article
  • Citations26

Refined Path Planning for Emergency Rescue Vehicles on Congested Urban Arterial Roads via Reinforcement Learning Approach

  • Aug 31, 2021
  • Journal of Advanced Transportation
  • Longhao Yan +5
Cactus Communications logo

Copyright 2026 Cactus Communications. All rights reserved.