• Home
  • Search
  • Research on multi-objective path planning and dynamic obstacle avoidance algorithm of manipulator based on reinforcement learning
  • https://doi.org/10.56028/aetr.14.1.1405.2025Copy DOI Icon

Research on multi-objective path planning and dynamic obstacle avoidance algorithm of manipulator based on reinforcement learning

Show More
  • Abstract
  • Literature Map
  • References
  • Similar Papers
Abstract

Aiming at the problem of multi-objective path planning (MOPP) and dynamic obstacle avoidance of manipulator in dynamic environment, this paper proposes a solution based on hierarchical reinforcement learning (HRL) framework. Traditional path planning methods have some problems in dynamic scenes, such as poor real-time performance, difficult to balance multi-objective conflicts and insufficient adaptability to environmental changes. Therefore, this paper designs a two-tier architecture including global path planning layer and local obstacle avoidance layer, which is trained by Proximal policy optimization (PPO) and Soft Actor-Critic (SAC) algorithms respectively, and achieves collaborative optimization through an efficient information exchange mechanism between layers. At the same time, a multi-objective reward function based on dynamic weight adjustment strategy is introduced. Fuzzy logic is used to adaptively balance the relationship among path length, obstacle avoidance safety and energy consumption according to environmental complexity. Combined with Long Short-Term Memory (LSTM), the trajectory of obstacles is predicted, and the potential field method is further introduced to modify the obstacle avoidance reward, which improves the real-time response ability and robustness of the algorithm in dynamic environment. The experimental results show that the HRL-SAC-PPO method proposed in this paper shows superior performance in both static and dynamic scenarios. In the static scene, the success rate of this method reaches 100%, the average path length is shortened to 2.13m, no collision occurs, and the energy consumption is reduced to 1.12 kJ, which shows a good multi-objective optimization effect. In the dynamic scene, the trajectory error of obstacles predicted by LSTM is only 4.2%, and the safe distance between the robot arm and obstacles is improved by 35%, which significantly enhances the reliability of obstacle avoidance. In addition, the average decision delay of this method is only 11.3ms, and the peak delay is 23ms, which is much lower than that of the contrast algorithm, showing stronger real-time response ability. The ablation experiment further verified the key role of LSTM trajectory prediction, dynamic weight adjustment and layered structure on the overall performance. In the welding task verification of the real UR5 manipulator, the success rate of the system in dynamic environment is 95.3%, the average path length is 3.41m, and the maximum joint acceleration is only 0.87 radian/s², which is far below the safety threshold, indicating that the algorithm has good stability and obstacle avoidance ability in practical application. The comprehensive performance comparison shows that this method performs well in different industrial scenarios, and has stronger environmental adaptability and comprehensive path planning ability.

Similar Papers
  • Book Chapter

A Brief Insight into Nanorobotics

  • Jan 01, 2017
  • Sanchita Paul
  • Research Article
  • Citations63

Reliable Path Planning Algorithm Based on Improved Artificial Potential Field Method

  • Jan 01, 2022
  • IEEE Access
  • Jie Luo +2
  • Research Article

Three Dimensional Path Planning System for Unmanned Aerial Vehicles Based on Reinforcement Learning Algorithm

  • Jan 01, 2024
  • Academic Journal of Engineering and Technology Science
  • Zhengxiang Huang +1
  • Research Article

Risk-Aware Reinforcement Learning with Dynamic Safety Filter for Collision Risk Mitigation in Mobile Robot Navigation

  • Sep 03, 2025
  • Sensors (Basel, Switzerland)
  • Bingbing Guo +4
  • Research Article

Hybrid path planning for UAV using rapidly-expanding random trees and proximal policy optimization in complex environments

  • Jan 07, 2026
  • Proceedings of the Institution of Mechanical Engineers Part I Journal of Systems and Control Engineering
  • Yuquan Fang +4
  • PDF
  • Research Article

Review of unmanned aerial vehicle obstacle avoidance planning based on artificial potential field

  • Sep 25, 2023
  • Applied and Computational Engineering
  • Hanyu Xing
  • Book Chapter

Design of Group Animation Algorithm for Virtual Reality Constraint Control

  • Aug 28, 2025
  • Yuan Sun
  • Research Article

Enhancing autonomous agriculture control systems in greenhouses for sustainable resource usage using deep learning techniques.

  • Jan 01, 2026
  • PloS one
  • Iman Hindi +3
  • Research Article
  • Citations1

An Autonomous Operation Path Planning Method for Wheat Planter Based on Improved Particle Swarm Algorithm

  • Sep 03, 2025
  • Sensors (Basel, Switzerland)
  • Shuangshuang Du +3
  • Research Article
  • Citations1

Application of improved sparrow search algorithm and dynamic window method in mobile robot path planning and real‐time obstacle avoidance

  • Jan 01, 2025
  • The Journal of Engineering
  • Lingyu Shen +1
  • Conference Article

An improved joint space three-dimensional Astar algorithm with pre-planning strategy

  • Jan 10, 2025
  • Chuangxin He +5
  • Research Article
  • Citations4

An improved joint space Astar algorithm for a 6-DOF manipulator with pre-planning strategy

  • May 25, 2025
  • Scientific Reports
  • Chuangxin He +3
  • PDF
  • Book Chapter
  • Citations2

Mapping-Based Navigation

  • Oct 27, 2017
  • Mordechai Ben-Ari +1
  • Research Article
  • Citations5

AgriPath: a robust multi-objective path planning framework for agricultural robots in dynamic field environments

  • Nov 27, 2025
  • Frontiers in Plant Science
  • Chenghan Yang +5
  • PDF
  • Research Article
  • Citations28

Research on path planning of three-neighbor search A* algorithm combined with artificial potential field

  • May 01, 2021
  • International Journal of Advanced Robotic Systems
  • Jiqing Chen +5
Cactus Communications logo

Copyright 2026 Cactus Communications. All rights reserved.