• Home
  • Search
  • Twin Delayed Multi-Agent Deep Deterministic Policy Gradient
  • Cite Icon10
  • https://doi.org/10.1109/pic53636.2021.9687069Copy DOI Icon

Twin Delayed Multi-Agent Deep Deterministic Policy Gradient

  • Dec 17, 2021
  • Mengying Zhan +3 more
Show More
  • Abstract
  • Literature Map
  • References
  • Citations
  • Similar Papers
Abstract

Recently, reinforcement learning has made remarkable achievements in the fields of natural science, engineering, medicine and operational research. Reinforcement learning addresses sequence problems and considers long-term returns. This long-term view of reinforcement learning is critical to find the optimal solution of many problems. The existing multi- agent reinforcement learning algorithms have the problem of overestimation in estimating the Q value. Unfortunately, there have not been many studies on overestimation of agent reinforcement learning, which will affect the learning efficiency of reinforcement learning. Based on the traditional multi-agent reinforcement learning algorithm, this paper improves the actor network and critic network, optimizes the overestimation of Q value and adopts the update delayed method to make the actor training more stable. In order to test the effectiveness of the algorithm structure, the modified method is compared with the traditional MADDPG, DDPG and DQN methods in the simulation environment.

Similar Papers
  • Research Article
  • Citations75

Sliding-mode surface-based decentralized event-triggered control of partially unknown interconnected nonlinear systems via reinforcement learning

  • May 09, 2023
  • Information Sciences
  • Tengda Wang +4
  • PDF
  • Research Article
  • Citations7

Improving Model-Based Deep Reinforcement Learning with Learning Degree Networks and Its Application in Robot Control

  • Mar 04, 2022
  • Journal of Robotics
  • Guoqing Ma +3
  • Conference Article
  • Citations8

Delay-Oriented Knowledge-Driven Resource Allocation in SAGIN-Based Vehicular Networks

  • Mar 01, 2023
  • Lei Huang +4
  • Book Chapter
  • Citations4

Automatic Curriculum Generation by Hierarchical Reinforcement Learning

  • Jan 01, 2020
  • Zhenghua He +3
  • Research Article
  • Citations18

Federated deep reinforcement learning for task offloading and resource allocation in mobile edge computing-assisted vehicular networks

  • Jun 25, 2024
  • Journal of Network and Computer Applications
  • Xu Zhao +4
  • Research Article
  • Citations27

A fuzzy-based potential field hierarchical reinforcement learning approach for target hunting by multi-AUV in 3-D underwater environments

  • Aug 05, 2019
  • International Journal of Control
  • Xiang Cao +1
  • Conference Article
  • Citations3

Optimal Control of Blank Holder Force Based on Deep Reinforcement Learning

  • Dec 01, 2019
  • Peng Guo +1
  • Research Article

Design of a consumer behavior prediction model integrating reinforcement learning and time series analysis in online e-commerce reviews

  • Jan 12, 2026
  • PeerJ Computer Science
  • Zongping Lin +6
  • Research Article
  • Citations1

Multi-Factor Ablation Study in Dynamic Pricing for Electronic Products: Comparing Deep Reinforcement Learning with Traditional Optimization

  • Dec 18, 2025
  • Advances in Economics, Management and Political Sciences
  • Xinyue Zhang
  • PDF
  • Research Article
  • Citations3

UAV Confrontation and Evolutionary Upgrade Based on Multi-Agent Reinforcement Learning

  • Aug 01, 2024
  • Drones
  • Xin Deng +2
  • Research Article
  • Citations30

Case-based myopic reinforcement learning for satisfying target service level in supply chain

  • Jul 17, 2007
  • Expert Systems with Applications
  • I Kwon +3
  • Research Article
  • Citations13

Reinforcement Learning for Clinical Applications.

  • Feb 08, 2023
  • Clinical Journal of the American Society of Nephrology
  • Kia Khezeli +5
  • Book Chapter
  • Citations4

A Parallel Evolutionary Algorithm with Value Decomposition for Multi-agent Problems

  • Jan 01, 2020
  • Gao Li +2
  • Research Article
  • Citations7

Path planning of mobile robot based on improved double deep Q-network algorithm.

  • Feb 13, 2025
  • Frontiers in neurorobotics
  • Zhenggang Wang +2
  • Conference Article
  • Citations24

Automatic task decomposition and state abstraction from demonstration

  • Jun 04, 2012
  • Luis C Cobo +2
Cactus Communications logo

Copyright 2026 Cactus Communications. All rights reserved.