• Home
  • Search
  • Improved Cooperative Multi-agent Reinforcement Learning Algorithm Augmented by Mixing Demonstrations from Centralized Policy
  • Cite Icon18
  • https://doi.org/10.65109/hxwh4898Copy DOI Icon

Improved Cooperative Multi-agent Reinforcement Learning Algorithm Augmented by Mixing Demonstrations from Centralized Policy

  • May 8, 2019
  • Hyun-Rok Lee +1 more
Show More
  • Abstract
  • Literature Map
  • References
  • Citations
  • Similar Papers
Abstract

Many decision problems for complex systems that involve multiple decision makers can be formulated as a decentralized partially observable markov decision process (dec-POMDP) problem. Due to the computational difficulty with obtaining optimal policies, recent approaches to dec-POMDP often use a multi-agent reinforcement learning (MARL) algorithm. We propose a method to improve the existing cooperative MARL algorithms by adopting an imitation learning technique. For a reference policy in the imitation learning part, we use a centralized policy from a multi-agent MDP or a multi-agent POMDP model reduced from the original dec-POMDP model. In the proposed method, during the training process, we mix demonstrations from the reference policy by using a demonstration buffer. Demonstration samples from the buffer are used in the augmented policy gradient function for policy updates. We assess the performance of the proposed method for three well-known dec-POMDP benchmark problems -- Mars rover, co-operative box pushing, and dec-tiger. Experimental results indicate that augmenting the baseline MARL algorithm by mixing the demonstrations significantly improves the quality of policy solutions. With these results, we conclude that the imitation learning can enhance MARL algorithms and that policy solutions from MMDP and MPOMDP models are a reasonable reference policy to use in the proposed algorithm.

Similar Papers
  • Research Article

Multi‐Agent Reinforcement Learning Algorithm Based on Local Observation Imitation Learning

  • Jan 01, 2025
  • IET Control Theory & Applications
  • Hui Zhang +3
  • Research Article
  • Citations1

Solving Action Semantic Conflict in Physically Heterogeneous Multi-Agent Reinforcement Learning with Generalized Action-Prediction Optimization

  • Feb 27, 2025
  • Applied Sciences
  • Xiaoyang Yu +3
  • Research Article
  • Citations3

An Effective Training Method for Counterfactual Multi-Agent Policy Network Based on Differential Evolution Algorithm

  • Sep 18, 2024
  • Applied Sciences
  • Shaochun Qu +5
  • PDF
  • Research Article
  • Citations6

Robust Multiagent Reinforcement Learning for UAV Systems: Countering Byzantine Attacks

  • Nov 19, 2023
  • Information
  • Jishu K Medhi +3
  • Research Article
  • Citations53

Reinforcement Learning for Joint Control of Traffic Signals in a Transportation Network

  • Jan 10, 2020
  • IEEE Transactions on Vehicular Technology
  • Jincheol Lee +2
  • PDF
  • Research Article
  • Citations3

UAV Confrontation and Evolutionary Upgrade Based on Multi-Agent Reinforcement Learning

  • Aug 01, 2024
  • Drones
  • Xin Deng +2
  • PDF
  • Research Article
  • Citations6

GHQ: grouped hybrid Q-learning for cooperative heterogeneous multi-agent reinforcement learning

  • Apr 23, 2024
  • Complex & Intelligent Systems
  • Xiaoyang Yu +4
  • Conference Article
  • Citations4

Gas Source Localization using Improved Multi-Agent Reinforcement Learning

  • Nov 06, 2020
  • Zhi-Pu Wang +1
  • Conference Article
  • Citations5

A Tensor Factorization Approach to Generalization in Multi-agent Reinforcement Learning

  • Dec 01, 2012
  • Stefano Bromuri
  • Research Article
  • Citations114

Multi-Agent Learning with Policy Prediction

  • Jul 04, 2010
  • Proceedings of the AAAI Conference on Artificial Intelligence
  • Chongjie Zhang +1
  • Conference Article
  • Citations23

Adversarial attacks in consensus-based multi-agent reinforcement learning

  • May 25, 2021
  • Martin Figura +2
  • Research Article
  • Citations3

Joint Spectrum and Power Allocation in Wireless Network: A Two-Stage Multi-Agent Reinforcement Learning Method

  • Jun 01, 2024
  • IEEE Transactions on Emerging Topics in Computational Intelligence
  • Pengcheng Dai +4
  • Research Article
  • Citations58

Multi-agent reinforcement learning algorithm to solve a partially-observable multi-agent problem in disaster response

  • Sep 18, 2020
  • European Journal of Operational Research
  • Hyun-Rok Lee +1
  • Conference Article
  • Citations1

Domain-Aware Multiagent Reinforcement Learning in Navigation

  • Jul 18, 2021
  • Ifrah Saeed +3
  • Conference Article

MOSMAC: A Multi-agent Reinforcement Learning Benchmark on Sequential Multi-Objective Tasks

  • May 28, 2025
  • Minghong Geng +3
Cactus Communications logo

Copyright 2026 Cactus Communications. All rights reserved.