• Home
  • Search
  • Coordinating multi-agent reinforcement learning with limited communication
  • Cite Icon141
  • https://doi.org/10.5555/2484920.2485093Copy DOI Icon

Coordinating multi-agent reinforcement learning with limited communication

  • May 6, 2013
  • Chongjie Zhang +1 more
Show More
  • Abstract
  • Literature Map
  • References
  • Citations
  • Similar Papers
Abstract

Coordinated multi-agent reinforcement learning (MARL) provides a promising approach to scaling learning in large cooperative multi-agent systems. Distributed constraint optimization (DCOP) techniques have been used to coordinate action selection among agents during both the learning phase and the policy execution phase (if learning is off-line) to ensure good overall system performance. However, running DCOP algorithms for each action selection through the whole system results in significant communication among agents, which is not practical for most applications with limited communication bandwidth. In this paper, we develop a learning approach that generalizes previous coordinated MARL approaches that use DCOP algorithms and enables MARL to be conducted over a spectrum from independent learning (without communication) to fully coordinated learning depending on agents' communication bandwidth. Our approach defines an interaction measure that allows agents to dynamically identify their beneficial coordination set (i.e., whom to coordinate with) in different situations and to trade off its performance and communication cost. By limiting their coordination set, agents dynamically decompose the coordination network in a distributed way, resulting in dramatically reduced communication for DCOP algorithms without significantly affecting overall learning performance. Essentially, our learning approach conducts co-adaptation of agents' policy learning and coordination set identification, which outperforms approaches that sequence them.

Similar Papers
  • Conference Article
  • Citations30

Experimental analysis of privacy loss in DCOP algorithms

  • May 08, 2006
  • Rachel Greenstadt +3
  • Conference Article

Analyzing the performance of distributed algorithms

  • Aug 28, 2007
  • Robert N Lass +2
  • Conference Article
  • Citations10

A common gradient in multi-agent reinforcement learning

  • Jun 04, 2012
  • Michael Kaisers +2
  • Conference Article
  • Citations90

Best-Response Multiagent Learning in Non-Stationary Environments

  • Jul 19, 2004
  • Michael C Weinberg +1
  • Conference Article
  • Citations19

Decentralized Bayesian reinforcement learning for online agent collaboration

  • May 16, 2018
  • W T Luke Teacy +6
  • Conference Article
  • Citations98

Evaluating the performance of DCOP algorithms in a real world, dynamic problem

  • May 12, 2008
  • Robert Junges +1
  • Conference Article
  • Citations4

Facilitating Communication for First Responders Using Dynamic Distributed Constraint Optimization

  • May 01, 2008
  • Robert N Lass +4
  • Conference Article
  • Citations28

A few good agents: multi-agent social learning

  • May 12, 2008
  • Jean Oh +1
  • Research Article
  • Citations12

Speeding up distributed pseudo-tree optimization procedures with cross edge consistency to solve DCOPs

  • Oct 10, 2020
  • Applied Intelligence
  • Mashrur Rashik +5
  • Conference Article
  • Citations12

Effective Disaster Evacuation by Solving the Distributed Constraint Optimization Problem

  • Aug 01, 2013
  • Katsuya Kinoshita +2
  • Conference Instance
  • Citations4

A Near-Optimal Node-to-Agent Mapping Heuristic for GDL-Based DCOP Algorithms in Multi-Agent Systems

  • Jul 09, 2018
  • Md Mosaddek Khan +3
  • Book Chapter
  • Citations11

A Distributed Constraint Optimization Approach for Vessel Rotation Planning

  • Jan 01, 2014
  • Shijie Li +2
  • Research Article
  • Citations20

CoCoA: A Non-Iterative Approach to a Local Search (A)DCOP Solver

  • Feb 12, 2017
  • Proceedings of the AAAI Conference on Artificial Intelligence
  • Cornelis Jan Van Leeuwen +1
  • Conference Article

ER-DCOPs: A Framework for Distributed Constraint Optimization with Uncertainty in Constraint Utilities

  • May 09, 2016
  • Tiep Le +4
  • Book Chapter

A Study of Relaxation Approaches for Asymmetric Constraint Optimization Problems

  • Jan 01, 2018
  • Toshihiro Matsui +1
Cactus Communications logo

Copyright 2026 Cactus Communications. All rights reserved.