• Home
  • Search
  • Differentiable Arbitrating in Zero-sum Markov Games
  • https://doi.org/10.65109/xuxp9984Copy DOI Icon

Differentiable Arbitrating in Zero-sum Markov Games

  • May 30, 2023
  • Jing Wang +5 more
Show More
  • Abstract
  • Literature Map
  • Similar Papers
Abstract

We initiate the study of how to perturb the reward in a zero-sum Markov game with two players to induce a desirable Nash equilibrium, namely arbitrating. Such a problem admits a bi-level optimization formulation. The lower level requires solving the Nash equilibrium under a given reward function, which makes the overall problem challenging to optimize in an end-to-end way. We propose a backpropagation scheme that differentiates through the Nash equilibrium, which provides the gradient feedback for the upper level. In particular, our method only requires a black-box solver for the (regularized) Nash equilibrium (NE). We develop the convergence analysis for the proposed framework with proper black-box NE solvers and demonstrate the empirical successes in two multi-agent reinforcement learning (MARL) environments. Supplementary for all the proofs in this paper could be found in: https://arxiv.org/abs/2302.10058.

Similar Papers
  • Research Article
  • Citations28

Evaluating semi-cooperative Nash/Stackelberg Q-learning for traffic routes plan in a single intersection

  • Jun 30, 2020
  • Control Engineering Practice
  • Jian Guo +1
  • Conference Article
  • Citations90

Best-Response Multiagent Learning in Non-Stationary Environments

  • Jul 19, 2004
  • Michael C Weinberg +1
  • Dissertation
  • Citations3

Autonomous Network Defence Using Multi-Agent Reinforcement Learning and Self-Play

  • May 30, 2022
  • Roberto G Campbell
  • Research Article
  • Citations13

Multiagent Reinforcement Learning for Antijamming Game of Frequency-Agile Radar

  • Jan 01, 2024
  • IEEE Geoscience and Remote Sensing Letters
  • Jie Geng +5
  • Conference Article

Multi-agent reinforcement learning approach for large cooling water system control

  • Aug 24, 2025
  • Xiao Wang +2
  • Conference Article
  • Citations1

Multi-agent Robust Time Differential Reinforcement Learning Over Communicated Networks

  • Jul 01, 2018
  • Jiahong Li +2
  • Research Article
  • Citations72

Bi-Level Actor-Critic for Multi-Agent Coordination

  • Apr 03, 2020
  • Proceedings of the AAAI Conference on Artificial Intelligence
  • Haifeng Zhang +6
  • PDF
  • Research Article
  • Citations34

Maneuver Strategy Generation of UCAV for within Visual Range Air Combat Based on Multi-Agent Reinforcement Learning and Target Position Prediction

  • Jul 28, 2020
  • Applied Sciences
  • Weiren Kong +4
  • Conference Article
  • Citations1

Fictitious Cross-Play: Learning Global Nash Equilibrium in Mixed Cooperative-Competitive Games

  • May 30, 2023
  • Zelai Xu +4
  • Research Article
  • Citations1

CAGE challenge 4: A scalable multi‐agent reinforcement learning gym for autonomous cyber defence

  • Sep 01, 2025
  • AI Magazine
  • Mitchell Kiely +29
  • Video Transcripts

Reducing the Price of Anarchy in Multi-Agent Learning

  • Apr 20, 2022
  • Underline Science Inc.
  • David Balduzzi +6
  • Research Article
  • Citations65

Automated clash resolution for reinforcement steel design in concrete frames via Q-learning and Building Information Modeling

  • Jan 31, 2020
  • Automation in Construction
  • Jiepeng Liu +5
  • Research Article

Multi‐Agent Reinforcement Learning Algorithm Based on Local Observation Imitation Learning

  • Jan 01, 2025
  • IET Control Theory & Applications
  • Hui Zhang +3
  • Conference Article
  • Citations5

A Tensor Factorization Approach to Generalization in Multi-agent Reinforcement Learning

  • Dec 01, 2012
  • Stefano Bromuri
  • Conference Article
  • Citations14

Dynamic Multi-user Computation Offloading for Mobile Edge Computing using Game Theory and Deep Reinforcement Learning

  • May 16, 2022
  • Peyvand Teymoori +1
Cactus Communications logo

Copyright 2026 Cactus Communications. All rights reserved.