• Home
  • Search
  • Open-ended coordination for multi-agent systems using modular open policies
  • Cite Icon2
  • https://doi.org/10.1007/s10458-025-09723-7Copy DOI Icon

Open-ended coordination for multi-agent systems using modular open policies

Show More
  • Abstract
  • PDF
  • Literature Map
  • References
  • Citations
  • Similar Papers
Abstract

Abstract Significant multi-agent advances addressing the challenge of learning policies for acting in ad hoc teamwork have been made. In ad hoc teamwork, a team of agents must cooperate effectively without prior coordination or communication. Many existing approaches, however, struggle to perform well in open environments where the setting can change significantly during deployment. This paper presents a new reinforcement learning approach to tackle collaboration in open environments controlling one agent with a changing number of distinct other agents, each with an individual task. The approach uses policy blending based on an online goal inference module and a collection of learned policies modeling the individual interaction impact between the agent and populations of partners with different tasks. Blending is done using the estimated goals of others and a posterior-based action blending with entropy adjustment and regularization. Our approach addresses issues of existing policy blending mechanisms, such as handling conflicting modes in action distributions leading to oscillation and instability and adapting to uncertain states dynamically. In experiments in two collaborative open environments based on Overcooked and Level-based Foraging, our approach outperforms a baseline learner, trained with the joint reward of all agents, across changes to both agents and tasks. Ablation studies further highlight the importance of our posterior-based blending mechanism to achieve high rewards as well as the provided goal weighting. The proposed approach provides an important step towards the application of reinforcement learning to AI assistance beyond strictly closed worlds and towards more realistic scenarios.

Loading PDF

Similar Papers
  • Research Article
  • Citations13

Reinforcement Learning for Clinical Applications.

  • Feb 08, 2023
  • Clinical Journal of the American Society of Nephrology
  • Kia Khezeli +5
  • Research Article

Reinforcement Learning for Automated Literature Screening: Enhancing E-Learning and University Research Classification in Computer Science

  • Feb 03, 2026
  • International Journal of Information Technology and Computer Science
  • Enes Bajrami +2
  • Research Article
  • Citations134

Reinforcement learning applied to production planning and control

  • Aug 04, 2022
  • International Journal of Production Research
  • Ana Esteso +3
  • Dissertation

Decision-making with cooperative multi-agent reinforcement learning

  • Jan 01, 2025
  • Jing Sun
  • Research Article

Ethical Decision-Making in Autonomous Robots Using Reinforcement Learning: A Comprehensive Exploration

  • Jan 07, 2026
  • International Journal For Multidisciplinary Research
  • Sridev Anoop
  • Research Article
  • Citations37

Curriculum Reinforcement Learning From Avoiding Collisions to Navigating Among Movable Obstacles in Diverse Environments

  • May 01, 2023
  • IEEE Robotics and Automation Letters
  • Hsueh-Cheng Wang +7
  • Research Article
  • Citations4

Energy Demand Response in a Food-Processing Plant: A Deep Reinforcement Learning Approach

  • Dec 20, 2024
  • Energies
  • Philipp Wohlgenannt +4
  • Research Article
  • Citations18

Learning scheduling control knowledge through reinforcements

  • Mar 01, 2000
  • International Transactions in Operational Research
  • Kazuo Miyashita
  • Research Article
  • Citations15

Learning scheduling control knowledge through reinforcements

  • Mar 01, 2000
  • International Transactions in Operational Research
  • K Miyashita
  • Research Article
  • Citations25

Bootstrapping $Q$ -Learning for Robotics From Neuro-Evolution Results

  • Mar 23, 2017
  • IEEE Transactions on Cognitive and Developmental Systems
  • Matthieu Zimmer +1
  • Book Chapter
  • Citations3

Reinforcement Learning: An Industrial Perspective

  • Jan 01, 2021
  • Amit Surana
  • Research Article
  • Citations1

Perspectives for the Application of Reinforcement Learning for the Integrated Order-Dispatching and Maintenance Scheduling

  • Jan 01, 2024
  • IFAC PapersOnLine
  • Djonathan L.O Quadras +4
  • Research Article
  • Citations14

Application of Reinforcement Learning to Generate Non-linear Optimal Feedback Controller for Ship's Automatic Berthing System

  • Jan 01, 2023
  • IFAC PapersOnLine
  • Naoki Mizuno +1
  • Research Article

Application of Reinforcement Learning to Solve Rubrik’s Cube with Markov Decision Process

  • Oct 08, 2025
  • Jurnal RESTI (Rekayasa Sistem dan Teknologi Informasi)
  • Defni +6
  • Research Article
  • Citations13

Entropy regularized actor-critic based multi-agent deep reinforcement learning for stochastic games

  • Oct 20, 2022
  • Information Sciences
  • Dong Hao +3
Cactus Communications logo

Copyright 2026 Cactus Communications. All rights reserved.