• Home
  • Search
  • Multi-Objective Policy Gradients with Topological Constraints
  • Open Access IconOpen Access
  • Cite Icon2
  • https://doi.org/10.1109/iros47612.2022.9982278Copy DOI Icon

Multi-Objective Policy Gradients with Topological Constraints

  • Oct 23, 2022
  • Kyle Hollins Wray +3 more
Show More
  • Abstract
  • Literature Map
  • References
  • Citations
  • Similar Papers
Abstract

Multi-objective optimization models that encode ordered sequential constraints provide a solution to model various challenging problems including encoding preferences, modeling a curriculum, and enforcing measures of safety. A recently developed theory of topological Markov decision processes (TMDPs) captures this range of problems for the case of discrete states and actions. In this work, we extend TMDPs towards continuous spaces and unknown transition dynamics by formulating, proving, and implementing the policy gradient theorem for TMDPs. This theoretical result enables the creation of TMDP learning algorithms that use function approximators, and can generalize existing deep reinforcement learning (DRL) approaches. Specifically, we present a new algorithm for a policy gradient in TMDPs by a simple extension of the proximal policy optimization (PPO) algorithm. We demonstrate this on a real-world multiple-objective navigation problem with an arbitrary ordering of objectives both in simulation and on a real robot.

Similar Papers
  • Research Article
  • Citations22

An optimal solutions-guided deep reinforcement learning approach for online energy storage control

  • Mar 07, 2024
  • Applied Energy
  • Gaoyuan Xu +6
  • Research Article
  • Citations27

Multi-agent quantum-inspired deep reinforcement learning for real-time distributed generation control of 100% renewable energy systems

  • Jan 07, 2023
  • Engineering Applications of Artificial Intelligence
  • Dan Liu +6
  • Research Article
  • Citations54

Energy efficient task scheduling based on deep reinforcement learning in cloud environment: A specialized review

  • Oct 14, 2023
  • Future Generation Computer Systems
  • Huanhuan Hou +2
  • PDF
  • Research Article
  • Citations6

Deep Reinforcement Learning for Optimizing Restricted Access Window in IEEE 802.11ah MAC Layer

  • May 10, 2024
  • Sensors (Basel, Switzerland)
  • Xiaojun Jiang +4
  • Research Article
  • Citations33

Improving the interpretability of deep reinforcement learning in urban drainage system operation

  • Nov 23, 2023
  • Water Research
  • Wenchong Tian +4
  • Research Article
  • Citations122

Optimal scheduling of island integrated energy systems considering multi-uncertainties and hydrothermal simultaneous transmission: A deep reinforcement learning approach

  • Dec 29, 2022
  • Applied Energy
  • Yang Li +3
  • Research Article
  • Citations5

Automating the optimization of proton PBS treatment planning for head and neck cancers using policy gradient-based deep reinforcement learning.

  • Jan 31, 2025
  • Medical physics
  • Qingqing Wang +1
  • Research Article
  • Citations75

Dealing with Limited Backhaul Capacity in Millimeter-Wave Systems: A Deep Reinforcement Learning Approach

  • Dec 27, 2018
  • IEEE Communications Magazine
  • Mingjie Feng +1
  • Research Article

Spectrum Allocation for Covert Communications in Cellular-Enabled UAV Networks: A Deep Reinforcement Learning Approach

  • Jan 01, 2022
  • Journal of Networking and Network Applications
  • Xinzhe Pi +1
  • Research Article
  • Citations5

Distributed Routing and Data Scheduling in IPNs With GNN-Based Multiagent DRL

  • Jun 15, 2025
  • IEEE Internet of Things Journal
  • Xixuan Zhou +6
  • Research Article
  • Citations19

Deep reinforcement learning for cerebral anterior vessel tree extraction from 3D CTA images.

  • Feb 01, 2023
  • Medical Image Analysis
  • Jiahang Su +6
  • Conference Article
  • Citations1

Multi-Resource Scheduling for Multiple Service Function Chains with Deep Reinforcement Learning

  • Jan 01, 2023
  • Rui He +4
  • Research Article
  • Citations112

Dynamic Pricing for EV Charging Stations: A Deep Reinforcement Learning Approach

  • Jun 01, 2022
  • IEEE Transactions on Transportation Electrification
  • Zhonghao Zhao +1
  • Conference Article
  • Citations4

Learning Similar Tasks Based On PPO By Transferring Trajectory

  • May 01, 2019
  • An Guo +2
  • Research Article

A general algorithm for eliminating critical conditions for solving the problem of controlling a real walking robot based on deep reinforcement learning methods

  • Mar 01, 2025
  • Программные системы и вычислительные методы
  • Vasily Vasil'Evich Kashko +1
Cactus Communications logo

Copyright 2026 Cactus Communications. All rights reserved.