• Home
  • Search
  • Run Time Assured Reinforcement Learning for Safe Satellite Docking
  • Cite Icon25
  • https://doi.org/10.2514/1.i011126Copy DOI Icon

Run Time Assured Reinforcement Learning for Safe Satellite Docking

Show More
  • Abstract
  • Literature Map
  • References
  • Citations
  • Similar Papers
Abstract

Reinforcement learning promises high performance in complex tasks as well as low online storage and computation cost. However, the trial-and-error learning approach of reinforcement learning could explore unsafe behavior in the search for an optimal solution. Run time assurance (RTA) approaches can be applied to monitor behavior and ensure safety constraint satisfaction during reinforcement learning. This paper investigates the effect of RTA on reinforcement learning training performance in terms of training efficiency, safety constraint satisfaction, control efficiency, task efficiency, and training duration. For the purposes of demonstration, a custom reinforcement learning environment is created where the objective is to develop a policy that moves a satellite into docking position with another satellite in a two-dimensional relative-motion reference frame. Six different policies are trained. The first features no RTA, the second features no RTA but a higher penalty for safety violations, and four others use different RTA techniques to enforce a dynamic velocity constraint during training. The trained policies are analyzed with standardized test points. It is shown that the policies trained without RTA do not learn to adhere to the constraint, whereas all policies trained with RTA do learn to adhere to the constraint. Although more complex RTA frameworks can be better for operational use, it is found that a simple RTA framework provides the best overall results for reinforcement learning training.

Similar Papers
  • Research Article
  • Citations3

Safe Spacecraft Inspection via Deep Reinforcement Learning and Discrete Control Barrier Functions

  • Sep 06, 2024
  • Journal of Aerospace Information Systems
  • David Van Wijk +3
  • Research Article
  • Citations5

The effect of multi‐tasking training on performance, situation awareness, and workload in simulated air traffic control

  • Jul 01, 2022
  • Applied Cognitive Psychology
  • Stephanie C Black +4
  • Book Chapter
  • Citations13

Policy Transfer via Markov Logic Networks

  • Jan 01, 2010
  • Lisa Torrey +1
  • Single Report

Optimizing Dynamic Resource Allocation in Teamwork

  • Feb 01, 2008
  • Steve W Kozlowski +1
  • Research Article
  • Citations37

In-Memory Realization of Eligibility Traces Based on Conductance Drift of Phase Change Memory for Energy-Efficient Reinforcement Learning.

  • Dec 16, 2021
  • Advanced Materials
  • Yingming Lu +7
  • Research Article

Reinforcement Learning‐Based Fast Frequency Response Using Energy Storage for Remote Microgrids

  • Jan 01, 2026
  • IET Energy Systems Integration
  • Pooja Aslami +5
  • PDF
  • Research Article
  • Citations1

Reinforcement learning optimal control method for multi chiller HVAC system in an existing office building

  • Jul 01, 2024
  • IOP Conference Series: Earth and Environmental Science
  • H Y Wang +3
  • Research Article

Reward Model Evaluation via Automatically-Ranked Policy Alignment

  • Mar 14, 2026
  • Proceedings of the AAAI Conference on Artificial Intelligence
  • Aoran Wang +3
  • Research Article

Dual-Functional IRS Design for ISAC via Reinforcement Learning in Unknown Environment

  • Dec 01, 2025
  • IEEE Transactions on Aerospace and Electronic Systems
  • Weitong Zhai +5
  • Research Article
  • Citations1

A Novel Equivariant Self-Supervised Vector Network for Three-Dimensional Point Clouds

  • Mar 07, 2025
  • Algorithms
  • Kedi Shen +2
  • Conference Article
  • Citations19

Self-Adaptive Double Bootstrapped DDPG

  • Jul 01, 2018
  • Zhuobin Zheng +4
  • Book Chapter
  • Citations20

Active Finite Reward Automaton Inference and Reinforcement Learning Using Queries and Counterexamples

  • Jan 01, 2021
  • Zhe Xu +4
  • Conference Article

Robot Control Using Model-Based Reinforcement Learning With Inverse Kinematics

  • Sep 12, 2022
  • Dario Luipers +5
  • Research Article

Do gender, task complexity, and prior knowledge predict academic achievement?

  • Oct 01, 2025
  • Acta psychologica
  • Abdulrahim R Hakami
  • Research Article
  • Citations12

The effects of manipulating feedback upon children's motives and performance: A propositional statement and empirical evaluation

  • Mar 01, 1975
  • Behavioral Science
  • Charles G Mcclintock +1
Cactus Communications logo

Copyright 2026 Cactus Communications. All rights reserved.