• Home
  • Search
  • Dynamic Inverse Reinforcement Learning for Feedback-driven Reward Estimation in Brain Machine Interface Tasks.
  • Cite Icon1
  • https://doi.org/10.1109/embc53108.2024.10782800Copy DOI Icon

Dynamic Inverse Reinforcement Learning for Feedback-driven Reward Estimation in Brain Machine Interface Tasks.

  • Abstract
  • Literature Map
  • References
  • Citations
  • Similar Papers
Abstract

Reinforcement learning (RL)-based brain machine interfaces (BMIs) provide a promising solution for paralyzed people. Enhancing the decoding performance of RL-based BMIs relies on the design of effective reward signals. Inverse reinforcement learning (IRL) offers an approach to infer subjects' own evaluation from the observed behavior. However, applying IRL to extract reward information in complex BMI tasks requires consideration of the dynamics of subjects' goal during the control process. This dynamic nature of subjects' evaluation requires the IRL method to be able to estimate a time varying reward function. Previous IRL methods applied in BMI systems only estimated a static reward function. Existing IRL algorithms for dynamic reward estimation employ optimization methods to approximate the reward map for each state at each time, which demands substantial amounts of data to achieve convergence. In this paper, we propose a dynamic IRL method to estimate the feedback-driven reward of subjects during BMI tasks. We utilize a state-observation model to continuously infer the reward value for each state, with sensory feedback serving as the external input to model the transition process of the reward. We evaluate our proposed method on a simulated BMI fetch task, which is a multistep task with a time varying reward function. Our method demonstrates improved reward estimation close to the ground truth value, and it significantly outperforms the existing dynamic IRL method when the map size exceeds 25(p<0.01). These preliminary results suggests that the dynamic IRL method for feedback-driven reward estimation holds potential for improving the design of RL-based BMIs.

Similar Papers
  • Book Chapter

Estimation of the Change of Agents Behavior Strategy Using State-Action History

  • Jan 01, 2017
  • Shihori Uchida +2
  • Research Article

Initial Excitation‐Based Inverse Reinforcement Learning for Continuous‐Time Linear Non‐Zero‐Sum Games

  • May 26, 2025
  • International Journal of Robust and Nonlinear Control
  • Hongyang Li +2
  • Video Transcripts

Individual-Level Inverse Reinforcement Learning for Mean Field Games

  • Apr 20, 2022
  • Underline Science Inc.
  • Yang Chen
  • Research Article
  • Citations2

Human-like decision making for automated vehicles using driving style-guided inverse reinforcement learning method

  • Aug 05, 2025
  • Proceedings of the Institution of Mechanical Engineers, Part D: Journal of Automobile Engineering
  • Wenjie Liu +3
  • Supplementary Content

Sample-Efficient I-Projections for Robot Learning

  • Apr 19, 2021
  • TUbilio (Technical University of Darmstadt)
  • Oleg Arenz
  • Research Article
  • Citations12

Model-Free IRL Using Maximum Likelihood Estimation

  • Jul 17, 2019
  • Proceedings of the AAAI Conference on Artificial Intelligence
  • Vinamra Jain +2
  • Conference Article
  • Citations15

Gradient-based inverse risk-sensitive reinforcement learning

  • Dec 01, 2017
  • Eric Mazumdar +3
  • PDF
  • Research Article
  • Citations108

A survey of inverse reinforcement learning

  • Feb 08, 2022
  • Artificial Intelligence Review
  • Stephen Adams +2
  • PDF
  • Research Article
  • Citations2

Model-free inverse reinforcement learning with multi-intention, unlabeled, and overlapping demonstrations

  • Nov 30, 2022
  • Machine Learning
  • Ariyan Bighashdel +2
  • Book Chapter
  • Citations1

Inverse Reinforcement Learning and Imitation Learning

  • Jan 01, 2020
  • Matthew F Dixon +2
  • Conference Article

Online Inverse Reinforcement Learning Under Occlusion

  • May 08, 2019
  • Saurabh Arora +2
  • Research Article

A Mutual Information-Based Assessment of Reverse Engineering on Rewards of Reinforcement Learning

  • Oct 01, 2023
  • IEEE Transactions on Artificial Intelligence
  • Tong Chen +8
  • Research Article

Generating Malicious Demonstration Policies to Exploit Vulnerabilities in Inverse Reinforcement Learning

  • May 19, 2025
  • Arezoo Alipanah +1
  • Research Article
  • Citations56

Energy-efficient and damage-recovery slithering gait design for a snake-like robot based on reinforcement learning and inverse reinforcement learning

  • Jun 16, 2020
  • Neural Networks
  • Zhenshan Bing +4
  • Research Article
  • Citations306

Advanced planning for autonomous vehicles using reinforcement learning and deep inverse reinforcement learning

  • Jan 15, 2019
  • Robotics and Autonomous Systems
  • Changxi You +3
Cactus Communications logo

Copyright 2026 Cactus Communications. All rights reserved.