• Home
  • Search
  • Model-Free IRL Using Maximum Likelihood Estimation
  • Cite Icon12
  • https://doi.org/10.1609/aaai.v33i01.33013951Copy DOI Icon

Model-Free IRL Using Maximum Likelihood Estimation

Show More
  • Abstract
  • Literature Map
  • References
  • Citations
  • Similar Papers
Abstract

The problem of learning an expert’s unknown reward function using a limited number of demonstrations recorded from the expert’s behavior is investigated in the area of inverse reinforcement learning (IRL). To gain traction in this challenging and underconstrained problem, IRL methods predominantly represent the reward function of the expert as a linear combination of known features. Most of the existing IRL algorithms either assume the availability of a transition function or provide a complex and inefficient approach to learn it. In this paper, we present a model-free approach to IRL, which casts IRL in the maximum likelihood framework. We present modifications of the model-free Q-learning that replace its maximization to allow computing the gradient of the Q-function. We use gradient ascent to update the feature weights to maximize the likelihood of expert’s trajectories. We demonstrate on two problem domains that our approach improves the likelihood compared to previous methods.

Similar Papers
  • Book Chapter

Estimation of the Change of Agents Behavior Strategy Using State-Action History

  • Jan 01, 2017
  • Shihori Uchida +2
  • Supplementary Content

Sample-Efficient I-Projections for Robot Learning

  • Apr 19, 2021
  • TUbilio (Technical University of Darmstadt)
  • Oleg Arenz
  • Research Article
  • Citations306

Advanced planning for autonomous vehicles using reinforcement learning and deep inverse reinforcement learning

  • Jan 15, 2019
  • Robotics and Autonomous Systems
  • Changxi You +3
  • Conference Article
  • Citations15

Gradient-based inverse risk-sensitive reinforcement learning

  • Dec 01, 2017
  • Eric Mazumdar +3
  • Research Article

Data-Based Inverse Reinforcement Learning for Nonlinear Systems With Control Constraints

  • Jan 01, 2025
  • IEEE Transactions on Systems Man and Cybernetics Systems
  • Huaipin Zhang +4
  • Book Chapter
  • Citations15

Analysis of Inverse Reinforcement Learning with Perturbed Demonstrations

  • Jan 01, 2010
  • Frontiers in artificial intelligence and applications
  • Melo Francisco S +2
  • PDF
  • Research Article
  • Citations108

A survey of inverse reinforcement learning

  • Feb 08, 2022
  • Artificial Intelligence Review
  • Stephen Adams +2
  • Research Article
  • Citations1

Dynamic Inverse Reinforcement Learning for Feedback-driven Reward Estimation in Brain Machine Interface Tasks.

  • Jul 15, 2024
  • Annual International Conference of the IEEE Engineering in Medicine and Biology Society. IEEE Engineering in Medicine and Biology Society. Annual International Conference
  • Jieyuan Tan +1
  • Research Article

A Mutual Information-Based Assessment of Reverse Engineering on Rewards of Reinforcement Learning

  • Oct 01, 2023
  • IEEE Transactions on Artificial Intelligence
  • Tong Chen +8
  • Research Article

Generating Malicious Demonstration Policies to Exploit Vulnerabilities in Inverse Reinforcement Learning

  • May 19, 2025
  • Arezoo Alipanah +1
  • Research Article
  • Citations56

Energy-efficient and damage-recovery slithering gait design for a snake-like robot based on reinforcement learning and inverse reinforcement learning

  • Jun 16, 2020
  • Neural Networks
  • Zhenshan Bing +4
  • Conference Article

Online Inverse Reinforcement Learning Under Occlusion

  • May 08, 2019
  • Saurabh Arora +2
  • Research Article

Initial Excitation‐Based Inverse Reinforcement Learning for Continuous‐Time Linear Non‐Zero‐Sum Games

  • May 26, 2025
  • International Journal of Robust and Nonlinear Control
  • Hongyang Li +2
  • Book Chapter
  • Citations1

Inverse Reinforcement Learning and Imitation Learning

  • Jan 01, 2020
  • Matthew F Dixon +2
  • Conference Article
  • Citations71

Learning from Demonstration for Shaping through Inverse Reinforcement Learning

  • May 09, 2016
  • Halit Bener Suay +3
Cactus Communications logo

Copyright 2026 Cactus Communications. All rights reserved.