• Home
  • Search
  • An Algorithmic Perspective on Imitation Learning
  • Cite Icon380
  • https://doi.org/10.1561/2300000053Copy DOI Icon

An Algorithmic Perspective on Imitation Learning

Show More
  • Abstract
  • Literature Map
  • References
  • Citations
  • Similar Papers
Abstract

As robots and other intelligent agents move from simple environments and problems to more complex, unstructured settings, manually programming their behavior has become increasingly challenging and expensive. Often, it is easier for a teacher to demonstrate a desired behavior rather than attempt to manually engineer it. This process of learning from demonstrations, and the study of algorithms to do so, is called imitation learning. This work provides an introduction to imitation learning. It covers the underlying assumptions, approaches, and how they relate; the rich set of algorithms developed to tackle the problem; and advice on effective tools and implementation. We intend this paper to serve two audiences. First, we want to familiarize machine learning experts with the challenges of imitation learning, particularly those arising in robotics, and the interesting theoretical and practical distinctions between it and more familiar frameworks like statistical supervised learning theory and reinforcement learning. Second, we want to give roboticists and experts in applied artificial intelligence a broader appreciation for the frameworks and tools available for imitation learning. We organize our work by dividing imitation learning into directly replicating desired behavior (sometimes called behavioral cloning [Bain and Sammut, 1996]) and learning the hidden objectives of the desired behavior from demonstrations (called inverse optimal control [Kalman, 1964] or inverse reinforcement learning [Russell, 1998]). In addition to method analysis, we discuss the design decisions a practitioner must make when selecting an imitation learning approach. Moreover, application examples—such as robots that play table tennis Kober and Peters, 2009 and programs that play the game of Go Silver et al. 2016—illustrate the properties and motivations behind different forms of imitation learning. We conclude by presenting a set of open questions and point towards possible future research directions.

Similar Papers
  • Supplementary Content

Sample-Efficient I-Projections for Robot Learning

  • Apr 19, 2021
  • TUbilio (Technical University of Darmstadt)
  • Oleg Arenz
  • PDF
  • Research Article
  • Citations108

A survey of inverse reinforcement learning

  • Feb 08, 2022
  • Artificial Intelligence Review
  • Stephen Adams +2
  • Book Chapter
  • Citations1

Inverse Reinforcement Learning and Imitation Learning

  • Jan 01, 2020
  • Matthew F Dixon +2
  • Research Article

Generating Malicious Demonstration Policies to Exploit Vulnerabilities in Inverse Reinforcement Learning

  • May 19, 2025
  • Arezoo Alipanah +1
  • PDF
  • Research Article
  • Citations15

Generating stable molecules using imitation and reinforcement learning

  • Dec 10, 2021
  • Machine Learning: Science and Technology
  • Søren Ager Meldgaard +5
  • Research Article

Multi‐Agent Reinforcement Learning Algorithm Based on Local Observation Imitation Learning

  • Jan 01, 2025
  • IET Control Theory & Applications
  • Hui Zhang +3
  • Research Article
  • Citations2

A DDPG-based Path Following Control Strategy for Autonomous Vehicles by Integrated Imitation Learning and Feedforward Exploration

  • Sep 08, 2025
  • Chinese Journal of Mechanical Engineering
  • Qianjie Liu +6
  • Conference Article
  • Citations15

Gradient-based inverse risk-sensitive reinforcement learning

  • Dec 01, 2017
  • Eric Mazumdar +3
  • Book Chapter

Estimation of the Change of Agents Behavior Strategy Using State-Action History

  • Jan 01, 2017
  • Shihori Uchida +2
  • PDF
  • Research Article
  • Citations29

A UAV Pursuit-Evasion Strategy Based on DDPG and Imitation Learning

  • May 09, 2022
  • International Journal of Aerospace Engineering
  • Xiaowei Fu +4
  • Book Chapter
  • Citations1001

Robot Programming by Demonstration

  • Jan 01, 2008
  • Aude Billard +3
  • Conference Article
  • Citations13

Reinforced approximate robust nonlinear model predictive control

  • Jun 01, 2021
  • Benjamin Karg +1
  • PDF
  • Conference Article
  • Citations14

An Empirical Comparison on Imitation Learning and Reinforcement Learning for Paraphrase Generation

  • Jan 01, 2019
  • Wanyu Du +1
  • PDF
  • Research Article
  • Citations6

A New AI Approach by Acquisition of Characteristics in Human Decision-Making Process

  • Jun 24, 2024
  • Applied Sciences
  • Yuan Zhou +1
  • Book Chapter
  • Citations13

Deep Active Learning for Autonomous Navigation

  • Jan 01, 2016
  • Ahmed Hussein +2
Cactus Communications logo

Copyright 2026 Cactus Communications. All rights reserved.