• Home
  • Search
  • Measuring and characterizing generalization in deep reinforcement learning
  • Open Access IconOpen Access
  • Cite Icon36
  • https://doi.org/10.1002/ail2.45Copy DOI Icon

Measuring and characterizing generalization in deep reinforcement learning

Show More
  • Abstract
  • Literature Map
  • References
  • Citations
  • Similar Papers
Abstract

Abstract Deep reinforcement learning (RL) methods have achieved remarkable performance on challenging control tasks. Observations of the resulting behavior give the impression that the agent has constructed a generalized representation that supports insightful action decisions. We re‐examine what is meant by generalization in RL, and propose several definitions based on an agent's performance in on‐policy, off‐policy, and unreachable states. We propose a set of practical methods for evaluating agents with these definitions of generalization. We demonstrate these techniques on a common benchmark task for deep RL, and we show that the learned networks make poor decisions for states that differ only slightly from on‐policy states, even though those states are not selected adversarially. We focus our analyses on the deep Q‐networks (DQNs) that kicked off the modern era of deep RL. Taken together, these results call into question the extent to which DQNs learn generalized representations, and suggest that more experimentation and analysis is necessary before claims of representation learning can be supported.

Similar Papers
  • Research Article
  • Citations6

Break through the limits of learning by machines

  • Sep 20, 2016
  • Chinese Science Bulletin
  • Zhongzhi Shi
  • Research Article
  • Citations1

DDPG Agent to Swing Up and Balance Cart- Pole System

  • Apr 09, 2021
  • International Journal of Advanced Research in Science, Communication and Technology
  • Buvanesh Pandian V
  • Conference Article
  • Citations3

Attention-based Partial Decoupling of Policy and Value for Generalization in Reinforcement Learning

  • Dec 01, 2022
  • Nasik Muhammad Nafi +2
  • Conference Article
  • Citations16

Traffic Signal Control with Deep Reinforcement Learning

  • Dec 01, 2019
  • Tongyu Zhao +2
  • Research Article

Recommendation of deep reinforcement learning based on value function considering error reduction.

  • Oct 07, 2025
  • Scientific reports
  • Jinlian Zhou +4
  • Book Chapter
  • Citations2

Evaluation of DQN and Double DQN Algorithms in Flappy Bird Environment

  • Jan 01, 2024
  • Zhenyu Chen
  • Research Article
  • Citations194

Deep reinforcement learning based energy management for a hybrid electric vehicle

  • Apr 14, 2020
  • Energy
  • Guodong Du +5
  • Book Chapter
  • Citations4

Nitty-Gritty of Deep Reinforcement Learning for the Healthcare Sector

  • Oct 18, 2023
  • Vaishnavi Kumari +5
  • Conference Article
  • Citations1

Research on AI-driven personalized learning path planning and effectiveness under dual-system teaching mode

  • Apr 18, 2025
  • Ling Chen
  • Research Article
  • Citations22

A Deep Ensemble Method for Multi-Agent Reinforcement Learning: A Case Study on Air Traffic Control

  • May 17, 2021
  • Proceedings of the International Conference on Automated Planning and Scheduling
  • Supriyo Ghosh +4
  • Discussion
  • Citations1

Optimal control strategy for COVID-19 concerning both life and economy based on deep reinforcement learning* *Project supported by the National Natural Science Foundation of China (Grant No. 61873186) and the Tianjin Natural Science Foundation, China (Grant No. 17JCZDJC38300).

  • Oct 22, 2021
  • Chinese Physics B
  • Wei Deng +2
  • Research Article
  • Citations163

Reinforcement Learning and Deep Learning Based Lateral Control for Autonomous Driving [Application Notes

  • May 01, 2019
  • IEEE Computational Intelligence Magazine
  • Dong Li +3
  • Research Article
  • Citations1

Advances in deep reinforcement learning enable better predictions of human behavior in time-continuous tasks

  • Dec 04, 2025
  • PLOS One
  • Sabine Haberland +2
  • Research Article
  • Citations14

MASAC-based confrontation game method of UAV clusters

  • Dec 01, 2022
  • SCIENTIA SINICA Informationis
  • 健 薛 +6
  • Conference Article
  • Citations9

AdaPool

  • Nov 18, 2020
  • Marina Haliem +2
Cactus Communications logo

Copyright 2026 Cactus Communications. All rights reserved.