• Home
  • Search
  • Navigating complex decision spaces: Problems and paradigms in sequential choice.
  • Open Access IconOpen Access
  • Cite Icon39
  • https://doi.org/10.1037/a0033455Copy DOI Icon

Navigating complex decision spaces: Problems and paradigms in sequential choice.

Show More
  • Abstract
  • Highlights & Summary
  • Literature Map
  • References
  • Citations
  • Similar Papers
Abstract

To behave adaptively, we must learn from the consequences of our actions. Doing so is difficult when the consequences of an action follow a delay. This introduces the problem of temporal credit assignment. When feedback follows a sequence of decisions, how should the individual assign credit to the intermediate actions that comprise the sequence? Research in reinforcement learning provides 2 general solutions to this problem: model-free reinforcement learning and model-based reinforcement learning. In this review, we examine connections between stimulus-response and cognitive learning theories, habitual and goal-directed control, and model-free and model-based reinforcement learning. We then consider a range of problems related to temporal credit assignment. These include second-order conditioning and secondary reinforcers, latent learning and detour behavior, partially observable Markov decision processes, actions with distributed outcomes, and hierarchical learning. We ask whether humans and animals, when faced with these problems, behave in a manner consistent with reinforcement learning techniques. Throughout, we seek to identify neural substrates of model-free and model-based reinforcement learning. The former class of techniques is understood in terms of the neurotransmitter dopamine and its effects in the basal ganglia. The latter is understood in terms of a distributed network of regions including the prefrontal cortex, medial temporal lobes, cerebellum, and basal ganglia. Not only do reinforcement learning techniques have a natural interpretation in terms of human and animal behavior but they also provide a useful framework for understanding neural reward valuation and action selection.

Similar Papers
  • PDF
  • Research Article
  • Citations32

Parallel model-based and model-free reinforcement learning for card sorting performance

  • Sep 22, 2020
  • Scientific Reports
  • Alexander Steinke +2
  • Conference Article
  • Citations2

Context-dependent meta-control for reinforcement learning using a Dirichlet process Gaussian mixture model

  • Jan 01, 2018
  • Dongjae Kim +1
  • PDF
  • Research Article
  • Citations7

Improving Model-Based Deep Reinforcement Learning with Learning Degree Networks and Its Application in Robot Control

  • Mar 04, 2022
  • Journal of Robotics
  • Guoqing Ma +3
  • Conference Article
  • Citations19

An Overview of Robust Reinforcement Learning

  • Oct 30, 2020
  • Shiyu Chen +1
  • Conference Article
  • Citations1

Hierarchical Control Architecture Regulating Competition between Model-Based and Context-Dependent Model-Free Reinforcement Learning Strategies

  • Oct 01, 2018
  • Dongjae Kim +2
  • Peer Review Report

Reviewer #2 (Public review): Neural signatures of model-based and model-free reinforcement learning across prefrontal cortex and striatum

  • Feb 27, 2026
  • Bruno Miranda +5
  • Components

Effects of subclinical depression on prefrontal–striatal model-based and model-free learning

  • May 14, 2021
  • Samuel J Gershman +4
  • Research Article
  • Citations321

Habits, action sequences and reinforcement learning

  • Apr 01, 2012
  • European Journal of Neuroscience
  • Amir Dezfouli +1
  • Research Article
  • Citations185

Distributed Coding of Actual and Hypothetical Outcomes in the Orbital and Dorsolateral Prefrontal Cortex

  • May 01, 2011
  • Neuron
  • Hiroshi Abe +1
  • Research Article
  • Citations12

Fuzzy-based predictive deep reinforcement learning for robust and constrained optimal control of industrial solar thermal plants

  • Feb 24, 2024
  • Applied Soft Computing
  • Fitsum Bekele Tilahun
  • PDF
  • Supplementary Content
  • Citations3

Understanding cingulotomy’s therapeutic effect in OCD through computer models

  • Jan 10, 2023
  • Frontiers in Integrative Neuroscience
  • Mohamed A Sherif +3
  • Research Article
  • Citations379

Intelligent Multi-Microgrid Energy Management Based on Deep Neural Network and Model-Free Reinforcement Learning

  • Jul 30, 2019
  • IEEE Transactions on Smart Grid
  • Yan Du +1
  • Research Article

Diversity-Driven Model Ensemble Adaptive Trust Region Policy Optimization

  • Apr 01, 2026
  • IEEE Transactions on Systems, Man, and Cybernetics: Systems
  • Haotian Xu +4
  • Conference Article

Pi-DON: Physics-Informed Deep Operator Network for Control-Oriented Modeling of Thermal Systems in Cluster of Buildings

  • Aug 24, 2025
  • Muhammad Hafeez Saeed +2
  • Conference Article

Learning Control for Robotic Manipulator with Free Energy

  • Jul 01, 2020
  • Yazhou Hu +4
Cactus Communications logo

Copyright 2026 Cactus Communications. All rights reserved.