• Home
  • Search
  • Optimizing inventory management through reinforcement learning
  • https://doi.org/10.30574/ijsra.2023.8.1.0137Copy DOI Icon

Optimizing inventory management through reinforcement learning

Show More
  • Abstract
  • Literature Map
  • References
  • Similar Papers
Abstract

Inventory management remains a cornerstone of effective supply chain performance, directly influencing cost efficiency, service quality, and organizational agility. In today’s hypercompetitive and uncertain market environment, inventory decisions must account for complex variables such as fluctuating demand, supply disruptions, lead time variability, and market seasonality. Traditional inventory control models such as the Economic Order Quantity (EOQ), base-stock policies, and (s, S) strategies are often static in nature. They rely on pre-defined parameters and assume stationarity in demand and supply, limiting their ability to respond dynamically to real-time changes. In contrast, Reinforcement Learning (RL) offers a paradigm shift in how inventory decisions can be optimized. As a subfield of machine learning, RL enables agents to learn optimal strategies through repeated interactions with an environment, using trial-and-error exploration and reward-based feedback. RL agents can observe the system state (e.g., inventory levels, demand signals, lead time status), choose actions (e.g., place an order or wait), and receive feedback in the form of rewards (e.g., service level achievements or cost penalties), thus iteratively improving their policies. This study explores how RL can be applied to optimize inventory management in environments characterized by uncertainty and real-time decision-making needs. Specifically, we investigate how different RL algorithms such as Q-learning, Deep Q Networks (DQN), and Policy Gradient methods perform in various inventory scenarios. Additionally, we examine the computational and operational implications of deploying RL in real-world settings, including issues of model convergence, exploration-exploitation tradeoffs, data requirements, and scalability. We also discuss how RL can complement other AI techniques such as demand forecasting models and predictive analytics in creating end-to-end intelligent supply chain solutions. By bridging the gap between theoretical RL frameworks and practical inventory management applications, this paper contributes to both the academic literature and industrial practice. Our goal is to demonstrate that RL is not only a theoretically elegant solution but also a viable tool for achieving inventory efficiency and supply chain resilience.

Similar Papers
  • Research Article
  • Citations212

A Deep Q-Network for the Beer Game: Deep Reinforcement Learning for Inventory Optimization

  • Feb 23, 2021
  • Manufacturing & Service Operations Management
  • Afshin Oroojlooyjadid +3
  • Research Article
  • Citations289

A review on reinforcement learning algorithms and applications in supply chain management

  • Nov 04, 2022
  • International Journal of Production Research
  • Benjamin Rolf +5
  • Research Article
  • Citations134

Reinforcement learning applied to production planning and control

  • Aug 04, 2022
  • International Journal of Production Research
  • Ana Esteso +3
  • Research Article
  • Citations7

Integrated Double Estimator Architecture for Reinforcement Learning.

  • Jan 01, 2020
  • IEEE Transactions on Cybernetics
  • Pingli Lv +4
  • Research Article
  • Citations23

Deep Reinforcement Learning With Modulated Hebbian Plus Q-Network Architecture.

  • May 01, 2022
  • IEEE Transactions on Neural Networks and Learning Systems
  • Pawel Ladosz +7
  • Research Article
  • Citations5

Feasibility of reinforcement learning for UAV-based target searching in a simulated communication denied environment

  • Feb 27, 2020
  • SCIENTIA SINICA Informationis
  • Songlin Hou +6
  • Research Article
  • Citations30

Case-based myopic reinforcement learning for satisfying target service level in supply chain

  • Jul 17, 2007
  • Expert Systems with Applications
  • I Kwon +3
  • Video Transcripts

Active Screening for Recurrent Diseases: A Reinforcement Learning Approach

  • Apr 11, 2021
  • Underline Science Inc.
  • Milind Tambe +3
  • Research Article
  • Citations76

The flying sidekick traveling salesman problem with stochastic travel time: A reinforcement learning approach

  • Jun 28, 2022
  • Transportation Research Part E: Logistics and Transportation Review
  • Zeyu Liu +2
  • Research Article
  • Citations1

A-205 Evaluating the Cost and Importance of Supply Chain Resilience in the Clinical Laboratory

  • Oct 02, 2024
  • Clinical Chemistry
  • E Jurinic
  • Conference Article
  • Citations12

Solve the inverted pendulum problem base on DQN algorithm

  • Jun 01, 2019
  • Xiaoqian Li +2
  • PDF
  • Research Article
  • Citations1

Inventory decision-making by small Sowetan retailers

  • Sep 17, 2018
  • Journal of Transport and Supply Chain Management
  • Themari Eicker +1
  • Research Article

The Role of Reinforcement Learning in Advancing Artificial Intelligence: An Experimental Study with Q-Learning and DQN

  • Aug 24, 2025
  • The Asian Bulletin of Big Data Management
  • Maryam Gul +5
  • Research Article
  • Citations1

Supply chain collaboration, supply chain disruption and supply chain resilience: a fuzzy-set QCA approach

  • Oct 21, 2025
  • Nankai Business Review International
  • Ran Zhuo +2
  • Conference Article

On the Design of Safe Continual RL Methods for Control of Nonlinear Systems

  • Jun 24, 2025
  • Austin Coursey +2
Cactus Communications logo

Copyright 2026 Cactus Communications. All rights reserved.