• Home
  • Search
  • Efficient Resource Management for Real-time AI Systems in the Cloud using Reinforcement Learning
  • Cite Icon15
  • https://doi.org/10.1109/ic3i61595.2024.10828656Copy DOI Icon

Efficient Resource Management for Real-time AI Systems in the Cloud using Reinforcement Learning

  • Sep 18, 2024
  • Vinay Mallikarjunaradhya +5 more
Show More
  • Abstract
  • Literature Map
  • References
  • Citations
  • Similar Papers
Abstract

The advent of artificial intelligence (AI) has driven an emergence in applications demanding online responses to immense amounts of data. Typically, the cloud-based deployment of these applications requires resource management to be optimized for high performance while meeting cost efficiency. However, conventional static resource allocation methods may only partially apply to the context of real-time AI applications as they are dynamic, unpredictable, and require more efficient & adaptive ways of allocating resources. To address this challenge, and as a solution in this research, we introduce an RL-based method for online resource management of cloud-based AI systems. Reinforcement learning ($\mathbf{R L}$) is a machine-learning algorithm that allows systems to learn how best to perform tasks in changing environments based on feedback from those interactions with the environment. The third way our proposed approach utilizes the weapon of RL entails allowing for dynamic resource allocation so that resources are allocated when needed rather than from a static method as set at design time.

Similar Papers
  • Conference Article
  • Citations3

Event-Triggered and Time-Triggered Duration Calculus for Model-Free Reinforcement Learning

  • Dec 01, 2021
  • Kalyani Dole +4
  • Research Article
  • Citations100

Trajectory smoothing method using reinforcement learning for computer numerical control machine tools

  • Aug 13, 2019
  • Robotics and Computer-Integrated Manufacturing
  • Bingran Li +3
  • Conference Article
  • Citations6

An enhancement of relational reinforcement learning

  • Jun 01, 2008
  • Renato R Da Silva +2
  • Research Article
  • Citations6

Optimized Resource Allocation for Cloud-Native 6G Networks: Zero-Touch ML Models in Microservices-Based VNF Deployments

  • Jul 01, 2025
  • IEEE Network
  • Swarna Bindu Chetty +6
  • Research Article
  • Citations126

Deep Q-Learning With Q-Matrix Transfer Learning for Novel Fire Evacuation Environment

  • Dec 01, 2021
  • IEEE Transactions on Systems, Man, and Cybernetics: Systems
  • Jivitesh Sharma +3
  • Research Article
  • Citations1

A bio-inspired neuro-adaptive deep reinforcement learning approach for real-time solar tracking system to enhance photovoltaic efficiency

  • Jan 01, 2026
  • Energy Conversion and Management: X
  • Udit Mamodiya +4
  • Research Article
  • Citations3

Systematic Review of Reinforcement Learning in Process Industries: A Contextual and Taxonomic Approach

  • Dec 07, 2025
  • Applied Sciences
  • Marco Antonio Paz Ramos +1
  • Research Article

Toward robust, interactive, and human‐aligned AI systems

  • Aug 29, 2025
  • AI Magazine
  • Daniel S Brown
  • Research Article
  • Citations117

Bringing Deep Learning at the Edge of Information-Centric Internet of Things

  • Jan 01, 2019
  • IEEE Communications Letters
  • Hakima Khelifi +6
  • Research Article

An Advantage Actor–Critic-Based Quality of Service-Aware Routing Optimization Mechanism for Optical Satellite Network

  • Nov 22, 2025
  • Photonics
  • Wei Zhou +8
  • Research Article

Machine Learning and Optimization Techniques in IoT and Smart Systems: A Survey of Recent Trends

  • Sep 26, 2024
  • Science Management Design Journal
  • Anasuya
  • Research Article

A Hybrid PPO and DDPG Algorithm for Resource Aware Task Offloading in Edge-Cloud Computing Paradigm

  • Sep 18, 2025
  • Journal of Advanced College of Engineering and Management
  • Rojina Baral +2
  • PDF
  • Research Article
  • Citations13

Reinforcement Learning for Energy Community Management: A European-Scale Study

  • Mar 06, 2024
  • Energies
  • Giulia Palma +4
  • Research Article
  • Citations1

Using AI for Dynamic Resource Allocation and Performance Optimization in Software Systems

  • Nov 01, 2024
  • International Journal of Innovative Research in Computer Science and Technology
  • Mrinal Kumar +1
  • Research Article

ENERGY-AWARE DATA AGGREGATION IN WIRELESS SENSOR NETWORKS THROUGH HYBRID DEEP REINFORCEMENT LEARNING

  • Sep 01, 2025
  • ICTACT Journal on Communication Technology
  • Ramdas D Gore +1
Cactus Communications logo

Copyright 2026 Cactus Communications. All rights reserved.