• Home
  • Search
  • Efficient Use of GPU Memory for Large-Scale Deep Learning Model Training
  • Open Access IconOpen Access
  • Cite Icon22
  • https://doi.org/10.3390/app112110377Copy DOI Icon

Efficient Use of GPU Memory for Large-Scale Deep Learning Model Training

Show More
  • Abstract
  • Highlights & Summary
  • PDF
  • Literature Map
  • References
  • Citations
  • Similar Papers
Abstract

To achieve high accuracy when performing deep learning, it is necessary to use a large-scale training model. However, due to the limitations of GPU memory, it is difficult to train large-scale training models within a single GPU. NVIDIA introduced a technology called CUDA Unified Memory with CUDA 6 to overcome the limitations of GPU memory by virtually combining GPU memory and CPU memory. In addition, in CUDA 8, memory advise options are introduced to efficiently utilize CUDA Unified Memory. In this work, we propose a newly optimized scheme based on CUDA Unified Memory to efficiently use GPU memory by applying different memory advise to each data type according to access patterns in deep learning training. We apply CUDA Unified Memory technology to PyTorch to see the performance of large-scale learning models through the expanded GPU memory. We conduct comprehensive experiments on how to efficiently utilize Unified Memory by applying memory advises when performing deep learning. As a result, when the data used for deep learning are divided into three types and a memory advise is applied to the data according to the access pattern, the deep learning execution time is reduced by 9.4% compared to the default Unified Memory.

Loading PDF

Similar Papers
  • Research Article

Evaluasi Pelaksanaan Program Pelatihan Pembelajaran Mendalam (PM) bagi Guru SMPN 2 Praya Tengah Menggunakan Model Countenance Evaluation Stake

  • Feb 21, 2026
  • Jurnal Pendidikan, Sains, Geologi, dan Geofisika (GeoScienceEd Journal)
  • Nurul Hidayati +1
  • Conference Article
  • Citations215

ZeRO-infinity

  • Nov 13, 2021
  • Samyam Rajbhandari +4
  • Conference Article
  • Citations17

Characterizing CUDA Unified Memory (UM)-Aware MPI Designs on Modern GPU Architectures

  • Apr 13, 2019
  • K V Manian +5
  • Research Article
  • Citations3

TENSILE: A Tensor Granularity Dynamic GPU Memory Scheduling Method Toward Multiple Dynamic Workloads System

  • Jan 01, 2022
  • IEEE Transactions on Knowledge and Data Engineering
  • Kaixin Zhang +6
  • Conference Article
  • Citations1

Simulated Annealing for Timeliness and Energy aware Deep Learning Job Assignment

  • Oct 01, 2019
  • Dong-Ki Kang +1
  • Conference Article

Safe Deep Reinforcement Learning Based on Sample Value Evaluation

  • Dec 02, 2022
  • Rongjun Ye +2
  • Research Article
  • Citations43

Medical Image Classification Algorithm Based on Visual Attention Mechanism-MCNN.

  • Jan 01, 2021
  • Oxidative Medicine and Cellular Longevity
  • Fengping An +2
  • PDF
  • Research Article

Computing Quantiles of Functions of the Agent Distribution Using t-Digests

  • Sep 26, 2023
  • Computational Economics
  • Robert Kirkby
  • Research Article
  • Citations2

TBEM: Testing-Based GPU-Memory Consumption Estimation for Deep Learning

  • Jan 01, 2022
  • IEEE Access
  • Haiyi Liu +3
  • Research Article
  • Citations8

Development of training environment for deep learning with medical images on supercomputer system based on asynchronous parallel Bayesian optimization

  • Jan 20, 2020
  • The Journal of Supercomputing
  • Yukihiro Nomura +11
  • Conference Article
  • Citations16

Elan: Towards Generic and Efficient Elastic Training for Deep Learning

  • Nov 01, 2020
  • Lei Xie +6
  • Conference Article
  • Citations24

Coordinated Page Prefetch and Eviction for Memory Oversubscription Management in GPUs

  • May 01, 2020
  • Qi Yu +5
  • Research Article
  • Citations2

Corun: Concurrent Inference and Continuous Training at the Edge for Cost-Efficient AI-Based Mobile Image Sensing

  • Aug 14, 2024
  • Sensors (Basel, Switzerland)
  • Yu Liu +2
  • Research Article
  • Citations60

Traversing large graphs on GPUs with unified memory

  • Mar 01, 2020
  • Proceedings of the VLDB Endowment
  • Prasun Gera +4
  • Research Article
  • Citations2

Addressing GPU memory limitations for Graph Neural Networks in High-Energy Physics applications

  • Sep 18, 2024
  • Frontiers in High Performance Computing
  • Claire Songhyun Lee +7
Cactus Communications logo

Copyright 2026 Cactus Communications. All rights reserved.