• Cite Icon3
  • https://doi.org/10.1145/3148055.3148059Copy DOI Icon

Victream

  • Dec 5, 2017
  • Jun Suzuki +6 more
Show More
  • Abstract
  • Literature Map
  • References
  • Citations
  • Similar Papers
Abstract

In data-parallel computing that uses a graphic processing unit (GPU), processing of large data requires that multiple GPUs be used in the computer to increase its execution performance. Increasing processing performance by using multiple computing resources has been enabled by the development of computing frameworks based on a directed acyclic graph (DAG). However, their performance degrades in out-of-core processing, which often occurs in processing of large data on GPUs with limited memory capacity. The GPU data input/output (I/O) for data swapping between host memory and GPU memory during the execution of a user DAG is usually a performance bottleneck. A computing framework called is proposed to overcome this drawback. It uses a novel scheduler that involves two methods to minimize the total amount of GPU data I/O of data swapping. First, it performs locality-aware scheduling. When it schedules a task, it selects one that requires the minimum amount of data swapping and reuses as much of the data residing in GPU memory as possible. Second, it extends the locality-aware scheduling so that GPUs can execute data prefetching. Prefetching data that are swapped out from a GPU enables efficient use of bottleneck GPU I/O resources. To prefetch the input data of future tasks, it is required to determine the schedule of future tasks. Victream's scheduler (hereafter, the Victream scheduler) extends the locality-aware scheduling so that it can schedule future tasks to enable data prefetching that is executed in the way that minimizes the amount of data I/O of data swapping. Evaluation of a Victream prototype showed that the performance of Victream is better than that of conventional frameworks by up to 117%.

Similar Papers
  • Conference Article
  • Citations3

FlexGPU: A Flexible and Efficient Scheduler for GPU Sharing Systems

  • May 01, 2020
  • Qichen Chen +3
  • Research Article
  • Citations21

MoDNN: Memory Optimal Deep Neural Network Training on Graphics Processing Units

  • Mar 01, 2019
  • IEEE Transactions on Parallel and Distributed Systems
  • Xiaoming Chen +3
  • Conference Article
  • Citations3

Efficient GPU-Based Query Processing with Pruned List Caching in Search Engines

  • Dec 01, 2017
  • Dongdong Wang +6
  • Research Article
  • Citations13

TOAST: Automatic tiling for iterative stencil computations on GPUs

  • Feb 02, 2017
  • Concurrency and Computation: Practice and Experience
  • Rodrigo C O Rocha +3
  • Conference Article
  • Citations13

GLoP

  • Sep 09, 2014
  • Xavier J A Bellekens +4
  • Research Article
  • Citations2

Gpubased, microsecond latency, hectochannel mimo feedback control of magnetically confined plasmas

  • Jan 01, 2013
  • Columbia Academic Commons (Columbia University)
  • M E Mauel +1
  • Book Chapter
  • Citations1

A Study on a Method of Effective Memory Utilization on GPU Applied for Neighboring Filter on Image Processing

  • Jan 01, 2012
  • Yoshio Yanagihara +1
  • Research Article

3D Freehand Ultrasound Reconstruction of the Carotid Artery.

  • May 01, 2026
  • Ultrasound in medicine & biology
  • Yimeng Dou +6
  • Conference Article
  • Citations114

Stealing Webpages Rendered on Your Browser by Exploiting GPU Vulnerabilities

  • May 01, 2014
  • Sangho Lee +3
  • Conference Article
  • Citations20

ConVGPU: GPU Management Middleware in Container Based Virtualized Environment

  • Sep 01, 2017
  • Daeyoun Kang +4
  • Conference Article
  • Citations1

Automatic Parallelization of GPU Applications Using OpenCL

  • Jul 01, 2015
  • Lizandro D Solano-Quinde +2
  • Conference Article
  • Citations53

Improving the Performance of CA-GMRES on Multicores with Multiple GPUs

  • May 01, 2014
  • Ichitaro Yamazaki +4
  • PDF
  • Research Article
  • Citations23

Massive-Parallel Trajectory Calculations version 2.2 (MPTRAC-2.2): Lagrangian transport simulations on graphics processing units (GPUs)

  • Apr 05, 2022
  • Geoscientific Model Development
  • Lars Hoffmann +13
  • PDF
  • Peer Review Report

Comment on gmd-2021-382

  • Feb 08, 2022
  • Lars Hoffmann +13
  • Conference Article
  • Citations36

Efficient large Pearson correlation matrix computing using hybrid MPI/CUDA

  • May 01, 2011
  • Ekasit Kijsipongse +3
Cactus Communications logo

Copyright 2026 Cactus Communications. All rights reserved.