• Home
  • Search
  • Wait-free programming for general purpose computations on graphics processors
  • Cite Icon14
  • https://doi.org/10.1145/1400751.1400849Copy DOI Icon

Wait-free programming for general purpose computations on graphics processors

  • Aug 18, 2008
  • Phuong Hoai Ha +2 more
Show More
  • Abstract
  • Literature Map
  • References
  • Citations
  • Similar Papers
Abstract

The fact that graphics processors (GPUs) are today's most powerful computational hardware for the dollar has motivated researchers to utilize the ubiquitous and powerful GPUs for general-purpose computing. Recent GPUs feature the single-program multiple-data (SPMD) multicore architecture instead of the single-instruction multiple-data (SIMD). However, unlike CPUs, GPUs devote their transistors mainly to data processing rather than data caching and flow control, and consequently most of the powerful GPUs with many cores do not support any synchronization mechanisms between their cores. This prevents GPUs from being deployed more widely for general-purpose computing. This paper aims at bridging the gap between the lack of synchronization mechanisms in recent GPU architectures and the need of synchronization mechanisms in parallel applications. Based on the intrinsic features of recent GPU architectures, we construct strong synchronization objects like wait-free and t-resilient read-modify-write objects for a general model of recent GPU architectures without strong hardware synchronization primitives like test-and- set and compare-and-swap. Accesses to the wait-free objects have time complexity O(N), whether N is the number of processes. Our result demonstrates that it is possible to construct wait-free synchronization mechanisms for GPUs without the need of strong synchronization primitives in hardware and that wait-free programming is possible for GPUs.

Similar Papers
  • Research Article

Wait-Free Programming for General Purpose Computations on Graphics Processors

  • Aug 01, 2017
  • IEEE Transactions on Computers
  • Phuong Hoai Ha +2
  • Research Article
  • Citations11

Cache-conscious off-line real-time scheduling for multi-core platforms: algorithms and implementation

  • Mar 06, 2019
  • Real-Time Systems
  • Viet Anh Nguyen +2
  • Research Article
  • Citations22

Performance modeling for SPMD message-passing programs

  • Apr 25, 1998
  • Concurrency: Practice and Experience
  • Jürgen Brehm +2
  • Conference Article
  • Citations11

Formal Specifications for Java's Synchronisation Classes

  • Feb 01, 2014
  • Afshin Amighi +4
  • Conference Article
  • Citations2

Introducing Energy Efficiency into Graphics Processors

  • Dec 01, 2010
  • B V N Silpa +1
  • Conference Article
  • Citations3

Notice of Violation of IEEE Publication Principles: Communication Generation for Irregular Parallel Applications

  • Jan 01, 2006
  • Changjun Hu +5
  • Book Chapter
  • Citations1

A Fine-Granular Programming Scheme for Irregular Scientific Applications

  • Jan 01, 2016
  • Haowei Huang +5
  • Conference Article
  • Citations104

From SODA to scotch: The evolution of a wireless baseband processor

  • Nov 01, 2008
  • Mark Woh +10
  • Book Chapter
  • Citations21

Fault Tolerant Wide-Area Parallel Computing

  • Jan 01, 2000
  • Jon B Weissman
  • Conference Article
  • Citations9

A new multi-core pipelined architecture for executing sequential programs for parallel geospatial computing

  • Jun 21, 2010
  • Duoduo Liao +1
  • Research Article

TACO: A Scheduling Scheme for Parallel Applications on Multicore Architectures

  • Jan 01, 2014
  • Scientific Programming
  • Jan Hendrik Schönherr +2
  • PDF
  • Research Article

TACO: A Scheduling Scheme for Parallel Applications on Multicore Architectures

  • Jan 01, 2014
  • Scientific Programming
  • Jan H Schönherr +2
  • Research Article
  • Citations6

Integrated performance models for SPMD applications and MIMD architectures

  • Jul 01, 2002
  • IEEE Transactions on Parallel and Distributed Systems
  • P Cremonesi +1
  • Conference Article
  • Citations749

Mars

  • Oct 25, 2008
  • Bingsheng He +4
  • Conference Article
  • Citations9

Pipelined MIPS processor with cache controller using VHDL implementation for educational purposes

  • Dec 01, 2013
  • Hadeel Sh Mahmood +1
Cactus Communications logo

Copyright 2026 Cactus Communications. All rights reserved.