• Home
  • Search
  • Lock-based synchronization for GPU architectures
  • Cite Icon26
  • https://doi.org/10.1145/2903150.2903155Copy DOI Icon

Lock-based synchronization for GPU architectures

  • May 16, 2016
  • Yunlong Xu +5 more
Show More
  • Abstract
  • Literature Map
  • References
  • Citations
  • Similar Papers
Abstract

Modern GPUs have shown promising results in accelerating compute-intensive and numerical workloads with limited data sharing. However, emerging GPU applications manifest ample amount of data sharing among concurrently executing threads. Often data sharing requires mutual exclusion mechanism to ensure data integrity in multithreaded environment. Although modern GPUs provide atomic primitives that can be leveraged to construct fine-grained locks, the existing GPU lock implementations either incur frequent concurrency bugs, or lead to extremely low hardware utilization due to the Single Instruction Multiple Threads (SIMT) execution paradigm of GPUs. To make more applications with data sharing benefit from GPU acceleration, we propose a new locking scheme for GPU architectures. The proposed locking scheme allows lock stealing within individual warps to avoid the concurrency bugs due to the SMIT execution of GPUs. Moreover, it adopts lock virtualization to reduce the memory cost of fine-grain GPU locks. To illustrate the usage and the benefit of GPU locks, we apply the proposed GPU locking scheme to Delaunay mesh refinement (DMR), an application involving massive data sharing among threads. Our lock-based implementation can achieve 1.22x speedup over an algorithmic optimization based implementation (which uses a synchronization mechanism tailored for DMR) with 94% less memory cost.

Similar Papers
  • Conference Article
  • Citations43

Unified on-chip memory allocation for SIMT architecture

  • Jun 10, 2014
  • Ari B Hayes +1
  • Conference Article
  • Citations3

Methods to utilize SIMT and SIMD instruction level parallelism in tridiagonal solvers

  • Jul 01, 2014
  • Endre Laszlo +3
  • Research Article

단일 명령 다중 스레드 병렬 플랫폼을 위한 무작위 부분적 Haar 웨이블릿 변환

  • Oct 31, 2015
  • Journal of Digital Contents Society
  • Taejung Park
  • Conference Article
  • Citations22

Processing of synthetic Aperture Radar data with GPGPU

  • Oct 01, 2009
  • Carmine Clemente +4
  • Conference Article
  • Citations14

BatchSizer

  • Jan 18, 2021
  • Seyed Morteza Nabavinejad +2
  • Research Article
  • Citations3

Four styles of parallel and net programming

  • May 16, 2009
  • Frontiers of Computer Science in China
  • Zhiwei Xu +3
  • Research Article
  • Citations6

Cluster‐based approach for improving graphics processing unit performance by inter streaming multiprocessors locality

  • Sep 01, 2015
  • IET Computers & Digital Techniques
  • Mohammad Mahdi Keshtegar +2
  • Research Article
  • Citations18

Locality based warp scheduling in GPGPUs

  • Feb 24, 2017
  • Future Generation Computer Systems
  • Yang Zhang +4
  • Research Article

An Alpha-Tree Algorithm for Massively Parallel Architectures.

  • Jan 01, 2025
  • IEEE transactions on image processing : a publication of the IEEE Signal Processing Society
  • Edwin Carlinet +2
  • Book Chapter
  • Citations15

A Hoare Logic for SIMT Programs

  • Jan 01, 2013
  • Kensuke Kojima +1
  • Conference Article
  • Citations4

CUDA-based H.264/AVC deblocking filtering

  • Nov 01, 2010
  • Ting Liu +3
  • Conference Article
  • Citations3

A GPU Algorithm for IC Floorplanning: Specification, Analysis and Optimization

  • Jan 01, 2011
  • Yiding Han +3
  • Research Article
  • Citations3

PYSEQM 2.0: Accelerated Semiempirical Excited-State Calculations on Graphical Processing Units.

  • Oct 14, 2025
  • Journal of chemical theory and computation
  • Vishikh Athavale +5
  • Research Article
  • Citations70

Coordinated Batching and DVFS for DNN Inference on GPU Accelerators

  • Oct 01, 2022
  • IEEE Transactions on Parallel and Distributed Systems
  • Seyed Morteza Nabavinejad +2
  • PDF
  • Research Article
  • Citations57

PyGeNN: A Python Library for GPU-Enhanced Neural Networks.

  • Apr 22, 2021
  • Frontiers in Neuroinformatics
  • James C Knight +2
Cactus Communications logo

Copyright 2026 Cactus Communications. All rights reserved.