• Home
  • Search
  • A comparative study of GPU programming models and architectures using neural networks
  • Cite Icon39
  • https://doi.org/10.1007/s11227-011-0631-3Copy DOI Icon

A comparative study of GPU programming models and architectures using neural networks

Show More
  • Abstract
  • Literature Map
  • References
  • Citations
  • Similar Papers
Abstract

Recently, General Purpose Graphical Processing Units (GP-GPUs) have been identified as an intriguing technology to accelerate numerous data-parallel algorithms. Several GPU architectures and programming models are beginning to emerge and establish their niche in the High-Performance Computing (HPC) community. New massively parallel architectures such as the Nvidia’s Fermi and AMD/ATi’s Radeon pack tremendous computing power in their large number of multiprocessors. Their performance is unleashed using one of the two GP-GPU programming models: Compute Unified Device Architecture (CUDA) and Open Computing Language (OpenCL). Both of them offer constructs and features that have direct bearing on the application runtime performance. In this paper, we compare the two GP-GPU architectures and the two programming models using a two-level character recognition network. The two-level network is developed using four different Spiking Neural Network (SNN) models, each with different ratios of computation-to-communication requirements. To compare the architectures, we have chosen the two extremes of the SNN models for implementation of the aforementioned two-level network. An architectural performance comparison of the SNN application running on Nvidia’s Fermi and AMD/ATi’s Radeon is done using the OpenCL programming model exhausting all of the optimization strategies plausible for the two architectures. To compare the programming models, we implement the two-level network on Nvidia’s Tesla C2050 based on the Fermi architecture. We present a hierarchy of implementations, where we successively add optimization techniques associated with the two programming models. We then compare the two programming models at these different levels of implementation and also present the effect of the network size (problem size) on the performance. We report significant application speed-up, as high as 1095× for the most computation intensive SNN neuron model, against a serial implementation on the Intel Core 2 Quad host. A comprehensive study presented in this paper establishes connections between programming models, architectures and applications.

Similar Papers
  • Research Article
  • Citations6

Exploring compiler optimization opportunities for the OpenMP 4.x accelerator model on a POWER8+GPU platform

  • Nov 13, 2016
  • Akira Hayashi +4
  • Conference Article
  • Citations3

Knowledge Distillation between DNN and SNN for Intelligent Sensing Systems on Loihi Chip

  • Apr 05, 2023
  • Shiya Liu +1
  • Conference Article
  • Citations34

Pipelined Parallel LZSS for Streaming Data Compression on GPGPUs

  • Dec 01, 2012
  • Adnan Ozsoy +2
  • Conference Article
  • Citations12

Evaluation of GPU Architectures Using Spiking Neural Networks

  • Jul 01, 2011
  • Vivek K Pallipuram +2
  • Conference Article
  • Citations9

Database processing by Linear Regression on GPU using CUDA

  • Jul 01, 2011
  • Jyoti B Kulkarni +2
  • Conference Article

Evaluation of modern GPGPU technologies for image processing

  • Apr 27, 2020
  • Joachim Meyer
  • Conference Article
  • Citations50

Encoding, model, and architecture

  • Nov 02, 2020
  • Haowen Fang +5
  • Conference Article
  • Citations3

Signature of an anticipatory response in area VI as modeled by a probabilistic model and a spiking neural network

  • Jan 01, 2014
  • Bernhard A Kaplan +3
  • Research Article
  • Citations4

Fault Detection in Electrical Equipment by Infrared Thermography Images Using Spiking Neural Network Through Hybrid Feature Selection

  • Dec 08, 2022
  • Journal of Circuits, Systems and Computers
  • Shanmugam Chellamuthu +3
  • Book Chapter

Optimization of Two Bottleneck Programs in SAR System on GPGPU

  • Jan 01, 2016
  • Yang Zhang +5
  • Conference Article
  • Citations5

Application of GPGPU to Accelerate CFD Simulation

  • Jun 17, 2018
  • Shafiul A Mintu +1
  • Research Article
  • Citations1

MiC

  • Apr 29, 2019
  • ACM Journal on Emerging Technologies in Computing Systems
  • Qixiao Liu +2
  • Conference Article
  • Citations29

SparkXD: A Framework for Resilient and Energy-Efficient Spiking Neural Network Inference using Approximate DRAM

  • Dec 05, 2021
  • Rachmad Vidya Wicaksana Putra +2
  • Book Chapter
  • Citations9

Automatic CUDA Code Synthesis Framework for Multicore CPU and GPU Architectures

  • Jan 01, 2012
  • Hanwoong Jung +2
  • Conference Article
  • Citations31

SpikeDyn: A Framework for Energy-Efficient Spiking Neural Networks with Continual and Unsupervised Learning Capabilities in Dynamic Environments

  • Dec 05, 2021
  • arXiv (Cornell University)
  • Rachmad Vidya Wicaksana Putra +1
Cactus Communications logo

Copyright 2026 Cactus Communications. All rights reserved.