• Home
  • Search
  • Sparse Matrix-Vector Multiplication on GPGPUs
  • Open Access IconOpen Access
  • Cite Icon135
  • https://doi.org/10.1145/3017994Copy DOI Icon

Sparse Matrix-Vector Multiplication on GPGPUs

Show More
  • Abstract
  • Highlights & Summary
  • Literature Map
  • References
  • Citations
  • Similar Papers
Abstract

The multiplication of a sparse matrix by a dense vector (SpMV) is a centerpiece of scientific computing applications: it is the essential kernel for the solution of sparse linear systems and sparse eigenvalue problems by iterative methods. The efficient implementation of the sparse matrix-vector multiplication is therefore crucial and has been the subject of an immense amount of research, with interest renewed with every major new trend in high-performance computing architectures. The introduction of General-Purpose Graphics Processing Units (GPGPUs) is no exception, and many articles have been devoted to this problem.With this article, we provide a review of the techniques for implementing the SpMV kernel on GPGPUs that have appeared in the literature of the last few years. We discuss the issues and tradeoffs that have been encountered by the various researchers, and a list of solutions, organized in categories according to common features. We also provide a performance comparison across different GPGPU models and on a set of test matrices coming from various application domains.

Similar Papers
  • Single Report

Solving large-scale sparse eigenvalue problems and linear systems of equations for accelerator modeling

  • Mar 30, 2009
  • Gene Golub +1
  • Research Article
  • Citations41

Efficient approximate solution of sparse linear systems

  • Nov 01, 1998
  • Computers & Mathematics with Applications
  • J.H Reif
  • Book Chapter
  • Citations27

The Efficient Parallel Iterative Solution of Large Sparse Linear Systems

  • Jan 01, 1993
  • Mark T Jones +1
  • Conference Article
  • Citations4

Parallel Sub-structuring Methods for Solving Sparse Linear Systems on a Cluster of GPUs

  • Aug 01, 2014
  • Abal- Kassim Cheik Ahamed +1
  • Research Article
  • Citations1

Mixed precision iterative refinement with adaptive precision sparse approximate inverse preconditioning

  • Aug 14, 2025
  • Engineering with Computers
  • Noaman Khan +1
  • Research Article

Parallel Implementation of the Coordinates-partitioning Based Aggregation-type Algebraic Multigrid Preconditioners

  • Mar 07, 2018
  • DEStech Transactions on Computer Science and Engineering
  • Jian-Ping Wu +2
  • Research Article
  • Citations29

A parallel preconditioned conjugate gradient package for solving sparse linear systems on a Cray Y-MP

  • Sep 01, 1991
  • Applied Numerical Mathematics
  • Michael A Heroux +2
  • Research Article
  • Citations25

Analysis of Coppersmith’s block Wiedemann algorithm for the parallel solution of sparse linear systems

  • Jan 01, 1995
  • Mathematics of Computation
  • Erich Kaltofen
  • Research Article
  • Citations7

Adaptive solution of linear systems of equations based on a posteriori error estimators

  • Jul 05, 2019
  • Numerical Algorithms
  • A Anciaux-Sedrakian +4
  • Book Chapter
  • Citations7

Process scheduling in DSC and the large sparse linear systems challenge

  • Sep 15, 1993
  • A Diaz +4
  • Research Article
  • Citations73

Adaptive precision in block‐Jacobi preconditioning for iterative sparse linear system solvers

  • Mar 12, 2018
  • Concurrency and Computation: Practice and Experience
  • Hartwig Anzt +4
  • Research Article
  • Citations8

Comparisons of two implementations for the solution of sparse linear systems - Part II

  • Apr 01, 1987
  • Advances in Engineering Software (1978)
  • Paolino Di Felice +1
  • Research Article
  • Citations29

Numerical recovery strategies for parallel resilient Krylov linear solvers

  • Aug 03, 2016
  • Numerical Linear Algebra with Applications
  • Emmanuel Agullo +4
  • Conference Article

An Experimental Study of Two-Level Schwarz Domain Decomposition Preconditioners on GPUs

  • May 01, 2023
  • Ichitaro Yamazaki +2
  • Research Article

Parallel ACO Algorithm on GPU for Fast Solution of QAPs

  • Jan 01, 2013
  • IEEJ Transactions on Electronics, Information and Systems
  • Shigeyoshi Tsutsui
Cactus Communications logo

Copyright 2026 Cactus Communications. All rights reserved.