• Home
  • Search
  • Fast Sparse Matrix and Sparse Vector Multiplication Algorithm on the GPU
  • Open Access IconOpen Access
  • Cite Icon46
  • https://doi.org/10.1109/ipdpsw.2015.77Copy DOI Icon

Fast Sparse Matrix and Sparse Vector Multiplication Algorithm on the GPU

  • May 1, 2015
  • Carl Yang +2 more
Show More
  • Abstract
  • Literature Map
  • References
  • Citations
  • Similar Papers
Abstract

We implement a promising algorithm for sparse-matrix sparse-vector multiplication (SpMSpV) on the GPU. An efficient k-way merge lies at the heart of finding a fast parallel SpMSpV algorithm. We examine the scalability of three approaches -- no sorting, merge sorting, and radix sorting -- in solving this problem. For breadth-first search (BFS), we achieve a 1.26x speedup over state-of-the-art sparse-matrix dense-vector (SpMV) implementations. The algorithm seems generalize able for single-source shortest path (SSSP) and sparse-matrix sparse-matrix multiplication, and other core graph primitives such as maximal independent set and bipartite matching.

Similar Papers
  • Research Article
  • Citations35

On external-memory MST, SSSP and multi-way planar graph separation

  • May 21, 2004
  • Journal of Algorithms
  • Lars Arge +2
  • Conference Article

Techniques for Practical Parallel BFS and SSSP

  • Mar 01, 2025
  • Quinten De Man +2
  • Research Article
  • Citations8

I/O-Optimal Algorithms for Outerplanar Graphs

  • Jan 01, 2004
  • Journal of Graph Algorithms and Applications
  • Anil Maheshwari +1
  • Conference Article

Comparison of MSP-EXP432P401R Program Performances by Using Different Development Environments

  • Jun 02, 2025
  • Amar Hasanović +4
  • PDF
  • Research Article
  • Citations11

A Comparative Study on MMDBM Classifier Incorporating Various Sorting Procedure

  • May 18, 2015
  • Indian Journal of Science and Technology
  • P Ganesan +2
  • Conference Article

H-BFT: A fast Breadth-First Traversal algorithm for Sparse graphs and its GPU implementation

  • Dec 01, 2016
  • Dinali R Dabarera +3
  • Research Article
  • Citations33

Ordered vs. unordered

  • Feb 12, 2011
  • ACM SIGPLAN Notices
  • Muhammad Amber Hassaan +2
  • Conference Article
  • Citations340

Popular Conjectures Imply Strong Lower Bounds for Dynamic Problems

  • Oct 01, 2014
  • Amir Abboud +1
  • Research Article
  • Citations74

Maximum weight bipartite matching in matrix multiplication time

  • Jul 19, 2009
  • Theoretical Computer Science
  • Piotr Sankowski
  • Conference Article
  • Citations7

How Well do CPU, GPU and Hybrid Graph Processing Frameworks Perform?

  • May 01, 2018
  • Tanuj Kr Aasawat +2
  • Conference Article
  • Citations41

GraVF: A vertex-centric distributed graph processing framework on FPGAs

  • Aug 01, 2016
  • Nina Engelhardt +1
  • Research Article
  • Citations997

A fast and simple randomized parallel algorithm for the maximal independent set problem

  • Dec 01, 1986
  • Journal of Algorithms
  • Noga Alon +2
  • Conference Article
  • Citations47

Deploying Graph Algorithms on GPUs: An Adaptive Solution

  • May 01, 2013
  • Da Li +1
  • Research Article
  • Citations292

A fast parallel algorithm for the maximal independent set problem

  • Oct 01, 1985
  • Journal of the ACM
  • Richard M Karp +1
  • Conference Article
  • Citations5

Fine-Grained Task Migration for Graph Algorithms Using Processing in Memory

  • May 01, 2016
  • Paula Aguilera +3
Cactus Communications logo

Copyright 2026 Cactus Communications. All rights reserved.