• Home
  • Search
  • Performance Analysis and Optimization of Sparse Matrix-Vector Multiplication on Modern Multi- and Many-Core Processors
  • Open Access IconOpen Access
  • Cite Icon39
  • https://doi.org/10.1109/icpp.2017.38Copy DOI Icon

Performance Analysis and Optimization of Sparse Matrix-Vector Multiplication on Modern Multi- and Many-Core Processors

  • Aug 1, 2017
  • Athena Elafrou +2 more
Show More
  • Abstract
  • Literature Map
  • References
  • Citations
  • Similar Papers
Abstract

This paper presents a low-overhead optimizer for the ubiquitous sparse matrix-vector multiplication (SpMV) kernel. Architectural diversity among different processors together with structural diversity among different sparse matrices lead to bottleneck diversity. This justifies an SpMV optimizer that is both matrix- and architecture-adaptive through runtime specialization. To this direction, we present an approach that first identifies the performance bottlenecks of SpMV for a given sparse matrix on the target platform either through profiling or by matrix property inspection, and then selects suitable optimizations to tackle those bottlenecks. Our optimization pool is based on the widely used Compressed Sparse Row (CSR) sparse matrix storage format and has low preprocessing overheads, making our overall approach practical even in cases where fast decision making and optimization setup is required. We evaluate our optimizer on three x86-based computing platforms and demonstrate that it is able to distinguish and appropriately optimize SpMV for the majority of matrices in a representative test suite, leading to significant speedups over the CSR and Inspector-Executor CSR SpMV kernels available in the latest release of the Intel MKL library.

Similar Papers
  • Research Article
  • Citations33

SMAT

  • Jun 16, 2013
  • ACM SIGPLAN Notices
  • Jiajia Li +3
  • Conference Article
  • Citations73

Sparse Matrix Format Selection with Multiclass SVM for SpMV on GPU

  • Aug 01, 2016
  • Akrem Benatia +3
  • Conference Article
  • Citations45

Structural Agnostic SpMV: Adapting CSR-Adaptive for Irregular Matrices

  • Dec 01, 2015
  • Mayank Daga +1
  • Conference Article
  • Citations81

Towards a Universal FPGA Matrix-Vector Multiplication Architecture

  • Apr 01, 2012
  • Srinidhi Kestur +2
  • Conference Article
  • Citations8

Optimizing Sparse Matrix Vector Multiplication Using Diagonal Storage Matrix Format

  • Sep 01, 2010
  • Liang Yuan +3
  • Research Article
  • Citations30

A hybrid format for better performance of sparse matrix-vector multiplication on a GPU

  • Jul 14, 2015
  • The International Journal of High Performance Computing Applications
  • Dahai Guo +2
  • Conference Article
  • Citations6

Exploiting dynamic sparse matrices for performance portable linear algebra operations

  • Nov 01, 2022
  • Christodoulos Stylianou +1
  • Research Article
  • Citations4

A new AXT format for an efficient SpMV product using AVX-512 instructions and CUDA

  • Apr 18, 2021
  • Advances in Engineering Software
  • E Coronado-Barrientos +2
  • Conference Article
  • Citations68

Merge-based parallel sparse matrix-vector multiplication

  • Nov 13, 2016
  • Duane Merrill +1
  • Conference Article
  • Citations21

Cache-aware sparse matrix formats for Kepler GPU

  • Dec 01, 2014
  • Yusuke Nagasaka +2
  • PDF
  • Research Article
  • Citations6

Parallel Implementation of Large-Scale Linear Scaling Density Functional Theory Calculations With Numerical Atomic Orbitals in HONPAS.

  • Nov 26, 2020
  • Frontiers in Chemistry
  • Zhaolong Luo +4
  • Conference Article
  • Citations3

On Shared-Memory Parallelization of a Sparse Matrix Scaling Algorithm

  • Sep 01, 2012
  • Umit V Catalyurek +2
  • Conference Article
  • Citations24

Speeding Up SpMV for Power-Law Graph Analytics by Enhancing Locality & Vectorization

  • Nov 01, 2020
  • Serif Yesil +3
  • Research Article
  • Citations356

Sparsity: Optimization Framework for Sparse Matrix Kernels

  • Feb 01, 2004
  • The International Journal of High Performance Computing Applications
  • Eun-Jin Im +2
  • Research Article
  • Citations13

Effect of the storage format of sparse linear systems on parallel CFD computations

  • Jul 01, 2000
  • Computer Methods in Applied Mechanics and Engineering
  • Laura C Dutto +2
Cactus Communications logo

Copyright 2026 Cactus Communications. All rights reserved.