• Home
  • Search
  • Parallel Breadth First Search on GPU clusters
  • Cite Icon28
  • https://doi.org/10.1109/bigdata.2014.7004219Copy DOI Icon

Parallel Breadth First Search on GPU clusters

  • Oct 1, 2014
  • Zhisong Fu +4 more
Show More
  • Abstract
  • Literature Map
  • References
  • Citations
  • Similar Papers
Abstract

Fast, scalable, low-cost, and low-power execution of parallel graph algorithms is important for a wide variety of commercial and public sector applications. Breadth First Search (BFS) imposes an extreme burden on memory bandwidth and network communications and has been proposed as a benchmark that may be used to evaluate current and future parallel computers. Hardware trends and manufacturing limits strongly imply that many-core devices, such as NVIDIA® GPUs and the Intel® Xeon Phi®, will become central components of such future systems. GPUs are well known to deliver the highest FLOPS/watt and enjoy a very significant memory bandwidth advantage over CPU architectures. Recent work has demonstrated that GPUs can deliver high performance for parallel graph algorithms and, further, that it is possible to encapsulate that capability in a manner that hides the low level details of the GPU architecture and the CUDA language but preserves the high throughput of the GPU. We extend previous research on GPUs and on scalable graph processing on supercomputers and demonstrate that a high-performance parallel graph machine can be created using commodity GPUs and networking hardware.

Similar Papers
  • Conference Article
  • Citations8

IPUG for Multiple Graphcore IPUs: Optimizing Performance and Scalability of Parallel Breadth-First Search

  • Dec 01, 2021
  • Luk Burchard +2
  • Conference Article
  • Citations4

Vectorization of Hybrid Breadth First Search on the Intel Xeon Phi

  • May 15, 2017
  • Mireya Paredes +2
  • Conference Article
  • Citations3

Applications of many-core technologies to on-line event reconstruction in High Energy Physics experiments

  • Oct 01, 2013
  • A Gianelle +12
  • Research Article
  • Citations49

Regularizing graph centrality computations

  • Aug 07, 2014
  • Journal of Parallel and Distributed Computing
  • Ahmet Erdem Sarıyüce +3
  • Conference Article
  • Citations7

ICE: A General and Validated Energy Complexity Model for Multithreaded Algorithms

  • Dec 01, 2016
  • Vi Ngoc-Nha Tran +1
  • Conference Article

H-BFT: A fast Breadth-First Traversal algorithm for Sparse graphs and its GPU implementation

  • Dec 01, 2016
  • Dinali R Dabarera +3
  • Conference Article
  • Citations14

Implementation of parallel graph algorithms on a massively parallel SIMD computer with virtual processing

  • Apr 25, 1995
  • Tsan-Sheng Hsu +2
  • Conference Article
  • Citations14

Analysis and Optimization of Financial Analytics Benchmark on Modern Multi- and Many-core IA-Based Architectures

  • Nov 01, 2012
  • Mikhail Smelyanskiy +12
  • Research Article
  • Citations1

Cooperative and out-of-core execution of the irregular wavefront propagation pattern on hybrid machines with IntelⓇ Xeon Phi™.

  • Jan 24, 2018
  • Concurrency and Computation: Practice and Experience
  • Jeremias Gomes +5
  • Conference Article
  • Citations48

Parallel external memory graph algorithms

  • Jan 01, 2010
  • Lars Arge +2
  • Research Article

RandMScan: accelerating parallel scan via matrix computation and random-jump strategy.

  • Dec 24, 2025
  • Scientific reports
  • Shujun Peng +4
  • Research Article
  • Citations1

Parallelized Simulation of a Finite Element Method in Many Integrated Core Architecture

  • Feb 07, 2017
  • Journal of Engineering Materials and Technology
  • Moonho Tak +1
  • Research Article
  • Citations16

From Distributed Memory Cycle Detection to Parallel LTL Model Checking

  • May 01, 2005
  • Electronic Notes in Theoretical Computer Science
  • J Barnat +2
  • Research Article

Exploiting fine-grained parallelism in graph traversal algorithms via lock virtualization on multi-core architecture

  • Jun 26, 2014
  • The Journal of Supercomputing
  • Jie Yan +2
  • Conference Article
  • Citations6

The Exploration of Pervasive and Fine-Grained Parallel Model Applied on Intel Xeon Phi Coprocessor

  • Oct 01, 2013
  • Christophe Calvin +2
Cactus Communications logo

Copyright 2026 Cactus Communications. All rights reserved.