• Home
  • Search
  • A Design Framework for Mapping Vectorized Synchronous Dataflow Graphs onto CPU-GPU Platforms
  • Cite Icon13
  • https://doi.org/10.1145/2906363.2906374Copy DOI Icon

A Design Framework for Mapping Vectorized Synchronous Dataflow Graphs onto CPU-GPU Platforms

  • May 23, 2016
  • Shuoxin Lin +3 more
Show More
  • Abstract
  • Literature Map
  • References
  • Citations
  • Similar Papers
Abstract

Heterogeneous computing platforms with multicore central processing units (CPUs) and graphics processing units (GPUs) are of increasing interest to designers of embedded signal processing systems since they offer the potential for significant performance boost while maintaining the flexibility of software-based design flows. Developing optimized implementations for CPU-GPU platforms is challenging due to complex, inter-related design issues, including task scheduling, interprocessor communication, memory management, and modeling and exploitation of different forms of parallelism. In this paper, we present an automated, dataflow based, design framework called DIF-GPU for application mapping and software synthesis on heterogeneous CPU-GPU platforms. DIF-GPU is based on novel extensions to the dataflow interchange format (DIF) package, which is a software environment for developing and experimenting with dataflow-based design methods and synthesis techniques for embedded signal processing systems. DIF-GPU exploits multiple forms of parallelism by deeply incorporating efficient vectorization and scheduling techniques for synchronous dataflow specifications, and incorporating techniques for streamlining interprocessor communication. DIF-GPU also provides software synthesis capabilities to help accelerate the process of moving from high-level application models to optimized implementations.

Similar Papers
  • PDF
  • Research Article
  • Citations6

Parallelizing Multiple Flow Accumulation Algorithm using CUDA and OpenACC

  • Sep 03, 2019
  • ISPRS International Journal of Geo-Information
  • Natalija Stojanovic +1
  • Research Article
  • Citations39

HEVC Encoding Optimization Using Multicore CPUs and GPUs

  • Nov 01, 2015
  • IEEE Transactions on Circuits and Systems for Video Technology
  • Wei Xiao +4
  • Book Chapter

Efficient Prefix Scan for the GPU-Based Implementation of Random Forest

  • Jan 01, 2015
  • Bojan Novak
  • Research Article
  • Citations17

Mesh–particle interpolations on graphics processing units and multicore central processing units

  • Jun 13, 2011
  • Philosophical Transactions of the Royal Society A: Mathematical, Physical and Engineering Sciences
  • Diego Rossinelli +2
  • PDF
  • Research Article
  • Citations3

Porting the Meso-NH atmospheric model on different GPU architectures for the next generation of supercomputers (version MESONH-v55-OpenACC)

  • May 14, 2025
  • Geoscientific Model Development
  • Juan Escobar +6
  • Research Article
  • Citations13

Practical parallel AES algorithms on cloud for massive users and their performance evaluation

  • Dec 17, 2015
  • Concurrency and Computation: Practice and Experience
  • Xiongwei Fei +3
  • Research Article
  • Citations26

Efficient methods for implementation of multi-level nonrigid mass-preserving image registration on GPUs and multi-threaded CPUs

  • Jan 06, 2016
  • Computer Methods and Programs in Biomedicine
  • Nathan D Ellingwood +3
  • Research Article
  • Citations21

IMPLEMENTATION OF FDTD-COMPATIBLE GREEN'S FUNCTION ON HETEROGENEOUS CPU-GPU PARALLEL PROCESSING SYSTEM

  • Jan 01, 2013
  • Progress In Electromagnetics Research
  • Tomasz P Stefanski
  • PDF
  • Research Article
  • Citations25

DdcMD: A fully GPU-accelerated molecular dynamics program for the Martini force field.

  • Jul 23, 2020
  • The Journal of Chemical Physics
  • Xiaohua Zhang +8
  • Conference Article
  • Citations9

COMP: Compiler Optimizations for Manycore Processors

  • Dec 01, 2014
  • Linhai Song +4
  • Research Article

Improving MSAProbs Algorithm performance and Parallel Computing using GPU

  • Apr 18, 2023
  • International Journal of Computer Applications
  • Sally Zaki El-Hadary +2
  • Research Article
  • Citations1

Fast GPU Algorithm for Analyzing Effective Connectivity in Functional Brain Imaging

  • Jan 01, 2013
  • IFAC Proceedings Volumes
  • Lawrence Chan +4
  • Research Article
  • Citations4

A systematic parallel strategy for generating contours from large-scale DEM data using collaborative CPUs and GPUs

  • Feb 18, 2021
  • Cartography and Geographic Information Science
  • Chen Zhou +1
  • Research Article
  • Citations30

Multi-core-CPU and GPU-accelerated radiative transfer models based on the discrete ordinate method

  • Aug 09, 2014
  • Computer Physics Communications
  • Dmitry S Efremenko +3
  • Research Article
  • Citations31

GPU acceleration of MPAS microphysics WSM6 using OpenACC directives: Performance and verification

  • Oct 14, 2020
  • Computers & Geosciences
  • Jae Youp Kim +2
Cactus Communications logo

Copyright 2026 Cactus Communications. All rights reserved.