• Home
  • Search
  • Skeletons and Asynchronous RPC for Embedded Data and Task Parallel Image Processing
  • Cite Icon11
  • https://doi.org/10.1093/ietisy/e89-d.7.2036Copy DOI Icon

Skeletons and Asynchronous RPC for Embedded Data and Task Parallel Image Processing

  • Abstract
  • Literature Map
  • References
  • Citations
  • Similar Papers
Abstract

Developing embedded parallel image processing applications is usually a very hardware-dependent process, often using the single instruction multiple data (SIMD) paradigm, and requiring deep knowledge of the processors used. Furthermore, the application is tailored to a specific hardware platform, and if the chosen hardware does not meet the requirements, it must be rewritten for a new platform. We have proposed the use of design space exploration [9] to find the most suitable hardware platform for a certain application. This requires a hardware-independent program, and we use algorithmic skeletons [5] to achieve this, while exploiting the data parallelism inherent to low-level image processing. However, since different operations run best on different kinds of processors, we need to exploit task parallelism as well. This paper describes how we exploit task parallelism using an asynchronous remote procedure call (RPC) system, optimized for low-memory and sparsely connected systems such as smart cameras. It uses a futures [16]-like model to present a normal imperative C-interface to the user in which the skeleton calls are implicitly parallelized and pipelined. Simulation provides the task dependency graph and performance numbers for the mapping, which can be done at run time to facilitate data dependent branching. The result is an easy to program, platform independent framework which shields the user from the parallel implementation and mapping of his application, while efficiently utilizing on-chip memory and interconnect bandwidth.

Similar Papers
  • Conference Article
  • Citations4

Towards fully user transparent task and data parallel image processing

  • Sep 01, 2009
  • J Lemeire +5
  • Conference Article
  • Citations9

ITAP: An Incremental Task Graph Partitioner for Task-parallel Static Timing Analysis

  • Jan 20, 2025
  • Boyang Zhang +10
  • Conference Article
  • Citations17

G-PASTA: GPU-Accelerated Partitioning Algorithm for Static Timing Analysis

  • Jun 23, 2024
  • Boyang Zhang +8
  • Conference Article

A data and task parallel image processing environment for distributed memory systems

  • Sep 03, 2001
  • C Nicolescu +1
  • Conference Article

A pre-run-time scheduling algorithm for object-based distributed real-time systems

  • Apr 01, 1997
  • I Santhoshkumar +2
  • Conference Article
  • Citations16

EASY-PIPE - An "Easy to use" parallel image processing environment based on algorithmic skeletons

  • Apr 23, 2001
  • C Nicolescu +1
  • Conference Article
  • Citations4

Parallel image processing on a transputer-based system

  • Mar 11, 1990
  • A Peng +1
  • Conference Article
  • Citations35

CloudRAMSort

  • May 20, 2012
  • Changkyu Kim +5
  • Conference Article
  • Citations3

Methods to utilize SIMT and SIMD instruction level parallelism in tridiagonal solvers

  • Jul 01, 2014
  • Endre Laszlo +3
  • Conference Article
  • Citations10

Abnormal motion detection in a real-time smart camera system

  • Aug 01, 2009
  • Mona Akbarniai Tehrani +3
  • Conference Article
  • Citations6

Implementation of an H.264 motion estimation algorithm on a VLIW programmable digital signal processor

  • Jan 01, 2005
  • Hyunchang Im +2
  • Research Article
  • Citations30

Analysis of a model for parallel image processing

  • Jan 01, 1985
  • Pattern Recognition
  • S Yalamanchili +1
  • Conference Article
  • Citations3

Reduction of Complexity and Automation of Parallel Execution through Loop Level Parallelism

  • Jan 01, 2007
  • Robert A Tefft +1
  • Conference Article
  • Citations71

Scheduling task parallelism on multi-socket multicore systems

  • May 31, 2011
  • Stephen L Olivier +3
  • Research Article
  • Citations18

A high-throughput hybrid task and data parallel Poisson solver for large-scale simulations of incompressible turbulent flows on distributed GPUs

  • Apr 02, 2021
  • Journal of Computational Physics
  • Hadi Zolfaghari +1
Cactus Communications logo

Copyright 2026 Cactus Communications. All rights reserved.