• Home
  • Search
  • Coarse-grained thread pipelining: a speculative parallel execution model for shared-memory multiprocessors
  • Cite Icon20
  • https://doi.org/10.1109/71.954629Copy DOI Icon

Coarse-grained thread pipelining: a speculative parallel execution model for shared-memory multiprocessors

Show More
  • Abstract
  • Literature Map
  • References
  • Citations
  • Similar Papers
Abstract

This paper presents a new parallelization model, called coarse-grained thread pipelining, for exploiting speculative coarse-grained parallelism from general-purpose application programs in shared-memory multiprocessor systems. This parallelization model, which is based on the fine-grained thread pipelining model proposed for the superthreaded architecture, allows concurrent execution of loop iterations in a pipelined fashion with runtime data-dependence checking and control speculation. The speculative execution combined with the runtime dependence checking allows the parallelization of a variety of program constructs that cannot be parallelized with existing runtime parallelization algorithms. The pipelined execution of loop iterations in this new technique results in lower parallelization overhead than in other existing techniques. We evaluated the performance of this new model using some real applications and a synthetic benchmark. These experiments show that programs with a sufficiently large grain size compared to the parallelization overhead obtain significant speedup using this model. The results from the synthetic benchmark provide a means for estimating the performance that can be obtained from application programs that will be parallelized with this model. The library routines developed for this thread pipelining model are also useful for evaluating the correctness of the codes generated by the superthreaded compiler and in debugging and verifying the simulator for the superthreaded processor.

Similar Papers
  • Conference Article
  • Citations7

Loop level speculation in a task based programming model

  • Dec 01, 2013
  • Rahulkumar Gayatri +2
  • PDF
  • Research Article

Using Heuristic Value Prediction and Dynamic Task Granularity Resizing to Improve Software Speculation

  • Jan 01, 2014
  • The Scientific World Journal
  • Fan Xu +5
  • Conference Article
  • Citations10

I2SEMS: Interconnects-Independent Security Enhanced Shared Memory Multiprocessor Systems

  • Sep 01, 2007
  • Manhee Lee +2
  • Research Article
  • Citations198

Rsim: simulating shared-memory multiprocessors with ILP processors

  • Jan 01, 2002
  • Computer
  • C.J Hughes +3
  • Book Chapter
  • Citations4

Data distribution and loop parallelization for shared-memory multiprocessors

  • Jan 01, 1997
  • Eduard Ayguadé +3
  • Conference Article
  • Citations3

High-speed image reconstruction based on CBP and Fourier inversion methods

  • Apr 28, 1997
  • K Rajan +2
  • Research Article
  • Citations150

A comparison of phenological models of leaf bud burst and flowering of boreal trees using independent observations

  • Oct 01, 2008
  • Tree Physiology
  • T Linkosalo +2
  • Conference Article
  • Citations11

Parallel Architecture Benchmarking: From Embedded Computing to HPC, a FiPS Project Perspective

  • Aug 01, 2014
  • Yves Lhuillier +4
  • Research Article
  • Citations29

Parallel computation of a highly nonlinear Boussinesq equation model through domain decomposition

  • Jun 03, 2005
  • International Journal for Numerical Methods in Fluids
  • Khairil Irfan Sitanggang +1
  • Book Chapter
  • Citations5

Performance Evaluation of Thread-Level Speculation in Off-the-Shelf Hardware Transactional Memories

  • Jan 01, 2017
  • Juan Salamanca +2
  • Research Article
  • Citations12

An efficient parallel model for coastal transport process simulation

  • Apr 27, 2000
  • Advances in Water Resources
  • Onyx Wing-Hong Wai +1
  • Conference Article
  • Citations11

Quantifying and reducing the effects of wrong-path memory references in cache-coherent multiprocessor systems

  • Jan 01, 2006
  • R Sendag +3
  • Research Article
  • Citations7

The impact of wrong-path memory references in cache-coherent multiprocessor systems

  • Mar 24, 2007
  • Journal of Parallel and Distributed Computing
  • Resit Sendag +3
  • Book Chapter

Compiler Directed Parallelization of Loops in Scale for Shared-Memory Multiprocessors

  • Jan 01, 2003
  • Gregory S Johnson +1
  • PDF
  • Research Article
  • Citations4

To Estimate Performance of Artificial Neural Network Model Based on Terahertz Spectrum: Gelatin Identification as an Example

  • Jul 14, 2022
  • Frontiers in Nutrition
  • Yizhang Li +8
Cactus Communications logo

Copyright 2026 Cactus Communications. All rights reserved.