• Home
  • Search
  • A combinatorial distributed architecture for Exascale computing
  • Cite Icon3
  • https://doi.org/10.1109/icoac.2012.6416853Copy DOI Icon

A combinatorial distributed architecture for Exascale computing

  • Dec 1, 2012
  • Ganapathy Mani +2 more
Show More
  • Abstract
  • Literature Map
  • References
  • Citations
  • Similar Papers
Abstract

Computer architectures are expected to change to support Exascale computing in the near future. As energy and cooling constraints limit increases in microprocessor clock speeds and number of cores, computer companies are turning to parallel programming. Nowadays, parallel programming is achieved by increasing the number of processing elements in processor cores, increasing the number of processor cores itself and complicated parallel programming where programmer has the responsibility of allocating memory and synchronizing the communication between the processing elements as well as processor cores. It becomes increasingly difficult and expensive to design and produce shared memory machines with ever increasing number of processors. Increase in the number of processors is a major disadvantage when it comes to energy consumption. In this work, we present a new architecture for processor design based on pairwise balanced combinatorial interconnection of processing and memory elements. The proposed processor uses two operand instructions, so that the set of executable machine instructions is partitioned by these pairs. This kind of partition allows parallel processing of data-independent instructions. Since this partition is done at the compile time, the architecture extracts the instruction level parallelism without run-time overheads. We analyze and confirm the performance improvements through simulations. The suggested combinatorial arrangement gives set of architectures with various degrees of performance enhancement.

Similar Papers
  • Conference Article
  • Citations29

Data parallel programming in an adaptive environment

  • Apr 25, 1995
  • G Edjlali +3
  • Conference Article
  • Citations4

Programmable Dictionary Code Compression for Instruction Stream Energy Efficiency

  • Oct 01, 2020
  • Joonas Multanen +2
  • Conference Article
  • Citations3

Reconfigurable Network-on-chip design for heterogeneous multi-core system architecture

  • Jul 01, 2014
  • Jih-Sheng Shen +2
  • Conference Article
  • Citations1

The fast Fourier transform as a test case for a systolic data flow machine

  • Oct 10, 1988
  • D Tal +2
  • Conference Article
  • Citations4

Accuracy and Performance Testing of Three-Dimensional Unsaturated Flow Finite Element Groundwater Programs on the Cray XT3 Using Analytical Solutions

  • Jun 01, 2006
  • Fred T Tracy
  • Research Article
  • Citations2

Load scheduling schemes using inter—PE network

  • Jan 01, 1987
  • Systems and Computers in Japan
  • Kei Hiraki +2
  • Research Article

最先端画像映像生成処理技法 並列計算機のためのレイトレーシングアルゴリズム

  • Jan 01, 1995
  • The Journal of the Institute of Television Engineers of Japan
  • Tsuyoshi Abe +1
  • Conference Article
  • Citations1

Compiler techniques for increasing CU/PE overlap in SIMD machines

  • Apr 25, 1995
  • G Saghi +1
  • Research Article
  • Citations6

Applying an improved fast ant system to DAG scheduling for computational grids

  • Oct 01, 2008
  • Journal of the Chinese Institute of Engineers
  • Chuan‐Wen Chiang
  • Conference Article
  • Citations14

Cooling control based on model predictive control using temperature information of IT equipment for modular data center utilizing fresh-air

  • Oct 01, 2013
  • Masatoshi Ogawa +7
  • Conference Article
  • Citations5

An Implementation of a Fully Implicit Reservoir Simulator on an ICL Distributed Array Processor

  • Jan 31, 1982
  • A J Scott +7
  • Research Article
  • Citations4

A Real-Time 2D/3D Perception Visual Vector Processor for 1920 × 1080 High-Resolution High-Speed Intelligent Vision Chips

  • Feb 01, 2024
  • IEEE Transactions on Circuits and Systems I: Regular Papers
  • Siyuan Wei +14
  • Book Chapter

Synaptic Plasticity

  • Nov 12, 1998
  • Christof Koch
  • Research Article
  • Citations9

Requirements-preserving design automation for multiprocessor embedded system applications

  • Jun 05, 2020
  • Journal of Ambient Intelligence and Humanized Computing
  • Md Al Maruf +1
  • Conference Article
  • Citations10

Fast cycle-accurate simulation and instruction set generation for constraint-based descriptions of programmable architectures

  • Jan 01, 2004
  • Scott J Weber +4
Cactus Communications logo

Copyright 2026 Cactus Communications. All rights reserved.