• Home
  • Search
  • Performance Analysis of Existing SIMD Architectures
  • https://doi.org/10.1007/978-981-15-1850-8_4Copy DOI Icon

Performance Analysis of Existing SIMD Architectures

  • Jan 1, 2019
  • Chao Cui +2 more
Show More
  • Abstract
  • Literature Map
  • References
  • Similar Papers
Abstract

SIMD (Single Instruction Multiple Data) architectures are widely used in application domains like the wireless communication, video and audio processing, and control engineering. The abundant data parallelism makes the SIMD architecture the proper match in data processing and performance improvement. However, there are also critical inefficiencies in current SIMD architectures. To understand such inefficiency, we carry out a deep investigation in the main components of Long Term Evolution (LTE) protocol, which is an important wireless communication protocol. Performance investigation is taken on a cycle-accurate simulator, featuring the main characteristics of existing SIMD architectures. Based on the investigation, we locate the inefficiencies in two aspects: the data communication operations among different processing units and the support for matrix-style computations. We have also carried out studies with enhanced SIMD architectures in the above two aspects. The overall performance of SIMD architectures can be greatly improved.

Similar Papers
  • Conference Article

Accelearation of Full-Search Algorithm on SIMD Architectures by Using Eight-Bit Partial Sums of Four Luminance Values

  • Dec 01, 2006
  • C J Duanmu
  • Research Article
  • Citations11

A Low-Energy Wide SIMD Architecture with Explicit Datapath

  • Sep 16, 2014
  • Journal of Signal Processing Systems
  • Luc Waeijen +3
  • Conference Article

Optimization of H.264 video decoder from Android OS for MIPS32 DSP ASE architectures

  • Nov 01, 2015
  • Branimir Vasić +3
  • Conference Article
  • Citations13

A configurable SIMD architecture with explicit datapath for intelligent learning

  • Jul 01, 2016
  • Yifan He +7
  • Book Chapter
  • Citations2

Design and Implementation of a General Purpose Neural Network Processor

  • Jun 03, 2007
  • Yi Qian +2
  • Research Article
  • Citations64

A Fast Iterative Method for Solving the Eikonal Equation on Triangulated Surfaces

  • Jan 01, 2011
  • SIAM Journal on Scientific Computing
  • Zhisong Fu +4
  • Conference Article

PHAFS: A parallel hardware accelerator for switch level fault simulation

  • Apr 04, 1993
  • C.A Ryan +1
  • Conference Article
  • Citations6

Using a commercial graphical processing unit and the CUDA programming language to accelerate scientific image processing applications

  • Jan 23, 2011
  • Proceedings of SPIE, the International Society for Optical Engineering/Proceedings of SPIE
  • Randy P Broussard +1
  • Conference Article
  • Citations3

On the scalability of SIMD processing for software defined radio algorithms

  • Jul 01, 2010
  • Peter Westermann +1
  • Conference Article
  • Citations104

From SODA to scotch: The evolution of a wireless baseband processor

  • Nov 01, 2008
  • Mark Woh +10
  • Conference Article
  • Citations6

Customized SIMD unit synthesis for system on programmable chip

  • Jan 01, 2006
  • Muhammad Omer Cheema +1
  • Conference Article
  • Citations1

Outer-Loop Auto-Vectorization for SIMD Architectures Based on Open64 Compiler

  • Dec 01, 2016
  • Wang Dong +3
  • Research Article
  • Citations2

An Efficient Matrix-Based 2-D DCT Splitter and Merger for SIMD Instructions

  • Jul 01, 2005
  • IEICE Transactions on Information and Systems
  • Y.-J Chuang
  • Research Article
  • Citations2

GRAPHIC: Gather and Process Harmoniously in the Cache With High Parallelism and Flexibility

  • Jan 01, 2024
  • IEEE Transactions on Emerging Topics in Computing
  • Yiming Chen +11
  • Conference Article

Efficient Architecture for Controlled Accurate Computation using AVX

  • May 02, 2018
  • Diaaeldin M Osman +3
Cactus Communications logo

Copyright 2026 Cactus Communications. All rights reserved.