• Home
  • Search
  • An Uninterrupted Processing Technique-Based High-Throughput and Energy-Efficient Hardware Accelerator for Convolutional Neural Networks
  • Cite Icon8
  • https://doi.org/10.1109/tvlsi.2022.3210963Copy DOI Icon

An Uninterrupted Processing Technique-Based High-Throughput and Energy-Efficient Hardware Accelerator for Convolutional Neural Networks

Show More
  • Abstract
  • Literature Map
  • References
  • Citations
  • Similar Papers
Abstract

This article proposes an uninterrupted processing technique for the convolutional neural network (CNN) accelerator. It primarily allows the CNN accelerator to simultaneously perform both processing element (PE) operation and data fetching that reduces its latency and enhances the achievable throughput. Corresponding to the suggested technique, this work also presents a low latency VLSI-architecture of the CNN accelerator using the new random access line-buffer (RALB)-based design of PE array. Subsequently, the proposed CNN-accelerator architecture has been further optimized by reusing the local data in PE array, incurring better energy conservation. Our CNN accelerator has been hardware implemented on Zynq-UltraScale + MPSoC-ZCU102 FPGA board, and it operates at a maximum clock frequency of 340 MHz, consuming 4.11 W of total power. In addition, the suggested CNN accelerator with 864 PEs delivers a peak throughput of 587.52 GOPs and an adequate energy efficiency of 142.95 GOPs/W. Comparison of aforementioned implementation results with the literature has shown that our CNN accelerator delivers 33.42% higher throughput and <inline-formula xmlns:mml="http://www.w3.org/1998/Math/MathML" xmlns:xlink="http://www.w3.org/1999/xlink"> <tex-math notation="LaTeX">$6.24\times $ </tex-math></inline-formula> better energy efficiency than the state-of-the-art work. Eventually, the field-programmable gate array (FPGA) prototype of the proposed CNN accelerator has been functionally validated using the real-world test setup for the detection of object from input image, using the GoogLeNet neural network.

Similar Papers
  • Conference Article
  • Citations7

Process Variation Mitigation on Convolutional Neural Network Accelerator Architecture

  • Nov 01, 2019
  • Maodi Ma +3
  • Conference Article
  • Citations54

An Experimental Study of Reduced-Voltage Operation in Modern FPGAs for Neural Network Acceleration

  • Jun 01, 2020
  • Behzad Salami +8
  • Research Article
  • Citations20

AdaPrune: An Accelerator-Aware Pruning Technique for Sustainable CNN Accelerators

  • Jan 01, 2022
  • IEEE Transactions on Sustainable Computing
  • Jiajun Li +1
  • Research Article
  • Citations86

STICKER: An Energy-Efficient Multi-Sparsity Compatible Accelerator for Convolutional Neural Networks in 65-nm CMOS

  • Nov 22, 2019
  • IEEE Journal of Solid-State Circuits
  • Zhe Yuan +8
  • Research Article
  • Citations10

CNN Accelerator at the Edge With Adaptive Zero Skipping and Sparsity-Driven Data Flow

  • Dec 01, 2023
  • IEEE Transactions on Circuits and Systems for Video Technology
  • Min Liu +4
  • Research Article
  • Citations10

ERDSE: efficient reinforcement learning based design space exploration method for CNN accelerator on resource limited platform

  • May 09, 2021
  • Graphics and Visual Computing
  • Kaijie Feng +7
  • Conference Article
  • Citations23

Safety Design of a Convolutional Neural Network Accelerator with Error Localization and Correction

  • Nov 01, 2019
  • Zheng Xu +1
  • Conference Article
  • Citations22

An energy-efficient and high-throughput bitwise CNN on sneak-path-free digital ReRAM crossbar

  • Jul 01, 2017
  • Leibin Ni +6
  • Research Article

Design and Optimization of Hardware Accelerators for Convolutional Neural Networks

  • Dec 31, 2024
  • Science and Technology of Engineering, Chemistry and Environmental Protection
  • Jialun Liang
  • Research Article

A Design of Network Reconfigurable Universal CNN Accelerator Based on FPGA

  • Jan 23, 2026
  • ACM Transactions on Embedded Computing Systems
  • Wenhua Ye +4
  • Conference Article
  • Citations2

A Low-cost High-performance 2D-Convolution Accelerator for Deep Neural Networks in IoT

  • Dec 20, 2022
  • Hung K Nguyen +1
  • Conference Article
  • Citations8

HPIPE NX: Boosting CNN Inference Acceleration Performance with AI-Optimized FPGAs

  • Dec 05, 2022
  • Marius Stan +3
  • Conference Article
  • Citations11

Squeezing the Last MHz for CNN Acceleration on FPGAs

  • Sep 01, 2019
  • Li Li +6
  • PDF
  • Research Article
  • Citations6

Accelerating convolutional neural networks: Exploring FPGA-based architectures and challenges

  • Jun 01, 2024
  • Journal of Physics: Conference Series
  • Haobo Ye
  • Research Article
  • Citations44

An Accelerator for High Efficient Vision Processing

  • Feb 01, 2017
  • IEEE Transactions on Computer-Aided Design of Integrated Circuits and Systems
  • Zidong Du +10
Cactus Communications logo

Copyright 2026 Cactus Communications. All rights reserved.