• Home
  • Search
  • Optimized Compression for Implementing Convolutional Neural Networks on FPGA
  • Open Access IconOpen Access
  • Cite Icon75
  • https://doi.org/10.3390/electronics8030295Copy DOI Icon

Optimized Compression for Implementing Convolutional Neural Networks on FPGA

Show More
  • Abstract
  • Highlights & Summary
  • PDF
  • Literature Map
  • References
  • Citations
  • Similar Papers
Abstract

Field programmable gate array (FPGA) is widely considered as a promising platform for convolutional neural network (CNN) acceleration. However, the large numbers of parameters of CNNs cause heavy computing and memory burdens for FPGA-based CNN implementation. To solve this problem, this paper proposes an optimized compression strategy, and realizes an accelerator based on FPGA for CNNs. Firstly, a reversed-pruning strategy is proposed which reduces the number of parameters of AlexNet by a factor of 13× without accuracy loss on the ImageNet dataset. Peak-pruning is further introduced to achieve better compressibility. Moreover, quantization gives another 4× with negligible loss of accuracy. Secondly, an efficient storage technique, which aims for the reduction of the whole overhead cache of the convolutional layer and the fully connected layer, is presented respectively. Finally, the effectiveness of the proposed strategy is verified by an accelerator implemented on a Xilinx ZCU104 evaluation board. By improving existing pruning techniques and the storage format of sparse data, we significantly reduce the size of AlexNet by 28×, from 243 MB to 8.7 MB. In addition, the overall performance of our accelerator achieves 9.73 fps for the compressed AlexNet. Compared with the central processing unit (CPU) and graphics processing unit (GPU) platforms, our implementation achieves 182.3× and 1.1× improvements in latency and throughput, respectively, on the convolutional (CONV) layers of AlexNet, with an 822.0× and 15.8× improvement for energy efficiency, separately. This novel compression strategy provides a reference for other neural network applications, including CNNs, long short-term memory (LSTM), and recurrent neural networks (RNNs).

Loading PDF

Similar Papers
  • Research Article
  • Citations94

Modern deep learning in bioinformatics.

  • Jun 23, 2020
  • Journal of Molecular Cell Biology
  • Haoyang Li +8
  • Conference Article
  • Citations54

An Experimental Study of Reduced-Voltage Operation in Modern FPGAs for Neural Network Acceleration

  • Jun 01, 2020
  • Behzad Salami +8
  • Research Article

A Design of Network Reconfigurable Universal CNN Accelerator Based on FPGA

  • Jan 23, 2026
  • ACM Transactions on Embedded Computing Systems
  • Wenhua Ye +4
  • Conference Article
  • Citations40

Novel Casestudy and Benchmarking of AlexNet for Edge AI: From CPU and GPU to FPGA

  • Aug 30, 2020
  • Firas Al-Ali +4
  • Research Article
  • Citations27

Efficient CNN Accelerator on FPGA

  • Sep 24, 2020
  • IETE Journal of Research
  • S Kala +1
  • Research Article
  • Citations8

An Uninterrupted Processing Technique-Based High-Throughput and Energy-Efficient Hardware Accelerator for Convolutional Neural Networks

  • Dec 01, 2022
  • IEEE Transactions on Very Large Scale Integration (VLSI) Systems
  • Md Najrul Islam +2
  • Research Article
  • Citations41

OMNI: A Framework for Integrating Hardware and Software Optimizations for Sparse CNNs

  • Sep 14, 2020
  • IEEE Transactions on Computer-Aided Design of Integrated Circuits and Systems
  • Yun Liang +2
  • Conference Article

Binarized Neural Network Acceleration with Field Programmable Gate Array and Its Application in Robot Vision

  • Oct 12, 2021
  • Po-Chiang Lin +1
  • PDF
  • Research Article
  • Citations6

Accelerating convolutional neural networks: Exploring FPGA-based architectures and challenges

  • Jun 01, 2024
  • Journal of Physics: Conference Series
  • Haobo Ye
  • Conference Article
  • Citations3

A Comparison of CNNs and LSTMs for EEG Signal Classification

  • Apr 27, 2022
  • Albert Ting +4
  • Conference Article

Design of MSSTP-SA Net Hardware Accelerator Based on FPGA

  • Dec 01, 2021
  • Yudong Zhang
  • Research Article
  • Citations10

Share Price Trend Prediction Using CRNN with LSTM Structure

  • Apr 18, 2019
  • Smart Science
  • Shyr-Shen Yu +3
  • Conference Article
  • Citations8

HPIPE NX: Boosting CNN Inference Acceleration Performance with AI-Optimized FPGAs

  • Dec 05, 2022
  • Marius Stan +3
  • Research Article
  • Citations44

Separation of water and fat signal in whole-body gradient echo scans using convolutional neural networks.

  • Apr 29, 2019
  • Magnetic Resonance in Medicine
  • Jonathan Andersson +2
  • Conference Article

Comparative Analysis of CNN, RNN and LSTM for Synthetic Grip Strength Prediction

  • Aug 27, 2025
  • N Ahmad Syazwan +5
Cactus Communications logo

Copyright 2026 Cactus Communications. All rights reserved.