• Home
  • Search
  • Target Capacity Filter Pruning Method for Optimized Inference Time Based on YOLOv5 in Embedded Systems
  • Cite Icon15
  • https://doi.org/10.1109/access.2022.3188323Copy DOI Icon

Target Capacity Filter Pruning Method for Optimized Inference Time Based on YOLOv5 in Embedded Systems

Show More
  • Abstract
  • Literature Map
  • References
  • Citations
  • Similar Papers
Abstract

Recently, convolutional neural networks (CNNs), which exhibit excellent performance in the field of computer vision, have been in the spotlight. However, as the networks become wider for higher accuracy, the number of parameters and the computational costs increase exponentially. Therefore, it is challenging to use deep learning networks in embedded environments with limited resources, computational performance, and power. Moreover, CNNs consume a great deal of time for inference. To solve this problem, we propose a practical method for filter pruning to provide an optimal network architecture for target capacity and inference acceleration. After revealing the correlation between the inference time and the FLOPs, we proposed a method to generate a network with the desired inference time. Various object detection datasets were used to evaluate the performance of the proposed filter pruning method. The inference time of the pruned network was measured and analyzed using the NVIDIA Jetson Xavier NX platform. As a result of pruning the number of parameters and FLOPs of the YOLOv5 network in the PASCAL VOC dataset by 30%, 40%, and 50%, the mAP decreased by 0.6%, 2.3%, and 2.9%, respectively, while the inference time was improved by 14.3%, 26.4%, and 34.5%, respectively.

Similar Papers
  • Research Article
  • Citations49

FP-AGL: Filter Pruning With Adaptive Gradient Learning for Accelerating Deep Convolutional Neural Networks

  • Jan 01, 2023
  • IEEE Transactions on Multimedia
  • Nam Joon Kim +1
  • Conference Article
  • Citations1101

Soft Filter Pruning for Accelerating Deep Convolutional Neural Networks

  • Jul 01, 2018
  • Yang He +4
  • Conference Article
  • Citations2

Accelerate CNN Models via Filter Pruning and Sparse Tensor Core

  • Nov 01, 2021
  • An-Tai Chen +3
  • PDF
  • Research Article
  • Citations9

Filter Pruning and Re-Initialization via Latent Space Clustering

  • Jan 01, 2020
  • IEEE Access
  • Seunghyun Lee +3
  • Conference Article
  • Citations1

A Convolutional Hierarchical Neural Network Classifier

  • Dec 06, 2021
  • Ismail Gadzhiev +1
  • Conference Article
  • Citations6

Inference Time Reduction of Deep Neural Networks on Embedded Devices: A Case Study

  • Aug 01, 2022
  • Isma-Ilou Sadou +3
  • Research Article
  • Citations100

Adaptive deep learning model selection on embedded systems

  • Jun 19, 2018
  • ACM SIGPLAN Notices
  • Ben Taylor +4
  • Research Article
  • Citations8

RT3D: Achieving Real-Time Execution of 3D Convolutional Neural Networks on Mobile Devices

  • May 18, 2021
  • Proceedings of the AAAI Conference on Artificial Intelligence
  • Wei Niu +9
  • Conference Article
  • Citations6

Split Convolutional Neural Networks for Distributed Inference on Concurrent IoT Sensors

  • Dec 01, 2021
  • Jiale Chen +3
  • Research Article

Evaluasi Kinerja CNN dan Vision Transformer pada Klasifikasi Citra Resolusi Tinggi Berbasis Deep Learning

  • Oct 31, 2025
  • METHOMIKA Jurnal Manajemen Informatika dan Komputerisasi Akuntansi
  • Humuntal Rumapea
  • Conference Article
  • Citations270

Accelerating Convolutional Networks via Global & Dynamic Filter Pruning

  • Jul 01, 2018
  • Shaohui Lin +5
  • PDF
  • Research Article
  • Citations88

WaveCRN: An Efficient Convolutional Recurrent Neural Network for End-to-End Speech Enhancement

  • Jan 01, 2020
  • IEEE Signal Processing Letters
  • Tsun-An Hsieh +3
  • Research Article
  • Citations89

Toward Performing Image Classification and Object Detection With Convolutional Neural Networks in Autonomous Driving Systems: A Survey

  • Jan 01, 2022
  • IEEE Access
  • Tolga Turay +1
  • Research Article
  • Citations20

QSFM: Model Pruning Based on Quantified Similarity Between Feature Maps for AI on Edge

  • Dec 01, 2022
  • IEEE Internet of Things Journal
  • Zidu Wang +6
  • PDF
  • Research Article
  • Citations89

Deep Convolutional Neural Networks Object Detector for Real-Time Waste Identification

  • Oct 19, 2020
  • Applied Sciences
  • Daniel Octavian Melinte +2
Cactus Communications logo

Copyright 2026 Cactus Communications. All rights reserved.