• Home
  • Search
  • Corun: Concurrent Inference and Continuous Training at the Edge for Cost-Efficient AI-Based Mobile Image Sensing
  • Cite Icon2
  • https://doi.org/10.3390/s24165262Copy DOI Icon

Corun: Concurrent Inference and Continuous Training at the Edge for Cost-Efficient AI-Based Mobile Image Sensing

Show More
  • Abstract
  • Literature Map
  • References
  • Citations
  • Similar Papers
Abstract

Intelligent mobile image sensing powered by deep learning analyzes images captured by cameras from mobile devices, such as smartphones or smartwatches. It supports numerous mobile applications, such as image classification, face recognition, and camera scene detection. Unfortunately, mobile devices often lack the resources necessary for deep learning, leading to increased inference latency and rapid battery consumption. Moreover, the inference accuracy may decline over time due to potential data drift. To address these issues, we introduce a new cost-efficient framework, called Corun, designed to simultaneously handle multiple inference queries and continual model retraining/fine-tuning of a pre-trained model on a single commodity GPU in an edge server to significantly improve the inference throughput, upholding the inference accuracy. The scheduling method of Corun undertakes offline profiling to find the maximum number of concurrent inferences that can be executed along with a retraining job on a single GPU without incurring an out-of-memory error or significantly increasing the latency. Our evaluation verifies the cost-effectiveness of Corun. The inference throughput provided by Corun scales with the number of concurrent inference queries. However, the latency of inference queries and the length of a retraining epoch increase at substantially lower rates. By concurrently processing multiple inference and retraining tasks on one GPU instead of using a separate GPU for each task, Corun could reduce the number of GPUs and cost required to deploy mobile image sensing applications based on deep learning at the edge.

Similar Papers
  • Research Article
  • Citations2

MCP-MedSAM: A Powerful Lightweight Medical Segment Anything Model Trained with a Single GPU in Just One Day

  • May 12, 2025
  • Machine Learning for Biomedical Imaging
  • Donghang Lyu +2
  • Research Article
  • Citations4

Facial age recognition based on deep manifold learning.

  • Jan 01, 2024
  • Mathematical Biosciences and Engineering
  • Huiying Zhang +4
  • PDF
  • Research Article
  • Citations22

Efficient Use of GPU Memory for Large-Scale Deep Learning Model Training

  • Nov 04, 2021
  • Applied Sciences
  • Hyeonseong Choi +1
  • Research Article
  • Citations20

Optimizing makespan and resource utilization for multi-DNN training in GPU cluster

  • Jun 24, 2021
  • Future Generation Computer Systems
  • Zhongjin Li +5
  • Research Article
  • Citations2

EvaSurf: Efficient View-Aware Implicit Textured Surface Reconstruction.

  • Sep 01, 2025
  • IEEE transactions on visualization and computer graphics
  • Jingnan Gao +6
  • Research Article
  • Citations39

Application of Face Recognition Method Under Deep Learning Algorithm in Embedded Systems

  • Jan 01, 2021
  • Microprocessors and Microsystems
  • Xue Lv +2
  • Book Chapter
  • Citations8

Recognition of Facial Expression Using Haar Cascade Classifier and Deep Learning

  • Jan 01, 2022
  • J Premaladha +4
  • Conference Article
  • Citations124

Deep learning on mobile devices: a review

  • May 13, 2019
  • Yunbin Deng
  • Research Article
  • Citations40

An Improvement of Data Classification Using Random Multimodel Deep Learning (RMDL)

  • Aug 01, 2018
  • International Journal of Machine Learning and Computing
  • Mojtaba Heidarysafa +4
  • Research Article
  • Citations14

MTFC: A Multi-GPU Training Framework for Cube-CNN-Based Hyperspectral Image Classification

  • Aug 17, 2020
  • IEEE Transactions on Emerging Topics in Computing
  • Ye Lu +5
  • Research Article
  • Citations89

Image Classification Algorithm Based on Deep Learning-Kernel Function

  • Jan 31, 2020
  • Scientific Programming
  • Jun-E Liu +1
  • Conference Article

Fast, Scalable, and Memory-Efficient High-Performance Graph Neural Networks for Subsurface in Porous Media Simulations using Multi-GPU Parallelism

  • Nov 03, 2025
  • Zeeshan Tariq +1
  • Research Article

YOLO-CIANNA: Galaxy detection with deep learning in radio data. II. Winning the SKA SDC2 using a generalized 3D-YOLO network

  • Feb 08, 2026
  • Astronomy & Astrophysics
  • D Cornu +10
  • Research Article
  • Citations8

Black hole weather forecasting with deep learning: a pilot study

  • Mar 22, 2022
  • Monthly Notices of the Royal Astronomical Society
  • Roberta Duarte +2
  • Research Article
  • Citations31

Flexible protein-protein docking with a multitrack iterative transformer.

  • Jan 23, 2024
  • Protein science : a publication of the Protein Society
  • Lee-Shin Chu +3
Cactus Communications logo

Copyright 2026 Cactus Communications. All rights reserved.