Computation-Efficient Deep Learning for Computer Vision: A Survey

Show Author's Information Hide Author's Information Yulin Wang^¹, Yizeng Han^¹, Chaofei Wang^¹, Shiji Song^¹, Qi Tian^², Gao Huang^¹(

)

¹ Department of Automation, BNRist, Tsinghua University, Beijing 100084, China

² Huawei Inc., China

Keywords:

Computer vision, Efficient deep learning, Deep neural networks

Cite this article:

Wang Y, Han Y, Wang C, et al. Computation-Efficient Deep Learning for Computer Vision: A Survey. Cybernetics and Intelligence, 2023, https://doi.org/10.26599/CAI.2024.9390002

Download citation

EndNote(RIS)

BibTeX

120

Views

Downloads

Citations

Crossref

N/A

WoS

N/A

Scopus

N/A

CSCD

Abstract About this article

Abstract

Over the past decade, deep learning models have exhibited considerable advancements, reaching or even exceeding human-level performance in a range of visual perception tasks. This remarkable progress has sparked interest in applying deep networks to real-world applications, such as autonomous vehicles, mobile devices, robotics, and edge computing. However, the challenge remains that state-of-the-art models usually demand significant computational resources, leading to impractical power consumption, latency, or carbon emissions in real-world scenarios. This trade-off between effectiveness and efficiency has catalyzed the emergence of a new research focus: computationally efficient deep learning, which strives to achieve satisfactory performance while minimizing the computational cost during inference. This review offers an extensive analysis of this rapidly evolving field by examining four key areas: 1) the development of static or dynamic light-weighted backbone models for the efficient extraction of discriminative deep representations; 2) the specialized network architectures or algorithms tailored for specific computer vision tasks; 3) the techniques employed for compressing deep learning models; and 4) the strategies for deploying efficient deep networks on hardware platforms. Additionally, we provide a systematic discussion on the critical challenges faced in this domain, such as network architecture design, training schemes, practical efficiency, and more realistic model compression approaches, as well as potential future research directions.

About this article

Publication history

Rights and permissions

Publication history

Received: 03 April 2023

Accepted: 18 August 2023

Available online: 29 December 2023

Copyright

Rights and permissions

The articles published in this open access journal are distributed under the terms of the Creative Commons Attribution 4.0 International License (http://creativecommons.org/licenses/by/4.0/).