AI Chat Paper
Note: Please note that the following content is generated by AMiner AI. SciOpen does not take any responsibility related to this content.
{{lang === 'zh_CN' ? '文章概述' : 'Summary'}}
{{lang === 'en_US' ? '中' : 'Eng'}}
Chat more with AI
PDF (1.5 MB)
Collect
Submit Manuscript AI Chat Paper
Show Outline
Outline
Show full outline
Hide outline
Outline
Show full outline
Hide outline
Article | Open Access

CloudViT: A Lightweight Ground-Based Cloud Image Classification Model with the Ability to Capture Global Features

Daoming Wei1Fangyan Ge2Bopeng Zhang1Zhiqiang Zhao3Dequan Li3( )Lizong Xi4Jinrong Hu5( )Xin Wang6
National Key Laboratory of Intelligent Spatial Information, Beijing, 100029, China
School of Artificial Intelligence, Neijiang Normal University, Neijiang, 641100, China
CMA Cloud-Precipitation Physics and Weather Modification Key Laboratory, Beijing, 100081, China
Gansu Weather Modification Office, Lanzhou, 730020, China
School of Computer Science, Chengdu University of Information Technology, Chengdu, 610225, China
Department of Epidemiology and Biostatistics, School of Public Health, University at Albany, State University of New York, New York, NY 12144, USA
Show Author Information

Abstract

Accurate cloud classification plays a crucial role in aviation safety, climate monitoring, and localized weather forecasting. Current research has been focusing on machine learning techniques, particularly deep learning based model, for the types identification. However, traditional approaches such as convolutional neural networks (CNNs) encounter difficulties in capturing global contextual information. In addition, they are computationally expensive, which restricts their usability in resource-limited environments. To tackle these issues, we present the Cloud Vision Transformer (CloudViT), a lightweight model that integrates CNNs with Transformers. The integration enables an effective balance between local and global feature extraction. To be specific, CloudViT comprises two innovative modules: Feature Extraction (E_Module) and Downsampling (D_Module). These modules are able to significantly reduce the number of model parameters and computational complexity while maintaining translation invariance and enhancing contextual comprehension. Overall, the CloudViT includes 0.93 × 106 parameters, which decreases more than ten times compared to the SOTA (State-of-the-Art) model CloudNet. Comprehensive evaluations conducted on the HBMCD and SWIMCAT datasets showcase the outstanding performance of CloudViT. It achieves classification accuracies of 98.45% and 100%, respectively. Moreover, the efficiency and scalability of CloudViT make it an ideal candidate for deployment in mobile cloud observation systems, enabling real-time cloud image classification. The proposed hybrid architecture of CloudViT offers a promising approach for advancing ground-based cloud image classification. It holds significant potential for both optimizing performance and facilitating practical deployment scenarios.

References

【1】
【1】
 
 
Computers, Materials & Continua
Pages 5729-5746

{{item.num}}

Comments on this article

Go to comment

< Back to all reports

Review Status: {{reviewData.commendedNum}} Commended , {{reviewData.revisionRequiredNum}} Revision Required , {{reviewData.notCommendedNum}} Not Commended Under Peer Review

Review Comment

Close
Close
Cite this article:
Wei D, Ge F, Zhang B, et al. CloudViT: A Lightweight Ground-Based Cloud Image Classification Model with the Ability to Capture Global Features. Computers, Materials & Continua, 2025, 83(3): 5729-5746. https://doi.org/10.32604/cmc.2025.061402

115

Views

2

Downloads

1

Crossref

0

Web of Science

0

Scopus

Received: 23 November 2024
Accepted: 20 March 2025
Published: 19 May 2025
© The Author 2025.

This work is licensed under a Creative Commons Attribution 4.0 International License, which permits unrestricted use, distribution, and reproduction in any medium, provided the original work is properly cited.