MMST: A Multi-Modal Ground-Based Cloud Image Classification Method-Reference-Cited by-同舟云学术

MMST: A Multi-Modal Ground-Based Cloud Image Classification Method

Published:2023-04-23 Issue:9 Volume:23 Page:4222
ISSN:1424-8220
Container-title:Sensors
language:en
Short-container-title:Sensors

Author:

Wei Liang¹,Zhu Tingting¹,Guo Yiren¹,Ni Chao¹

Affiliation:

1. College of Mechanical and Electronic Engineering, Nanjing Forestry University, Nanjing 210037, China

Abstract

In recent years, convolutional neural networks have been in the leading position for ground-based cloud image classification tasks. However, this approach introduces too much inductive bias, fails to perform global modeling, and gradually tends to saturate the performance effect of convolutional neural network models as the amount of data increases. In this paper, we propose a novel method for ground-based cloud image recognition based on the multi-modal Swin Transformer (MMST), which discards the idea of using convolution to extract visual features and mainly consists of an attention mechanism module and linear layers. The Swin Transformer, the visual backbone network of MMST, enables the model to achieve better performance in downstream tasks through pre-trained weights obtained from the large-scale dataset ImageNet and can significantly shorten the transfer learning time. At the same time, the multi-modal information fusion network uses multiple linear layers and a residual structure to thoroughly learn multi-modal features, further improving the model’s performance. MMST is evaluated on the multi-modal ground-based cloud public data set MGCD. Compared with the state-of-art methods, the classification accuracy rate reaches 91.30%, which verifies its validity in ground-based cloud image classification and proves that in ground-based cloud image recognition, models based on the Transformer architecture can also achieve better results.

Funder

National Natural Science Foundation of China

Graduate Research Practice Innovation Plan of Jiangsu in 2021

Publisher

MDPI AG

Subject

Electrical and Electronic Engineering,Biochemistry,Instrumentation,Atomic and Molecular Physics, and Optics,Analytical Chemistry

Link

https://www.mdpi.com/1424-8220/23/9/4222/pdf

Reference34 articles.

1. Cloud Classification of Ground-Based Cloud Images Based on Convolutional Neural Network;Zhu;J. Phys. Conf. Ser.,2021

2. Cloud Classification Based on Structure Features of Infrared Images;Liu;J. Atmos. Ocean. Technol.,2011

3. Automatic Cloud Classification of Whole Sky Images;Heinle;Atmos. Meas. Tech.,2010

4. A Local Binary Pattern Classification Approach for Cloud Types Derived from All-Sky Imagers;Oikonomou;Int. J. Remote Sens.,2019

5. MCLOUD: A Multiview Visual Feature Extraction Mechanism for Ground-Based Cloud Image Categorization;Xiao;J. Atmos. Ocean. Technol.,2016

Cited by 1 articles. 订阅此论文施引文献订阅此论文施引文献，注册后可以免费订阅5篇论文的施引文献，订阅后可以查看论文全部施引文献

1. CloudFU-Net: A Fine-Grained Segmentation Method for Ground-Based Cloud Images Based on an Improved Encoder–Decoder Structure;IEEE Transactions on Geoscience and Remote Sensing;2024