Comparison of Machine Learning Approaches for Reconstructing Sea Subsurface Salinity Using Synthetic Data

Author:

Tian Tian,Leng Hongze,Wang Gongjie,Li Guancheng,Song Junqiang,Zhu Jiang,An Yuzhu

Abstract

There is a growing interest in using sparse in situ salinity data to reconstruct high-resolution three-dimensional subsurface salinity with global coverage. However, in areas with no observations, there is a lack of observation data for comparison with reconstructed fields, leading to challenges in assessing the quality and improving the accuracy of the reconstructed data. To address these issues, this study adopted the ‘resampling test’ method to establish the ‘synthetic data’ to test the performance of different machine learning algorithms. The Centre National de Recherches Meteorologiques Climate Model Version 6, and its high-resolution counterpart (CNRM-CM6-1-HR) model data was used. The key advantage of the CNRM-CM6-1-HR is that the true values for salinity are known across the entire ocean at every point in time, and thus we can compare the reconstruction result to this data. The ‘synthetic dataset’ was established by resampling the model data according to the location of in situ observations. This synthetic dataset was then used to prepare two datasets: an ‘original synthetic dataset’ with no noise added to the resampled truth value and a ‘noised synthetic dataset’ with observation error perturbation added to the resampled truth value. The resampled salinity values of the model were taken as the ‘truth values’, and the feed-forward neural network (FFNN) and light gradient boosting machine (LightGBM) approaches were used to design four reconstruction experiments and build multiple sets of reconstruction data. Finally, the advantages and disadvantages of the different reconstruction schemes were compared through multi-dimensional evaluation of the reconstructed data, and the applicability of the FFNN and LightGBM approaches for reconstructing global salinity data from sparse data was discussed. The results showed that the best-performing scheme has low root-mean-square errors (~0.035 psu) and high correlation coefficients (~0.866). The reconstructed dataset from this experiment accurately reflected the geographical pattern and vertical structure of salinity fields, and also performed well on the noised synthetic dataset. This reconstruction scheme has good generalizability and robustness, which indicates its potential as a solution for reconstructing high-resolution subsurface salinity data with global coverage in practical applications.

Funder

National Natural Science Foundation of China

Strategic Priority Research Program of the Chinese Academy of Sciences

Publisher

MDPI AG

Subject

General Earth and Planetary Sciences

Reference68 articles.

1. Estimating Global Ocean Heat Content Changes in the Upper 1800 m since 1950 and the Influence of Climatology Choice;Lyman;J. Clim.,2014

2. Quantifying Underestimates of Long-Term Upper-Ocean Warming;Durack;Nat. Clim. Chang.,2014

3. Ciais, P., Sabine, C., Bala, G., Bopp, L., Brovkin, V., Canadell, J., Chhabra, A., DeFries, R., Galloway, J., Heimann, M., Contribution of Working Group I to the Fifth Assessment Report of the Intergovernmental Panel on Climate Change. Climate Change 2013—The Physical Science Basis, 2014.

4. Improved Estimates of Upper-Ocean Warming and Multi-Decadal Sea-Level Rise;Domingues;Nature,2008

5. 20th Century Cooling of the Deep Ocean Contributed to Delayed Acceleration of Earth’s Energy Imbalance;Bagnell;Nat. Commun.,2021

Cited by 1 articles. 订阅此论文施引文献 订阅此论文施引文献,注册后可以免费订阅5篇论文的施引文献,订阅后可以查看论文全部施引文献

同舟云学术

1.学者识别学者识别

2.学术分析学术分析

3.人才评估人才评估

"同舟云学术"是以全球学者为主线,采集、加工和组织学术论文而形成的新型学术文献查询和分析系统,可以对全球学者进行文献检索和人才价值评估。用户可以通过关注某些学科领域的顶尖人物而持续追踪该领域的学科进展和研究前沿。经过近期的数据扩容,当前同舟云学术共收录了国内外主流学术期刊6万余种,收集的期刊论文及会议论文总量共计约1.5亿篇,并以每天添加12000余篇中外论文的速度递增。我们也可以为用户提供个性化、定制化的学者数据。欢迎来电咨询!咨询电话:010-8811{复制后删除}0370

www.globalauthorid.com

TOP

Copyright © 2019-2024 北京同舟云网络信息技术有限公司
京公网安备11010802033243号  京ICP备18003416号-3