Simple Optimal Sampling Algorithm to Strengthen Digital Soil Mapping Using the Spatial Distribution of Machine Learning Predictive Uncertainty: A Case Study for Field Capacity Prediction-Reference-Cited by-同舟云学术

Simple Optimal Sampling Algorithm to Strengthen Digital Soil Mapping Using the Spatial Distribution of Machine Learning Predictive Uncertainty: A Case Study for Field Capacity Prediction

Published:2022-11-21 Issue:11 Volume:11 Page:2098
ISSN:2073-445X
Container-title:Land
language:en
Short-container-title:Land

Author:

Yang Hyunje^ORCID,Lim Honggeun,Moon Haewon,Li Qiwen,Nam Sooyoun,Kim Jaehoon,Choi Hyung Tae

Abstract

Machine learning models are now capable of delivering coveted digital soil mapping (DSM) benefits (e.g., field capacity (FC) prediction); therefore, determining the optimal sample sites and sample size is essential to maximize the training efficacy. We solve this with a novel optimal sampling algorithm that allows the authentic augmentation of insufficient soil features using machine learning predictive uncertainty. Nine hundred and fifty-three forest soil samples and geographically referenced forest information were used to develop predictive models, and FCs in South Korea were estimated with six predictor set hierarchies. Random forest and gradient boosting models were used for estimation since tree-based models had better predictive performance than other machine learning algorithms. There was a significant relationship between model predictive uncertainties and training data distribution, where higher uncertainties were distributed in the data scarcity area. Further, we confirmed that the predictive uncertainties decreased when additional sample sites were added to the training data. Environmental covariate information of each grid cell in South Korea was then used to select the sampling sites. Optimal sites were coordinated at the cell having the highest predictive uncertainty, and the sample size was determined using the predictable rate. This intuitive method can be generalized to improve global DSM.

Publisher

MDPI AG

Subject

Nature and Landscape Conservation,Ecology,Global and Planetary Change

Link

https://www.mdpi.com/2073-445X/11/11/2098/pdf

Reference43 articles.

1. Deriving hydrological signatures from soil moisture data;Hydrol. Process.,2020

2. Multiscale investigations in a mesoscale catchment—Hydrological modelling in the Gera catchment;Adv. Geosci.,2006

3. A quantitative Australian approach to medium and small scale surveys based on soil stratigraphy and environmental correlation;Geoderma,1993

4. Perspectives on validation in digital soil mapping of continuous attributes—A review;Soil Use Manag.,2021

5. Data Evaluation and Enhancement for Quality Improvement of Machine Learning;IEEE Trans. Reliab.,2021

Cited by 3 articles. 订阅此论文施引文献订阅此论文施引文献，注册后可以免费订阅5篇论文的施引文献，订阅后可以查看论文全部施引文献

1. Innovative graph neural network approach for predicting soil heavy metal pollution in the Pearl River Basin, China;Scientific Reports;2024-07-17

2. Probing the randomness of the local current distributions of 316 L stainless steel corrosion in NaCl solution;Corrosion Science;2023-06

3. Identifying the Minimum Number of Flood Events for Reasonable Flood Peak Prediction of Ungauged Forested Catchments in South Korea;Forests;2023-05-30