Adaptive estimation of multi-regional soil salinization using extreme gradient boosting with Bayesian TPE optimization

土壤盐分 环境科学 特征选择 计算机科学 均方误差 土壤科学 随机森林 水文学(农业) 土壤水分 数学 统计 地质学 机器学习 岩土工程
作者
Baili Chen,Hongwei Zheng,Geping Luo,Chunbo Chen,Anming Bao,Tie Liu,Xi Chen
出处
期刊:International Journal of Remote Sensing [Informa]
卷期号:43 (3): 778-811 被引量:22
标识
DOI:10.1080/01431161.2021.2009589
摘要

Soil salinization endangers the development of ecological agriculture. As soil salinization is often heavily affected by regional environments, difficulties arise when constructing an adaptive multi-regional soil salinity estimation model. In this study, we proposed an extreme gradient boosting (XGBoost) model based on the Tree-structure Parzen Estimator (TPE) optimization algorithm to apply to four study areas with different environments (TPE-XGBoost). The four areas are the Weigan-Kuqa Oasis (Weiku), the Sangong River Basin (Sgr) and the Qitai Oasis in Xinjiang, China, and the middle and lower reaches of the Syr Darya Basin in Kazakhstan. Most previous soil salinity studies did not pay much attention to the impact of feature selection and hyper-parameter tuning on the performance of machine learning models, and the complex dependence and interaction between input features and hyper-parameters. In order to improve the performance of XGBoost model in estimating soil salinity, we proposed for the first time to use TPE algorithm to jointly optimize feature selection and hyper-parameter tuning, and verified it in four areas. Coefficient of determination (R2) and Root Mean Square Error (RMSE) were used to evaluate the model performance. First, we calculated 55 environmental features from Landsat and terrain data. Then, in order to reduce the computational complexity of the TPE-XGBoost model, we used Pearson correlation analysis between surface soil salinity content (SSC) and features to initially filter out the features that were not significantly related (P > 0.05). Finally, the TPE algorithm was used to jointly optimize the parameter space composed of features and hyper-parameters. The results showed that (1) TPE joint optimization algorithm significantly improved the performance of the XGBoost model, achieving high accuracy in the four areas, and had powerful generalization. R2 values of test sets for Weiku Oasis, Qitai Oasis, Sgr Basin, and the Syr Basin were 0.95, 0.95, 0.80, and 0.81, respectively. (2) There is no universal feature can be applied to soil salinity inversion in different environments. TPE algorithm adaptively selected different types and numbers of features for four areas, 19, 11, 25, and 15 features were selected in Weiku Oasis, Qitai Oasis, Sgr Basin, and the Syr Basin, respectively. This showed that the optimal model parameters should not be fixed parameters, but should be re-determined locally according to different environmental conditions. The TPE algorithm can capture the features that reflect environmental differences. (3) The XGBoost model can provide feature importance ranking, which improves the interpretability of machine learning model. The importance analysis results showed that the features had different contributions in different areas. The TPE-XGBoost model proposed in this study has great potential in multi-regional soil salt estimation research.

科研通智能强力驱动
Strongly Powered by AbleSci AI
科研通是完全免费的文献互助平台,具备全网最快的应助速度,最高的求助完成率。 对每一个文献求助,科研通都将尽心尽力,给求助人一个满意的交代。
实时播报
科研通AI2S应助孙雅欣采纳,获得10
1秒前
果汁发布了新的文献求助10
2秒前
Zoe013发布了新的文献求助10
3秒前
周周完成签到,获得积分10
4秒前
科研通AI6.2应助yayaj采纳,获得10
5秒前
矮小的笑旋完成签到,获得积分10
5秒前
清爽的夏瑶关注了科研通微信公众号
5秒前
新明完成签到,获得积分10
5秒前
谨慎冰薇发布了新的文献求助10
5秒前
邓佳鑫Alan应助专注乐荷采纳,获得10
6秒前
哈基米哈吉完成签到,获得积分10
8秒前
州神完成签到 ,获得积分20
9秒前
10秒前
谨慎冰薇完成签到,获得积分10
10秒前
12秒前
13秒前
Sciolto发布了新的文献求助10
13秒前
lqx完成签到,获得积分10
14秒前
14秒前
甜美傲蕾完成签到,获得积分10
15秒前
15秒前
QDU发布了新的文献求助10
15秒前
zhzhzh发布了新的文献求助10
16秒前
常温发布了新的文献求助10
17秒前
17秒前
chen完成签到,获得积分10
18秒前
19秒前
不吃菠菜关注了科研通微信公众号
19秒前
19秒前
积木123完成签到,获得积分10
19秒前
科研通AI6.3应助梦惊禅采纳,获得10
20秒前
温乘云完成签到,获得积分10
20秒前
英姑应助巴啦啦采纳,获得10
22秒前
英姑应助dg_fisher采纳,获得10
22秒前
23秒前
24秒前
乐乐应助snowman采纳,获得10
26秒前
GPTea举报唐华若求助涉嫌违规
27秒前
啊啊啊哦哦哦完成签到,获得积分10
27秒前
科研通AI6.3应助帅气冰蓝采纳,获得10
28秒前
高分求助中
(应助此贴封号)【重要!!请各用户(尤其是新用户)详细阅读】【科研通的精品贴汇总】 10000
Molecular Biology of Cancer: Mechanisms, Targets, and Therapeutics 3000
Les Mantodea de guyane 2500
Feldspar inclusion dating of ceramics and burnt stones 1000
What is the Future of Psychotherapy in a Digital Age? 801
The Psychological Quest for Meaning 800
Digital and Social Media Marketing 600
热门求助领域 (近24小时)
化学 材料科学 生物 医学 工程类 计算机科学 有机化学 物理 生物化学 纳米技术 复合材料 内科学 化学工程 人工智能 催化作用 遗传学 数学 基因 量子力学 物理化学
热门帖子
关注 科研通微信公众号,转发送积分 5968736
求助须知:如何正确求助?哪些是违规求助? 7268509
关于积分的说明 15981227
捐赠科研通 5106138
什么是DOI,文献DOI怎么找? 2742370
邀请新用户注册赠送积分活动 1707235
关于科研通互助平台的介绍 1620886