Predicting risk of obesity in overweight adults using interpretable machine learning algorithms

利用可解释的机器学习算法预测超重成年人的肥胖风险

阅读：1

作者：Lin,Wei,Shi,Songchang,Huang,Huibin,Wen,Junping,Chen,Gang

期刊：	Frontiers in Endocrinology	影响因子：	4.600
时间：	2023	起止号：	2023;14:1292167
doi：	10.3389/fendo.2023.1292167	研究方向：	代谢
疾病类型：	肥胖

Abstract

OBJECTIVE: To screen for predictive obesity factors in overweight populations using an optimal and interpretable machine learning algorithm. METHODS: This cross-sectional study was conducted between June 2011 and January 2012. The participants were randomly selected using a simple random sampling technique. Seven commonly used machine learning methods were employed to construct obesity risk prediction models. A total of 5,236 Chinese participants from Ningde City, Fujian Province, Southeast China, participated in this study. The best model was selected through appropriate verification and validation and suitably explained. Subsequently, a minimal set of significant predictors was identified. The Shapley additive explanation force plot was used to illustrate the model at the individual level. RESULTS: Machine learning models for predicting obesity have demonstrated strong performance, with CatBoost emerging as the most effective in both model validity and net clinical benefit. Specifically, the CatBoost algorithm yielded the highest scores, registering 0.91 in the training set and an impressive 0.83 in the test set. This was further corroborated by the area under the curve (AUC) metrics, where CatBoost achieved 0.95 for the training set and 0.87 for the test set. In a rigorous five-fold cross-validation, the AUC for the CatBoost model ranged between 0.84 and 0.91, with an average AUC of ROC at 0.87 ± 0.022. Key predictors identified within these models included waist circumference, hip circumference, female gender, and systolic blood pressure. CONCLUSION: CatBoost may be the best machine learning method for prediction. Combining Shapley's additive explanation and machine learning methods can be effective in identifying disease risk factors for prevention and control.

特别声明

1、本页面内容包含部分的内容是基于公开信息的合理引用；引用内容仅为补充信息，不代表本站立场。

2、若认为本页面引用内容涉及侵权，请及时与本站联系，我们将第一时间处理。

3、其他媒体/个人如需使用本页面原创内容，需注明“来源：[生知库]”并获得授权；使用引用内容的，需自行联系原作者获得许可。

4、投稿及合作请联系：info@biocloudy.com。