维普中文期刊产品整合服务

Comparison and development of advanced machine learning tools to predict nonalcoholic fatty liver disease:An extended study

查看全文 作  者:Yuan-Xing [1]Liu;Xi [1]Liu;Chao [1]Cen;Xin [2]Li;Ji-Min [3]Liu;Zhao-Yan [4]Ming;Song-Feng [1]Yu;Xiao-Feng [1]Tang;Lin [1]Zhou;Jun [1]Yu;Ke-Jie [2]Huang;Shu-Sen [1]Zheng 高影响力作者 机构地区:[1]Department of Hepatobiliary and Pancreatic Surgery,the First Affiliated Hospital,Zhejiang University School of Medicine,Hangzhou 310003,China;[2]College of Information Science and Electrical Engineering,Zhejiang University,Hangzhou 310027,China;[3]Department of Pathology and Molecular Medicine,Faculty of Health Sciences,McMaster University,1200 Main Street West,Hamilton L8S 4K1,Canada;[4]College of Computer Science and Technology,Zhejiang University,Hangzhou 310027,China高影响力机构 出  处:《Hepatobiliary & Pancreatic Diseases International》索引2021年第20卷第5期,共7页高影响力期刊 基  金:supported by grants from the National Natural Science Foundation of China(81970543 and 81570591);Zhejiang Provincial Medical&Hygienic Science and Technology Project of China(2018KY385);Zhejiang Provincial Natural Science Foundation of China(LY20H160023)。 摘  要:Background:Nonalcoholic fatty liver disease(NAFLD)is a public health challenge and significant cause of morbidity and mortality worldwide.Early identification is crucial for disease intervention.We recently proposed a nomogram-based NAFLD prediction model from a large population cohort.We aimed to explore machine learning tools in predicting NAFLD.Methods:A retrospective cross-sectional study was performed on 15315 Chinese subjects(10373 training and 4942 testing sets).Selected clinical and biochemical factors were evaluated by different types of machine learning algorithms to develop and validate seven predictive models.Nine evaluation indicators including area under the receiver operating characteristic curve(AUROC),area under the precision-recall curve(AUPRC),accuracy,positive predictive value,sensitivity,F1 score,Matthews correlation coefficient(MCC),specificity and negative prognostic value were applied to compare the performance among the models.The selected clinical and biochemical factors were ranked according to the importance in prediction ability.Results:Totally 4018/10373(38.74%)and 1860/4942(37.64%)subjects had ultrasound-proven NAFLD in the training and testing sets,respectively.Seven machine learning based models were developed and demonstrated good performance in predicting NAFLD.Among these models,the XGBoost model revealed the highest AUROC(0.873),AUPRC(0.810),accuracy(0.795),positive predictive value(0.806),F1 score(0.695),MCC(0.557),specificity(0.909),demonstrating the best prediction ability among the built models.Body mass index was the most valuable indicator to predict NAFLD according to the feature ranking scores.Conclusions:The XGBoost model has the best overall prediction ability for diagnosing NAFLD.The novel machine learning tools provide considerable beneficial potential in NAFLD screening. 关 键 词:Nonalcoholic fatty liver disease Machine learning Population screening Prediction model Body mass index
相关文献

参考文献(30)

引证文献(3)

网站首页 | 关于我们 | 联系我们 | 产品服务 | 客服中心 | 广告服务 | 版权声明 | 网站联盟 | 友情链接 | 售卡网点

版权所有© 渝B2-20050021-1 渝公网安备 50019002500403号 违法和不良信息举报中心

互联网出版许可证 新出网证(渝)字10号 全国400电话 - 免长途话费