4.5 Article

Comparing data mining methods with logistic regression in childhood obesity prediction

期刊

INFORMATION SYSTEMS FRONTIERS
卷 11, 期 4, 页码 449-460

出版社

SPRINGER
DOI: 10.1007/s10796-009-9157-0

关键词

Medical data mining; Machine learning; Public health; Prediction; Accuracy

资金

  1. ESRC [ES/F029721/1] Funding Source: UKRI
  2. Economic and Social Research Council [ES/F029721/1] Funding Source: researchfish

向作者/读者索取更多资源

The epidemiological question of concern here is can young children at risk of obesity be identified from their early growth records? Pilot work using logistic regression to predict overweight and obese children demonstrated relatively limited success. Hence we investigate the incorporation of non-linear interactions to help improve accuracy of prediction; by comparing the result of logistic regression with those of six mature data mining techniques. The contributions of this paper are as follows: a) a comparison of logistic regression with six data mining techniques: specifically, for the prediction of overweight and obese children at 3 years using data recorded at birth, 6 weeks, 8 months and 2 years respectively; b) improved accuracy of prediction: prediction at 8 months accuracy is improved very slightly, in this case by using neural networks, whereas for prediction at 2 years obtained accuracy is improved by over 10%, in this case by using Bayesian methods. It has also been shown that incorporation of non-linear interactions could be important in epidemiological prediction, and that data mining techniques are becoming sufficiently well established to offer the medical research community a valid alternative to logistic regression.

作者

我是这篇论文的作者
点击您的名字以认领此论文并将其添加到您的个人资料中。

评论

主要评分

4.5
评分不足

次要评分

新颖性
-
重要性
-
科学严谨性
-
评价这篇论文

推荐

暂无数据
暂无数据