Background <p>Malaria remains a serious public health challenge in sub-Saharan Africa, and Nigeria accounts for almost 30% of global child malaria deaths. This study employs machine learning (ML) to improve prediction efficiency and identify the most significant risk factors associated with childhood malaria in high-burden populations.</p> Methods <p>A cross-sectional study was conducted among 693 under-5 children from Nigerian Internally Displaced Persons (IDP) camps. Sociodemographic data, household living conditions, and malaria knowledge were collected in addition to Rapid Diagnostic Test (RDT) outcomes. The dataset was split 70:30 to train and evaluate four ML models: Logistic Regression (LR), Decision Tree (DT), Random Forest (RF), and Gradient Boosting Machine (GBM). The performance of the models was measured by AUC, precision, recall, F1-score, and variable importance.</p> Results <p>Malaria prevalence was 68.5%. Key risk factors were a caregiver with no education (aOR = 3.23, <i>p</i> = 0.026), while female caregivers were significantly associated (aOR = 0.53, <i>p</i> = 0.024). The Random Forest model performed best (AUC = 0.892), where caregiver occupation and residential camp were the most significant predictors. A vast knowledge-practice gap was observed, where 60.3% of caregivers had knowledge of prevention but low bed net usage (2%).</p> Conclusion <p>Random Forest machine learning greatly improves the precision of malaria risk prediction. The results underscore the importance of extensive, modifiable factors such as caregiver occupation and education. Integration of these ML models into surveillance can enable precision public health interventions, including enhanced vector control and focused health education, to combat malaria effectively in high-burden populations.</p>

错误:搜索内容不能为空,请输入英文关键词
错误:关键词超出字数限制,请精简
高级检索

From risk factors to predictive modelling: applying machine learning to childhood malaria surveillance in resource-limited settings

  • Joseph Opeolu Ashaolu,
  • Taiwo S. Akanji,
  • Victoria I. Ayansola,
  • Olajumoke O. Olawale-Succes,
  • Agbolade J. Sunday,
  • Sylvain Y. M. Some

摘要

Background

Malaria remains a serious public health challenge in sub-Saharan Africa, and Nigeria accounts for almost 30% of global child malaria deaths. This study employs machine learning (ML) to improve prediction efficiency and identify the most significant risk factors associated with childhood malaria in high-burden populations.

Methods

A cross-sectional study was conducted among 693 under-5 children from Nigerian Internally Displaced Persons (IDP) camps. Sociodemographic data, household living conditions, and malaria knowledge were collected in addition to Rapid Diagnostic Test (RDT) outcomes. The dataset was split 70:30 to train and evaluate four ML models: Logistic Regression (LR), Decision Tree (DT), Random Forest (RF), and Gradient Boosting Machine (GBM). The performance of the models was measured by AUC, precision, recall, F1-score, and variable importance.

Results

Malaria prevalence was 68.5%. Key risk factors were a caregiver with no education (aOR = 3.23, p = 0.026), while female caregivers were significantly associated (aOR = 0.53, p = 0.024). The Random Forest model performed best (AUC = 0.892), where caregiver occupation and residential camp were the most significant predictors. A vast knowledge-practice gap was observed, where 60.3% of caregivers had knowledge of prevention but low bed net usage (2%).

Conclusion

Random Forest machine learning greatly improves the precision of malaria risk prediction. The results underscore the importance of extensive, modifiable factors such as caregiver occupation and education. Integration of these ML models into surveillance can enable precision public health interventions, including enhanced vector control and focused health education, to combat malaria effectively in high-burden populations.