IDEAS home Printed from https://ideas.repec.org/a/plo/pone00/0299386.html
   My bibliography  Save this article

Predicting malaria outbreak in The Gambia using machine learning techniques

Author

Listed:
  • Ousman Khan
  • Jimoh Olawale Ajadi
  • M Pear Hossain

Abstract

Malaria is the most common cause of death among the parasitic diseases. Malaria continues to pose a growing threat to the public health and economic growth of nations in the tropical and subtropical parts of the world. This study aims to address this challenge by developing a predictive model for malaria outbreaks in each district of The Gambia, leveraging historical meteorological data. To achieve this objective, we employ and compare the performance of eight machine learning algorithms, including C5.0 decision trees, artificial neural networks, k-nearest neighbors, support vector machines with linear and radial kernels, logistic regression, extreme gradient boosting, and random forests. The models are evaluated using 10-fold cross-validation during the training phase, repeated five times to ensure robust validation. Our findings reveal that extreme gradient boosting and decision trees exhibit the highest prediction accuracy on the testing set, achieving 93.3% accuracy, followed closely by random forests with 91.5% accuracy. In contrast, the support vector machine with a linear kernel performs less favorably, showing a prediction accuracy of 84.8% and underperforming in specificity analysis. Notably, the integration of both climatic and non-climatic features proves to be a crucial factor in accurately predicting malaria outbreaks in The Gambia.

Suggested Citation

  • Ousman Khan & Jimoh Olawale Ajadi & M Pear Hossain, 2024. "Predicting malaria outbreak in The Gambia using machine learning techniques," PLOS ONE, Public Library of Science, vol. 19(5), pages 1-22, May.
  • Handle: RePEc:plo:pone00:0299386
    DOI: 10.1371/journal.pone.0299386
    as

    Download full text from publisher

    File URL: https://journals.plos.org/plosone/article?id=10.1371/journal.pone.0299386
    Download Restriction: no

    File URL: https://journals.plos.org/plosone/article/file?id=10.1371/journal.pone.0299386&type=printable
    Download Restriction: no

    File URL: https://libkey.io/10.1371/journal.pone.0299386?utm_source=ideas
    LibKey link: if access is restricted and if your library uses this service, LibKey will redirect you to where you can use your library subscription to access this item
    ---><---

    References listed on IDEAS

    as
    1. Kuhn, Max, 2008. "Building Predictive Models in R Using the caret Package," Journal of Statistical Software, Foundation for Open Access Statistics, vol. 28(i05).
    2. repec:osf:socarx:hkc96_v1 is not listed on IDEAS
    3. Jae Hun Kim & Juyeon Kim & Gunwoo Lee & Juneyoung Park, 2021. "Machine Learning-Based Models for Accident Prediction at a Korean Container Port," Sustainability, MDPI, vol. 13(16), pages 1-14, August.
    4. Thackway, William & Ng, Matthew Kok Ming & Lee, Chyi Lin & Pettit, Christopher, 2021. "Building a predictive machine learning model of gentrification in Sydney," SocArXiv hkc96, Center for Open Science.
    5. Ousmane Diao & P.-A. Absil & Mouhamadou Diallo, 2023. "Generalized Linear Models to Forecast Malaria Incidence in Three Endemic Regions of Senegal," IJERPH, MDPI, vol. 20(13), pages 1-27, July.
    Full references (including those not matched with items on IDEAS)

    Most related items

    These are the items that most often cite the same works as this one and are cited by the same works as this one.
    1. Prabal Das & D. A. Sachindra & Kironmala Chanda, 2022. "Machine Learning-Based Rainfall Forecasting with Multiple Non-Linear Feature Selection Algorithms," Water Resources Management: An International Journal, Published for the European Water Resources Association (EWRA), Springer;European Water Resources Association (EWRA), vol. 36(15), pages 6043-6071, December.
    2. Paulo Infante & Gonçalo Jacinto & Anabela Afonso & Leonor Rego & Pedro Nogueira & Marcelo Silva & Vitor Nogueira & José Saias & Paulo Quaresma & Daniel Santos & Patrícia Góis & Paulo Rebelo Manuel, 2023. "Factors That Influence the Type of Road Traffic Accidents: A Case Study in a District of Portugal," Sustainability, MDPI, vol. 15(3), pages 1-16, January.
    3. Ephrem Habyarimana & Faheem S Baloch, 2021. "Machine learning models based on remote and proximal sensing as potential methods for in-season biomass yields prediction in commercial sorghum fields," PLOS ONE, Public Library of Science, vol. 16(3), pages 1-23, March.
    4. Crespo, Cristian, 2020. "Two become one: improving the targeting of conditional cash transfers with a predictive model of school dropout," LSE Research Online Documents on Economics 123139, London School of Economics and Political Science, LSE Library.
    5. Alexander Wettstein & Gabriel Jenni & Ida Schneider & Fabienne Kühne & Martin grosse Holtforth & Roberto La Marca, 2023. "Predictors of Psychological Strain and Allostatic Load in Teachers: Examining the Long-Term Effects of Biopsychosocial Risk and Protective Factors Using a LASSO Regression Approach," IJERPH, MDPI, vol. 20(10), pages 1-20, May.
    6. Tang, Kayu & Parsons, David J. & Jude, Simon, 2019. "Comparison of automatic and guided learning for Bayesian networks to analyse pipe failures in the water distribution system," Reliability Engineering and System Safety, Elsevier, vol. 186(C), pages 24-36.
    7. Daifeng Xiang & Gangsheng Wang & Jing Tian & Wanyu Li, 2023. "Global patterns and edaphic-climatic controls of soil carbon decomposition kinetics predicted from incubation experiments," Nature Communications, Nature, vol. 14(1), pages 1-14, December.
    8. Joel Podgorski & Oliver Kracht & Luis Araguas-Araguas & Stefan Terzer-Wassmuth & Jodie Miller & Ralf Straub & Rolf Kipfer & Michael Berg, 2024. "Groundwater vulnerability to pollution in Africa’s Sahel region," Nature Sustainability, Nature, vol. 7(5), pages 558-567, May.
    9. Tranos, Emmanouil & Incera, Andre Carrascal & Willis, George, 2022. "Using the web to predict regional trade flows: data extraction, modelling, and validation," OSF Preprints 9bu5z, Center for Open Science.
    10. Štefan Lyócsa & Petra Vašaničová & Branka Hadji Misheva & Marko Dávid Vateha, 2022. "Default or profit scoring credit systems? Evidence from European and US peer-to-peer lending markets," Financial Innovation, Springer;Southwestern University of Finance and Economics, vol. 8(1), pages 1-21, December.
    11. Marcos Rodrigues & Fermín Alcasena & Pere Gelabert & Cristina Vega‐García, 2020. "Geospatial Modeling of Containment Probability for Escaped Wildfires in a Mediterranean Region," Risk Analysis, John Wiley & Sons, vol. 40(9), pages 1762-1779, September.
    12. Siyu Han & Shixiang Yu & Mengya Shi & Makoto Harada & Jianhong Ge & Jiesheng Lin & Cornelia Prehn & Agnese Petrera & Ying Li & Flora Sam & Giuseppe Matullo & Jerzy Adamski & Karsten Suhre & Christian , 2025. "LEOPARD: missing view completion for multi-timepoint omics data via representation disentanglement and temporal knowledge transfer," Nature Communications, Nature, vol. 16(1), pages 1-20, December.
    13. Natalia Pardo-Lorente & Anestis Gkanogiannis & Luca Cozzuto & Antoni Gañez Zapater & Lorena Espinar & Ritobrata Ghose & Jacqueline Severino & Laura García-López & Rabia Gül Aydin & Laura Martin & Mari, 2024. "Nuclear localization of MTHFD2 is required for correct mitosis progression," Nature Communications, Nature, vol. 15(1), pages 1-23, December.
    14. Andrea Lazzari & Simone Giovinazzo & Giovanni Cabassi & Massimo Brambilla & Carlo Bisaglia & Elio Romano, 2025. "Evaluating Urban Sewage Sludge Distribution on Agricultural Land Using Interpolation and Machine Learning Techniques," Agriculture, MDPI, vol. 15(2), pages 1-13, January.
    15. Giovanny Pillajo-Quijia & Blanca Arenas-Ramírez & Camino González-Fernández & Francisco Aparicio-Izquierdo, 2020. "Influential Factors on Injury Severity for Drivers of Light Trucks and Vans with Machine Learning Methods," Sustainability, MDPI, vol. 12(4), pages 1-28, February.
    16. Zander S. Venter & Adam Sadilek & Charlotte Stanton & David N. Barton & Kristin Aunan & Sourangsu Chowdhury & Aaron Schneider & Stefano Maria Iacus, 2021. "Mobility in Blue-Green Spaces Does Not Predict COVID-19 Transmission: A Global Analysis," IJERPH, MDPI, vol. 18(23), pages 1-12, November.
    17. G. Brooke Anderson & Keith W. Oleson & Bryan Jones & Roger D. Peng, 2018. "Classifying heatwaves: developing health-based models to predict high-mortality versus moderate United States heatwaves," Climatic Change, Springer, vol. 146(3), pages 439-453, February.
    18. Van Belle, Jente & Guns, Tias & Verbeke, Wouter, 2021. "Using shared sell-through data to forecast wholesaler demand in multi-echelon supply chains," European Journal of Operational Research, Elsevier, vol. 288(2), pages 466-479.
    19. Jun Wang & Jinyong Huang & Yunlong Hu & Qianwen Guo & Shasha Zhang & Jinglin Tian & Yanqin Niu & Ling Ji & Yuzhong Xu & Peijun Tang & Yaqin He & Yuna Wang & Shuya Zhang & Hao Yang & Kang Kang & Xinchu, 2024. "Terminal modifications independent cell-free RNA sequencing enables sensitive early cancer detection and classification," Nature Communications, Nature, vol. 15(1), pages 1-13, December.
    20. Ali Al-Ramini & Mohammad A Takallou & Daniel P Piatkowski & Fadi Alsaleem, 2022. "Quantifying changes in bicycle volumes using crowdsourced data," Environment and Planning B, , vol. 49(6), pages 1612-1630, July.

    More about this item

    Statistics

    Access and download statistics

    Corrections

    All material on this site has been provided by the respective publishers and authors. You can help correct errors and omissions. When requesting a correction, please mention this item's handle: RePEc:plo:pone00:0299386. See general information about how to correct material in RePEc.

    If you have authored this item and are not yet registered with RePEc, we encourage you to do it here. This allows to link your profile to this item. It also allows you to accept potential citations to this item that we are uncertain about.

    If CitEc recognized a bibliographic reference but did not link an item in RePEc to it, you can help with this form .

    If you know of missing items citing this one, you can help us creating those links by adding the relevant references in the same way as above, for each refering item. If you are a registered author of this item, you may also want to check the "citations" tab in your RePEc Author Service profile, as there may be some citations waiting for confirmation.

    For technical questions regarding this item, or to correct its authors, title, abstract, bibliographic or download information, contact: plosone (email available below). General contact details of provider: https://journals.plos.org/plosone/ .

    Please note that corrections may take a couple of weeks to filter through the various RePEc services.

    IDEAS is a RePEc service. RePEc uses bibliographic data supplied by the respective publishers.