Kayo Henrique de Carvalho Monteiro, Élisson da Silva Rocha, Luis Augusto Morais, Elton Gino Santos, Sebastião Rogerio da S Neto, Vanderson Sampaio, Patricia Takako Endo
{"title":"Integrating machine learning and spatial clustering for malaria case prediction in Brazil's Legal Amazon.","authors":"Kayo Henrique de Carvalho Monteiro, Élisson da Silva Rocha, Luis Augusto Morais, Elton Gino Santos, Sebastião Rogerio da S Neto, Vanderson Sampaio, Patricia Takako Endo","doi":"10.1186/s12879-025-11193-x","DOIUrl":null,"url":null,"abstract":"<p><p>Malaria remains a major global health challenge, particularly in Brazil's Legal Amazon region, where environmental and socioeconomic conditions foster favorable conditions for disease transmission. Traditional control measures have shown limited effectiveness, emphasizing the need for better predictive approaches to support timely and targeted public health interventions. This study evaluates the performance of six computational models-Long Short-Term Memory (LSTM), Gated Recurrent Units (GRU), Support Vector Regression (SVR), Random Forest (RF), eXtreme Gradient Boosting (XGBoost), and Autoregressive Integrated Moving Average (ARIMA)-for forecasting weekly malaria cases across multiple states in the Legal Amazon. The results demonstrate that the RF model consistently outperformed the other models, achieving the lowest Root Mean Squared Error (RMSE) and Mean Absolute Error (MAE) values in most cases, such as in cluster 02 of the state of Acre, with RMSE of 0.00203 and MAE of 0.00133. The integration of K-means clustering further improved the model predictive accuracy by accounting for spatial heterogeneity and capturing localized transmission dynamics. This hybrid modeling approach, combining machine learning models with spatial clustering, offers a promising tool for enhancing malaria surveillance and guiding more effective public health strategies, especially for malaria control efforts in high-risk regions.</p>","PeriodicalId":8981,"journal":{"name":"BMC Infectious Diseases","volume":"25 1","pages":"802"},"PeriodicalIF":3.4000,"publicationDate":"2025-06-08","publicationTypes":"Journal Article","fieldsOfStudy":null,"isOpenAccess":false,"openAccessPdf":"https://www.ncbi.nlm.nih.gov/pmc/articles/PMC12147289/pdf/","citationCount":"0","resultStr":null,"platform":"Semanticscholar","paperid":null,"PeriodicalName":"BMC Infectious Diseases","FirstCategoryId":"3","ListUrlMain":"https://doi.org/10.1186/s12879-025-11193-x","RegionNum":3,"RegionCategory":"医学","ArticlePicture":[],"TitleCN":null,"AbstractTextCN":null,"PMCID":null,"EPubDate":"","PubModel":"","JCR":"Q2","JCRName":"INFECTIOUS DISEASES","Score":null,"Total":0}
引用次数: 0
Abstract
Malaria remains a major global health challenge, particularly in Brazil's Legal Amazon region, where environmental and socioeconomic conditions foster favorable conditions for disease transmission. Traditional control measures have shown limited effectiveness, emphasizing the need for better predictive approaches to support timely and targeted public health interventions. This study evaluates the performance of six computational models-Long Short-Term Memory (LSTM), Gated Recurrent Units (GRU), Support Vector Regression (SVR), Random Forest (RF), eXtreme Gradient Boosting (XGBoost), and Autoregressive Integrated Moving Average (ARIMA)-for forecasting weekly malaria cases across multiple states in the Legal Amazon. The results demonstrate that the RF model consistently outperformed the other models, achieving the lowest Root Mean Squared Error (RMSE) and Mean Absolute Error (MAE) values in most cases, such as in cluster 02 of the state of Acre, with RMSE of 0.00203 and MAE of 0.00133. The integration of K-means clustering further improved the model predictive accuracy by accounting for spatial heterogeneity and capturing localized transmission dynamics. This hybrid modeling approach, combining machine learning models with spatial clustering, offers a promising tool for enhancing malaria surveillance and guiding more effective public health strategies, especially for malaria control efforts in high-risk regions.
期刊介绍:
BMC Infectious Diseases is an open access, peer-reviewed journal that considers articles on all aspects of the prevention, diagnosis and management of infectious and sexually transmitted diseases in humans, as well as related molecular genetics, pathophysiology, and epidemiology.