Assessing dengue forecasting methods: A comparative study of statistical models and machine learning techniques in Rio de Janeiro, Brazil

Author:

Chen XiangORCID,Moraga Paula

Abstract

AbstractBackgroundDengue is a mosquito-borne viral disease that poses a significant public health threat in tropical and subtropical regions worldwide. Accurate forecasting of dengue outbreaks is crucial for effective public health planning and intervention. This study aims to assess the predictive performance and computational efficiency of a number of statistical models and machine learning techniques for dengue forecasting, both with and without the inclusion of climate factors, to inform the design of dengue surveillance systems.MethodsThe study considers dengue cases in Rio de Janeiro, Brazil, as well as climate factors known to affect disease transmission. Employing a dynamic window approach, various statistical methods and machine learning techniques were used to generate weekly forecasts at several time horizons. Error measures, uncertainty intervals, and computational efficiency obtained with each method were compared. Statistical models considered were Autoregressive (AR), Moving Average (MA), Autoregressive Integrated Moving Average (ARIMA), and Exponential Smoothing State Space Model (ETS). Additionally, models incorporating temperature and humidity as covariates, such as Vector Autoregression (VAR) and Seasonal ARIMAX (SARIMAX), were employed. Machine learning techniques evaluated were Random Forest, XGBoost, Support Vector Machine (SVM), Long Short-Term Memory (LSTM) networks, and Prophet. Ensemble approaches that integrated the top performing methods were also considered. The evaluated methods also incorporated lagged climatic variables to account for delayed effects.ResultsAmong the statistical models, ARIMA demonstrated the best performance using only historical case data, while SARIMAX significantly improved predictive accuracy by incorporating climate covariates. In general, the LSTM model, particularly when combined with climate covariates, proved to be the most accurate machine learning model, despite being slower to train and predict. For long-term forecasts, Prophet with climate covariates was the most effective. Ensemble models, such as the combination of LSTM and ARIMA, showed substantial improvements over individual models.ConclusionThis study demonstrates the strengths and limitations of various methods for dengue forecasting across multiple timeframes. It highlights the best-performing statistical and machine learning methods, including their computational efficiency, underscoring the significance of machine learning techniques and the integration of climate covariates to improve forecasts. These findings offer valuable insights for public health officials, facilitating the development of dengue surveillance systems for more accurate forecasting and timely allocation of resources to mitigate dengue outbreaks.Author summaryDengue is a mosquito-borne viral disease that poses a significant public health threat in tropical and subtropical regions worldwide. Accurate forecasting of dengue can significantly aid in public health planning and response. In this study, we compared the performance of various statistical models and machine learning techniques to predict dengue cases across several timeframes. In the evaluation, we used historical dengue case data in Rio de Janeiro, Brazil, as well as climate factors such as temperature and humidity known to affect transmission. Methods considered included traditional statistical models like ARIMA and SARIMAX, and advanced machine learning approaches like Random Forest, XGBoost, SVM, LSTM, and Prophet. We found that integrating climate data significantly improved the accuracy of forecasts. Specifically, the LSTM model combined with climate covariates provided the most accurate predictions overall, while Prophet was particularly effective for long-term forecasts. Additionally, ensemble approaches that combined multiple models outperformed individual models. This work demonstrates the potential of machine learning techniques to provide timely and accurate predictions, and emphasizes the importance of climate data in dengue forecasting. The study aims to support public health officials in developing dengue surveillance systems to enable informed decision-making for mitigating the impact of dengue outbreaks.

Publisher

Cold Spring Harbor Laboratory

Reference54 articles.

1. World Health Organization. Dengue and severe dengue – Fact Sheet; 2024. Available from: https://www.who.int/news-room/fact-sheets/detail/dengue-and-severe-dengue.

2. Dengue

3. Gavi, the Vaccine Alliance. Why we need integrated dengue management to achieve zero deaths; 2024. Available from: https://www.gavi.org/vaccineswork/why-we-need-integrated-dengue-management-achieve-zero-deaths.

4. Temporal patterns and forecast of dengue infection in Northeastern Thailand;Southeast Asian journal of tropical medicine and public health,2008

5. Luz PM , Mendes BV , Codeço CT , Struchiner CJ , Galvani AP , et al. Time series analysis of dengue incidence in Rio de Janeiro, Brazil. American Society of Tropical Medicine and Hygiene. 2008;.

同舟云学术

1.学者识别学者识别

2.学术分析学术分析

3.人才评估人才评估

"同舟云学术"是以全球学者为主线,采集、加工和组织学术论文而形成的新型学术文献查询和分析系统,可以对全球学者进行文献检索和人才价值评估。用户可以通过关注某些学科领域的顶尖人物而持续追踪该领域的学科进展和研究前沿。经过近期的数据扩容,当前同舟云学术共收录了国内外主流学术期刊6万余种,收集的期刊论文及会议论文总量共计约1.5亿篇,并以每天添加12000余篇中外论文的速度递增。我们也可以为用户提供个性化、定制化的学者数据。欢迎来电咨询!咨询电话:010-8811{复制后删除}0370

www.globalauthorid.com

TOP

Copyright © 2019-2024 北京同舟云网络信息技术有限公司
京公网安备11010802033243号  京ICP备18003416号-3