科研速览 · Science Skim继续刷下去 · Keep skimming →
◆ PLoS computational biology2026-08-01

Leveraging synthetic and genetic data to improve epidemic forecasting.

Dave Osthus, Alexander C Murph, Emma E Goldberg, Lauren J Beesley, William M Fischer, Nidhi Parikh, Lauren A Castro

原始摘要(英文原文)· Original abstract
Forecasting infectious disease outbreaks is hard. Forecasting emerging infectious diseases with limited historical data is even harder. In this paper, we investigate ways to improve emerging infectious disease forecasting when little pathogen-specific training data are available. Specifically, we explore two sources of information that may be available near the start of an emerging disease outbreak: synthetic data and genetic information. For this investigation, we conducted an experiment where we trained deep learning models on different combinations of real and synthetic data, both with and without genetic information, to explore how these models compare when forecasting COVID-19 cases for US states. All models are developed with an eye towards forecasting the next pandemic. We find that models trained with synthetic data have better forecast accuracy than models trained on real data alone, and models that use genetic variants have better forecast accuracy compared to those that do not. All models outperformed a baseline persistence model, a benchmark that proved challenging for many real-time COVID-19 case forecasting models, and multiple models outperformed the COVIDHub-4_week_ensemble. This paper demonstrates the value of these underutilized sources of information and provides a blueprint for forecasting future pandemics.
读原文 · Read the paper ↗

AI 追问PRO

登录后使用 AI 追问

讨论区

登录后参与讨论

相关论文 · Related

Leveraging synthetic and genetic data to improve epidemic forecasting. — 科研速览 Science Skim