학술논문

Unified real-time environmental-epidemiological data for multiscale modeling of the COVID-19 pandemic
Document Type
article
Source
Scientific Data, Vol 10, Iss 1, Pp 1-14 (2023)
Subject
Science
Language
English
ISSN
2052-4463
Abstract
Abstract An impressive number of COVID-19 data catalogs exist. However, none are fully optimized for data science applications. Inconsistent naming and data conventions, uneven quality control, and lack of alignment between disease data and potential predictors pose barriers to robust modeling and analysis. To address this gap, we generated a unified dataset that integrates and implements quality checks of the data from numerous leading sources of COVID-19 epidemiological and environmental data. We use a globally consistent hierarchy of administrative units to facilitate analysis within and across countries. The dataset applies this unified hierarchy to align COVID-19 epidemiological data with a number of other data types relevant to understanding and predicting COVID-19 risk, including hydrometeorological data, air quality, information on COVID-19 control policies, vaccine data, and key demographic characteristics.