TransCode: Uncovering COVID-19 transmission patterns via deep learning
Bibliographic record
Abstract
BACKGROUND: The heterogeneity of COVID-19 spread dynamics is determined by complex spatiotemporal transmission patterns at a fine scale, especially in densely populated regions. In this study, we aim to discover such fine-scale transmission patterns via deep learning. METHODS: We introduce the notion of TransCode to characterize fine-scale spatiotemporal transmission patterns of COVID-19 caused by metapopulation mobility and contact behaviors. First, in Hong Kong, China, we construct the mobility trajectories of confirmed cases using their visiting records. Then we estimate the transmissibility of individual cases in different locations based on their temporal infectiousness distribution. Integrating the spatial and temporal information, we represent the TransCode via spatiotemporal transmission networks. Further, we propose a deep transfer learning model to adapt the TransCode of Hong Kong, China to achieve fine-scale transmission characterization and risk prediction in six densely populated metropolises: New York City, San Francisco, Toronto, London, Berlin, and Tokyo, where fine-scale data are limited. All the data used in this study are publicly available. RESULTS: The TransCode of Hong Kong, China derived from the spatial transmission information and temporal infectiousness distribution of individual cases reveals the transmission patterns (e.g., the imported and exported transmission intensities) at the district and constituency levels during different COVID-19 outbreaks waves. By adapting the TransCode of Hong Kong, China to other data-limited densely populated metropolises, the proposed method outperforms other representative methods by more than 10% in terms of the prediction accuracy of the disease dynamics (i.e., the trend of case numbers), and the fine-scale spatiotemporal transmission patterns in these metropolises could also be well captured due to some shared intrinsically common patterns of human mobility and contact behaviors at the metapopulation level. CONCLUSIONS: The fine-scale transmission patterns due to the metapopulation level mobility (e.g., travel across different districts) and contact behaviors (e.g., gathering in social-economic centers) are one of the main contributors to the rapid spread of the virus. Characterization of the fine-scale transmission patterns using the TransCode will facilitate the development of tailor-made intervention strategies to effectively contain disease transmission in the targeted regions.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame distilled prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.
Codex and Gemma teacher scores by category
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.000 | 0.005 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.001 | 0.000 |
| Bibliometrics | 0.000 | 0.000 |
| Science and technology studies | 0.000 | 0.000 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.000 | 0.000 |
| Research integrity | 0.000 | 0.000 |
| Insufficient payload (model declined to judge) | 0.000 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one teacher head, not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".