Risk assessment and genomic characterization of Zika virus in China and its surrounding areas
Bibliographic record
Abstract
BACKGROUND: Zika virus (ZIKV) has emerged as a global pathogen causing significant public health concerns. China has reported several imported cases where ZIKV were carried by travelers who frequently travel between China and ZIKV-endemic regions. To fully characterize the ZIKV strains isolated from the cases reported in China and assess the risk of ZIKV transmission in China, comprehensive phylogenetic and genetic analyses were performed both on all ZIKV sequences of China and on a group of scientifically selected ZIKV sequences reported in some of the top interested destinations for Chinese travelers. METHODS: ZIKV genomic sequences were retrieved from the National Center for Biotechnology Information database through stratified sampling. Recombination event detection, maximum likelihood (ML) phylogenetic analysis, molecular clock analysis, selection pressure analysis, and amino acid substitution analysis were used to reconstruct the epidemiology and molecular transmission of ZIKV. RESULTS: The present study investigated 18 ZIKV sequences from China and 70 sequences from 16 selected countries. Recombination events rarely happens in all ZIKV Asian lineage. ZIKV genomes were generally undergone episodic positive selection (17 sites), and only one site was under pervasive positive selection. All ZIKV imported into China were Asian lineage and were assigned into two clusters: Venezuela-origin (cluster A) and Samoa-origin cluster (cluster B) with common ancestor from French Polynesia. The time of most recent common ancestors of Cluster A dated to approximately 2013/11 (95% highest posterior density [HPD] 2013/06, 2014/03) and cluster B dated to 2014/08 (95% HPD 2014/02, 2015/01). Cluster B is more variable than Cluster A in comparison with other clusters, but no varied site of biological significance was revealed. ZIKV strains in Southeast Asia countries are independent from strains in America epidemics. CONCLUSIONS: The genetic evolution of ZIKV is conservative. There are two independent introductions of ZIKV into China and China is in danger of autochthonous transmission of ZIKV because of high-risk surrounding areas. Southeast Asia areas have high risk of originating the next large-scale epidemic ZIKV strains.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame distilled prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.
Codex and Gemma teacher scores by category
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.001 | 0.000 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.000 | 0.000 |
| Bibliometrics | 0.000 | 0.000 |
| Science and technology studies | 0.000 | 0.000 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.000 | 0.000 |
| Research integrity | 0.000 | 0.000 |
| Insufficient payload (model declined to judge) | 0.001 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one teacher head, not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".