Author response: Evidence for transmission of COVID-19 prior to symptom onset
Bibliographic record
Abstract
The first cases of COVID-19 were identified in Wuhan, a city in Central China, in December 2019. The virus quickly spread within the country and then across the globe. By the third week in January, the first cases were confirmed in Tianjin, a city in Northern China, and in Singapore, a city country in Southeast Asia. By late February, Tianjin had 135 cases and Singapore had 93 cases. In both cities, public health officials immediately began identifying and quarantining the contacts of infected people. The information collected in Tianjin and Singapore about COVID-19 is very useful for scientists. It makes it possible to determine the disease’s incubation period, which is how long it takes to develop symptoms after virus exposure. It can also show how many days pass between an infected person developing symptoms and a person they infect developing symptoms. This period is called the serial interval. Scientists use this information to determine whether individuals infect others before showing symptoms themselves and how often this occurs. Using data from Tianjin and Singapore, Tindale, Stockdale et al. now estimate the incubation period for COVID-19 is between five and eight days and the serial interval is about four days. About 40% to 80% of the novel coronavirus transmission occurs two to four days before an infected person has symptoms. This transmission from apparently healthy individuals means that staying home when symptomatic is not enough to control the spread of COVID-19. Instead, broad-scale social distancing measures are necessary. Understanding how COVID-19 spreads can help public health officials determine how to best contain the virus and stop the outbreak. The new data suggest that public health measures aimed at preventing asymptomatic transmission are essential. This means that even people who appear healthy need to comply with preventive measures like mask use and social distancing.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame machine prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. The Gemma side is a direct model label for every work in the frame, read from the title-only record. The Codex side is a classifier learned from the 10,348 direct Codex labels and calibrated to design-weighted sample rates; fields without enough sample support carry no Codex call. Candidate is the union of the two sides; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels.
Distilled classifier scores by category (both heads)
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.005 | 0.064 |
| Meta-epidemiology (narrow) | 0.001 | 0.001 |
| Meta-epidemiology (broad) | 0.001 | 0.001 |
| Bibliometrics | 0.001 | 0.001 |
| Science and technology studies | 0.001 | 0.001 |
| Scholarly communication | 0.002 | 0.002 |
| Open science | 0.002 | 0.002 |
| Research integrity | 0.010 | 0.005 |
| Insufficient payload (model declined to judge) | 0.183 | 0.084 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one source (direct Gemma or distilled Codex), not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".