Syndromic Surveillance Using Structured Telehealth Data: Case Study of the First Wave of COVID-19 in Brazil
Bibliographic record
Abstract
BACKGROUND: Telehealth has been widely used for new case detection and telemonitoring during the COVID-19 pandemic. It safely provides access to health care services and expands assistance to remote, rural areas and underserved communities in situations of shortage of specialized health professionals. Qualified data are systematically collected by health care workers containing information on suspected cases and can be used as a proxy of disease spread for surveillance purposes. However, the use of this approach for syndromic surveillance has yet to be explored. Besides, the mathematical modeling of epidemics is a well-established field that has been successfully used for tracking the spread of SARS-CoV-2 infection, supporting the decision-making process on diverse aspects of public health response to the COVID-19 pandemic. The response of the current models depends on the quality of input data, particularly the transmission rate, initial conditions, and other parameters present in compartmental models. Telehealth systems may feed numerical models developed to model virus spread in a specific region. OBJECTIVE: Herein, we evaluated whether a high-quality data set obtained from a state-based telehealth service could be used to forecast the geographical spread of new cases of COVID-19 and to feed computational models of disease spread. METHODS: We analyzed structured data obtained from a statewide toll-free telehealth service during 4 months following the first notification of COVID-19 in the Bahia state, Brazil. Structured data were collected during teletriage by a health team of medical students supervised by physicians. Data were registered in a responsive web application for planning and surveillance purposes. The data set was designed to quickly identify users, city, residence neighborhood, date, sex, age, and COVID-19-like symptoms. We performed a temporal-spatial comparison of calls reporting COVID-19-like symptoms and notification of COVID-19 cases. The number of calls was used as a proxy of exposed individuals to feed a mathematical model called "susceptible, exposed, infected, recovered, deceased." RESULTS: For 181 (43%) out of 417 municipalities of Bahia, the first call to the telehealth service reporting COVID-19-like symptoms preceded the first notification of the disease. The calls preceded, on average, 30 days of the notification of COVID-19 in the municipalities of the state of Bahia, Brazil. Additionally, data obtained by the telehealth service were used to effectively reproduce the spread of COVID-19 in Salvador, the capital of the state, using the "susceptible, exposed, infected, recovered, deceased" model to simulate the spatiotemporal spread of the disease. CONCLUSIONS: Data from telehealth services confer high effectiveness in anticipating new waves of COVID-19 and may help understand the epidemic dynamics.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame machine prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. The Gemma side is a direct model label for every work in the frame, read from the title-only record. The Codex side is a classifier learned from the 10,348 direct Codex labels and calibrated to design-weighted sample rates; fields without enough sample support carry no Codex call. Candidate is the union of the two sides; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels.
Distilled classifier scores by category (both heads)
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.001 | 0.006 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.000 | 0.000 |
| Bibliometrics | 0.001 | 0.001 |
| Science and technology studies | 0.001 | 0.001 |
| Scholarly communication | 0.001 | 0.000 |
| Open science | 0.001 | 0.001 |
| Research integrity | 0.001 | 0.000 |
| Insufficient payload (model declined to judge) | 0.000 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one source (direct Gemma or distilled Codex), not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".