MétaCan
Menu
Back to cohort
Record W3089592695 · doi:10.2196/21434

COVID-19 Surveillance in a Primary Care Sentinel Network: In-Pandemic Development of an Application Ontology

2020· article· en· W3089592695 on OpenAlexvenueno aff
Simon de Lusignan, Harshana Liyanage, Dylan McGagh, Bhautesh Jani, Jorgen Bauwens, Rachel Byford, Dai Evans, Tom Fahey, Trisha Greenhalgh, Nicholas Jones, Frances S Mair, Cecilia Okusi, Vaishnavi Parimalanathan, Jill P. Pell, Julian Sherlock, Oscar Tamburis, Manasa Tripathy, Filipa Ferreira, John Williams, Richard Hobbs

Bibliographic record

VenueJMIR Public Health and Surveillance · 2020
Typearticle
Languageen
FieldMedicine
TopicData-Driven Disease Surveillance
Canadian institutionsnot available
FundersPublic Health EnglandNational Institute for Health and Care ResearchRoyal College of General PractitionersWellcome Trust
KeywordsPandemicCoronavirus disease 2019 (COVID-19)Primary care2019-20 coronavirus outbreakSevere acute respiratory syndrome coronavirus 2 (SARS-CoV-2)OntologyMedical emergencyMedicineVirologyComputer scienceOutbreakFamily medicineInfectious disease (medical specialty)Pathology

Abstract

fetched live from OpenAlex

BACKGROUND: Creating an ontology for COVID-19 surveillance should help ensure transparency and consistency. Ontologies formalize conceptualizations at either the domain or application level. Application ontologies cross domains and are specified through testable use cases. Our use case was an extension of the role of the Oxford Royal College of General Practitioners (RCGP) Research and Surveillance Centre (RSC) to monitor the current pandemic and become an in-pandemic research platform. OBJECTIVE: This study aimed to develop an application ontology for COVID-19 that can be deployed across the various use-case domains of the RCGP RSC research and surveillance activities. METHODS: We described our domain-specific use case. The actor was the RCGP RSC sentinel network, the system was the course of the COVID-19 pandemic, and the outcomes were the spread and effect of mitigation measures. We used our established 3-step method to develop the ontology, separating ontological concept development from code mapping and data extract validation. We developed a coding system-independent COVID-19 case identification algorithm. As there were no gold-standard pandemic surveillance ontologies, we conducted a rapid Delphi consensus exercise through the International Medical Informatics Association Primary Health Care Informatics working group and extended networks. RESULTS: Our use-case domains included primary care, public health, virology, clinical research, and clinical informatics. Our ontology supported (1) case identification, microbiological sampling, and health outcomes at an individual practice and at the national level; (2) feedback through a dashboard; (3) a national observatory; (4) regular updates for Public Health England; and (5) transformation of a sentinel network into a trial platform. We have identified a total of 19,115 people with a definite COVID-19 status, 5226 probable cases, and 74,293 people with possible COVID-19, within the RCGP RSC network (N=5,370,225). CONCLUSIONS: The underpinning structure of our ontological approach has coped with multiple clinical coding challenges. At a time when there is uncertainty about international comparisons, clarity about the basis on which case definitions and outcomes are made from routine data is essential.

Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.

How this classification was reachedexpand

Full frame machine prediction

Teacher imitation

Not calibrated prevalence, not ground truth. Human validation pending. The Gemma side is a direct model label for every work in the frame, read from the title-only record. The Codex side is a classifier learned from the 10,348 direct Codex labels and calibrated to design-weighted sample rates; fields without enough sample support carry no Codex call. Candidate is the union of the two sides; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels.

metaresearch head score (Codex)0.013
metaresearch head score (Gemma)0.013
Version: metacan-v3-hybrid-931329e0061cValidation status: machine_predicted_unvalidated
Candidate categoriesnone
Consensus categoriesnone
DomainCandidate signal: none · Consensus signal: none
Study designCandidate signal: Not applicable · Consensus signal: none
GenreCandidate signal: Methods · Consensus signal: Methods
Teacher disagreement score0.017
Threshold uncertainty score0.069

Distilled classifier scores by category (both heads)

CategoryCodexGemma
Metaresearch0.0130.013
Meta-epidemiology (narrow)0.0010.001
Meta-epidemiology (broad)0.0000.001
Bibliometrics0.0030.002
Science and technology studies0.0020.002
Scholarly communication0.0040.006
Open science0.0020.005
Research integrity0.0020.002
Insufficient payload (model declined to judge)0.0020.001

Machine scores (provisional)

The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.

Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.

Opus teacher head0.043
GPT teacher head0.336
Teacher spread0.294 · how far apart the two teachers sit on this one work
Validation statusscore_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from it

Classification

machine, unvalidated

Machine predicted; a candidate call from one source (direct Gemma or distilled Codex), not a consensus.

The models applied no category: nothing in the taxonomy fit this work.
Study designNot applicable
Domainnot available
GenreMethods

How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".

Quick stats

Citations63
Published2020
Admission routes1
Has abstractyes

Explore more

Same venueJMIR Public Health and SurveillanceSame topicData-Driven Disease SurveillanceFrench-language works237,207