Defining and Risk-Stratifying Immunosuppression (the DESTINIES Study): Protocol for an Electronic Delphi Study
Bibliographic record
Abstract
BACKGROUND: Globally, there are marked inconsistencies in how immunosuppression is characterized and subdivided into clinical risk groups. This is detrimental to the precision and comparability of disease surveillance efforts-which has negative implications for the care of those who are immunosuppressed and their health outcomes. This was particularly apparent during the COVID-19 pandemic; despite collective motivation to protect these patients, conflicting clinical definitions created international rifts in how those who were immunosuppressed were monitored and managed during this period. We propose that international clinical consensus be built around the conditions that lead to immunosuppression and their gradations of severity concerning COVID-19. Such information can then be formalized into a digital phenotype to enhance disease surveillance and provide much-needed intelligence on risk-prioritizing these patients. OBJECTIVE: We aim to demonstrate how electronic Delphi objectives, methodology, and statistical approaches will help address this lack of consensus internationally and deliver a COVID-19 risk-stratified phenotype for "adult immunosuppression." METHODS: Leveraging existing evidence for heterogeneous COVID-19 outcomes in adults who are immunosuppressed, this work will recruit over 50 world-leading clinical, research, or policy experts in the area of immunology or clinical risk prioritization. After 2 rounds of clinical consensus building and 1 round of concluding debate, these panelists will confirm the medical conditions that should be classed as immunosuppressed and their differential vulnerability to COVID-19. Consensus statements on the time and dose dependencies of these risks will also be presented. This work will be conducted iteratively, with opportunities for panelists to ask clarifying questions between rounds and provide ongoing feedback to improve questionnaire items. Statistical analysis will focus on levels of agreement between responses. RESULTS: This protocol outlines a robust method for improving consensus on the definition and meaningful subdivision of adult immunosuppression concerning COVID-19. Panelist recruitment took place between April and May of 2024; the target set for over 50 panelists was achieved. The study launched at the end of May and data collection is projected to end in July 2024. CONCLUSIONS: This protocol, if fully implemented, will deliver a universally acceptable, clinically relevant, and electronic health record-compatible phenotype for adult immunosuppression. As well as having immediate value for COVID-19 resource prioritization, this exercise and its output hold prospective value for clinical decision-making across all diseases that disproportionately affect those who are immunosuppressed. INTERNATIONAL REGISTERED REPORT IDENTIFIER (IRRID): PRR1-10.2196/56271.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame machine prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. The Gemma side is a direct model label for every work in the frame, read from the title-only record. The Codex side is a classifier learned from the 10,348 direct Codex labels and calibrated to design-weighted sample rates; fields without enough sample support carry no Codex call. Candidate is the union of the two sides; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels.
Distilled classifier scores by category (both heads)
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.172 | 0.115 |
| Meta-epidemiology (narrow) | 0.003 | 0.003 |
| Meta-epidemiology (broad) | 0.003 | 0.004 |
| Bibliometrics | 0.005 | 0.003 |
| Science and technology studies | 0.004 | 0.004 |
| Scholarly communication | 0.005 | 0.006 |
| Open science | 0.004 | 0.006 |
| Research integrity | 0.005 | 0.008 |
| Insufficient payload (model declined to judge) | 0.062 | 0.016 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one source (direct Gemma or distilled Codex), not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".