Landscape analysis for a neonatal disease progression model of bronchopulmonary dysplasia: Leveraging clinical trial experience and real-world data
Bibliographic record
Abstract
Century Cures Act requires FDA to expand its use of real-world evidence (RWE) to support approval of previously approved drugs for new disease indications and post-marketing study requirements. To address this need in neonates, the FDA and the Critical Path Institute (C-Path) established the International Neonatal Consortium (INC) to advance regulatory science and expedite neonatal drug development. FDA recently provided funding for INC to generate RWE to support regulatory decision making in neonatal drug development. One study is focused on developing a validated definition of bronchopulmonary dysplasia (BPD) in neonates. BPD is difficult to diagnose with diverse disease trajectories and few viable treatment options. Despite intense research efforts, limited understanding of the underlying disease pathobiology and disease projection continues in the context of a computable phenotype. It will be important to determine if: 1) a large, multisource aggregation of real-world data (RWD) will allow identification of validated risk factors and surrogate endpoints for BPD, and 2) the inclusion of these simulations will identify risk factors and surrogate endpoints for studies to prevent or treat BPD and its related long-term complications. The overall goal is to develop qualified, fit-for-purpose disease progression models which facilitate credible trial simulations while quantitatively capturing mechanistic relationships relevant for disease progression and the development of future treatments. The extent to which neonatal RWD can inform these models is unknown and its appropriateness cannot be guaranteed. A component of this approach is the critical evaluation of the various RWD sources for context-of use (COU)-driven models. The present manuscript defines a landscape of the data including targeted literature searches and solicitation of neonatal RWD sources from international stakeholders; analysis plans to develop a family of models of BPD in neonates, leveraging previous clinical trial experience and real-world patient data is also described.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame machine prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. The Gemma side is a direct model label for every work in the frame, read from the title-only record. The Codex side is a classifier learned from the 10,348 direct Codex labels and calibrated to design-weighted sample rates; fields without enough sample support carry no Codex call. Candidate is the union of the two sides; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels.
Distilled classifier scores by category (both heads)
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.010 | 0.031 |
| Meta-epidemiology (narrow) | 0.001 | 0.001 |
| Meta-epidemiology (broad) | 0.001 | 0.001 |
| Bibliometrics | 0.002 | 0.001 |
| Science and technology studies | 0.000 | 0.001 |
| Scholarly communication | 0.002 | 0.001 |
| Open science | 0.002 | 0.002 |
| Research integrity | 0.002 | 0.002 |
| Insufficient payload (model declined to judge) | 0.004 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one source (direct Gemma or distilled Codex), not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".