A Delphi Process to Optimize Quality and Performance of Drug Evaluation in Neonates
Bibliographic record
Abstract
BACKGROUND: Neonatal trials remain difficult to conduct for several reasons: in particular the need for study sites to have an existing infrastructure in place, with trained investigators and validated quality procedures to ensure good clinical, laboratory practices and a respect for high ethical standards. The objective of this work was to identify the major criteria considered necessary for selecting neonatal intensive care units that are able to perform drug evaluations competently. METHODOLOGY AND MAIN FINDINGS: This Delphi process was conducted with an international multidisciplinary panel of 25 experts from 13 countries, selected to be part of two committees (a scientific committee and an expert committee), in order to validate criteria required to perform drug evaluation in neonates. Eighty six items were initially selected and classified under 7 headings: "NICUs description-Level of care" (21), "Ability to perform drug trials: NICU organization and processes (15), "Research Experience" (12), "Scientific competencies and area of expertise" (8), "Quality Management" (16), "Training and educational capacity" (8) and "Public involvement" (6). Sixty-one items were retained and headings were rearranged after the first round, 34 were selected after the second round. A third round was required to validate 13 additional items. The final set includes 47 items divided under 5 headings. CONCLUSION: A set of 47 relevant criteria will help to NICUs that want to implement, conduct or participate in drug trials within a neonatal network identify important issues to be aware of. SUMMARY POINTS: 1) Neonatal trials remain difficult to conduct for several reasons: in particular the need for study sites to have an existing infrastructure in place, with trained investigators and validated quality procedures to ensure good clinical, laboratory practices and a respect for high ethical standards. 2) The present Delphi study was conducted with an international multidisciplinary panel of 25 experts from 13 countries and aims to identify the major criteria considered necessary for selecting neonatal intensive care units (NICUs) that are able to perform drug evaluations competently. 3) Of the 86 items initially selected and classified under 7 headings--"NICUs description-Level of care" (21), "Ability to perform drug trials: NICU organization and processes (15), "Research Experience" (12), "Scientific competencies and area of expertise" (8), "Quality Management" (16), "Training and educational capacity" (8) and "Public involvement" (6)--47 items were selected following a three rounds Delphi process. 4) The present consensus will help NICUs to implement, conduct or participate in drug trials within a neonatal network.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame distilled prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.
Codex and Gemma teacher scores by category
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.006 | 0.002 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.000 | 0.000 |
| Bibliometrics | 0.000 | 0.000 |
| Science and technology studies | 0.000 | 0.000 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.000 | 0.000 |
| Research integrity | 0.000 | 0.000 |
| Insufficient payload (model declined to judge) | 0.000 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one teacher head, not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".