Sensitivity and specificity of double-blinded penicillin skin testing in relation to oral provocation with amoxicillin in children
Bibliographic record
Abstract
Current recommendations for the management of penicillin allergy are to perform penicillin skin testing (PST) with penicilloyl-polylysine (PPL) and benzylpenicillin (BP) prior to drug challenge with amoxicillin. However, the role of PST is increasingly questioned in the pediatric setting. To resolve the question of PST's diagnostic accuracy, consecutive children with a history of non-life-threatening penicillin allergy referred to a tertiary-care allergy center were recruited to undergo double-blinded PST with PPL and BP prior to drug provocation to amoxicillin. Five of 158 participants (3.2%) presented with an immediate or accelerated reaction upon amoxicillin challenge, none of which were severe. Only one of these had positive PST (20%), compared to 15 of 153 amoxicillin tolerant participants (9.8%). The sensitivity and specificity of PST with PPL and BP for reacting upon amoxicillin challenge were 20% (95% CI: 0.5-71.6%) and 90% (95% CI: 84.4-94.4%), respectively. These results argue against the routine use of PST as a preliminary step to drug provocation with amoxicillin in this population, as it is unlikely to significantly alter pre-test probability of reacting to challenge.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame distilled prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.
Codex and Gemma teacher scores by category
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.000 | 0.001 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.000 | 0.000 |
| Bibliometrics | 0.000 | 0.000 |
| Science and technology studies | 0.000 | 0.000 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.000 | 0.000 |
| Research integrity | 0.000 | 0.001 |
| Insufficient payload (model declined to judge) | 0.000 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one teacher head, not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".