A systematic review of evaluation methods for neonatal brachial plexus palsy
Bibliographic record
Abstract
OBJECT: Neonatal brachial plexus palsy (NBPP) affects 0.4-2.6 newborns per 1000 live births in the US. Many infants recover spontaneously, but for those without spontaneous recovery, nerve and/or secondary musculoskeletal reconstruction can restore function to the affected arm. This condition not only manifests in a paretic/paralyzed arm, but also affects the overall health and psychosocial condition of the children and their parents. Currently, measurement instruments for NBPP focus primarily on physical ability, with limited information regarding the effect of the disablement on activities of daily living and the child's psychosocial well-being. It is also difficult to assess and compare overall treatment efficacy among medical (conservative) or surgical management strategies without consistent use of evaluation instruments. The purpose of this study is to review the reported measurement evaluation methods for NBPP in an attempt to provide recommendations for future measurement usage and development. METHODS: The authors systematically reviewed the literature published between January 1980 and February 2012 using multiple databases to search the keywords "brachial plexus" and "obstetric" or "pediatrics" or "neonatal" or "congenital." Original articles with primary patient outcomes were included in the data summary. Four types of evaluation methods (classification, diagnostics, physical assessment, and functional outcome) were distinguished among treatment management groups. Descriptive statistics and 1-way ANOVA were applied to compare the data summaries among specific groups. RESULTS: Of 2836 articles initially identified, 307 were included in the analysis, with 198 articles (9646 patients) reporting results after surgical treatment, 70 articles (4434 patients) reporting results after medical treatment, and 39 articles (4247 patients) reporting results after combined surgical and medical treatment. Among medical practitioners who treat NBPP, there was equivalence in usage of classification, diagnostic, and physical assessment tools (that focused on the Body Function and Structures measure of the International Classification of Functioning, Disability, and Health [ICF]). However, there was discordance in the functional outcome measures that focus on ICF levels of Activity and Participation. Of the 126 reported evaluation methods, only a few (the Active Movement Scale, Toronto Scale Score, Mallet Scale, Assisting Hand Assessment, and Pediatric Outcomes Data Collection Instrument) are specifically validated for evaluating the NBPP population. CONCLUSIONS: In this review, the authors demonstrate disparities in the use of NBPP evaluation instruments in the current literature. Additionally, valid and reliable evaluation instruments specifically for the NBPP population are significantly lacking, manifesting in difficulties with evaluating the overall impact and effectiveness of clinical treatments in a consistent and comparative manner, extending across the various subspecialties that are involved in the treatment of patients with NBPP. The authors suggest that all ICF domains should be considered, and future efforts should include consideration of spontaneous (not practitioner-elicited) use of the affected arm in activities of daily living with attention to the psychosocial impact of the disablement.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame machine prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. The Gemma side is a direct model label for every work in the frame, read from the title-only record. The Codex side is a classifier learned from the 10,348 direct Codex labels and calibrated to design-weighted sample rates; fields without enough sample support carry no Codex call. Candidate is the union of the two sides; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels.
Distilled classifier scores by category (both heads)
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.023 | 0.108 |
| Meta-epidemiology (narrow) | 0.002 | 0.001 |
| Meta-epidemiology (broad) | 0.009 | 0.008 |
| Bibliometrics | 0.031 | 0.028 |
| Science and technology studies | 0.001 | 0.002 |
| Scholarly communication | 0.004 | 0.004 |
| Open science | 0.003 | 0.002 |
| Research integrity | 0.002 | 0.001 |
| Insufficient payload (model declined to judge) | 0.006 | 0.001 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one source (direct Gemma or distilled Codex), not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".