The Americleft Project: A Comparison of Short- and Longer-Term Secondary Alveolar Bone Graft Outcomes in Two Centers Using the Standardized Way to Assess Grafts Scale
Bibliographic record
Abstract
OBJECTIVE: To compare length of follow-up and cleft site dental management on bone graft ratings from two centers. DESIGN: Blind retrospective analysis of cleft site radiographs and chart reviews for determination of cleft-site lateral incisor management. PATIENTS: A total of 78 consecutively grafted patients with complete clefts from two major cleft/craniofacial centers (43 from Center 1 and 35 from Center 2). INTERVENTIONS: Secondary iliac crest alveolar bone grafting, at a mean age of 9 years 9 months (Center 1: 9 years 7 months; Center 2: 10 years 0 month). MAIN OUTCOME MEASURES: The Americleft Standardized Way to Assess Grafts scale from 0 (failed graft) to 6 (ideal) was used to rate graft outcome at two time points (T1, T2). Average T1 was 11 years 1 month of age, 1 year 3 months postgraft. Average T2 was 17 years 11 months of age, 8 years 0 months postgraft. Six trained and calibrated raters scored each radiograph twice. Reliability was calculated at T1 and T2 using weighted kappa. A paired Wilcoxon signed rank test (P < .05) tested T1 and T2 differences for each center. A Kruskal-Wallis test was used to determine the significance of differences between centers at T1 and T2. Correlation tested whether T1 ratings predicted T2. Linear regression determined possible factors that might contribute to graft rating changes over time. RESULTS: Reliability was good at T1 and T2 (interrater = .713 and .701, respectively; intrarater = .790 and .805, respectively). Center 1 scores were significantly better than those from Center 2 at both T1 (5.21 versus 3.29) and T2 (5.18 versus 3.44). There was no statistical difference between T1 and T2 scores for either center; although, there was a greater chance of bone graft score improving with completion of canine eruption and substitution for missing lateral incisors. CONCLUSIONS: Short-term ratings of graft outcomes identified significant differences between centers that persisted over time. Dental cleft-site management influenced final graft outcome.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame distilled prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.
Codex and Gemma teacher scores by category
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.002 | 0.000 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.000 | 0.000 |
| Bibliometrics | 0.000 | 0.000 |
| Science and technology studies | 0.001 | 0.000 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.001 | 0.000 |
| Research integrity | 0.000 | 0.001 |
| Insufficient payload (model declined to judge) | 0.000 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one teacher head, not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".