Standardized Diagnostic Criteria for Developmental Dysplasia of the Hip in Early Infancy
Bibliographic record
Abstract
BACKGROUND: Clinicians use various criteria to diagnose developmental dysplasia of the hip (DDH) in early infancy, but the importance of these various criteria for a definite diagnosis is controversial. The lack of uniform, widely agreed-on diagnostic criteria for DDH in patients in this age group may result in a delay in diagnosis of some patients. QUESTIONS/PURPOSES: Our purpose was to establish a consensus among pediatric orthopaedic surgeons worldwide regarding the most relevant criteria for diagnosis of DDH in infants younger than 9 weeks. MATERIAL AND METHODS: We identified 212 potential criteria relevant for diagnosing DDH in infants by surveying 467 professionals. We used the Delphi technique to reach a consensus regarding the most important criteria. We then sent the survey to 261 orthopaedic surgeons from 34 countries. RESULTS: The response rate was 75%. Thirty-seven items were identified by surgeons as most relevant to diagnose DDH in patients in this age group. Of these, 10 of 37 (27%) related to patient characteristics and history, 13 of 37 (35%) to clinical examination, 11 of 37 (30%) to ultrasound, and three of 37 (8%) to radiography. A Cronbach alpha of 0.9 for both iterations suggested consensus among the panelists. CONCLUSION: We established a consensus regarding the most relevant criteria for the diagnosis of DDH in early infancy and established their relative importance on an international basis. The highest ranked clinical criteria included the Ortolani/Barlow test, asymmetry in abduction of 20° or greater, breech presentation, leg-length discrepancy, and first-degree relative treated for DDH. LEVEL OF EVIDENCE: Level IV, diagnostic study. See the Guidelines for Authors for a complete description of levels of evidence.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame distilled prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.
Codex and Gemma teacher scores by category
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.002 | 0.004 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.000 | 0.000 |
| Bibliometrics | 0.000 | 0.000 |
| Science and technology studies | 0.000 | 0.001 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.000 | 0.000 |
| Research integrity | 0.000 | 0.001 |
| Insufficient payload (model declined to judge) | 0.000 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one teacher head, not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".