Digital blinding of radiographs to mask allocation in a randomized control trial
Bibliographic record
Abstract
AIM: To demonstrate the effectiveness of a digital radiographic altering technique in concealing treatment allocation to blind outcome assessment of distal femur fracture fixation. METHODS: Digital postoperative anteroposterior and lateral radiographs from a sample of 33 randomly-selected patients with extra-articular distal femur fractures treated by surgical fixation at a Level 1 trauma center were included. Using commercially available digital altering software, we devised a technique to blind the radiographs by overlaying black boxes over the implant hardware while preserving an exposed fracture site for assessment of fracture healing. Three fellowship-trained surgeons evaluated a set of blinded radiographs twice and a control set of unblinded radiographs once. Each set of radiographs were reviewed independently and in a randomly-assigned order. The degrees of agreement and disagreement among evaluators in identifying implant type while reviewing both blinded and unblinded radiographs were assessed using the Bang Blinding Index and James Blinding Index. The degree of agreement in fracture union was assessed using kappa statistics. RESULTS: The assessment of blinded radiographs with both the Bang Blinding Index (BBI) and James Blinding Index (JBI) demonstrated a low degree of evaluator success at identifying implant type (Mean BBI, far cortical locking: -0.03, SD: 0.04; Mean BBI, standard screw: 0, SD: 0; JBI: 0.98, SD: 0), suggesting near perfect blinding. The assessment of unblinded radiographs with both blinding indices demonstrated a high degree of evaluator success at identifying implant type (Mean BBI, far cortical locking: 0.89, SD: 0.19; Mean BBI, standard screw: 0.87, SD: 0.04; JBI: 0.26, SD: 0.12), as expected. There was moderate agreement with regard to assessment of fracture union among the evaluators in both the blinded (Kappa: 0.38, 95%CI: 0.25-0.52) and unblinded (Kappa: 0.35, 95%CI: 0.25-0.45) arms of the study. There was no statistically significant difference in fracture union agreement between the blinded and unblinded groups. CONCLUSION: The digital blinding technique successfully masked the surgeons to the type of implant used for surgical treatment of distal femur fractures but did not interfere with the surgeons' ability to reliably evaluate radiographic healing at the fracture site.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame distilled prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.
Codex and Gemma teacher scores by category
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.001 | 0.001 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.001 | 0.000 |
| Bibliometrics | 0.001 | 0.000 |
| Science and technology studies | 0.000 | 0.000 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.000 | 0.000 |
| Research integrity | 0.000 | 0.000 |
| Insufficient payload (model declined to judge) | 0.000 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one teacher head, not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".