Predictors of the accuracy of quotation of references in peer-reviewed orthopaedic literature in relation to publications on the scaphoid
Bibliographic record
Abstract
Using inaccurate quotations can propagate misleading information, which might affect the management of patients. The aim of this study was to determine the predictors of quotation inaccuracy in the peer-reviewed orthopaedic literature related to the scaphoid. We randomly selected 100 papers from ten orthopaedic journals. All references were retrieved in full text when available or otherwise excluded. Two observers independently rated all quotations from the selected papers by comparing the claims made by the authors with the data and expressed opinions of the reference source. A statistical analysis determined which article-related factors were predictors of quotation inaccuracy. The mean total inaccuracy rate of the 3840 verified quotes was 7.6%. There was no correlation between the rate of inaccuracy and the impact factor of the journal. Multivariable analysis identified the journal and the type of study (clinical, biomechanical, methodological, case report or review) as important predictors of the total quotation inaccuracy rate. We concluded that inaccurate quotations in the peer-reviewed orthopaedic literature related to the scaphoid were common and slightly more so for certain journals and certain study types. Authors, reviewers and editorial staff play an important role in reducing this inaccuracy.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame distilled prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.
Codex and Gemma teacher scores by category
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.005 | 0.010 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.001 | 0.000 |
| Bibliometrics | 0.001 | 0.001 |
| Science and technology studies | 0.000 | 0.000 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.000 | 0.000 |
| Research integrity | 0.000 | 0.000 |
| Insufficient payload (model declined to judge) | 0.000 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one teacher head, not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".