Level of evidence of free papers presented at the European Society of Sports Traumatology, Knee Surgery and Arthroscopy congress from 2008 to 2016
Bibliographic record
Abstract
PURPOSE: The European Society of Sports Traumatology, Knee Surgery and Arthroscopy (ESSKA) congress is an important venue, and the research presented can be a critical source of information used to impact clinical decisions and health policies. The purpose of this study was to evaluate the level of evidence of clinical free papers presented at the ESSKA congress from 2008 to 2016. Moreover, this study evaluated whether there were any changes in the distribution of level of evidence over time. METHODS: Two reviewers screened the free papers presented at the ESSKA biannual congresses 2008-2016 for clinical evidence. Clinical papers included observational studies and trials involving direct interaction between an investigator and human subjects. Biomechanical studies, technique demonstrations, cadaveric studies, and panel discussions were excluded. The reviewers independently graded their level of evidence from level I (e.g. high-quality randomized trials) to level IV (e.g. case series and reports) using the classification system published by the American Academy of Orthopaedic Surgeons. RESULTS: Of 1036 free papers that were identified, 729 met the inclusion criteria and were evaluated. Overall, 18% of studies were level I, 24% level II, 25% level III, and 33% level IV evidence. There was a significant improvement in level of evidence over time (p < 0.0001), with the proportion of level I studies increasing most dramatically (9% in 2008, 20% in 2012, 24% in 2016). Free papers studying the knee had higher levels of evidence than those evaluating other joints (p = 0.002). CONCLUSION: The level of evidence of clinical free papers presented at the ESSKA congress between 2008 and 2016 is high relative to other orthopaedic meetings. Moreover, there has been a significant improvement in the level of evidence over time. LEVEL OF EVIDENCE: Systematic review, Level IV.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame machine prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. The Gemma side is a direct model label for every work in the frame, read from the title-only record. The Codex side is a classifier learned from the 10,348 direct Codex labels and calibrated to design-weighted sample rates; fields without enough sample support carry no Codex call. Candidate is the union of the two sides; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels.
Distilled classifier scores by category (both heads)
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.067 | 0.340 |
| Meta-epidemiology (narrow) | 0.003 | 0.002 |
| Meta-epidemiology (broad) | 0.010 | 0.009 |
| Bibliometrics | 0.021 | 0.019 |
| Science and technology studies | 0.003 | 0.004 |
| Scholarly communication | 0.016 | 0.007 |
| Open science | 0.005 | 0.005 |
| Research integrity | 0.010 | 0.005 |
| Insufficient payload (model declined to judge) | 0.048 | 0.009 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one source (direct Gemma or distilled Codex), not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".