Quality and completeness of, and spin in reporting of, pilot and feasibility studies in hip and knee arthroplasty: a protocol for a methodological survey
Bibliographic record
Abstract
INTRODUCTION: Pilot or feasibility trials examine the feasibility, viability and recruitment potential of larger, main trials. Specifically, a pilot trial can be instrumental in identifying methodological issues essential to the development of an effective research protocol. However, numerous studies published as pilot or feasibility studies have demonstrated notable inconsistencies in the nature of information reported, resulting in poor-quality and incomplete reporting. It is unclear whether such low quality or incompleteness of reporting is also prevalent in arthroplasty pilot trials. METHODS AND ANALYSIS: This protocol outlines a methodological survey examining the completeness of reporting among hip and knee arthroplasty pilot trials in accordance with the Consolidated Standards of Reporting Trials (CONSORT) 2010 extension to pilot trials. Secondary objectives include: (1) determining the prevalence of 'spin' practices, defined as: (a) placing a focus on statistical significance rather than feasibility, (b) presenting results that show the trial to be non-feasible as feasible or (c) emphasising the effectiveness or potential intervention benefits rather than feasibility; (2) determining factors associated with incomplete reporting, and 'spin'. A search of PubMed will be conducted for pilot trials in hip or knee arthroplasty published between 01 January 2017 and 31 December 2023. Following screening, appropriate data will be extracted from eligible publications and reported as descriptive statistics, encompassing elements of the CONSORT checklist associated with completeness of reporting. Logistic regression analysis and Poisson regression will be used to analyse factors associated with completeness of reporting and spin. ETHICS AND DISSEMINATION: This methodological review does not require formal ethical approval, as it will solely involve the use of published and publicly reported literature. The results of this study will be disseminated through submission to peer-reviewed journals and academic conference presentations. Study details will be sent to McMaster University's media coordinators to be shared through the institution's research-focused platforms.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame machine prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. The Gemma side is a direct model label for every work in the frame, read from the title-only record. The Codex side is a classifier learned from the 10,348 direct Codex labels and calibrated to design-weighted sample rates; fields without enough sample support carry no Codex call. Candidate is the union of the two sides; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels.
Distilled classifier scores by category (both heads)
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.437 | 0.470 |
| Meta-epidemiology (narrow) | 0.004 | 0.005 |
| Meta-epidemiology (broad) | 0.006 | 0.012 |
| Bibliometrics | 0.016 | 0.013 |
| Science and technology studies | 0.005 | 0.008 |
| Scholarly communication | 0.007 | 0.009 |
| Open science | 0.004 | 0.007 |
| Research integrity | 0.010 | 0.010 |
| Insufficient payload (model declined to judge) | 0.031 | 0.015 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; the direct Gemma label and the distilled Codex classifier agree on what is shown here.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".