MétaCan
Menu
Back to cohort
Record W4310919995 · doi:10.1136/bmjopen-2022-067656

Reporting of and explanations for under-recruitment and over-recruitment in pragmatic trials: a secondary analysis of a database of primary trial reports published from 2014 to 2019

2022· article· en· W4310919995 on OpenAlexafffund
Pascale Nevins, Stuart G. Nicholls, Yongdong Ouyang, Kelly Carroll, Karla Hemming, Charles Weijer, Monica Taljaard

Bibliographic record

VenueBMJ Open · 2022
Typearticle
Languageen
FieldMedicine
TopicEthics in Clinical Research
Canadian institutionsWestern UniversityOttawa HospitalUniversity of Ottawa
FundersCanadian Institutes of Health ResearchNational Institute on AgingNational Institutes of Health
KeywordsMedicineSample size determinationClinical trialPatient recruitmentPsychological interventionLogistic regressionMultinomial logistic regressionCluster (spacecraft)Family medicineDemographyInternal medicineNursing

Abstract

fetched live from OpenAlex

OBJECTIVES: To describe the extent to which pragmatic trials underachieved or overachieved their target sample sizes, examine explanations and identify characteristics associated with under-recruitment and over-recruitment. STUDY DESIGN AND SETTING: Secondary analysis of an existing database of primary trial reports published during 2014-2019, registered in ClinicalTrials.gov, self-labelled as pragmatic and with target and achieved sample sizes available. RESULTS: Of 372 eligible trials, the prevalence of under-recruitment (achieving <90% of target sample size) was 71 (19.1%) and of over-recruitment (>110% of target) was 87 (23.4%). Under-recruiting trials commonly acknowledged that they did not achieve their targets (51, 71.8%), with the majority providing an explanation, but only 11 (12.6%) over-recruiting trials acknowledged recruitment excess. The prevalence of under-recruitment in individually randomised versus cluster randomised trials was 41 (17.0%) and 30 (22.9%), respectively; prevalence of over-recruitment was 39 (16.2%) vs 48 (36.7%), respectively. Overall, 101 025 participants were recruited to trials that did not achieve at least 90% of their target sample size. When considering trials with over-recruitment, the total number of participants recruited in excess of the target was a median (Q1-Q3) 319 (75-1478) per trial for an overall total of 555 309 more participants than targeted. In multinomial logistic regression, cluster randomisation and lower journal impact factor were significantly associated with both under-recruitment and over-recruitment, while using exclusively routinely collected data and educational/behavioural interventions were significantly associated with over-recruitment; we were unable to detect significant associations with obtaining consent, publication year, country of recruitment or public engagement. CONCLUSIONS: A clear explanation for under-recruitment or over-recruitment in pragmatic trials should be provided to encourage transparency in research, and to inform recruitment to future trials with comparable designs. The issues and ethical implications of over-recruitment should be more widely recognised by trialists, particularly when designing cluster randomised trials.

Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.

How this classification was reachedexpand

Full frame machine prediction

Teacher imitation

Not calibrated prevalence, not ground truth. Human validation pending. The Gemma side is a direct model label for every work in the frame, read from the title-only record. The Codex side is a classifier learned from the 10,348 direct Codex labels and calibrated to design-weighted sample rates; fields without enough sample support carry no Codex call. Candidate is the union of the two sides; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels.

metaresearch head score (Codex)0.490
metaresearch head score (Gemma)0.810
Version: metacan-v3-hybrid-931329e0061cValidation status: machine_predicted_unvalidated
Candidate categoriesMetaresearch
Consensus categoriesMetaresearch
DomainCandidate signal: Reporting · Consensus signal: none
Study designCandidate signal: Observational · Consensus signal: none
GenreCandidate signal: Empirical · Consensus signal: Empirical
Teacher disagreement score0.510
Threshold uncertainty score0.629

Distilled classifier scores by category (both heads)

CategoryCodexGemma
Metaresearch0.4900.810
Meta-epidemiology (narrow)0.0020.002
Meta-epidemiology (broad)0.0050.007
Bibliometrics0.0210.022
Science and technology studies0.0020.004
Scholarly communication0.0050.007
Open science0.0030.008
Research integrity0.0030.002
Insufficient payload (model declined to judge)0.0040.001

Machine scores (provisional)

The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.

Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.

Opus teacher head0.758
GPT teacher head0.642
Teacher spread0.116 · how far apart the two teachers sit on this one work
Validation statusscore_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from it

Classification

machine, unvalidated

Machine predicted; the direct Gemma label and the distilled Codex classifier agree on what is shown here.

Study designObservational
DomainReporting
GenreEmpirical

How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".

Quick stats

Citations5
Published2022
Admission routes2
Has abstractyes

Explore more

Same venueBMJ OpenSame topicEthics in Clinical ResearchFrench-language works237,207