Stepped-wedge trials should be classified as research for the purpose of ethical review
Bibliographic record
Abstract
BACKGROUND: All studies classified as research involving human participants require research ethics review. Most regulation and guidance on ethical oversight of research involving human participants was written for pharmacotherapy interventions. Interpretation of such guidance for cluster-randomized trials and stepped-wedge trials, which commonly evaluate complex non-therapeutic interventions such as knowledge translation, public health, or health service delivery interventions, can pose challenges to researchers and regulators. CURRENT GUIDANCE: provides guidance on the ethical oversight and consent procedures for cluster-randomized trials, and while not explicit, this includes stepped-wedge trials. Yet, stepped-wedge trials have unique characteristics that differentiate them from standard cluster-randomized trials. In particular, they can be used to evaluate knowledge translation interventions within the context of a routine health system rollout; they may have a non-randomized design; and the decision to implement the intervention is not always made by the researcher. Many stepped-wedge trials do not undergo ethical review and do not report trial registration. This suggests that those undertaking these studies and research ethics committees perceive them as non-research activities. RECOMMENDATIONS: Through an ethical analysis of two case studies, we argue that stepped-wedge trials, like parallel arm cluster trials, are systematic investigations designed to produce generalizable knowledge. We contend that stepped-wedge trials usually include human research participants, which may be patients, health care providers, or both. Stepped-wedge trials are therefore research involving human participants for the purpose of ethical review. Nevertheless, the use of a waiver or alteration of consent may be appropriate in many stepped-wedge trials due to the infeasibility of obtaining informed consent and the low-risk nature of the interventions. To ensure that traditional ethical principles such as respect for persons are upheld, these studies must undergo research ethics review.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame distilled prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.
Codex and Gemma teacher scores by category
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.738 | 0.959 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.007 | 0.004 |
| Bibliometrics | 0.000 | 0.001 |
| Science and technology studies | 0.000 | 0.002 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.002 | 0.001 |
| Research integrity | 0.004 | 0.015 |
| Insufficient payload (model declined to judge) | 0.004 | 0.001 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; both teacher heads agree on what is shown here.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".