MétaCan
Menu
Back to cohort
Record W2060971566 · doi:10.1186/1471-2288-13-116

Inconsistency in the items included in tools used in general health research and physical therapy to evaluate the methodological quality of randomized controlled trials: a descriptive analysis

2013· review· en· W2060971566 on OpenAlexafffund
Susan Armijo‐Olivo, Jorge Fuentes, Maria B. Ospina, Humam Saltaji, Lisa Hartling

Bibliographic record

VenueBMC Medical Research Methodology · 2013
Typereview
Languageen
FieldDecision Sciences
TopicMeta-analysis and systematic reviews
Canadian institutionsInstitute of Health EconomicsUniversity of Alberta
FundersAlberta InnovatesKillam TrustsUniversity of AlbertaCanadian Institutes of Health ResearchPhysiotherapy Foundation of CanadaAlberta Innovates - Health SolutionsWomen and Children's Health Research InstituteChildren's Health Research Institute
KeywordsBlindingComparabilityRandomized controlled trialQuality (philosophy)MedicineDescriptive statisticsResearch designInclusion and exclusion criteriaMedical physicsStatisticsAlternative medicineMathematicsSurgery

Abstract

fetched live from OpenAlex

BACKGROUND: Assessing the risk of bias of randomized controlled trials (RCTs) is crucial to understand how biases affect treatment effect estimates. A number of tools have been developed to evaluate risk of bias of RCTs; however, it is unknown how these tools compare to each other in the items included. The main objective of this study was to describe which individual items are included in RCT quality tools used in general health and physical therapy (PT) research, and how these items compare to those of the Cochrane Risk of Bias (RoB) tool. METHODS: We used comprehensive literature searches and a systematic approach to identify tools that evaluated the methodological quality or risk of bias of RCTs in general health and PT research. We extracted individual items from all quality tools. We calculated the frequency of quality items used across tools and compared them to those in the RoB tool. Comparisons were made between general health and PT quality tools using Chi-squared tests. RESULTS: In addition to the RoB tool, 26 quality tools were identified, with 19 being used in general health and seven in PT research. The total number of quality items included in general health research tools was 130, compared with 48 items across PT tools and seven items in the RoB tool. The most frequently included items in general health research tools (14/19, 74%) were inclusion and exclusion criteria, and appropriate statistical analysis. In contrast, the most frequent items included in PT tools (86%, 6/7) were: baseline comparability, blinding of investigator/assessor, and use of intention-to-treat analysis. Key items of the RoB tool (sequence generation and allocation concealment) were included in 71% (5/7) of PT tools, and 63% (12/19) and 37% (7/19) of general health research tools, respectively. CONCLUSIONS: There is extensive item variation across tools that evaluate the risk of bias of RCTs in health research. Results call for an in-depth analysis of items that should be used to assess risk of bias of RCTs. Further empirical evidence on the use of individual items and the psychometric properties of risk of bias tools is needed.

Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.

How this classification was reachedexpand

Full frame distilled prediction

Teacher imitation

Not calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.

metaresearch head score (Codex)0.991
metaresearch head score (Gemma)0.986
Version: codex-gemma-dda1882f352aValidation status: machine_predicted_unvalidated
Candidate categoriesMetaresearch, Meta-epidemiology (narrow), Meta-epidemiology (broad), Bibliometrics, Science and technology studies, Open science, Research integrity, Insufficient payload (model declined to judge)
Consensus categoriesMetaresearch, Meta-epidemiology (broad)
DomainCandidate signal: Methods · Consensus signal: Methods
Study designCandidate signal: Other design · Consensus signal: none
GenreCandidate signal: Review · Consensus signal: Review
Teacher disagreement score0.935
Threshold uncertainty score1.000

Codex and Gemma teacher scores by category

CategoryCodexGemma
Metaresearch0.9910.986
Meta-epidemiology (narrow)0.0010.000
Meta-epidemiology (broad)0.1130.013
Bibliometrics0.0090.024
Science and technology studies0.0000.004
Scholarly communication0.0010.000
Open science0.0090.002
Research integrity0.0010.006
Insufficient payload (model declined to judge)0.0030.000

Machine scores (provisional)

The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.

Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.

Opus teacher head0.996
GPT teacher head0.820
Teacher spread0.176 · how far apart the two teachers sit on this one work
Validation statusscore_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from it

Classification

machine, unvalidated

Machine predicted; both teacher heads agree on what is shown here.

Study designOther design
DomainMethods
GenreReview

How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".

Quick stats

Citations58
Published2013
Admission routes2
Has abstractyes

Explore more

Same venueBMC Medical Research MethodologySame topicMeta-analysis and systematic reviewsFrench-language works237,207