MétaCan
Menu
Back to cohort
Record W2316013226 · doi:10.1097/ede.0b013e3181f56fc0

Accounting for Center Effects in Multicenter Trials

2010· letter· en· W2316013226 on OpenAlexafffund
Navdeep Tangri, Georgios D. Kitsios, Shi Su, David M. Kent

Bibliographic record

VenueEpidemiology · 2010
Typeletter
Languageen
FieldMathematics
TopicAdvanced Causal Inference Techniques
Canadian institutionsBC Studies
FundersCanadian Institutes of Health Research
KeywordsConfoundingMedicineRandomized controlled trialPopulationMEDLINEObservational studySample size determinationClinical trialFamily medicineInternal medicineStatisticsEnvironmental health

Abstract

fetched live from OpenAlex

To the Editors: Individual patients in multicenter trials may not represent truly independent observations. Center characteristics and practice patterns, particularly in cases with substantial between-center variability in outcome rates, can lead to potentially misleading conclusions, if ignored. More specifically, failure to consider the center can lead to incorrect standard errors and P values (due to clustering), biased estimates (from uncontrolled confounding), and unrecognized heterogeneity across centers (from effect modification).1–3 Although statistical methods for dealing with center-level clustering in the analysis of randomized controlled trials (RCTs) have been well-described and advocated,1–5 there is some evidence that the application of these methods is limited.6,7 We believe that accounting for center effects is of crucial importance in RCTs of medicinal products, particularly in large multicenter studies.2,3 Because the extent of center-effect adjustment has not previously been described for medicinal products, we performed an empirical evaluation of adjustment for center-level clustering in reports of multicenter RCTs in 4 major medical journals. A systematic search for RCTs published during the year 2007 in 4 prominent medical journals (British Medical Journal, Journal of the American Medical Association, Lancet, and New England Journal of Medicine) was conducted in PubMed. We evaluated the retrieved articles for the following inclusion criteria: adult human study population enrolled, use of randomized design, multicenter enrollment, and testing for efficacy/effectiveness of medicinal products. The main characteristics of the 101 included RCTs are shown in the Table. The majority of the trials used a superiority study design; cardiovascular and oncology disorders were the most commonly studied conditions (32% and 25%, respectively); and binary and time-to-event outcomes were examined in almost equal proportions.TABLE: Characteristics of RCTs Included in the Analysis (n = 101)The number of centers included in the RCTs ranged from 2–707 (median = 64, interquartile range = 22–117), representing 70–22,949 patients. Of the 101 studies, 36 (36%) performed random allocation stratified by center. Statistical analysis adjusting for the clustering of patients by center was present in 18% of the reports. Of these, only 1 used a random term for the center-effect, whereas the remaining used a fixed-effects model. Thus, a total of 82% did not adjust for center effects. Previous investigators have studied RCTs of surgical interventions; they reported similar results for allocation stratification on center (38%) and adjusted statistical analysis (6%).6–8 Our literature sample is more contemporary and focused on 4 major medical journals publishing RCTs of medical interventions. Studies published in these journals are expected to be of high quality and have the potential to disproportionately influence clinical practice. Nonetheless, the similarities between our results and those of previous investigators highlight the lack of accounting for center effects in RCTs as a widespread problem. Our analysis has some limitations. Our results may not be generalizable to the entire medical literature. (The proportion of RCTs accounting for center effects is likely to be even lower among the less prestigious journals.) Second, our conclusions regarding center effect accounting were based on the published statistical methods. It is possible that appropriate accounting was performed, but not reported due to space constraints. In summary, using a contemporary sample of RCTs from 4 major medical journals, we find that center effects are not accounted for in the recruitment stage or in the statistical analysis of the majority of RCTs evaluating medical interventions. The recent extension of the CONSORT statement advocates center-effect reporting and adjustment for trials, involving nonpharmacologic interventions. Our analysis highlights the need for similar recommendations in trials of medical therapies. Navdeep Tangri Georgios D. Kitsios Shi Hann Su David M. Kent Tufts Clinical and Translational Science Institute Institute for Clinical Research and Health Policy Studies Tufts Medical Center Boston, MA [email protected]

Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.

How this classification was reachedexpand

Full frame distilled prediction

Teacher imitation

Not calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.

metaresearch head score (Codex)0.011
metaresearch head score (Gemma)0.171
Version: codex-gemma-dda1882f352aValidation status: machine_predicted_unvalidated
Candidate categoriesMetaresearch, Meta-epidemiology (narrow), Research integrity
Consensus categoriesResearch integrity
DomainCandidate signal: none · Consensus signal: none
Study designCandidate signal: Not applicable · Consensus signal: none
GenreCandidate signal: Commentary · Consensus signal: Commentary
Teacher disagreement score0.504
Threshold uncertainty score1.000

Codex and Gemma teacher scores by category

CategoryCodexGemma
Metaresearch0.0110.171
Meta-epidemiology (narrow)0.0010.000
Meta-epidemiology (broad)0.0030.001
Bibliometrics0.0000.000
Science and technology studies0.0000.000
Scholarly communication0.0000.000
Open science0.0010.000
Research integrity0.0040.005
Insufficient payload (model declined to judge)0.0000.000

Machine scores (provisional)

The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.

Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.

Opus teacher head0.368
GPT teacher head0.517
Teacher spread0.149 · how far apart the two teachers sit on this one work
Validation statusscore_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from it

Classification

machine, unvalidated

Machine predicted; both teacher heads agree on what is shown here.

Study designNot applicable
Domainnot available
GenreCommentary

How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".

Quick stats

Citations10
Published2010
Admission routes2
Has abstractyes

Explore more

Same venueEpidemiologySame topicAdvanced Causal Inference TechniquesFrench-language works237,207