Accuracy of Cardiovascular Trial Outcome Ascertainment and Treatment Effect Estimates from Routine Health Data: A Systematic Review and Meta-Analysis
Bibliographic record
Abstract
BACKGROUND: Registry-based randomized controlled trials allow for outcome ascertainment using routine health data (RHD). While this method provides a potential solution to the rising cost and complexity of clinical trials, comparative analyses of outcome ascertainment by clinical end point committee (CEC) adjudication compared with RHD sources are sparse. Among cardiovascular trials, we set out to systematically compare the incidence of cardiovascular events and estimated randomized treatment effects ascertained from RHD versus traditional clinical evaluation and adjudication. METHODS: We searched MEDLINE (1976 to August 2020) for studies where outcome ascertainment was performed by both RHD and CEC adjudication to compare the incidence of cardiovascular events and treatment effects. We derived ratios of hazard ratios to compare treatment effects from RHD and CEC adjudication. We pooled ratios of hazard ratios using an inverse variance random-effects meta-analysis. RESULTS: Nine studies (1988-2020; 32 156 patients) involving 10 randomized control trials compared outcome ascertainment with RHD and CEC in patients with or at risk of cardiovascular disease. There was a high degree of agreement and interrater reliability between CEC and RHD outcome determination for all-cause mortality (agreement percentage: 98.4%-100% and κ: 0.95-1.0) and cardiovascular mortality (agreement percentage: 97.8%-99.9% and κ: 0.66-0.99). For myocardial infarction, the κ values ranged from 0.67-0.98, and for stroke the values ranged from 0.52-0.89. In contrast, the κ value for peripheral artery disease was low (κ: 0.27). There was little difference in the randomized treatment effect derived from CEC and RHD ascertainment of events based on the ratios of hazard ratio, with pooled ratios of hazard ratios ranging from 0.93 (95% CI, 0.63-1.39) for cardiovascular mortality to 1.27 (95% CI, 0.67-2.41) for stroke. CONCLUSIONS: Clinical outcome ascertainment using retrospectively acquired RHD displayed high levels of agreement with CEC adjudication for identifying all-cause mortality and cardiovascular outcomes. Importantly, cardiovascular treatment effects in randomized control trials determined from RHD and CEC resulted in similar point estimates. Overall, our review supports the use of RHD as a potential alternative source for clinical outcome ascertainment in cardiovascular trials. Validation studies with prospectively planned linkage are warranted.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame distilled prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.
Codex and Gemma teacher scores by category
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.011 | 0.007 |
| Meta-epidemiology (narrow) | 0.001 | 0.001 |
| Meta-epidemiology (broad) | 0.037 | 0.013 |
| Bibliometrics | 0.000 | 0.001 |
| Science and technology studies | 0.000 | 0.000 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.000 | 0.000 |
| Research integrity | 0.000 | 0.000 |
| Insufficient payload (model declined to judge) | 0.000 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; both teacher heads agree on what is shown here.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".