MétaCan
Menu
Back to cohort
Record W4231273027 · doi:10.1001/jamacardio.2020.0459

Error in Figure 3

2020· erratum· en· W4231273027 on OpenAlexafffund
Nariman Sepehrvand, Wendimagegn Alemayehu, Justin A. Ezekowitz

Bibliographic record

VenueJAMA Cardiology · 2020
Typeerratum
Languageen
FieldEconomics, Econometrics and Finance
TopicHealth Systems, Economic Evaluations, Quality of Life
Canadian institutionsCanadian VIGOUR CentreUniversity of AlbertaCanadiana.org
FundersNational Institutes of HealthAlberta InnovatesAmerican RegentSanofiZOLL Medical CorporationBristol-Myers SquibbAstraZenecaZOLL FoundationAmgenNational Heart, Lung, and Blood InstitutePfizerCanadian Institutes of Health ResearchUniversity of Washington
KeywordsMedicineMEDLINE

Abstract

fetched live from OpenAlex

In Reply We thank Fernandes et al for their interest in our study 1 and agree that this field requires further exploration.Explanatory trials are primed to maximize the likelihood of finding efficacy of an intervention by testing it in an ideal setting, whereas pragmatic trials aim to test effectiveness of an intervention in a more generalizable setting.Hence, they are expected to generate more generalizable results, with the risk understood that there may be more variation in less tightly controlled environments, which may result in differing results.In our study, 1 380 of 616 randomized clinical trials (61.7%) were positive for the primary end point, 56 (9.1%) were neutral for the primary end point but positive for at least one secondary end point, and 180 (29.2%) were neutral for both the primary and secondary end points.The proportion of trials with positive results was fairly stable over time, with 113 of 172 (65.7%), 104 of 168 (61.9%), 76 of 137 (55.5%), and 87 of 139 (62.6%) in 2000, 2005, 2010, and 2015, respectively.Compared with trials with neutral findings, randomized clinical trials that were positive for the primary end point had lower mean [SD] Pragmatic Explanatory Continuum Index Summary (PRECIS)-2 scores (3.17 [0.70] vs 3.42 [0.66]; P < .001);however, the Cohen d effect size of 0.36 denotes a small difference in the level of pragmatism between trials with positive and neutral findings. 1However, we would caution against the interpretation that the level of pragmatism is the root cause for the neutral results in these trials.Many other factors can play a role in the neutral findings, including the lack of an actual effect, the type of question being addressed, operational considerations, and patient or health system factors, among others.This is analogous to the considerations to trials using surrogate end points (eg, biomarkers), as trials with surrogate markers often yield positive results compared with trials that are focused on clinical end points. 2Pragmatic trials complement explanatory trials, as the intent is different, and we need to be comfortable that not all interventions work as hypothesized.

Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.

How this classification was reachedexpand

Full frame machine prediction

Teacher imitation

Not calibrated prevalence, not ground truth. Human validation pending. The Gemma side is a direct model label for every work in the frame, read from the title-only record. The Codex side is a classifier learned from the 10,348 direct Codex labels and calibrated to design-weighted sample rates; fields without enough sample support carry no Codex call. Candidate is the union of the two sides; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels.

metaresearch head score (Codex)0.006
metaresearch head score (Gemma)0.088
Version: metacan-v3-hybrid-931329e0061cValidation status: machine_predicted_unvalidated
Candidate categoriesInsufficient payload (model declined to judge)
Consensus categoriesnone
DomainCandidate signal: none · Consensus signal: none
Study designCandidate signal: Not applicable · Consensus signal: Not applicable
GenreCandidate signal: Other · Consensus signal: none
Teacher disagreement score0.308
Threshold uncertainty score0.987

Distilled classifier scores by category (both heads)

CategoryCodexGemma
Metaresearch0.0060.088
Meta-epidemiology (narrow)0.0010.001
Meta-epidemiology (broad)0.0010.002
Bibliometrics0.0030.003
Science and technology studies0.0030.002
Scholarly communication0.0050.003
Open science0.0040.003
Research integrity0.0090.009
Insufficient payload (model declined to judge)0.3080.235

Machine scores (provisional)

The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.

Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.

Opus teacher head0.311
GPT teacher head0.416
Teacher spread0.105 · how far apart the two teachers sit on this one work
Validation statusscore_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from it

Classification

machine, unvalidated

Machine predicted; a candidate call from one source (direct Gemma or distilled Codex), not a consensus.

Study designNot applicable
Domainnot available
GenreOther

How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".

Quick stats

Citations0
Published2020
Admission routes2
Has abstractyes

Explore more

Same venueJAMA CardiologySame topicHealth Systems, Economic Evaluations, Quality of LifeFrench-language works237,207