MétaCan
Menu
Back to cohort
Record W2312469707 · doi:10.1097/dcr.0000000000000273

Have We Progressed in the Surgical Literature? Thirty-Year Trends in Clinical Studies in 3 Surgical Journals

2014· article· en· W2312469707 on OpenAlexaff
R.R. Shawhan, Quinton Hatch, Jason Bingham, Daniel W. Nelson, Emile B. Fitzpatrick, Robin S. McLeod, Eric K. Johnson, Justin A. Maykel, Scott R. Steele

Bibliographic record

VenueDiseases of the Colon & Rectum · 2014
Typearticle
Languageen
FieldDecision Sciences
TopicMeta-analysis and systematic reviews
Canadian institutionsUniversity of TorontoMount Sinai Hospital
Fundersnot available
KeywordsMedicineColorectal surgeryClinical trialSample size determinationPopulationMEDLINEEvidence-based medicineSurgeryGeneral surgeryInternal medicineAlternative medicineAbdominal surgeryPathology

Abstract

fetched live from OpenAlex

BACKGROUND: We practice in an era of evidence-based medicine. In 1993, Solomon and McLeod published an article examining study designs in 3 surgical journals from 1980 and 1990. OBJECTIVE: The purpose of this study was to evaluate subsequent 30-year trends in the quality of selected literature. DESIGN: All of the articles from Diseases of the Colon & Rectum, Surgery, and the British Journal of Surgery during 2000 and 2010 were classified by study design. Nonclinical studies were substratified by animal/laboratory, surgical technique, editorial/review, or miscellaneous articles. Clinical articles were categorized as case or comparative studies, further categorized by study design, and rated on a 10-point scale to determine strength. We compared interobserver reliability using a random sample. SETTING: This study was conducted at 3 North American medical centers. PATIENTS: Patients described in the scope of the literature were included in this study. MAIN OUTCOME MEASURES: Frequency, type, and strength of study design were measured. RESULTS: We evaluated 1911 articles (967 clinical; 17% comparative). There was a significant increase in multicenter clinical studies (from 12% to 27%; p < 0.0001) and mean study population (from 326 to 6775; p < 0.05). Studies using administrative data increased from 14% to 43% (p < 0.0001). Case reports decreased from 16% to 7% of all clinical studies (p < 0.001), whereas the percentage of comparative studies increased from 14% to 21% (p = 0.001). The percentage of randomized controlled trials did not increase significantly (8.5% in 2000; 10.0% in 2010; p = 0.44). The mean 10-point score for comparative studies was 6.7 for both years (p = 0.50). There was good interobserver agreement in the classification of studies (κ = 0.70) and moderate agreement in scoring comparative studies (κ = 0.47). LIMITATIONS: This descriptive study cannot fully account for the reasons behind the identified differences. CONCLUSIONS: Comparative and multicenter studies, mean study population, and the use of administrative data increased from 2000 to 2010. This suggests that increased use of administrative databases has allowed larger populations of patients from more institutions to be studied and may be more generalizable. Researchers should strive toward improving the level of evidence (see Video, Supplemental Digital Content 1, http://links.lww.com/DCR/A167).

Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.

How this classification was reachedexpand

Full frame machine prediction

Teacher imitation

Not calibrated prevalence, not ground truth. Human validation pending. The Gemma side is a direct model label for every work in the frame, read from the title-only record. The Codex side is a classifier learned from the 10,348 direct Codex labels and calibrated to design-weighted sample rates; fields without enough sample support carry no Codex call. Candidate is the union of the two sides; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels.

metaresearch head score (Codex)0.088
metaresearch head score (Gemma)0.265
Version: metacan-v3-hybrid-931329e0061cValidation status: machine_predicted_unvalidated
Candidate categoriesMetaresearch, Bibliometrics
Consensus categoriesnone
DomainCandidate signal: Reporting · Consensus signal: none
Study designCandidate signal: Observational · Consensus signal: Observational
GenreCandidate signal: Empirical · Consensus signal: Empirical
Teacher disagreement score0.967
Threshold uncertainty score0.464

Distilled classifier scores by category (both heads)

CategoryCodexGemma
Metaresearch0.0880.265
Meta-epidemiology (narrow)0.0000.001
Meta-epidemiology (broad)0.0010.002
Bibliometrics0.0330.029
Science and technology studies0.0020.003
Scholarly communication0.0060.006
Open science0.0010.004
Research integrity0.0020.001
Insufficient payload (model declined to judge)0.0010.000

Machine scores (provisional)

The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.

Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.

Opus teacher head0.673
GPT teacher head0.594
Teacher spread0.080 · how far apart the two teachers sit on this one work
Validation statusscore_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from it

Classification

machine, unvalidated

Machine predicted; a candidate call from one source (direct Gemma or distilled Codex), not a consensus.

Study designObservational
DomainReporting
GenreEmpirical

How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".

Quick stats

Citations8
Published2014
Admission routes1
Has abstractyes

Explore more

Same venueDiseases of the Colon & RectumSame topicMeta-analysis and systematic reviewsFrench-language works237,207