Outcomes reported in high-impact surgical journals
Bibliographic record
Abstract
BACKGROUND: With advances in operative technique and perioperative care, traditional endpoints such as morbidity and mortality provide an incomplete description of surgical outcomes. There is increasing emphasis on the need for patient-reported outcomes (PROs) to evaluate fully the effectiveness and quality of surgical interventions. The objective of this study was to identify the outcomes reported in clinical studies published in high-impact surgical journals and the frequency with which PROs are used. METHODS: Electronic versions of material published between 2008 and 2012 in the four highest-impact non-subspecialty surgical journals (Annals of Surgery, British Journal of Surgery (BJS), Journal of the American College of Surgeons (JACS), Journal of the American Medical Association (JAMA) Surgery) were hand-searched. Clinical studies of adult patients undergoing planned abdominal, thoracic or vascular surgery were included. Reported outcomes were classified into five categories using Wilson and Cleary's conceptual model. RESULTS: A total of 893 articles were assessed, of which 770 were included in the analysis. Some 91·6 per cent of studies reported biological and physiological outcomes, 36·0 per cent symptoms, 13·4 per cent direct indicators of functional status, 10·6 per cent general health perception and 14·8 per cent overall quality of life (QoL). The proportion of studies with at least one PRO was 38·7 per cent overall and 73·4 per cent in BJS (P < 0·001). The proportion of studies using a formal measure of health-related QoL ranged from 8·9 per cent (JAMA Surgery) to 33·8 per cent (BJS). CONCLUSION: The predominant reporting of clinical endpoints and the inconsistent use of PROs underscore the need for further research and education to enhance the applicability of these measures in specific surgical settings.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame distilled prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.
Codex and Gemma teacher scores by category
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.315 | 0.151 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.007 | 0.005 |
| Bibliometrics | 0.001 | 0.002 |
| Science and technology studies | 0.000 | 0.000 |
| Scholarly communication | 0.002 | 0.001 |
| Open science | 0.001 | 0.000 |
| Research integrity | 0.000 | 0.000 |
| Insufficient payload (model declined to judge) | 0.015 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; both teacher heads agree on what is shown here.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".