Physician experience and outcomes among patients admitted to general internal medicine teaching wards
Bibliographic record
Abstract
BACKGROUND: Physician scores on examinations decline with time after graduation. However, whether this translates into declining quality of care is unknown. Our objective was to determine how physician experience is associated with negative outcomes for patients admitted to hospital. METHODS: We conducted a retrospective cohort study involving all patients admitted to general internal medicine wards over a 2-year period at all 7 teaching hospitals in Alberta, Canada. We used files from the Alberta College of Physicians and Surgeons to determine the number of years since medical school graduation for each patient's most responsible physician. Our primary outcome was the composite of in-hospital death, or readmission or death within 30 days postdischarge. RESULTS: We identified 10 046 patients who were cared for by 149 physicians. Patient characteristics were similar across physician experience strata, as were primary outcome rates (17.4% for patients whose care was managed by physicians in the highest quartile of experience, compared with 18.8% in those receiving care from the least experienced physicians; adjusted odds ratio [OR] 0.88, 95% confidence interval [CI] 0.72-1.06). Outcomes were similar between experience quartiles when further stratified by physician volume, most responsible diagnosis or complexity of the patient's condition. Although we found substantial variability in length of stay between individual physicians, there were no significant differences between physician experience quartiles (mean adjusted for patient covariates and accounting for intraphysician clustering: 7.90 [95% CI 7.39-8.42] d for most experienced quartile; 7.63 [95% CI 7.13-8.14] d for least experienced quartile). INTERPRETATION: For patients admitted to general internal medicine teaching wards, we saw no negative association between physician experience and outcomes commonly used as proxies for quality of inpatient care.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame distilled prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.
Codex and Gemma teacher scores by category
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.002 | 0.012 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.000 | 0.000 |
| Bibliometrics | 0.000 | 0.000 |
| Science and technology studies | 0.001 | 0.000 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.000 | 0.000 |
| Research integrity | 0.000 | 0.002 |
| Insufficient payload (model declined to judge) | 0.001 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one teacher head, not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".