Development of Multivariable Prediction Models for the Identification of Patients Admitted to Hospital with an Exacerbation of COPD and the Prediction of Risk of Readmission: A Retrospective Cohort Study using Electronic Medical Record Data
Bibliographic record
Abstract
BACKGROUND: Approximately 20% of patients who are discharged from hospital for an acute exacerbation of COPD (AECOPD) are readmitted within 30 days. To reduce this, it is important both to identify all individuals admitted with AECOPD and to predict those who are at higher risk for readmission. OBJECTIVES: To develop two clinical prediction models using data available in electronic medical records: 1) identifying patients admitted with AECOPD and 2) predicting 30-day readmission in patients discharged after AECOPD. METHODS: Two datasets were created using all admissions to General Internal Medicine from 2012 to 2018 at two hospitals: one cohort to identify AECOPD and a second cohort to predict 30-day readmissions. We fit and internally validated models with four algorithms. RESULTS: Of the 64,609 admissions, 3,620 (5.6%) were diagnosed with an AECOPD. Of those discharged, 518 (15.4%) had a readmission to hospital within 30 days. For identification of patients with a diagnosis of an AECOPD, the top-performing models were LASSO and a four-variable regression model that consisted of specific medications ordered within the first 72 hours of admission. For 30-day readmission prediction, a two-variable regression model was the top performing model consisting of number of COPD admissions in the previous year and the number of non-COPD admissions in the previous year. CONCLUSION: We generated clinical prediction models to identify AECOPDs during hospitalization and to predict 30-day readmissions after an acute exacerbation from a dataset derived from available EMR data. Further work is needed to improve and externally validate these models.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame distilled prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.
Codex and Gemma teacher scores by category
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.002 | 0.001 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.000 | 0.000 |
| Bibliometrics | 0.000 | 0.000 |
| Science and technology studies | 0.000 | 0.000 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.000 | 0.000 |
| Research integrity | 0.000 | 0.000 |
| Insufficient payload (model declined to judge) | 0.000 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one teacher head, not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".