Variation in length of stay within and between hospitals
Bibliographic record
Abstract
Background and objective: Variation in the delivery of health care services and the lack of association between greater utilization and higher quality care signal inefficient, low value care. The extent to which patient and hospital variables can explain variation in hospital length of stay is unclear. Methods: We examined hospital inpatient length of stay using data from 684 hospitals and 5.4 million discharges in the 2007 Healthcare Cost and Utilization Project’s Nationwide Inpatient Sample. We used a mixed effects model with a random effect for hospitals to quantify variation in length of stay due to differences within and between hospitals. Results: The interquartile range of hospital mean LOS was 3.4 days (3.3-6.7). Fifty-nine percent of the overall variation in length of stay remained unexplained after adjustment for discharge-level disease status, illness-severity, regional poverty, hospital-level contextual factors (e.g. proportion of patients from low-income ZIP-codes, proportion uninsured), and structural variables (e.g. teaching status, urban or rural location). Seventy-seven percent of the explainable variation was due to differences between hospitals. Conclusion: These findings indicate that wide variability in length of stay persists after adjustment for patient and hospital variables, signaling an opportunity for improved productivity and efficiency in the delivery of health care.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame distilled prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.
Codex and Gemma teacher scores by category
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.001 | 0.000 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.000 | 0.000 |
| Bibliometrics | 0.000 | 0.000 |
| Science and technology studies | 0.000 | 0.000 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.000 | 0.000 |
| Research integrity | 0.000 | 0.000 |
| Insufficient payload (model declined to judge) | 0.000 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one teacher head, not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".