Beyond Traditional Inefficiency Measures: Quantifying Health System Waste Through a Hierarchical Model of Inherited and Self-Generated Persistent Inefficiency
Bibliographic record
Abstract
This study develops a hierarchical inefficiency model to quantify how persistent technical inefficiency is generated and transmitted through multi-tiered healthcare systems. Departing from conventional hospital-centric assessments, the model decomposes inefficiency into inherited (propagated from higher governance levels) and self-generated (locally produced) components across three administrative tiers: province, region, and hospital. By leveraging the nested structure of Canadian healthcare governance, the framework captures system-level inefficiencies embedded in institutional design rather than isolated provider-level performance. Applied to panel data from hospitals in Alberta, Nova Scotia, and Ontario, the analysis shows that persistent inefficiency consistently originates at the top of the hierarchy, within provincial governance, where it is highest: 7.25% in Alberta, 7.02% in Nova Scotia, and 6.97% in Ontario. It then compounds as it flows downward, yielding total system inefficiencies of 16.45%, 14.63%, and 14.88%, respectively. These results demonstrate that hospital-level inefficiencies often reflect upstream structural constraints rather than solely local mismanagement. While the decomposition provides a tractable diagnostic of how inefficiency propagates, it also raises a policy dilemma: Should reforms equip hospitals and regions to absorb inherited burdens, or should they target the persistent sources of inefficiency embedded in provincial governance? These findings challenge standard health economics approaches that localise inefficiency at the point of care. By reframing inefficiency as a cascading, structural phenomenon, this study offers a system-aligned perspective for identifying where meaningful reform must begin.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame distilled prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.
Codex and Gemma teacher scores by category
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.000 | 0.000 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.001 | 0.000 |
| Bibliometrics | 0.001 | 0.001 |
| Science and technology studies | 0.000 | 0.000 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.000 | 0.000 |
| Research integrity | 0.000 | 0.000 |
| Insufficient payload (model declined to judge) | 0.000 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one teacher head, not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".