Mortality Risk Prediction Models for People With Kidney Failure
Bibliographic record
Abstract
Importance: People with kidney failure have a high risk of death and poor quality of life. Mortality risk prediction models may help them decide which form of treatment they prefer. Objective: To systematically review the quality of existing mortality prediction models for people with kidney failure and assess whether they can be applied in clinical practice. Evidence Review: MEDLINE, Embase, and the Cochrane Library were searched for studies published between January 1, 2004, and September 30, 2024. Studies were included if they created or evaluated mortality prediction models for people who developed kidney failure, whether treated or not treated with kidney replacement with hemodialysis or peritoneal dialysis. Studies including exclusively kidney transplant recipients were excluded. Two reviewers independently extracted data and graded each study at low, high, or unclear risk of bias and applicability using recommended checklists and tools. Reviewers used the Prediction Model Risk of Bias Assessment Tool and followed prespecified questions about study design, prediction framework, modeling algorithm, performance evaluation, and model deployment. Analyses were completed between January and October 2024. Findings: A total of 7184 unique abstracts were screened for eligibility. Of these, 77 were selected for full-text review, and 50 studies that created all-cause mortality prediction models were included, with 2 963 157 total participants, who had a median (range) age of 64 (52-81) years. Studies had a median (range) proportion of women of 42% (2%-54%). Included studies were at high risk of bias due to inadequate selection of study population (27 studies [54%]), shortcomings in methods of measurement of predictors (15 [30%]) and outcome (12 [24%]), and flaws in the analysis strategy (50 [100%]). Concerns for applicability were also high, as study participants (31 [62%]), predictors (17 [34%]), and outcome (5 [10%]) did not fit the intended target clinical setting. One study (2%) reported decision curve analysis, and 15 (30%) included a tool to enhance model usability. Conclusions and Relevance: According to this systematic review of 50 studies, published mortality prediction models were at high risk of bias and had applicability concerns for clinical practice. New mortality prediction models are needed to inform treatment decisions in people with kidney failure.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame distilled prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.
Codex and Gemma teacher scores by category
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.001 | 0.000 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.002 | 0.001 |
| Bibliometrics | 0.000 | 0.001 |
| Science and technology studies | 0.000 | 0.000 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.000 | 0.000 |
| Research integrity | 0.000 | 0.000 |
| Insufficient payload (model declined to judge) | 0.000 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one teacher head, not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".