Empirical estimation of life expectancy from a linked health database of adults who entered care for HIV
Bibliographic record
Abstract
BACKGROUND: While combination antiretroviral therapy (cART) has significantly improved survival times for persons diagnosed with HIV, estimation of life expectancy (LE) for this cohort remains a challenge, as mortality rates are a function of both time since diagnosis and age, and mortality rates for the oldest age groups may not be available. METHODS: A validated case-finding algorithm for HIV was used to update the cohort of HIV-positive adults who had entered care in Ontario, Canada as of 2012. The Chiang II abridged life table algorithm was modified to use mortality rates stratified by time since entering the cohort and to include various methods for extrapolation of the excess HIV mortality rates to older age groups. RESULTS: As of 2012, there were approximately 15,000 adults in care for HIV in Ontario. The crude all-cause mortality rate declined from 2.6% (95%CI 2.3, 2.9) per year in 2000 to 1.3% (1.2, 1.5) in 2012. Mortality rates were elevated for the first year of care compared to subsequent years (rate ratio of 2.6 (95% CI 2.3, 3.1)). LE for a 20-year old living in Ontario was 62 years (expected age at death is 82), while LE for a 20-year old with HIV was estimated to be reduced to 47 years, for a loss of 15 years of life. Ignoring the higher mortality rates among new cases introduced a modest bias of 1.5 additional years of life lost. In comparison, using 55+ as the open-ended age group was a major source of bias, adding 11 years to the calculated LE. CONCLUSIONS: Use of age limits less than the expected age at death for the open-ended age group significantly overstates the estimated LE and is not recommended. The Chiang II method easily accommodated input of stratified mortality rates and extrapolation of excess mortality rates.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame machine prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. The Gemma side is a direct model label for every work in the frame, read from the title-only record. The Codex side is a classifier learned from the 10,348 direct Codex labels and calibrated to design-weighted sample rates; fields without enough sample support carry no Codex call. Candidate is the union of the two sides; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels.
Distilled classifier scores by category (both heads)
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.006 | 0.047 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.001 | 0.001 |
| Bibliometrics | 0.003 | 0.003 |
| Science and technology studies | 0.000 | 0.000 |
| Scholarly communication | 0.001 | 0.001 |
| Open science | 0.001 | 0.001 |
| Research integrity | 0.001 | 0.001 |
| Insufficient payload (model declined to judge) | 0.002 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one source (direct Gemma or distilled Codex), not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".