The Incidence Patterns Model to Estimate the Distribution of New HIV Infections in Sub-Saharan Africa: Development and Validation of a Mathematical Model
Bibliographic record
Abstract
BACKGROUND: Programmatic planning in HIV requires estimates of the distribution of new HIV infections according to identifiable characteristics of individuals. In sub-Saharan Africa, robust routine data sources and historical epidemiological observations are available to inform and validate such estimates. METHODS AND FINDINGS: We developed a predictive model, the Incidence Patterns Model (IPM), representing populations according to factors that have been demonstrated to be strongly associated with HIV acquisition risk: gender, marital/sexual activity status, geographic location, "key populations" based on risk behaviours (sex work, injecting drug use, and male-to-male sex), HIV and ART status within married or cohabiting unions, and circumcision status. The IPM estimates the distribution of new infections acquired by group based on these factors within a Bayesian framework accounting for regional prior information on demographic and epidemiological characteristics from trials or observational studies. We validated and trained the model against direct observations of HIV incidence by group in seven rounds of cohort data from four studies ("sites") conducted in Manicaland, Zimbabwe; Rakai, Uganda; Karonga, Malawi; and Kisesa, Tanzania. The IPM performed well, with the projections' credible intervals for the proportion of new infections per group overlapping the data's confidence intervals for all groups in all rounds of data. In terms of geographical distribution, the projections' credible intervals overlapped the confidence intervals for four out of seven rounds, which were used as proxies for administrative divisions in a country. We assessed model performance after internal training (within one site) and external training (between sites) by comparing mean posterior log-likelihoods and used the best model to estimate the distribution of HIV incidence in six countries (Gabon, Kenya, Malawi, Rwanda, Swaziland, and Zambia) in the region. We subsequently inferred the potential contribution of each group to transmission using a simple model that builds on the results from the IPM and makes further assumptions about sexual mixing patterns and transmission rates. In all countries except Swaziland, individuals in unions were the single group contributing to the largest proportion of new infections acquired (39%-77%), followed by never married women and men. Female sex workers accounted for a large proportion of new infections (5%-16%) compared to their population size. Individuals in unions were also the single largest contributor to the proportion of infections transmitted (35%-62%), followed by key populations and previously married men and women. Swaziland exhibited different incidence patterns, with never married men and women accounting for over 65% of new infections acquired and also contributing to a large proportion of infections transmitted (up to 56%). Between- and within-country variations indicated different incidence patterns in specific settings. CONCLUSIONS: It is possible to reliably predict the distribution of new HIV infections acquired using data routinely available in many countries in the sub-Saharan African region with a single relatively simple mathematical model. This tool would complement more specific analyses to guide resource allocation, data collection, and programme planning.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame distilled prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.
Codex and Gemma teacher scores by category
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.001 | 0.002 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.000 | 0.000 |
| Bibliometrics | 0.000 | 0.000 |
| Science and technology studies | 0.000 | 0.000 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.000 | 0.000 |
| Research integrity | 0.000 | 0.000 |
| Insufficient payload (model declined to judge) | 0.000 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one teacher head, not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".