What is the empirical basis for converting banded ordinal data on numbers of sex partners among MSM into a continuous scale level variable? A secondary analysis of 13 surveys across 17 countries
Bibliographic record
Abstract
BACKGROUND: To provide empirically based guidance for substituting partner number categories in large MSM surveys with mean numbers of sexual and condomless anal intercourse (CAI) partners in a secondary analysis of survey data. METHODS: We collated data on numbers of sexual and CAI partners reported in a continuous scale (write-in number) in thirteen MSM surveys on sexual health and behaviour across 17 countries. Pooled descriptive statistics for the number of sexual and CAI partners during the last twelve (N = 55,180) and 6 months (N = 31,759) were calculated for two sets of categories commonly used in reporting numbers of sexual partners in sexual behaviour surveys. RESULTS: The pooled mean number of partners in the previous 12 months for the total sample was 15.8 partners (SD = 36.6), while the median number of partners was 5 (IQR = 2-15). Means for number of partners in the previous 12 months for the first set of categories were: 16.4 for 11-20 partners (SD = 3.3); 27.8 for 21-30 (SD = 2.8); 38.6 for 31-40 (SD = 2.4); 49.6 for 41-50 (SD = 1.5); and 128.2 for 'more than 50' (SD = 98.1). Alternative upper cut-offs: 43.4 for 'more than 10' (SD = 57.7); 65.3 for 'more than 20' (SD = 70.3). Self-reported partner numbers for both time frames consistently exceeded 200 or 300. While there was substantial variation of overall means across surveys, the means for all chosen categories were very similar. Partner numbers above nine mainly clustered at multiples of tens, regardless of the selected time frame. The overall means for CAI partners were lower than those for sexual partners; however, such difference was completely absent from all categories beyond ten sexual and CAI partners. CONCLUSIONS: Clustering of reported partner numbers confirm common MSM sexual behaviour surveys' questionnaire piloting feedback indicating that responses to numbers of sexual partners beyond 10 are best guesses rather than precise counts, but large partner numbers above typical upper cut-offs are common.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame distilled prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.
Codex and Gemma teacher scores by category
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.122 | 0.104 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.002 | 0.001 |
| Bibliometrics | 0.001 | 0.002 |
| Science and technology studies | 0.001 | 0.003 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.002 | 0.002 |
| Research integrity | 0.000 | 0.002 |
| Insufficient payload (model declined to judge) | 0.013 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; both teacher heads agree on what is shown here.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".