MétaCan
Menu
Back to cohort
Record W2134411641 · doi:10.1186/1471-2288-13-93

Understanding recruitment: outcomes associated with alternate methods for seed selection in respondent driven sampling

2013· article· en· W2134411641 on OpenAlexafffund
John Wylie, Ann Jolly

Bibliographic record

VenueBMC Medical Research Methodology · 2013
Typearticle
Languageen
FieldMedicine
TopicHIV, Drug Use, Sexual Risk
Canadian institutionsPublic Health Agency of CanadaUniversity of ManitobaManitoba Health
FundersCanadian Institutes of Health Research
KeywordsRespondentPopulationDemographyLogistic regressionHomophilySelection (genetic algorithm)Sampling (signal processing)Multinomial logistic regressionMedicinePsychologyStatisticsSocial psychologyMathematics

Abstract

fetched live from OpenAlex

BACKGROUND: Respondent driven sampling (RDS) was designed for sampling "hidden" populations and intended as a means of generating unbiased population estimates. Its widespread use has been accompanied by increasing scrutiny as researchers attempt to understand the extent to which the population estimates produced by RDS are, in fact, generalizable to the actual population of interest. In this study we compare two different methods of seed selection to determine whether this may influence recruitment and RDS measures. METHODS: Two seed groups were established. One group was selected as per a standard RDS approach of study staff purposefully selecting a small number of individuals to initiate recruitment chains. The second group consisted of individuals self-presenting to study staff during the time of data collection. Recruitment was allowed to unfold from each group and RDS estimates were compared between the groups. A comparison of variables associated with HIV was also completed. RESULTS: Three analytic groups were used for the majority of the analyses-RDS recruits originating from study staff-selected seeds (n = 196); self-presenting seeds (n = 118); and recruits of self-presenting seeds (n = 264). Multinomial logistic regression demonstrated significant differences between the three groups across six of ten sociodemographic and risk behaviours examined. Examination of homophily values also revealed differences in recruitment from the two seed groups (e.g. in one arm of the study sex workers and solvent users tended not to recruit others like themselves, while the opposite was true in the second arm of the study). RDS estimates of population proportions were also different between the two recruitment arms; in some cases corresponding confidence intervals between the two recruitment arms did not overlap. Further differences were revealed when comparisons of HIV prevalence were carried out. CONCLUSIONS: RDS is a cost-effective tool for data collection, however, seed selection has the potential to influence which subgroups within a population are accessed. Our findings indicate that using multiple methods for seed selection may improve access to hidden populations. Our results further highlight the need for a greater understanding of RDS to ensure appropriate, accurate and representative estimates of a population can be obtained from an RDS sample.

Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.

How this classification was reachedexpand

Full frame machine prediction

Teacher imitation

Not calibrated prevalence, not ground truth. Human validation pending. The Gemma side is a direct model label for every work in the frame, read from the title-only record. The Codex side is a classifier learned from the 10,348 direct Codex labels and calibrated to design-weighted sample rates; fields without enough sample support carry no Codex call. Candidate is the union of the two sides; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels.

metaresearch head score (Codex)0.550
metaresearch head score (Gemma)0.751
Version: metacan-v3-hybrid-931329e0061cValidation status: machine_predicted_unvalidated
Candidate categoriesMetaresearch
Consensus categoriesMetaresearch
DomainCandidate signal: Methods · Consensus signal: Methods
Study designCandidate signal: Observational · Consensus signal: Observational
GenreCandidate signal: Empirical · Consensus signal: none
Teacher disagreement score0.450
Threshold uncertainty score0.555

Distilled classifier scores by category (both heads)

CategoryCodexGemma
Metaresearch0.5500.751
Meta-epidemiology (narrow)0.0010.001
Meta-epidemiology (broad)0.0010.002
Bibliometrics0.0020.003
Science and technology studies0.0020.010
Scholarly communication0.0070.009
Open science0.0030.007
Research integrity0.0040.004
Insufficient payload (model declined to judge)0.0030.001

Machine scores (provisional)

The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.

Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.

Opus teacher head0.877
GPT teacher head0.652
Teacher spread0.224 · how far apart the two teachers sit on this one work
Validation statusscore_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from it

Classification

machine, unvalidated

Machine predicted; the direct Gemma label and the distilled Codex classifier agree on what is shown here.

Study designObservational
DomainMethods
GenreEmpirical

How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".

Quick stats

Citations30
Published2013
Admission routes2
Has abstractyes

Explore more

Same venueBMC Medical Research MethodologySame topicHIV, Drug Use, Sexual RiskFrench-language works237,207