Recommendations From a Descriptive Evaluation to Improve Screening Procedures for Web-Based Studies With Couples: Cross-Sectional Study
Notice bibliographique
Résumé
BACKGROUND: Although there are a number of advantages to using the internet to recruit and enroll participants into Web-based research studies, these advantages hinge on data validity. In response to this concern, researchers have provided recommendations for how best to screen for fraudulent survey entries and to handle potentially invalid responses. Yet, the majority of this previous work focuses on screening (ie, verification that individual met the inclusion criteria) and validating data from 1 individual, and not from 2 people who are in a dyadic relationship with one another (eg, same-sex male couple; mother and daughter). Although many of the same data validation and screening recommendations for Web-based studies with individual participants can be used with dyads, there are differences and challenges that need to be considered. OBJECTIVE: This paper aimed to describe the methods used to verify and validate couples' relationships and data from a Web-based research study, as well as the associated lessons learned for application toward future Web-based studies involving the screening and enrollment of couples with dyadic data collection. METHODS: We conducted a descriptive evaluation of the procedures and associated benchmarks (ie, decision rules) used to verify couples' relationships and validate whether data uniquely came from each partner of the couple. Data came from a large convenience sample of same-sex male couples in the United States, who were recruited through social media venues for a Web-based, mixed methods HIV prevention research study. RESULTS: Among the 3815 individuals who initiated eligibility screening, 1536 paired individuals (ie, data from both partners of a dyad) were assessed for relationship verification; all passed this benchmark. For data validation, 450 paired individuals (225 dyads) were identified as fraudulent and failed this benchmark, resulting in a total sample size of 1086 paired participants representing 543 same-sex male couples who were enrolled. The lessons learned from the procedures used to screen couples for this Web-based research study have led us to identify and describe four areas that warrant careful attention: (1) creation of new and replacement of certain relationship verification items, (2) identification of resources needed relative to using a manual or electronic approach for screening, (3) examination of approaches to link and identify both partners of the couple, and (4) handling of bots. CONCLUSIONS: The screening items and associated rules used to verify and validate couples' relationships and data worked yet required extensive resources to implement. New or updating some items to verify a couple's relationship may be beneficial for future studies. The procedures used to link and identify whether both partners were coupled also worked, yet they call into question whether new approaches are possible to help increase linkage, suggesting the need for further inquiry.
Récupéré en direct depuis OpenAlex et désinversé. Les résumés ne sont pas conservés dans cette base de données : les index inversés représentent 8,6 Go des 9,3 Go de texte de la base, et le serveur dispose de 13 Go libres.
Comment cette classification a été obtenuedéplier
Prédiction machine sur la base complète
Imitation des enseignantsNi prévalence calibrée, ni vérité terrain. Validation humaine à venir. Le volet Gemma est une étiquette directe du modèle pour chaque travail de la base, lue sur la notice réduite au titre. Le volet Codex est un classifieur appris des 10 348 étiquettes directes de Codex et calibré sur les taux pondérés de l'échantillon; les champs sans appui suffisant ne portent aucun appel Codex. Le mode candidate est l'union des deux volets; le consensus est leur intersection. Ces sorties portent le statut machine_predicted_unvalidated et ne sont pas des étiquettes humaines.
Scores du classifieur distillé par catégorie (deux têtes)
| Catégorie | Codex | Gemma |
|---|---|---|
| Métarecherche | 0,725 | 0,783 |
| Méta-épidémiologie (sens strict) | 0,003 | 0,004 |
| Méta-épidémiologie (sens large) | 0,005 | 0,010 |
| Bibliométrie | 0,013 | 0,012 |
| Études des sciences et des technologies | 0,006 | 0,005 |
| Communication savante | 0,012 | 0,016 |
| Science ouverte | 0,009 | 0,006 |
| Intégrité de la recherche | 0,006 | 0,006 |
| Charge utile insuffisante (le modèle a refusé de juger) | 0,008 | 0,005 |
Scores machine (provisoires)
Les deux têtes enseignantes du modèle étudiant, lues sur ce travail. Un score ordonne la base pour la relecture; il n'affirme jamais une catégorie, et le statut de validation accompagne chaque rangée tel quel.
Scores de référence d'un modèle non mature (critères de maturité non atteints, 7 itérations). Un score ordonne; il n'affirme jamais une catégorie.
score_only:v0-immature-baseline · tel quel depuis la passe de notation : score_only signifie que le nombre peut ordonner les travaux, et qu'aucune étiquette de catégorie n'en découleClassification
machine, non validéePrédiction automatique; l’étiquette directe de Gemma et le classifieur distillé Codex s’accordent sur ce qui est montré ici.
Le détail, modèle par modèle et score par score, se trouve en fin de page sous « Comment cette classification a été obtenue ».