MétaCan
Menu
Retour à la cohorte
Enregistrement W4406838304 · doi:10.2196/63252

Introducing Novel Methods to Identify Fraudulent Responses (Sampling With Sisyphus): Web-Based LGBTQ2S+ Mixed-Methods Study

2025· article· en· W4406838304 sur OpenAlexafffundabout
Kinnon R. MacKinnon, Naail Khan, Katherine M. Newman, Wren Ariel Gould, Gin Marshall, Travis Salway, Annie Pullen Sansfaçon, Hannah Kia, June Sing Hong Lam

Notice bibliographique

RevueJournal of Medical Internet Research · 2025
Typearticle
Langueen
DomainePsychology
ThématiqueLGBTQ Health, Identity, and Policy
Établissements canadiensCentre for Addiction and Mental HealthUniversité de MontréalSimon Fraser UniversityUniversity of British ColumbiaPublic Health OntarioYork UniversityUniversity of Toronto
Organismes subventionnairesSocial Sciences and Humanities Research Council of Canada
Mots-clésPreprintWorld Wide WebSampling (signal processing)The InternetComputer scienceData scienceInternet privacyTelecommunications

Résumé

récupéré en direct d'OpenAlex

BACKGROUND: The myth of Sisyphus teaches about resilience in the face of life challenges. Detransition after an initial gender transition is an emerging experience that requires sensitive and community-driven research. However, there are significant complexities and costs that researchers must confront to collect reliable data to better understand this phenomenon, including the lack of a uniform definition and challenges with recruitment. OBJECTIVE: This paper presents the sampling and recruitment methods of a new study on detransition-related phenomena among lesbian, gay, bisexual, transgender, queer, and 2-spirit (LGBTQ2S+) populations. It introduces a novel protocol for identifying and removing bot, scam, and ineligible responses from survey datasets and presents preliminary descriptive sociodemographic results of the sample. This analysis does not present gender-affirming health care outcomes. METHODS: To attract a large and heterogeneous sample, 3 different study flyers were created in English, French, and Spanish. Between December 1, 2023, and May 1, 2024, these flyers were distributed to >615 sexual and gender minority organizations and gender care providers in the United States and Canada, and paid advertisements totaling >CAD $7400 (US $5551) were promoted on 5 different social media platforms. Although many social media promotions were rejected or removed, the advertisements reached >7.7 million accounts. Study website visitors were directed from 35 different traffic sources, with the top 5 being Facebook (3,577,520/7,777,218, 46%), direct link (2,255,393/7,777,218, 29%), Reddit (1,011,038/7,777,218, 13%), Instagram (466,633/7,777,218, 6%), and X (formerly known as Twitter; 233,317/7,777,218, 3%). A systematic protocol was developed to identify scam, nonsense, and ineligible responses and to conduct web-based Zoom video platform screening with select participants. RESULTS: Of the 1377 completed survey responses, 957 (69.5%) were deemed eligible and included in the analytic dataset after applying the exclusion protocol and conducting 113 virtual screenings. The mean age of the sample was 25.87 (SD 7.77; median 24, IQR 21-29 years). A majority of the participants were White (Canadian, American, or of European descent; 748/950, 78.7%), living in the United States (704/957, 73.6%), and assigned female at birth (754/953, 79.1%). Many participants reported having a sexual minority identity, with more than half the sample (543/955, 56.8%) indicating plurisexual orientations, such as bisexual or pansexual identities. A minority of participants (108/955, 11.3%) identified as straight or heterosexual. When asked about their gender-diverse identities after stopping or reversing gender transition, 33.2% (318/957) reported being nonbinary, 43.2% (413/957) transgender, and 40.5% (388/957) identified as detransitioned. CONCLUSIONS: Despite challenges encountered during the study promotion and data collection phases, a heterogeneous sample of >950 eligible participants was obtained, presenting opportunities for future analyses to better understand these LGBTQ2S+ experiences. This study is among the first to introduce an innovative strategy to sample a hard-to-reach and equity-deserving group, and to present an approach to remove fraudulent responses.

Récupéré en direct depuis OpenAlex et désinversé. Les résumés ne sont pas conservés dans cette base de données : les index inversés représentent 8,6 Go des 9,3 Go de texte de la base, et le serveur dispose de 13 Go libres.

Comment cette classification a été obtenuedéplier

Prédiction machine sur la base complète

Imitation des enseignants

Ni prévalence calibrée, ni vérité terrain. Validation humaine à venir. Le volet Gemma est une étiquette directe du modèle pour chaque travail de la base, lue sur la notice réduite au titre. Le volet Codex est un classifieur appris des 10 348 étiquettes directes de Codex et calibré sur les taux pondérés de l'échantillon; les champs sans appui suffisant ne portent aucun appel Codex. Le mode candidate est l'union des deux volets; le consensus est leur intersection. Ces sorties portent le statut machine_predicted_unvalidated et ne sont pas des étiquettes humaines.

score de la tête « metaresearch » (Codex)0,165
score de la tête « metaresearch » (Gemma)0,207
Version: metacan-v3-hybrid-931329e0061cStatut de validation: machine_predicted_unvalidated
Catégories candidatesMétarecherche, Intégrité de la recherche
Catégories consensuellesaucune
DomaineSignal candidat: Méthodes · Signal consensuel: aucune
Devis d'étudeSignal candidat: Simulation ou modélisation · Signal consensuel: aucune
GenreSignal candidat: Méthodes · Signal consensuel: Méthodes
Score de désaccord entre enseignants0,997
Score d'incertitude au seuil0,872

Scores du classifieur distillé par catégorie (deux têtes)

CatégorieCodexGemma
Métarecherche0,1650,207
Méta-épidémiologie (sens strict)0,0020,002
Méta-épidémiologie (sens large)0,0010,002
Bibliométrie0,0040,003
Études des sciences et des technologies0,0040,003
Communication savante0,0040,003
Science ouverte0,0040,005
Intégrité de la recherche0,0030,003
Charge utile insuffisante (le modèle a refusé de juger)0,0120,003

Scores machine (provisoires)

Les deux têtes enseignantes du modèle étudiant, lues sur ce travail. Un score ordonne la base pour la relecture; il n'affirme jamais une catégorie, et le statut de validation accompagne chaque rangée tel quel.

Scores de référence d'un modèle non mature (critères de maturité non atteints, 7 itérations). Un score ordonne; il n'affirme jamais une catégorie.

Tête enseignante Opus0,178
Tête enseignante GPT0,640
Écart entre enseignants0,462 · la distance entre les deux têtes enseignantes sur ce seul travail
Statut de validationscore_only:v0-immature-baseline · tel quel depuis la passe de notation : score_only signifie que le nombre peut ordonner les travaux, et qu'aucune étiquette de catégorie n'en découle

Classification

machine, non validée

Prédiction automatique; un appel candidat d’une seule source (Gemma direct ou Codex distillé), pas un consensus.

Devis d'étudeSimulation ou modélisation
DomaineMéthodes
GenreMéthodes

Le détail, modèle par modèle et score par score, se trouve en fin de page sous « Comment cette classification a été obtenue ».

En bref

Citations5
Publié2025
Routes d'admission3
Résumé présentoui

Explorer davantage

Même revueJournal of Medical Internet ResearchMême sujetLGBTQ Health, Identity, and PolicyTravaux en français237 207