Name Recognition to Identify Patients of South Asian Ethnicity within the Cancer Registry
Notice bibliographique
Résumé
Objective:The goal of this project was to develop a list of forenames and surnames of South Asian (SA) women that could be used to identify SA breast cancer patients within the cancer registry. This list was compiled, evaluated, and validated to ensure comprehensiveness, accuracy, and applicability of SA names.Methods:This project was conducted by Canadian researchers who are immersed in conducting behavioral studies with SA women diagnosed with cancer in the province of British Columbia. Recruiting SA cancer patients for research can be a difficult task due to social and cultural factors. Methods used by other researchers to identify ethnicity related unique names were employed to filter surnames and forenames that were not common to this ethnic group. Co-author (Gurpreet Oshan) of SA ethnicity rigorously identified and deleted multiple lists and redundant entries along with common English forenames which resulted in a list of 16,888 SA forenames. All co-authors of Indian ethnicity (Gurpreet Oshan, Savitri Singh-Carlson, Harajit Lail) were involved in critiquing and manually reviewing the names list throughout this process. Comprehensive lists of SA surnames and women′s forenames were reviewed to identify those that were unique to SA ethnicity. Accuracy was ensured by constantly filtering the redundancy by using an Excel program which helped to illustrate the number of times each name was spelled in different ways.Results:The final lists included 9112 surnames and 16,888 forenames of SA ethnicity. On the basis of the surname linkage only, the sensitivity of the list was 76.6%, specificity was 62.9%, and the positive predictive value was 58.5%. On the basis of both the surname and forename linkage, the specificity of the list was 88.6%. These lists include variations in spelling forenames and surnames as well.Conclusions:The list of surnames and forenames can be useful tools to identify SA ethnic groups from large population database in healthcare-related research. Ethnicity-specific population research is important in order to help identify how cancer care should be delivered for the SA population, as well as for planning and provision of other related health services. We are willing to share this list upon request to the authors. The goal of this project was to develop a list of forenames and surnames of South Asian (SA) women that could be used to identify SA breast cancer patients within the cancer registry. This list was compiled, evaluated, and validated to ensure comprehensiveness, accuracy, and applicability of SA names. This project was conducted by Canadian researchers who are immersed in conducting behavioral studies with SA women diagnosed with cancer in the province of British Columbia. Recruiting SA cancer patients for research can be a difficult task due to social and cultural factors. Methods used by other researchers to identify ethnicity related unique names were employed to filter surnames and forenames that were not common to this ethnic group. Co-author (Gurpreet Oshan) of SA ethnicity rigorously identified and deleted multiple lists and redundant entries along with common English forenames which resulted in a list of 16,888 SA forenames. All co-authors of Indian ethnicity (Gurpreet Oshan, Savitri Singh-Carlson, Harajit Lail) were involved in critiquing and manually reviewing the names list throughout this process. Comprehensive lists of SA surnames and women′s forenames were reviewed to identify those that were unique to SA ethnicity. Accuracy was ensured by constantly filtering the redundancy by using an Excel program which helped to illustrate the number of times each name was spelled in different ways. The final lists included 9112 surnames and 16,888 forenames of SA ethnicity. On the basis of the surname linkage only, the sensitivity of the list was 76.6%, specificity was 62.9%, and the positive predictive value was 58.5%. On the basis of both the surname and forename linkage, the specificity of the list was 88.6%. These lists include variations in spelling forenames and surnames as well. The list of surnames and forenames can be useful tools to identify SA ethnic groups from large population database in healthcare-related research. Ethnicity-specific population research is important in order to help identify how cancer care should be delivered for the SA population, as well as for planning and provision of other related health services. We are willing to share this list upon request to the authors.
Récupéré en direct depuis OpenAlex et désinversé. Les résumés ne sont pas conservés dans cette base de données : les index inversés représentent 8,6 Go des 9,3 Go de texte de la base, et le serveur dispose de 13 Go libres.
Comment cette classification a été obtenuedéplier
Prédiction distillée sur la base complète
Imitation des enseignantsNi prévalence calibrée, ni vérité terrain. Validation humaine à venir. Apprise à partir de 10 348 étiquettes directes de Codex et de 10 348 étiquettes directes de Gemma. Le mode candidate est l'union des têtes enseignantes seuillées; le consensus est leur intersection. Ces sorties portent le statut machine_predicted_unvalidated et ne sont ni des étiquettes humaines ni des étiquettes directes de modèles de pointe.
Scores Codex et Gemma par catégorie
| Catégorie | Codex | Gemma |
|---|---|---|
| Métarecherche | 0,001 | 0,001 |
| Méta-épidémiologie (sens strict) | 0,000 | 0,000 |
| Méta-épidémiologie (sens large) | 0,000 | 0,000 |
| Bibliométrie | 0,000 | 0,000 |
| Études des sciences et des technologies | 0,000 | 0,000 |
| Communication savante | 0,000 | 0,000 |
| Science ouverte | 0,000 | 0,000 |
| Intégrité de la recherche | 0,000 | 0,000 |
| Charge utile insuffisante (le modèle a refusé de juger) | 0,000 | 0,000 |
Scores machine (provisoires)
Les deux têtes enseignantes du modèle étudiant, lues sur ce travail. Un score ordonne la base pour la relecture; il n'affirme jamais une catégorie, et le statut de validation accompagne chaque rangée tel quel.
Scores de référence d'un modèle non mature (critères de maturité non atteints, 7 itérations). Un score ordonne; il n'affirme jamais une catégorie.
score_only:v0-immature-baseline · tel quel depuis la passe de notation : score_only signifie que le nombre peut ordonner les travaux, et qu'aucune étiquette de catégorie n'en découleClassification
machine, non validéePrédiction automatique; un appel candidat d’une seule tête enseignante, pas un consensus.
Le détail, modèle par modèle et score par score, se trouve en fin de page sous « Comment cette classification a été obtenue ».