MétaCan
Menu
Retour à la cohorte
Enregistrement W2894992918 · doi:10.5041/rmmj.10351

Detection and Diagnostic Overall Accuracy Measures of Medical Tests

2018· article· en· W2894992918 sur OpenAlexaff
Gilat L. Grunau, Shai Linn

Notice bibliographique

RevueRambam Maimonides Medical Journal · 2018
Typearticle
Langueen
DomaineMathematics
ThématiqueStatistical Methods in Clinical Trials
Établissements canadiensUniversity of British Columbia
Organismes subventionnairesnon disponible
Mots-clésMedical diagnosisDiagnostic accuracyBayes' theoremPopulationMedicineTest (biology)Accuracy and precisionStatisticsContrast (vision)Diagnostic testSensitivity (control systems)Computer scienceMachine learningArtificial intelligenceBayesian probabilityPathologyPediatricsMathematicsRadiologyEnvironmental health

Résumé

récupéré en direct d'OpenAlex

BACKGROUND: Overall accuracy measures of medical tests are often used with unclear interpretations. OBJECTIVES: To develop methods of calculating the overall accuracy of medical tests in the patient population. METHODS: Algebraic equations based on Bayes' theorem. RESULTS: A new approach is proposed for calculating overall accuracy in the patient population. Examples and applications using published data are presented. CONCLUSIONS: The overall accuracy is the proportion of the correct test results. We introduce a clear distinction between the overall accuracy measures of medical tests that are aimed at the detection of a disease in a screening of populations for public health purposes in the general population and the overall accuracy measures of tests aimed at determining a diagnosis in individuals in a clinical setting. We show that the overall detection accuracy measure is obtained in a specific study that explores test accuracy among persons with known diagnoses and may be useful for public health screening tests. It is different from the overall diagnostic accuracy that could be calculated in the clinical setting for the evaluation of medical tests aimed at determining the individual patients' diagnoses. We show that the overall detection accuracy is constant and is not affected by the prevalence of the disease. In contrast, the overall diagnostic accuracy changes and is dependent on the prevalence. Moreover, it ranges according to the ratio between the sensitivity and specificity. Thus, when the sensitivity is greater than the specificity, the overall diagnostic accuracy increases with increasing prevalence, and vice versa, that is, when the sensitivity is lower than the specificity, the overall diagnostic accuracy decreases with increasing prevalence so that another test might be more useful for diagnostic procedures. Our paper suggests a new and more appropriate methodology for estimating the overall diagnostic accuracy of any medical test. This may be important for helping clinicians avoid errors.

Récupéré en direct depuis OpenAlex et désinversé. Les résumés ne sont pas conservés dans cette base de données : les index inversés représentent 8,6 Go des 9,3 Go de texte de la base, et le serveur dispose de 13 Go libres.

Comment cette classification a été obtenuedéplier

Prédiction distillée sur la base complète

Imitation des enseignants

Ni prévalence calibrée, ni vérité terrain. Validation humaine à venir. Apprise à partir de 10 348 étiquettes directes de Codex et de 10 348 étiquettes directes de Gemma. Le mode candidate est l'union des têtes enseignantes seuillées; le consensus est leur intersection. Ces sorties portent le statut machine_predicted_unvalidated et ne sont ni des étiquettes humaines ni des étiquettes directes de modèles de pointe.

score de la tête « metaresearch » (Codex)0,014
score de la tête « metaresearch » (Gemma)0,776
Version: codex-gemma-dda1882f352aStatut de validation: machine_predicted_unvalidated
Catégories candidatesMétarecherche, Charge utile insuffisante (le modèle a refusé de juger)
Catégories consensuellesaucune
DomaineSignal candidat: aucune · Signal consensuel: aucune
Devis d'étudeSignal candidat: Théorique ou conceptuel · Signal consensuel: aucune
GenreSignal candidat: Empirique · Signal consensuel: Empirique
Score de désaccord entre enseignants0,963
Score d'incertitude au seuil0,997

Scores Codex et Gemma par catégorie

CatégorieCodexGemma
Métarecherche0,0140,776
Méta-épidémiologie (sens strict)0,0000,000
Méta-épidémiologie (sens large)0,0010,000
Bibliométrie0,0000,000
Études des sciences et des technologies0,0000,001
Communication savante0,0000,000
Science ouverte0,0010,000
Intégrité de la recherche0,0010,001
Charge utile insuffisante (le modèle a refusé de juger)0,0040,000

Scores machine (provisoires)

Les deux têtes enseignantes du modèle étudiant, lues sur ce travail. Un score ordonne la base pour la relecture; il n'affirme jamais une catégorie, et le statut de validation accompagne chaque rangée tel quel.

Scores de référence d'un modèle non mature (critères de maturité non atteints, 7 itérations). Un score ordonne; il n'affirme jamais une catégorie.

Tête enseignante Opus0,377
Tête enseignante GPT0,538
Écart entre enseignants0,161 · la distance entre les deux têtes enseignantes sur ce seul travail
Statut de validationscore_only:v0-immature-baseline · tel quel depuis la passe de notation : score_only signifie que le nombre peut ordonner les travaux, et qu'aucune étiquette de catégorie n'en découle

Classification

machine, non validée

Prédiction automatique; un appel candidat d’une seule tête enseignante, pas un consensus.

Devis d'étudeThéorique ou conceptuel
Domainenon disponible
GenreEmpirique

Le détail, modèle par modèle et score par score, se trouve en fin de page sous « Comment cette classification a été obtenue ».

En bref

Citations14
Publié2018
Routes d'admission1
Résumé présentoui

Explorer davantage

Même revueRambam Maimonides Medical JournalMême sujetStatistical Methods in Clinical TrialsTravaux en français237 207