Optimal Statistical Decisions for Hospital Report Cards
Notice bibliographique
Résumé
PURPOSE: Hospital report cards provide information designed to help patients and providers to make decisions. The purpose of this study was to place the design of hospital report cards into a decision-theoretic framework. The authors' objectives were 2-fold: 1st, to determine what the choice of significance level implies about the relative value of the different types of misclassifications that can arise. Second, to determine optimal significance levels for specific cost functions describing the relative costs associated with different types of misclassifications. METHODS: Using a previously published theoretical model for hospital mortality, the authors computed false positive (i.e., falsely classified as providing poor-quality care) and false negative (falsely classified as providing good-quality care) rates. First, they determined the cost functions for false negatives and false positives that are implicitly associated with the use of significance levels of 0.05 and 0.01 for identifying hospitals with higher than average mortality. Second, they determined the levels of statistical significance that should be chosen to minimize predefined cost functions, thus minimizing costs associated with misclassifying hospitals. RESULTS: The lower the statistical significance level required for identifying hospitals with higher than average mortality, the lower the implicit cost of false negatives compared to false positives. For a given significance level, the greater the number of patients treated at each hospital or the greater the proportion of truly poorly performing hospitals, the lower the value of the implicit cost incurred by a false negative compared to that for a false positive. For cost functions that put a high relative penalty on false negatives compared to false positives, the use of significance levels of 0.05 or 0.01 does not result in optimal decisions across expected number of patients treated at each hospital or proportions of truly poor-quality care. CONCLUSIONS: Hospital report cards that use significance levels of either 0.05 or 0.01 to identify hospitals that have statistically significantly higher than average mortality make implicit assumptions about cost functions, and the values of the optimal cost function vary across scenarios.
Récupéré en direct depuis OpenAlex et désinversé. Les résumés ne sont pas conservés dans cette base de données : les index inversés représentent 8,6 Go des 9,3 Go de texte de la base, et le serveur dispose de 13 Go libres.
Comment cette classification a été obtenuedéplier
Prédiction distillée sur la base complète
Imitation des enseignantsNi prévalence calibrée, ni vérité terrain. Validation humaine à venir. Apprise à partir de 10 348 étiquettes directes de Codex et de 10 348 étiquettes directes de Gemma. Le mode candidate est l'union des têtes enseignantes seuillées; le consensus est leur intersection. Ces sorties portent le statut machine_predicted_unvalidated et ne sont ni des étiquettes humaines ni des étiquettes directes de modèles de pointe.
Scores Codex et Gemma par catégorie
| Catégorie | Codex | Gemma |
|---|---|---|
| Métarecherche | 0,003 | 0,048 |
| Méta-épidémiologie (sens strict) | 0,000 | 0,000 |
| Méta-épidémiologie (sens large) | 0,001 | 0,000 |
| Bibliométrie | 0,000 | 0,000 |
| Études des sciences et des technologies | 0,001 | 0,000 |
| Communication savante | 0,000 | 0,000 |
| Science ouverte | 0,000 | 0,000 |
| Intégrité de la recherche | 0,001 | 0,001 |
| Charge utile insuffisante (le modèle a refusé de juger) | 0,007 | 0,001 |
Scores machine (provisoires)
Les deux têtes enseignantes du modèle étudiant, lues sur ce travail. Un score ordonne la base pour la relecture; il n'affirme jamais une catégorie, et le statut de validation accompagne chaque rangée tel quel.
Scores de référence d'un modèle non mature (critères de maturité non atteints, 7 itérations). Un score ordonne; il n'affirme jamais une catégorie.
score_only:v0-immature-baseline · tel quel depuis la passe de notation : score_only signifie que le nombre peut ordonner les travaux, et qu'aucune étiquette de catégorie n'en découleClassification
machine, non validéePrédiction automatique; les deux têtes enseignantes s’accordent sur ce qui est montré ici.
Le détail, modèle par modèle et score par score, se trouve en fin de page sous « Comment cette classification a été obtenue ».