MétaCan
Menu
Retour à la cohorte
Enregistrement W7117716984 · doi:10.17605/osf.io/qan4m

Screening for Post-stroke Cognitive Impairment: a comparison of univariate and multivariate normative approaches

2025· other· W7117716984 sur OpenAlexaboutno aff
Céline Gillebert, Hanne Huygelier, Nele Demeyere, Sam Sappho Webb

Notice bibliographique

RevueOpen Science Framework · 2025
Typeother
Langue
Domaine
Thématique
Établissements canadiensnon disponible
Organismes subventionnairesnon disponible
Mots-clésNormativeCognitionCognitive impairmentUnivariateStroke (engine)AnosognosiaCognitive Assessment SystemNeuropsychologyLogistic regression

Résumé

récupéré en direct d'OpenAlex

Screening for domain-specific cognitive impairment after stroke is essential due to the heterogeneous patterns and presentation of cognitive impairments post-stroke and the effect of these impairments on daily life (Esmael et al., 2021; Jokinen et al., 2015; Kusec et al., 2023; Merriman et al., 2019; Milosevich et al., 2023; Mole & Demeyere, 2020; Nys et al., 2005, 2007; Oksala et al., 2009; Samuelsson et al., 2021; Sexton et al., 2019a; Stroke Association, 2018). Failure to detect impairments can lead to failure to target rehabilitation to address or attenuate these problems, and can increase the severity of disability and incidence of mortality (Lucka et al., 2022). To screen for post-stroke cognitive impairment (PSCI) many approaches are in use (Sexton et al., 2019b). Some researchers identify a subtest-specific “impairment” by comparing the patient’s score to a cut-off based on normative data (e.g., the 5th centile cut off, or a z-score > 1 SD from the mean) (e.g., (Demeyere et al., 2015; Sexton et al., 2019b), and then identify all patients with PSCI as those who show impairment on at least one test (e.g., Demeyere et al., 2016; Jokinen et al., 2015; Nys et al., 2007). An alternative approach is to combine subtest scores in one global score indicating PSCI, such as summing subtests of cognitive screens such as the Montreal Cognitive Assessment (MoCA) (Nasreddine et al., 2005). When identifying PSCI as “at least 1 impairment”, one may risk inflating false positives as the approach typically does not control for multiple comparisons. Indeed, in the seminal paper by Brooks and colleagues (Brooks et al., 2009), they highlight that, even in healthy adults, there is an expected high frequency of low scores when completing many tests. This frequency will increase if the threshold for impairments is liberal (set such that many are going to be impaired). Consequently, the number of individuals with at least 1 low score increases the more tests there are administered. In contrast, when identifying PSCI using total scores, one risks under-identifying domain-specific impairments (Demeyere et al., 2016). A second challenge is the fact that a patient may score subclinically on multiple tests, not meeting the per-test criterion of impairment, while their overall test profile may still deviate from healthy controls, and thus with current per-test criteria (i.e., univariate testing) lead to under identification of impairment (Huizenga et al., 2007). Thus, assessors face the challenge of assessing cognition sufficiently to identify potential domain-specific impairments, while: 1) avoiding false positives, and 2) avoiding false negatives resulting from univariate testing. Since the early 2000s, more attention has been given to offer solutions for this challenge. For instance, more neuropsychological tests are considering a multivariate base rate of performance in normative data (Kiselica et al., 2024) and examples of these base rates to aid interpretation of performance have been developed, including for the Wechsler Adult Intelligence Scale–Fourth Edition (Brooks et al., 2013). The multivariate base rates allow to control for false positives, while considering test associations in normative samples. In addition, a test statistic, the Multivariate Normative Comparison (MNC), has been validated to compare a patient’s test profile to normative data in a multivariate way. A detailed tutorial has been written regarding how to compute MNC for psychological test data in order to understand patterns of performance rather than just isolated subtest performance (Huizenga et al., 2007). Huizenga et al demonstrated that in situations where the number of (sub)tests is just exceeded by the number of controls the Bonferroni correction for multiple comparisons of tests is sufficient to see if a given patient is different to normative performance across multiple tests (Huizenga et al., 2007). Whereas, in situations with more healthy control data, where estimations can be more precise and patterns of performance across different (sub)tests can be estimated, MNC is more statistically powerful to identify cognitive impairment (Huizenga et al., 2007). Although the MNC approach may be theoretically more powerful to identify PSCI across multiple domains, it has not yet been empirically evaluated for identifying PSCI. It is thus essential to investigate whether multivariate methods of interpreting cognitive tests impact the identification of PSCI. To this end, we will investigate different methods of identifying across-domain PSCI using the Oxford Cognitive Screen (OCS). The OCS (Demeyere et al., 2015) is widely used in clinical practice as a first line screen for cognitive impairment (Murphy et al., 2023). The OCS briefly screens for impairments in language, memory, attention, executive function, praxis, and numerical cognition. At its base level, the OCS only provides impairment scores on subtests and does not officially have cut-offs for across-domain PSCI. In the current study, we will contrast four methods to identify PSCI: (1) the traditional approach of identifying PSCI based on an “at least 1 impairment” criterion and 5th centile cut-offs per test (not corrected for multiple comparisons) (Demeyere et al., 2016; Kusec et al., 2023; Sexton et al., 2019b), (2) an adjusted version of the “at least 1 impairment” approach where the per-test cut-offs were Bonferroni corrected, (3) a total score approach which combines all subtests in a single outcome measure and contrasts this to a normative group, and (4) the “MNC” approach which has the power to consider test-associations in the normative sample and identify multivariate deviations in test performance (Grasman et al., 2010). Our primary aim was to assess whether these methods result in different estimates of the prevalence of PSCI in a large stroke sample. In addition, we examined for how many patients’ diagnosis would be different depending on the approach used. To estimate the impact of multiple comparisons for the OCS, we applied these methods on the normative group as well. We predict that the “at least 1 impairment” approach without Bonferroni corrected cut-offs for subtest impairment will not control the false positive rate (i.e., maintain PSCI identification at the nominal significance level in the normative group). In a worst-case (unrealistic) scenario, PSCI would be identified in 45% of the normative group (if the 13 OCS subtests were fully independent) (Huizenga et al., 2007). We predict that all other methods will control this rate at the nominal significance level. For the stroke patients, we predict that the MNC approach will identify PSCI in more patients than the Bonferroni corrected “at least 1 impairment” approach (as it allows to consider test-associations) and in less patients than the uncorrected “at least 1 impairment” approach (as it controls for multiple comparisons).

Récupéré en direct depuis OpenAlex et désinversé. Les résumés ne sont pas conservés dans cette base de données : les index inversés représentent 8,6 Go des 9,3 Go de texte de la base, et le serveur dispose de 13 Go libres.

Comment cette classification a été obtenuedéplier

Prédiction distillée sur la base complète

Imitation des enseignants

Ni prévalence calibrée, ni vérité terrain. Validation humaine à venir. Apprise à partir de 10 348 étiquettes directes de Codex et de 10 348 étiquettes directes de Gemma. Le mode candidate est l'union des têtes enseignantes seuillées; le consensus est leur intersection. Ces sorties portent le statut machine_predicted_unvalidated et ne sont ni des étiquettes humaines ni des étiquettes directes de modèles de pointe.

score de la tête « metaresearch » (Codex)0,009
score de la tête « metaresearch » (Gemma)0,010
Version: codex-gemma-dda1882f352aStatut de validation: machine_predicted_unvalidated
Catégories candidatesMétarecherche, Méta-épidémiologie (sens strict), Études des sciences et des technologies, Communication savante, Science ouverte, Intégrité de la recherche, Charge utile insuffisante (le modèle a refusé de juger)
Catégories consensuellesMéta-épidémiologie (sens strict), Études des sciences et des technologies, Science ouverte
DomaineSignal candidat: aucune · Signal consensuel: aucune
Devis d'étudeSignal candidat: Qualitatif · Signal consensuel: aucune
GenreSignal candidat: Méthodes · Signal consensuel: aucune
Score de désaccord entre enseignants0,592
Score d'incertitude au seuil1,000

Scores Codex et Gemma par catégorie

CatégorieCodexGemma
Métarecherche0,0090,010
Méta-épidémiologie (sens strict)0,0020,002
Méta-épidémiologie (sens large)0,0040,000
Bibliométrie0,0030,006
Études des sciences et des technologies0,0030,009
Communication savante0,0030,003
Science ouverte0,0080,009
Intégrité de la recherche0,0010,002
Charge utile insuffisante (le modèle a refusé de juger)0,0010,000

Scores machine (provisoires)

Les deux têtes enseignantes du modèle étudiant, lues sur ce travail. Un score ordonne la base pour la relecture; il n'affirme jamais une catégorie, et le statut de validation accompagne chaque rangée tel quel.

Scores de référence d'un modèle non mature (critères de maturité non atteints, 7 itérations). Un score ordonne; il n'affirme jamais une catégorie.

Tête enseignante Opus0,086
Tête enseignante GPT0,383
Écart entre enseignants0,297 · la distance entre les deux têtes enseignantes sur ce seul travail
Statut de validationscore_only:v0-immature-baseline · tel quel depuis la passe de notation : score_only signifie que le nombre peut ordonner les travaux, et qu'aucune étiquette de catégorie n'en découle

Classification

machine, non validée

Prédiction automatique; les deux têtes enseignantes s’accordent sur ce qui est montré ici.

Devis d'étudeQualitatif
Domainenon disponible
GenreMéthodes

Le détail, modèle par modèle et score par score, se trouve en fin de page sous « Comment cette classification a été obtenue ».

En bref

Citations0
Publié2025
Routes d'admission1
Résumé présentoui

Explorer davantage

Même revueOpen Science FrameworkTravaux en français237 207