The psychometric properties of three self-report screening instruments for identifying frail older people in the community
Bibliographic record
Abstract
BACKGROUND: Frailty is highly prevalent in older people. Its serious adverse consequences, such as disability, are considered to be a public health problem. Therefore, disability prevention in community-dwelling frail older people is considered to be a priority for research and clinical practice in geriatric care. With regard to disability prevention, valid screening instruments are needed to identify frail older people in time. The aim of this study was to evaluate and compare the psychometric properties of three screening instruments: the Groningen Frailty Indicator (GFI), the Tilburg Frailty Indicator (TFI) and the Sherbrooke Postal Questionnaire (SPQ). For validation purposes the Groningen Activity Restriction Scale (GARS) was added. METHODS: A questionnaire was sent to 687 community-dwelling older people (> or = 70 years). Agreement between instruments, internal consistency, and construct validity of instruments were evaluated and compared. RESULTS: The response rate was 77%. Prevalence estimates of frailty ranged from 40% to 59%. The highest agreement was found between the GFI and the TFI (Cohen's kappa = 0.74). Cronbach's alpha for the GFI, the TFI and the SPQ was 0.73, 0.79 and 0.26, respectively. Scores on the three instruments correlated significantly with each other (GFI - TFI, r = 0.87; GFI - SPQ, r = 0.47; TFI - SPQ, r = 0.42) and with the GARS (GFI - GARS, r = 0.57; TFI - GARS, r = 0.61; SPQ - GARS, r = 0.46). The GFI and the TFI scores were, as expected, significantly related to age, sex, education and income. CONCLUSIONS: The GFI and the TFI showed high internal consistency and construct validity in contrast to the SPQ. Based on these findings it is not yet possible to conclude whether the GFI or the TFI should be preferred; data on the predictive values of both instruments are needed. The SPQ seems less appropriate for postal screening of frailty among community-dwelling older people.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame distilled prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.
Codex and Gemma teacher scores by category
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.008 | 0.004 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.000 | 0.000 |
| Bibliometrics | 0.000 | 0.001 |
| Science and technology studies | 0.001 | 0.000 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.001 | 0.000 |
| Research integrity | 0.000 | 0.001 |
| Insufficient payload (model declined to judge) | 0.000 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one teacher head, not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".