Increasing the Clinical Utility of the BESTest, Mini-BESTest, and Brief-BESTest: Normative Values in Canadian Adults Who Are Healthy and Aged 50 Years or Older
Bibliographic record
Abstract
BACKGROUND: Balance is a composite ability requiring the integration of multiple systems. The Balance Evaluation Systems Test (BESTest) and 2 abbreviated versions (the Mini-BESTest and the Brief-BESTest) are balance assessment tools that target these systems. To date, no normative data exist for any version of the BESTest. OBJECTIVE: The purpose of this study was to determine the age-related normative scores on the BESTest, Mini-BESTest, and Brief-BESTest for Canadians who are healthy and 50 to 89 years of age. DESIGN: A cross-sectional study design was used. METHODS: Seventy-nine adults who were healthy and aged 50 to 89 years (mean age=68.9 years; 50.6% women) participated. Normative scores were reported by age decade. RESULTS: Mean BESTest scores were 95.7 (95% confidence interval [CI]=94.4-97.1) for adults who were aged 50 to 59 years, 91.4 (95% CI=89.8-93.0) for those who were aged 60 to 69 years, 85.4 (95% CI=82.5-88.2) for those who were aged 70 to 79 years, and 79.4 (95% CI=74.3-84.5) for those who were aged 80 to 89 years. Similar results were obtained for the Mini-BESTest and the Brief-BESTest, and all 3 tests showed statistically significant differences in scores among the age cohorts. LIMITATIONS: Because only adults who were 50 to 89 years of age were tested, there are still no normative data for people outside this age range. Also, the scores presented may not be generalizable to all countries. CONCLUSIONS: These normative data enhance the clinical utility of the BESTest, Mini-BESTest, and Brief-BESTest by providing clinicians with reference points to guide treatment.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame distilled prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.
Codex and Gemma teacher scores by category
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.001 | 0.000 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.000 | 0.000 |
| Bibliometrics | 0.000 | 0.000 |
| Science and technology studies | 0.000 | 0.000 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.000 | 0.000 |
| Research integrity | 0.000 | 0.001 |
| Insufficient payload (model declined to judge) | 0.000 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one teacher head, not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".