Reliability, Validity, and Responsiveness of the Lower Extremity Functional Scale for Inpatients of an Orthopaedic Rehabilitation Ward
Bibliographic record
Abstract
STUDY DESIGN: Single-group, repeated-measures study. OBJECTIVE: To estimate the test-retest reliability, construct validity, and responsiveness of the Lower Extremity Functional Scale (LEFS) on inpatients attending an orthopaedic rehabilitation ward. BACKGROUND: The LEFS has acceptable validity on outpatients in assessing functional mobility, but it has not been tested for use on an inpatient orthopaedic ward. METHODS AND MEASURES: Inpatients in an orthopaedic ward (n = 142) completed the 20-item, self-report LEFS on admission, 7 to 10 days after admission, and on discharge. To test reliability, 24 patients had the LEFS repeated 1 day after the admission test, and the intraclass correlation (ICC) and the standard error of measurement (SEM) were calculated. Change scores of the LEFS were evaluated against patients' and therapists' rating of improvement, and change scores of comparison measures that included pain, functional performance, and the composite index created from scores of these comparison measures. The standardized response mean (SRM) of the LEFS was also computed. RESULTS: The ICC of the LEFS was 0.88, and the SEM was 4 LEFS points (LEFS score range, 0-80). The change in LEFS correlated with changes of comparison measures in the same direction of improvement. Patients rated as improved by both themselves and their therapists had significantly larger change in LEFS scores than subjects rated as no change. The SRM of the LEFS from admission to discharge was 1.76 on patients rated as improved. CONCLUSION: The LEFS is reliable and valid to assess group and individual change, and has large responsiveness. The LEFS and the comparison measures likely assess different constructs.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame distilled prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.
Codex and Gemma teacher scores by category
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.001 | 0.000 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.001 | 0.000 |
| Bibliometrics | 0.000 | 0.000 |
| Science and technology studies | 0.000 | 0.000 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.000 | 0.000 |
| Research integrity | 0.000 | 0.000 |
| Insufficient payload (model declined to judge) | 0.000 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one teacher head, not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".