The Prognostic Value of Functional Capacity Evaluation in Patients With Chronic Low Back Pain: Part 2
Bibliographic record
Abstract
In Brief Study Design. Historical cohort study. Objectives. We investigated the ability of the Isernhagen Work Systems’ Functional Capacity Evaluation to predict sustained recovery. Summary of Background Data. Functional Capacity Evaluation is commonly used to determine readiness or ability for safe return to work following musculoskeletal injury, implying a low risk of future recurrence or “reinjury.” However, this theoretical construct has not yet been tested. Methods. Workers’ compensation claimants who underwent Functional Capacity Evaluation following low back injury and subsequently demonstrated recovery in the form of suspension of total temporary disability benefits or claim closure were studied. The number of failed tasks and performance on the floor-to-waist lift task in the protocol were used as indicators of Functional Capacity Evaluation performance. Indicators of sustained recovery included whether or not total temporary disability benefits restarted, the claim was reopened, or a new back claim was filed. Logistic regression was used to determine the prognostic effect of Functional Capacity Evaluation alone and after controlling for suspected confounding variables. Results. Overall, 46 of 226 patients (20%) experienced a recurrent back-related event within the year following Functional Capacity Evaluation. Opposite to the initial hypothesis, a lower number of failed Functional Capacity Evaluation tasks was consistently associated with higher risk of recurrence after controlling for potential confounding variables. Performance on the floor-to-waist lift task was not related to future recurrence. Conclusions. Contrary to Functional Capacity Evaluation theory, better Functional Capacity Evaluation performance as indicated by a lower number of failed tasks was associated with higher risk of recurrence. The validity of Functional Capacity Evaluation’s purported ability to identify claimants who are “safe” to return to work is suspect. An historical cohort study was conducted of workers’ compensation claimants who underwent Functional Capacity Evaluation for low back injuries and subsequently demonstrated signs of recovery. Better Functional Capacity Evaluation performance was associated with higher risk of recurrence. The validity of Functional Capacity Evaluation's ability to identify claimants who are “safe” to return to work is suspect.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame machine prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. The Gemma side is a direct model label for every work in the frame, read from the title-only record. The Codex side is a classifier learned from the 10,348 direct Codex labels and calibrated to design-weighted sample rates; fields without enough sample support carry no Codex call. Candidate is the union of the two sides; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels.
Distilled classifier scores by category (both heads)
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.002 | 0.003 |
| Meta-epidemiology (narrow) | 0.001 | 0.000 |
| Meta-epidemiology (broad) | 0.000 | 0.001 |
| Bibliometrics | 0.001 | 0.001 |
| Science and technology studies | 0.000 | 0.000 |
| Scholarly communication | 0.001 | 0.001 |
| Open science | 0.000 | 0.000 |
| Research integrity | 0.001 | 0.000 |
| Insufficient payload (model declined to judge) | 0.003 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one source (direct Gemma or distilled Codex), not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".