Long-Term Trajectories of Patient-Reported Outcomes Following Total Knee Arthroplasty
Bibliographic record
Abstract
BACKGROUND: Although total knee arthroplasty (TKA) is known to improve patient-reported outcome measure (PROM) scores in the short term to midterm, the long-term trajectories of both disease-specific and generic PROM scores remain unclear. METHODS: We retrospectively analyzed the prospectively collected registry data of 1,264 patients (mean age, 68.5 years; 93.7% female) who underwent primary TKA for osteoarthritis between 2005 and 2013 and completed PROM assessments at baseline and 10 years postoperatively. Disease-specific PROMs were assessed using the Knee Society Knee Score (KSKS), Knee Society Function Score (KSFS), and the Western Ontario and McMaster Universities Osteoarthritis Index (WOMAC). Generic PROMs were assessed using the Short Form-36 Health Survey (SF-36). Assessments were performed preoperatively and at 6 months and 1, 2, 5, 10, and 15 years postoperatively. Generalized linear models and linear mixed-effects models were used to evaluate temporal changes and subgroup differences by age and sex. RESULTS: All PROM scores improved significantly within 6 months after TKA. Thereafter, disease-specific PROMs showed modest changes up to 1 year, with relative stability until 5 years, whereas generic PROMs demonstrated heterogeneous patterns across different domains. Between 5 and 10 years postoperatively, WOMAC pain and stiffness scores did not show significant changes, the KSKS decreased but not significantly so, and WOMAC physical function scores exhibited small but significant deterioration that was not clinically meaningful. SF-36 domains demonstrated varied trajectories: physical and mental component scores declined by more than the minimal clinically important difference after 5 years, whereas the social functioning score showed continuous improvement, although not all changes were significant. Octogenarians demonstrated lower physical functioning scores but higher social functioning scores in the long term compared with younger patients, and female patients demonstrated inferior functional and vitality scores compared with male patients. CONCLUSIONS: Both disease-specific and generic PROM scores after TKA improved significantly and remained superior to baseline scores over a 15-year period, although physical function scores tended to decline in the long term. In this large, predominantly female Korean cohort, the distinct age- and sex-specific trajectories highlight the importance of implementing individualized, time-adapted long-term management strategies to optimize patient outcomes. LEVEL OF EVIDENCE: Therapeutic Level III . See Instructions for Authors for a complete description of levels of evidence.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame machine prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. The Gemma side is a direct model label for every work in the frame, read from the title-only record. The Codex side is a classifier learned from the 10,348 direct Codex labels and calibrated to design-weighted sample rates; fields without enough sample support carry no Codex call. Candidate is the union of the two sides; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels.
Distilled classifier scores by category (both heads)
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.004 | 0.016 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.000 | 0.001 |
| Bibliometrics | 0.001 | 0.002 |
| Science and technology studies | 0.000 | 0.000 |
| Scholarly communication | 0.001 | 0.001 |
| Open science | 0.001 | 0.001 |
| Research integrity | 0.001 | 0.001 |
| Insufficient payload (model declined to judge) | 0.001 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one source (direct Gemma or distilled Codex), not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".