Wrist-Based Accelerometers and Visual Analog Scales as Outcome Measures for Shoulder Activity During Daily Living in Patients With Rotator Cuff Tendinopathy: Instrument Validation Study
Bibliographic record
Abstract
BACKGROUND: Shoulder pain secondary to rotator cuff tendinopathy affects a large proportion of patients in orthopedic surgery practices. Corticosteroid injections are a common intervention proposed for these patients. The clinical evaluation of a response to corticosteroid injections is usually based only on the patient's self-evaluation of his function, activity, and pain by multiple questionnaires with varying metrological qualities. Objective measures of upper extremity functions are lacking, but wearable sensors are emerging as potential tools to assess upper extremity function and activity. OBJECTIVE: This study aimed (1) to evaluate and compare test-retest reliability and sensitivity to change of known clinical assessments of shoulder function to wrist-based accelerometer measures and visual analog scales (VAS) of shoulder activity during daily living in patients with rotator cuff tendinopathy convergent validity and (2) to determine the acceptability and compliance of using wrist-based wearable sensors. METHODS: A total of 38 patients affected by rotator cuff tendinopathy wore wrist accelerometers on the affected side for a total of 5 weeks. Western Ontario Rotator Cuff (WORC) index; Short version of the Disability of the Arm, Shoulder, and Hand questionnaire (QuickDASH); and clinical examination (range of motion and strength) were performed the week before the corticosteroid injections, the day of the corticosteroid injections, and 2 and 4 weeks after the corticosteroid injections. Daily Single Assessment Numeric Evaluation (SANE) and VAS were filled by participants to record shoulder pain and activity. Accelerometer data were processed to extract daily upper extremity activity in the form of active time; activity counts; and ratio of low-intensity activities, medium-intensity activities, and high-intensity activities. RESULTS: Daily pain measured using VAS and SANE correlated well with the WORC and QuickDASH questionnaires (r=0.564-0.815) but not with accelerometry measures, amplitude, and strength. Daily activity measured with VAS had good correlation with active time (r=0.484, P=.02). All questionnaires had excellent test-retest reliability at 1 week before corticosteroid injections (intraclass correlation coefficient [ICC]=0.883-0.950). Acceptable reliability was observed with accelerometry (ICC=0.621-0.724), apart from low-intensity activities (ICC=0.104). Sensitivity to change was excellent at 2 and 4 weeks for all questionnaires (standardized response mean=1.039-2.094) except for activity VAS (standardized response mean=0.50). Accelerometry measures had low sensitivity to change at 2 weeks, but excellent sensitivity at 4 weeks (standardized response mean=0.803-1.032). CONCLUSIONS: Daily pain VAS and SANE had good correlation with the validated questionnaires, excellent reliability at 1 week, and excellent sensitivity to change at 2 and 4 weeks. Daily activity VAS and accelerometry-derived active time correlated well together. Activity VAS had excellent reliability, but moderate sensitivity to change. Accelerometry measures had moderate reliability and acceptable sensitivity to change at 4 weeks.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame machine prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. The Gemma side is a direct model label for every work in the frame, read from the title-only record. The Codex side is a classifier learned from the 10,348 direct Codex labels and calibrated to design-weighted sample rates; fields without enough sample support carry no Codex call. Candidate is the union of the two sides; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels.
Distilled classifier scores by category (both heads)
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.014 | 0.025 |
| Meta-epidemiology (narrow) | 0.001 | 0.000 |
| Meta-epidemiology (broad) | 0.001 | 0.001 |
| Bibliometrics | 0.001 | 0.001 |
| Science and technology studies | 0.000 | 0.001 |
| Scholarly communication | 0.001 | 0.001 |
| Open science | 0.001 | 0.001 |
| Research integrity | 0.001 | 0.001 |
| Insufficient payload (model declined to judge) | 0.001 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one source (direct Gemma or distilled Codex), not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".