Reliability, validity and responsiveness of physical activity monitors in patients with inflammatory myopathy
Bibliographic record
Abstract
OBJECTIVE: Idiopathic inflammatory myopathies (IIMs) cause proximal muscle weakness, which affects the ability to carry out the activities of daily living. Wearable physical activity monitors (PAMs) objectively assess continuous activity and potentially have clinical usefulness in the assessment of IIMs. We examined the psychometric characteristics for PAM outcomes in IIMs. METHODS: Adult IIM patients were prospectively evaluated (at baseline, 3 months and 6 months) in an observational study. A waist-worn PAM (ActiGraph GT3X-BT) assessed average step counts/minute, peak 1-minute cadence, and vector magnitude/minute. Validated myositis core set measures (CSMs) including manual muscle testing (MMT), physician global disease activity (MD global), patient global disease activity (Pt global), extramuscular disease activity (Ex-muscular global), HAQ-DI (HAQ disability index), muscle enzymes, and patient-reported physical function were evaluated. Test-retest reliability, construct validity, and responsiveness were determined for PAM measures and CSMs, using Pearson correlations and other appropriate analyses. RESULTS: A total of 50 adult IIM patients enrolled [mean (s.d.) age, 53.6 (14.6); 60% female, 94% Caucasian]. PAM measures showed strong test-retest reliability, moderate-to-strong correlations at baseline with MD global (r = -0.37 to -0.48), Pt global (r=-0.43 to -0.61), HAQ-DI (r = -0.47 to -0.59) and MMT (r = 0.37-0.52), and strong discriminant validity for categorical MMT and HAQ-DI. Longitudinal associations with MD global (r=-0.38 to -0.44), MMT (r = 0.50-0.57), HAQ-DI (r = -0.45 to -0.55) and functional tests (r = 0.30-0.65) were moderate to strong. PAM measures were responsive to MMT improvement ≥10% and moderate-to-major improvement on ACR/EULAR myositis response criteria. Peak 1-minute cadence had the largest effect size and standardized response means. CONCLUSION: PAM measures showed promising construct validity, reliability, and longitudinal responsiveness; especially peak 1-minute cadence. PAMs are able to provide valid outcome measures for future use in IIM clinical trials.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame distilled prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.
Codex and Gemma teacher scores by category
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.000 | 0.000 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.000 | 0.000 |
| Bibliometrics | 0.000 | 0.000 |
| Science and technology studies | 0.000 | 0.000 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.000 | 0.000 |
| Research integrity | 0.000 | 0.000 |
| Insufficient payload (model declined to judge) | 0.000 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one teacher head, not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".