Development and Testing of Reduced Versions of the Manual Muscle Test-8 in Juvenile Dermatomyositis
Bibliographic record
Abstract
Objective. To develop and test shortened versions of the Manual Muscle Test-8 (MMT-8) in juvenile dermatomyositis (JDM). Methods. Construction of reduced tools was based on a retrospective analysis of individual scores of MMT-8 muscle groups in 3 multinational datasets. The 4 and 6 most frequently impaired muscle groups were included in MMT-4 and MMT-6, respectively. Metrologic properties of reduced tools were assessed by evaluating construct validity, internal consistency, discriminant ability, and responsiveness to change. Results. Neck flexors, hip extensors, hip abductors, and shoulder abductors were included in MMT-4, whereas MMT-6 also included elbow flexors and hip flexors. Both shortened tools revealed strong correlations with MMT-8 and other muscle strength measures. Correlations with other JDM outcome measures were in line with predictions. Internal consistency was good (0.88–0.96) for both MMT-4 and MMT-6. Both reduced tools showed strong ability to discriminate between disease activity states, assessed by the caring physician or a parent ( P < 0.001), and between patients whose parents were satisfied or not satisfied with illness course ( P < 0.001). Responsiveness to change (assessed by both standardized response mean and relative efficiency) of MMT-4 and, to a lesser degree, MMT-6, was slightly superior to that of MMT-8. Conclusion. Overall, the metrologic performance of MMT-4 and MMT-6 was comparable to that of the other established muscle strength tools, which indicates that they may be suitable for use in clinical practice and research, including clinical trials. The measurement properties of these tools should be further tested in other patient populations and evaluated prospectively.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame distilled prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.
Codex and Gemma teacher scores by category
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.000 | 0.001 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.000 | 0.000 |
| Bibliometrics | 0.000 | 0.000 |
| Science and technology studies | 0.000 | 0.000 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.000 | 0.000 |
| Research integrity | 0.000 | 0.000 |
| Insufficient payload (model declined to judge) | 0.000 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one teacher head, not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".