Ultrasound measures of muscle thickness may be superior to strength testing in adults with knee osteoarthritis: a cross-sectional study
Bibliographic record
Abstract
BACKGROUND: Evaluation of muscle strength as performed routinely with a dynamometer may be limited by important factors such as pain during muscle contraction. Few studies have compared formal strength testing with ultrasound to measure muscle bulk in adults with knee osteoarthritis (OA). METHODS: We investigated the muscle bulk of lower limb muscles in adults with knee OA using quantitative ultrasound. We analyzed the relationship between patient reported function and the muscle bulk of hip adductors, hip abductors, knee extensors and ankle plantarflexors. We further correlated muscle bulk measures with joint torques calculated with a hand held dynamometer. We hypothesized that ultrasound muscle bulk would have high levels of interrater reliability and correlate more strongly with pain and function than strength measured by a dynamometer. 23 subjects with unilateral symptomatic knee OA completed baseline questionnaires including the Western Ontario and McMaster Universities Arthritis Index (WOMAC) and Lower Extremity Activity Scale. Joint torque was measured with a dynamometer and muscle bulk was assessed with ultrasound. RESULTS: Higher ultrasound measured muscle bulk was correlated with less pain in all muscle groups. When comparing muscle bulk and torque measures, ultrasound-measured muscle bulk of the quadriceps was more strongly correlated with measures of pain and function than quadriceps isometric strength measured with a dynamometer. CONCLUSIONS: Ultrasound is a feasible method to assess muscle bulk of lower limb muscles in adults with knee OA, with high levels of interrater reliability, and correlates negatively with patient reported function. Compared with use of a hand held dynamometer to measure muscle function, ultrasound may be a superior modality.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame distilled prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.
Codex and Gemma teacher scores by category
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.000 | 0.001 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.000 | 0.000 |
| Bibliometrics | 0.000 | 0.001 |
| Science and technology studies | 0.000 | 0.000 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.000 | 0.000 |
| Research integrity | 0.000 | 0.000 |
| Insufficient payload (model declined to judge) | 0.000 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one teacher head, not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".