On the reliability and validity of manual muscle testing: a literature review
Bibliographic record
Abstract
INTRODUCTION: A body of basic science and clinical research has been generated on the manual muscle test (MMT) since its first peer-reviewed publication in 1915. The aim of this report is to provide an historical overview, literature review, description, synthesis and critique of the reliability and validity of MMT in the evaluation of the musculoskeletal and nervous systems. METHODS: Online resources were searched including Pubmed and CINAHL (each from inception to June 2006). The search terms manual muscle testing or manual muscle test were used. Relevant peer-reviewed studies, commentaries, and reviews were selected. The two reviewers assessed data quality independently, with selection standards based on predefined methodologic criteria. Studies of MMT were categorized by research content type: inter- and intraexaminer reliability studies, and construct, content, concurrent and predictive validity studies. Each study was reviewed in terms of its quality and contribution to knowledge regarding MMT, and its findings presented. RESULTS: More than 100 studies related to MMT and the applied kinesiology chiropractic technique (AK) that employs MMT in its methodology were reviewed, including studies on the clinical efficacy of MMT in the diagnosis of patients with symptomatology. With regard to analysis there is evidence for good reliability and validity in the use of MMT for patients with neuromusculoskeletal dysfunction. The observational cohort studies demonstrated good external and internal validity, and the 12 randomized controlled trials (RCTs) that were reviewed show that MMT findings were not dependent upon examiner bias. CONCLUSION: The MMT employed by chiropractors, physical therapists, and neurologists was shown to be a clinically useful tool, but its ultimate scientific validation and application requires testing that employs sophisticated research models in the areas of neurophysiology, biomechanics, RCTs, and statistical analysis.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame machine prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. The Gemma side is a direct model label for every work in the frame, read from the title-only record. The Codex side is a classifier learned from the 10,348 direct Codex labels and calibrated to design-weighted sample rates; fields without enough sample support carry no Codex call. Candidate is the union of the two sides; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels.
Distilled classifier scores by category (both heads)
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.041 | 0.147 |
| Meta-epidemiology (narrow) | 0.002 | 0.002 |
| Meta-epidemiology (broad) | 0.007 | 0.005 |
| Bibliometrics | 0.034 | 0.031 |
| Science and technology studies | 0.001 | 0.003 |
| Scholarly communication | 0.004 | 0.007 |
| Open science | 0.004 | 0.002 |
| Research integrity | 0.003 | 0.002 |
| Insufficient payload (model declined to judge) | 0.003 | 0.001 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one source (direct Gemma or distilled Codex), not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".