Physical Examination Tests Are Not Valid for Diagnosing SLAP Tears: A Review
Bibliographic record
Abstract
OBJECTIVE: To critically evaluate the evidence for the use of physical examination procedures for diagnosing superior labrum anterior posterior (SLAP) lesions, by means of a systematic review. DATA SOURCES: MEDLINE, EMBASE, and The Cochrane data bases were searched for studies published between January 1970 and June 2004, in 3 stages, using SLAP lesion; arthroscopy, shoulder joint and athletic injuries combined with testing and physical examination; and arthroscopy, shoulder joint and athletic injuries combined with sensitivity and specificity (total yield, 260 articles). Additional studies were sought in the reference lists of relevant articles. STUDY SELECTION: Potentially relevant abstracts were selected from the 3 search strategies. Studies were included if they focused on physical examination of SLAP lesions and presented original data on the accuracy of the test. Of 29 potentially relevant studies, 15 were selected when the full text was reviewed. DATA EXTRACTION: Information on the number of participants, the study design, the physical test(s) evaluated, and the sensitivity, specificity, and positive and negative predictive value of the tests were extracted or calculated. Study validity was evaluated (1-5 points: independent, blind comparison with a reference standard; inclusion of an appropriate spectrum of patients; all participants were assessed by the reference standard; replicable description of the test; and likelihood ratios presented or calculable). MAIN RESULTS: The physical tests included from 1 to 6 of the anterior slide test, SLAPprehension test, biceps load tests, crank test, O'Brien test, active compression, compression rotation, Speed's test, Yergason's test, Jobe test, bicippital groove pain, and pain provocation. The only study that passed all 5 methods criteria found that Speed's test and Yergason's test had sensitivity of 32% and 43%, and specificity of 79% and 75%, respectively; thus, the positive and negative likelihood ratios for Speed's test were 1.2727 and 0.9091 and for Yergason's test were 2.000 and 0.7272. The confidence intervals for the likelihood ratios all included 1.0. Whereas the test descriptions in the other reports were generally clear, only 7 of the other 14 studies passed 1 further methods criterion. Nine of these studies reported sensitivities and specificities for the physical tests of >75%. CONCLUSION: The accuracy of Speed's and Yergason's tests for diagnosing a SLAP lesion was poor in the only methodologically robust study reviewed. The likelihood ratios for these tests could not rule in, or rule out, the presence of a SLAP lesion when compared with arthroscopic results. Assessments of numerous other tests could not be considered valid because of the serious shortcomings in the studies' methods.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame machine prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. The Gemma side is a direct model label for every work in the frame, read from the title-only record. The Codex side is a classifier learned from the 10,348 direct Codex labels and calibrated to design-weighted sample rates; fields without enough sample support carry no Codex call. Candidate is the union of the two sides; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels.
Distilled classifier scores by category (both heads)
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.024 | 0.155 |
| Meta-epidemiology (narrow) | 0.002 | 0.002 |
| Meta-epidemiology (broad) | 0.012 | 0.008 |
| Bibliometrics | 0.018 | 0.012 |
| Science and technology studies | 0.001 | 0.002 |
| Scholarly communication | 0.004 | 0.004 |
| Open science | 0.003 | 0.002 |
| Research integrity | 0.004 | 0.001 |
| Insufficient payload (model declined to judge) | 0.002 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one source (direct Gemma or distilled Codex), not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".