Clinical Pressure Pain Threshold Testing in Neck Pain: Comparing Protocols, Responsiveness, and Association With Psychological Variables
Bibliographic record
Abstract
BACKGROUND: Quantitative sensory testing, including pressure pain threshold (PPT), is seeing increased use in clinical practice. In order to facilitate clinical utility, knowledge of the properties of the tool and interpretation of results are required. OBJECTIVES: This observational study used a clinical sample of people with mechanical neck pain to determine: (1) the influence of number of testing repetitions on measurement properties, (2) reliability and minimum clinically important difference, and (3) associations between PPT and key psychological constructs. DESIGN: This study was observational with both cross-sectional and prospective elements. METHODS: Experienced clinicians measured PPT in patients with mechanical neck pain following a standardized protocol. Subcohorts also provided repeated measures and completed scales of key psychological constructs. RESULTS: The total sample was 206 participants, but not all participants provided data for all analyses. Interrater and 1-week test-retest reliability were excellent (intraclass correlation coefficients [2,1]=.75-.95). Potentially important differences in reliability and PPT scores were found when using only 1 or 2 repeated measures compared with all 3. The PPT over a distal location (tibialis anterior muscle) was not adequately responsive in this sample, but the local site (upper trapezius muscle) was responsive and may be useful as part of a protocol to evaluate clinical change. Sensitivity values (range=0.08-0.50) and specificity values (range=0.82-0.97) for a range of change scores are presented. Depression, catastrophizing, and kinesiophobia were able to explain small but statistically significant variance in local PPT (3.9%-5.9%), but only catastrophizing and kinesiophobia explained significant variance in the distal PPT (3.6% and 2.9%, respectively). LIMITATIONS: Limitations of the study include multiple raters, unknown recruitment rates, and unknown measurement properties at sites other than those tested here. CONCLUSIONS: The results suggest that PPT is adequately reliable and that 3 measurements should be taken to maximize measurement properties. The variance explained by the psychological variables was small but significant for 3 constructs related to catastrophizing, depression, and fear of movement. Clinical implications for application and interpretation of PPT are discussed.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame machine prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. The Gemma side is a direct model label for every work in the frame, read from the title-only record. The Codex side is a classifier learned from the 10,348 direct Codex labels and calibrated to design-weighted sample rates; fields without enough sample support carry no Codex call. Candidate is the union of the two sides; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels.
Distilled classifier scores by category (both heads)
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.094 | 0.163 |
| Meta-epidemiology (narrow) | 0.001 | 0.000 |
| Meta-epidemiology (broad) | 0.001 | 0.001 |
| Bibliometrics | 0.001 | 0.001 |
| Science and technology studies | 0.000 | 0.001 |
| Scholarly communication | 0.001 | 0.001 |
| Open science | 0.001 | 0.001 |
| Research integrity | 0.001 | 0.001 |
| Insufficient payload (model declined to judge) | 0.000 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one source (direct Gemma or distilled Codex), not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".