Deep tissue massage, strengthening and stretching exercises, and a combination of both compared with advice to stay active for subacute or persistent non-specific neck pain: A cost-effectiveness analysis of the Stockholm Neck trial (STONE)
Bibliographic record
Abstract
OBJECTIVE: To evaluate the cost-effectiveness of deep tissue massage ('massage'), strengthening and stretching exercises ('exercises') or a combination of both ('combined therapy') in comparison with advice to stay active ('advice') for subacute and persistent neck pain, from a societal perspective. METHODS: We conducted a cost-effectiveness analysis alongside a four-arm randomized controlled trial of 619 participants followed-up for one year. Health-related quality of life was measured using EQ-5D-3L and costs were calculated from baseline to one year. The interventions were ranked according to quality adjusted life years (QALYs) in a cost-consequence analysis. Thereafter, an incremental cost per QALY was calculated. RESULTS: In the cost-consequence analysis, in comparison with advice, exercises resulted in higher QALY gains, and massage and the combined therapy were more costly and less beneficial. Exercises may be a cost-effective treatment compared with advice to stay active if society is willing to pay 17 640 EUR per QALY. However, differences in QALY gains were minimal; on average, participants in the massage group, spent a year in a state of health valued at 0.88, exercises: 0.89, combined therapy: 0.88 and, advice: 0.88. CONCLUSIONS: Exercises are cost-effective compared to advice given that the societal willingness to pay is above 17 640 EUR per year in full health gained. Massage and a combined therapy are not cost-effective. While exercise appeared to have the best cost/benefit profile, even this treatment had only a modest benefit and treatment innovation is needed. Advice to stay active remains as a good therapeutic alternative from an economical perspective.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame machine prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. The Gemma side is a direct model label for every work in the frame, read from the title-only record. The Codex side is a classifier learned from the 10,348 direct Codex labels and calibrated to design-weighted sample rates; fields without enough sample support carry no Codex call. Candidate is the union of the two sides; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels.
Distilled classifier scores by category (both heads)
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.003 | 0.005 |
| Meta-epidemiology (narrow) | 0.002 | 0.001 |
| Meta-epidemiology (broad) | 0.005 | 0.008 |
| Bibliometrics | 0.001 | 0.001 |
| Science and technology studies | 0.000 | 0.001 |
| Scholarly communication | 0.001 | 0.001 |
| Open science | 0.001 | 0.001 |
| Research integrity | 0.002 | 0.002 |
| Insufficient payload (model declined to judge) | 0.007 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one source (direct Gemma or distilled Codex), not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".