Validation of a New Patient-Reported Outcome Measure of the Functional Impact of Essential Tremor on Activities of Daily Living
Bibliographic record
Abstract
Background: The Essential Tremor Rating Assessment Scale (TETRAS) is a popular scale for essential tremor (ET), but its activities of daily living (ADL) and performance (P) subscales are based on a structured interview and physical exam. No patient-reported outcome (PRO) scale for ET has been developed according to US regulatory guidelines. Objective: Develop and validate a TETRAS PRO subscale. Methods: Fourteen items, rated 0–4, were derived from TETRAS ADL and structured cognitive interviews of 18 ET patients. Convergent validity analyses of TETRAS PRO versus TETRAS ADL, TETRAS-P, and the Quality of Life in Essential Tremor Questionnaire (QUEST) were computed for 67 adults with ET or ET plus. Test-retest reliability was computed at intervals of 1 and 30 days. The influence of mood (Hospital Anxiety and Depression Scale, HADS) and coping behaviors (Essen Coping Questionnaire, ECQ) was examined with multiple linear regression. Results: TETRAS PRO was strongly correlated (r > 0.7) with TETRAS ADL, TETRAS-P, and QUEST and exhibited good to excellent reliability (Cronbach alpha 95%CI = 0.853–0.926; 30-day test-retest intraclass correlation 95%CI = 0.814–0.921). The 30-day estimate of minimum detectable change (MDC) was 6.6 (95%CI 5.2–8.0). TETRAS-P (rsemipartial = 0.607), HADS depression (rsemipartial = 0.384), and the coping strategy of information seeking and exchange of experiences (rsemipartial = 0.176) contributed statistically to TETRAS PRO in a multiple linear regression (R2 = 0.67). Conclusions: TETRAS PRO is a valid and reliable scale that is influenced strongly by tremor severity, moderately by mood (depression), and minimally by coping skills. The MDC for TETRAS PRO is probably sufficient to detect clinically important change.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame distilled prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.
Codex and Gemma teacher scores by category
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.000 | 0.000 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.000 | 0.000 |
| Bibliometrics | 0.000 | 0.000 |
| Science and technology studies | 0.000 | 0.000 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.000 | 0.000 |
| Research integrity | 0.000 | 0.000 |
| Insufficient payload (model declined to judge) | 0.000 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one teacher head, not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".