International consensus on the most useful assessments used by physical therapists to evaluate patients with temporomandibular disorders: A Delphi study
Bibliographic record
Abstract
OBJECTIVE: To identify assessment tools used to evaluate patients with temporomandibular disorders (TMD) considered to be clinically most useful by a panel of international experts in TMD physical therapy (PT). METHODS: A Delphi survey method administered to a panel of international experts in TMD PT was conducted over three rounds from October 2017 to June 2018. The initial contact was made by email. Participation was voluntary. An e-survey, according to the Checklist for Reporting Results of Internet E-Surveys (CHERRIES), was posted using SurveyMonkey for each round. Percentages of responses were analysed for each question from each round of the Delphi survey administrations. RESULTS: Twenty-three experts (completion rate: 23/25) completed all three rounds of the survey for three clinical test categories: 1) questionnaires, 2) pain screening tools and 3) physical examination tests. The following was the consensus-based decision regarding the identification of the clinically most useful assessments. (1) Four of 9 questionnaires were identified: Jaw Functional Limitation (JFL-8), Mandibular Function Impairment Questionnaire (MFIQ), Tampa Scale for Kinesiophobia for Temporomandibular disorders (TSK/TMD) and the neck disability index (NDI). (2) Three of 8 identified pain screening tests: visual analog scale (VAS), numeric pain rating scale (NRS) and pain during mandibular movements. (3) Eight of 18 identified physical examination tests: physiological temporomandibular joint (TMJ) movements, trigger point (TrP) palpation of the masticatory muscles, TrP palpation away from the masticatory system, accessory movements, articular palpation, noise detection during movement, manual screening of the cervical spine and the Neck Flexor Muscle Endurance Test. CONCLUSION: After three rounds in this Delphi survey, the results of the most used assessment tools by TMD PT experts were established. They proved to be founded on test construct, test psychometric properties (reliability/validity) and expert preference for test clusters. A concordance with the screening tools of the diagnostic criteria of TMD consortium was noted. Findings may be used to guide policymaking purposes and future diagnostic research.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame machine prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. The Gemma side is a direct model label for every work in the frame, read from the title-only record. The Codex side is a classifier learned from the 10,348 direct Codex labels and calibrated to design-weighted sample rates; fields without enough sample support carry no Codex call. Candidate is the union of the two sides; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels.
Distilled classifier scores by category (both heads)
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.158 | 0.141 |
| Meta-epidemiology (narrow) | 0.001 | 0.001 |
| Meta-epidemiology (broad) | 0.001 | 0.002 |
| Bibliometrics | 0.005 | 0.003 |
| Science and technology studies | 0.003 | 0.003 |
| Scholarly communication | 0.003 | 0.003 |
| Open science | 0.002 | 0.010 |
| Research integrity | 0.003 | 0.003 |
| Insufficient payload (model declined to judge) | 0.002 | 0.001 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one source (direct Gemma or distilled Codex), not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".