Competence of Senior Otolaryngology Residents with the Bedside Head Impulse Test—Has There Been Improvement After 5 Years of Competency By Design?
Bibliographic record
Abstract
Background The bedside head impulse test (bHIT) is a clinical method of assessing the vestibulo-ocular reflex. It is a critical component of the bedside assessment of dizzy patients and helps differentiate acute stroke from vestibular neuritis. A previous study on senior Otolaryngology residents showed poor competence in performing and interpreting the bHIT and called for specific evaluations in the Competency By Design (CBD) curriculum to remedy this. This study aimed to assess whether those competencies have improved after full implementation of CBD in residency programs. Methods Thirty post-graduate year 4 Otolaryngology residents in Canada were evaluated on the use of the bHIT using a written multiple-choice question (MCQ) examination, interpretation of bHIT videos, and performance of a bHIT. Ratings of bHIT performance were completed by 2 expert examiners (DT, DL) using the Ottawa Clinic Assessment Tool. Results Only 6.7% (rater DT) and 20% (rater DL) of residents were found able to perform the bHIT independently. Inter-rater reliability was moderate (0.55, intraclass correlation). Mean scores were 70% (13.4% standard deviation) for video interpretation and 59% (20.6% standard deviation) for multiple-choice questions. Video interpretation scores did not correlate with bHIT ratings (Pearson r = 0.11), but MCQs and bHIT ratings did correlate moderately (Pearson r = 0.52). Comparing to the prior study, residents performed worse on the bHIT (3.14 average score vs 3.64, P < .01) and fewer residents performed the bHIT independently (6.7% vs 22%—rater DT, 20% vs 39%—rater DL). Residents also performed worse on MCQs (58.7% vs 70.9%, P = 0.038), though similarly on video interpretation (70% vs 65%, P = .198). Conclusion Fourth year OTL-HNS residents in Canada are not competent in performing the bHIT. These findings have implications for refining competency-based curricula in the evaluation of critical physical exam skills.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame machine prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. The Gemma side is a direct model label for every work in the frame, read from the title-only record. The Codex side is a classifier learned from the 10,348 direct Codex labels and calibrated to design-weighted sample rates; fields without enough sample support carry no Codex call. Candidate is the union of the two sides; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels.
Distilled classifier scores by category (both heads)
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.002 | 0.010 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.000 | 0.000 |
| Bibliometrics | 0.001 | 0.000 |
| Science and technology studies | 0.000 | 0.001 |
| Scholarly communication | 0.001 | 0.001 |
| Open science | 0.000 | 0.001 |
| Research integrity | 0.001 | 0.000 |
| Insufficient payload (model declined to judge) | 0.001 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one source (direct Gemma or distilled Codex), not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".