Quality of written feedback given to medical students after introduction of real-time audio monitoring of clinical encounters
Bibliographic record
Abstract
BACKGROUND: Direct observation is necessary for specific and actionable feedback, however clinicians often struggle to integrate observation into their practice. Remotely audio-monitoring trainees for periods of time may improve the quality of written feedback given to them and may be a minimally disruptive task for a consultant to perform in a busy clinic. METHODS: Volunteer faculty used a wireless audio receiver during the second half of students' oncology rotations to listen to encounters during clinic in real time. They then gave written feedback as per usual practice, as did faculty who did not use the listening-in intervention. Feedback was de-identified and rated, using a rubric, as strong/medium/weak according to consensus of 2/3 rating investigators. RESULTS: Monitoring faculty indicated that audio monitoring made the feedback process easier and increased confidence in 95% of encounters. Most students (19/21 respondents) felt monitoring contributed positively to their learning and included more useful comments. 101 written evaluations were completed by 7 monitoring and 19 non-monitoring faculty. 22/23 (96%) of feedback after monitoring was rated as high quality, compared to 16/37 (43%) (p < 0.001) for monitoring faculty before using the equipment (and 20/78 (26%) without monitoring for all consultants (p < 0.001)). Qualitative analysis of student and faculty comments yielded prevalent themes of highly specific and actionable feedback given with greater frequency and more confidence on the part of the faculty if audio monitoring was used. CONCLUSIONS: Using live audio monitoring improved the quality of written feedback given to trainees, as judged by the trainees themselves and also using an exploratory grading rubric. The method was well received by both faculty and trainees. Although there are limitations compared to in-the-room observation (body language), the benefits of easy integration into clinical practice and a more natural patient encounter without the observer physically present lead the authors to now use this method routinely while teaching oncology students.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame machine prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. The Gemma side is a direct model label for every work in the frame, read from the title-only record. The Codex side is a classifier learned from the 10,348 direct Codex labels and calibrated to design-weighted sample rates; fields without enough sample support carry no Codex call. Candidate is the union of the two sides; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels.
Distilled classifier scores by category (both heads)
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.013 | 0.109 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.001 | 0.001 |
| Bibliometrics | 0.001 | 0.001 |
| Science and technology studies | 0.000 | 0.001 |
| Scholarly communication | 0.001 | 0.001 |
| Open science | 0.000 | 0.001 |
| Research integrity | 0.001 | 0.001 |
| Insufficient payload (model declined to judge) | 0.003 | 0.001 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one source (direct Gemma or distilled Codex), not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".