An Individualized Postoperative Pain Risk Communication Tool for Use in Pediatric Surgery: Co-Design and Usability Evaluation
Bibliographic record
Abstract
BACKGROUND: Risk identification and communication tools have the potential to improve health care by supporting clinician-patient or family discussion of treatment risks and benefits and helping patients make more informed decisions; however, they have yet to be tailored to pediatric surgery. User-centered design principles can help to ensure the successful development and uptake of health care tools. OBJECTIVE: We aimed to develop and evaluate the usability of an easy-to-use tool to communicate a child's risk of postoperative pain to improve informed and collaborative preoperative decision-making between clinicians and families. METHODS: With research ethics board approval, we conducted web-based co-design sessions with clinicians and family participants (people with lived surgical experience and parents of children who had recently undergone a surgical or medical procedure) at a tertiary pediatric hospital. Qualitative data from these sessions were analyzed thematically using NVivo (Lumivero) to identify design requirements to inform the iterative redesign of an existing prototype. We then evaluated the usability of our final prototype in one-to-one sessions with a new group of participants, in which we measured mental workload with the National Aeronautics and Space Administration (NASA) Task Load Index (TLX) and user satisfaction with the Post-Study System Usability Questionnaire (PSSUQ). RESULTS: A total of 12 participants (8 clinicians and 4 family participants) attended 5 co-design sessions. The 5 requirements were identified: (A) present risk severity descriptively and visually; (B) ensure appearance and navigation are user-friendly; (C) frame risk identification and mitigation strategies in positive terms; (D) categorize and describe risks clearly; and (E) emphasize collaboration and effective communication. A total of 12 new participants (7 clinicians and 5 family participants) completed a usability evaluation. Tasks were completed quickly (range 5-17 s) and accurately (range 11/12, 92% to 12/12, 100%), needing only 2 requests for assistance. The median (IQR) NASA TLX performance score of 78 (66-89) indicated that participants felt able to perform the required tasks, and an overall PSSUQ score of 2.1 (IQR 1.5-2.7) suggested acceptable user satisfaction with the tool. CONCLUSIONS: The key design requirements were identified, and that guided the prototype redesign, which was positively evaluated during usability testing. Implementing a personalized risk communication tool into pediatric surgery can enhance the care process and improve informed and collaborative presurgical preparation and decision-making between clinicians and families of pediatric patients.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame machine prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. The Gemma side is a direct model label for every work in the frame, read from the title-only record. The Codex side is a classifier learned from the 10,348 direct Codex labels and calibrated to design-weighted sample rates; fields without enough sample support carry no Codex call. Candidate is the union of the two sides; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels.
Distilled classifier scores by category (both heads)
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.026 | 0.049 |
| Meta-epidemiology (narrow) | 0.001 | 0.001 |
| Meta-epidemiology (broad) | 0.001 | 0.001 |
| Bibliometrics | 0.002 | 0.001 |
| Science and technology studies | 0.001 | 0.001 |
| Scholarly communication | 0.002 | 0.002 |
| Open science | 0.002 | 0.002 |
| Research integrity | 0.001 | 0.001 |
| Insufficient payload (model declined to judge) | 0.002 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one source (direct Gemma or distilled Codex), not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".