Point of care biliary ultrasound in the emergency department (BUSED) predicts final surgical management decisions
Bibliographic record
Abstract
Objectives: Gallstone disease is a common reason for emergency department (ED) presentation. Surgeons often prefer radiology department ultrasound (RUS) over point of care ultrasound (POCUS) because of perceived of unreliability. Our study was designed to test the hypothesis that POCUS is sufficient to guide the management of surgeons treating select cases of biliary disease as compared to RUS. Methods: This was a prospective cohort study. Patients who presented to the ED with abdominal pain and findings of biliary disease on POCUS were included. The surgeon was then presented the case with POCUS only and recorded their management decision. Patients then proceeded to RUS, were followed through their stay, and analysis was performed to analyze the proportion of patients where the introduction of the RUS changed the management plan. Results: 100 patients were included in this study, and all received both POCUS and RUS. Depending on the surgeons' POCUS based management decisions, the patients were divided into three groups: (1) surgery, (2) duct clearance, (3) no surgery. Total bilirubin was 34±22 mmol/L in the duct clearance group vs 8.4±6.5 mmol/L and 16±12 mmol/L in the surgery and no surgery groups, respectively (p<0.05). POCUS results showed 68 patients would have been offered surgery, 21 offered duct clearance, and 11 no surgery. In 90% of cases, the introduction of RUS did not change management. The acute care surgeons elected to operate on patients more frequently than other surgical subspecialties (p<0.05). Conclusions: This study showed that fewer than 10% of patients with biliary disease seen on POCUS had a change in surgical decision-making based on the addition of RUS imaging. In uncomplicated cases of biliary disease, relying on POCUS imaging for surgical decision-making has the potential to improve patient flow. Level of evidence: II Prospective Cohort Study.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame distilled prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.
Codex and Gemma teacher scores by category
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.001 | 0.000 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.001 | 0.000 |
| Bibliometrics | 0.000 | 0.001 |
| Science and technology studies | 0.000 | 0.000 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.001 | 0.000 |
| Research integrity | 0.000 | 0.000 |
| Insufficient payload (model declined to judge) | 0.003 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one teacher head, not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".