Point-of-care haemoglobin accuracy and transfusion outcomes in non-cardiac surgery at a Canadian tertiary academic hospital: protocol for the PREMISE observational study
Bibliographic record
Abstract
INTRODUCTION: Transfusions in surgery can be life-saving interventions, but inappropriate transfusions may lack clinical benefit and cause harm. Transfusion decision-making in surgery is complex and frequently informed by haemoglobin (Hgb) measurement in the operating room. Point-of-care testing for haemoglobin (POCT-Hgb) is increasingly relied on given its simplicity and rapid provision of results. POCT-Hgb devices lack adequate validation in the operative setting, particularly for Hgb values within the transfusion zone (60-100 g/L). This study aims to examine the accuracy of intraoperative POCT-Hgb instruments in non-cardiac surgery, and the association between POCT-Hgb measurements and transfusion decision-making. METHODS AND ANALYSIS: PREMISE is an observational prospective method comparison study. Enrolment will occur when adult patients undergoing major non-cardiac surgery require POCT-Hgb, as determined by the treating team. Three concurrent POCT-Hgb results, considered as index tests, will be compared with a laboratory analysis of Hgb (lab-Hgb), considered the gold standard. Participants may have multiple POCT-Hgb measurements during surgery. The primary outcome is the difference in individual Hgb measurements between POCT-Hgb and lab-Hgb, primarily among measurements that are within the transfusion zone. Secondary outcomes include POCT-Hgb accuracy within the entire cohort, postoperative morbidity, mortality and transfusion rates. The sample size is 1750 POCT-Hgb measurements to obtain a minimum of 652 Hgb measurements <100 g/L, based on an estimated incidence of 38%. The sample size was calculated to fit a logistic regression model to predict instances when POCT-Hgb are inaccurate, using 4 g/L as an acceptable margin of error. ETHICS AND DISSEMINATION: Institutional ethics approval has been obtained by the Ottawa Health Science Network-Research Ethics Board prior to initiating the study. Findings from this study will be published in peer-reviewed journals and presented at relevant scientific conferences. Social media will be leveraged to further disseminate the study results and engage with clinicians.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Direct model labels (unvalidated)
Per-model category and study-design labels from the labeling rounds. They are machine output, unvalidated, and the disagreement between models ships as data. No study design here is MEDLINE-validated yet.
| Model arm | Categories | Study design | Confidence |
|---|---|---|---|
| gemma | no category Domain: not available · Genre: Protocol About the Canadian research system: no · About a Canadian topic: no | Observational | high |
| gpt | no category Domain: not available · Genre: Protocol About the Canadian research system: no · About a Canadian topic: no | Observational | high |
Full frame machine prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. The Gemma side is a direct model label for every work in the frame, read from the title-only record. The Codex side is a classifier learned from the 10,348 direct Codex labels and calibrated to design-weighted sample rates; fields without enough sample support carry no Codex call. Candidate is the union of the two sides; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels.
Distilled classifier scores by category (both heads)
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.043 | 0.052 |
| Meta-epidemiology (narrow) | 0.003 | 0.002 |
| Meta-epidemiology (broad) | 0.004 | 0.005 |
| Bibliometrics | 0.003 | 0.005 |
| Science and technology studies | 0.005 | 0.003 |
| Scholarly communication | 0.004 | 0.002 |
| Open science | 0.003 | 0.002 |
| Research integrity | 0.003 | 0.003 |
| Insufficient payload (model declined to judge) | 0.031 | 0.004 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedLabeled directly by 2 models reading the full record.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".