Diagnostic criteria and symptom grading for delayed gastric conduit emptying after esophagectomy for cancer: international expert consensus based on a modified Delphi process
Bibliographic record
Abstract
Delayed gastric conduit emptying (DGCE) after esophagectomy for cancer is associated with adverse outcomes and troubling symptoms. Widely accepted diagnostic criteria and a symptom grading tool for DGCE are missing. This hampers the interpretation and comparison of studies. A modified Delphi process, using repeated web-based questionnaires, combined with live interim group discussions was conducted by 33 experts within the field, from Europe, North America, and Asia. DGCE was divided into early DGCE if present within 14 days of surgery and late if present later than 14 days after surgery. The final criteria for early DGCE, accepted by 25 of 27 (93%) experts, were as follows: >500 mL diurnal nasogastric tube output measured on the morning of postoperative day 5 or later or >100% increased gastric tube width on frontal chest x-ray projection together with the presence of an air-fluid level. The final criteria for late DGCE accepted by 89% of the experts were as follows: the patient should have 'quite a bit' or 'very much' of at least two of the following symptoms; early satiety/fullness, vomiting, nausea, regurgitation or inability to meet caloric need by oral intake and delayed contrast passage on upper gastrointestinal water-soluble contrast radiogram or on timed barium swallow. A symptom grading tool for late DGCE was constructed grading each symptom as: 'not at all', 'a little', 'quite a bit', or 'very much', generating 0, 1, 2, or 3 points, respectively. For the five symptoms retained in the diagnostic criteria for late DGCE, the minimum score would be 0, and the maximum score would be 15. The final symptom grading tool for late DGCE was accepted by 27 of 31 (87%) experts. For the first time, diagnostic criteria for early and late DGCE and a symptom grading tool for late DGCE are available, based on an international expert consensus process.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame distilled prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.
Codex and Gemma teacher scores by category
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.000 | 0.001 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.000 | 0.000 |
| Bibliometrics | 0.000 | 0.000 |
| Science and technology studies | 0.000 | 0.000 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.000 | 0.000 |
| Research integrity | 0.000 | 0.000 |
| Insufficient payload (model declined to judge) | 0.000 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one teacher head, not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".