Development of the ParaOesophageal hernia SympTom (POST) tool
Bibliographic record
Abstract
BACKGROUND: The aim of this study was to develop a symptom severity instrument (ParaOesophageal hernia SympTom (POST) tool) specific to para-oesophageal hernia (POH). METHODS: The POST tool was developed in four stages. The first was establishment of a Steering Committee. In the second stage, items were generated through a systematic review and online scoping survey of international experts. In the third stage, a three-round modified Delphi consensus process was conducted with a group of international experts who were asked to rate the importance of candidate items. An a priori threshold for inclusion was set at 80 per cent. The modified Delphi process culminated in a consensus meeting to develop the first iteration of the tool. In the final stage, two international patient workshops were held to assess the content validity and acceptability of the POST tool. RESULTS: The systematic review and scoping survey generated 64 symptoms, refined to 20 for inclusion in the modified Delphi consensus process. Twenty-six global experts participated in the Delphi consensus process. Five symptoms reached consensus across two rounds: difficulty getting solid foods down, chest pain after meals, difficulty getting liquids down, shortness of breath only after meals, and an early feeling of fullness after eating. The subsequent patient workshops deemed these five symptoms to be relevant and suggested that reflux should be included; these were taken forward to create the final POST tool. CONCLUSION: The POST tool is the first instrument designed to capture POH-specific symptoms. It will allow clinicians to standardize reporting of symptoms of POH and evaluate the response to surgical intervention.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame distilled prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.
Codex and Gemma teacher scores by category
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.001 | 0.000 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.003 | 0.002 |
| Bibliometrics | 0.000 | 0.000 |
| Science and technology studies | 0.000 | 0.000 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.000 | 0.000 |
| Research integrity | 0.000 | 0.001 |
| Insufficient payload (model declined to judge) | 0.001 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one teacher head, not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".