Levels of evidence available for techniques in antireflux surgery
Bibliographic record
Abstract
The objective of this study was to determine the levels of evidence and grades of recommendations available for techniques in antireflux surgery. Areas of technical controversy in antireflux surgery were identified and developed into eight answerable questions. The external evidence was surveyed using the databases Medline and EMBASE. Abstracts and appropriate articles were identified from January 1966 to December 2005. A set of search strategies was systematically employed to determine the levels of evidence available for each clinical question. Primary outcome measures included the determination of levels of evidence and grade of recommendation based on The Oxford Center for Evidence-Based Medicine. Secondary outcome measures included for randomized controlled trials were Jadad scores, noting the presence of a sample size calculation, and the determination of an effect estimate and the reporting of a confidence interval. Higher quality randomized controlled trials (mostly level 2b, occasional level 1b) existed to answer three questions: whether to complete a 360 degrees or partial wrap; whether or not to divide the short gastric vessels; and whether to perform laparoscopic or open surgery. Lower quality randomized controlled trials were available to determine whether the use of mesh was helpful, whether or not to use a bougie catheter for calibration of the wrap, and whether an anterior or posterior wrap results in a superior outcome. This was deemed to be of inferior grade of recommendation due to the lack (< 2) of trials available and the sole presence of level 2b evidence. The final two questions: whether to complete fundoplication using a thoracic or abdominal approach and whether to use intraoperative manometry relied exclusively upon level 4 evidence and thus received a lower grade of recommendation. A higher Jadad score seemed to be associated with studies having a higher level of evidence available to answer the question. Sample size calculations were given to answer three questions. Effect estimate was difficult to interpret given inconsistent findings, composite outcomes and lack of reported confidence intervals. In conclusion, antireflux surgery has many randomized controlled trials available upon which to base clinical practice. Unfortunately, these are generally of poor quality. We recommend that esophageal surgeons determine consistent outcome measures and endeavor to improve the quality of randomized controlled trials they perform.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame distilled prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.
Codex and Gemma teacher scores by category
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.000 | 0.001 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.000 | 0.000 |
| Bibliometrics | 0.000 | 0.000 |
| Science and technology studies | 0.000 | 0.000 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.000 | 0.000 |
| Research integrity | 0.000 | 0.000 |
| Insufficient payload (model declined to judge) | 0.000 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one teacher head, not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".