MétaCan
Menu
Back to cohort
Record W4387474353 · doi:10.1093/jcag/gwad039

Interobserver Reliability of the Paris Classification for Superficial Gastrointestinal Tract Neoplasms: A Systematic Review

2023· review· en· W4387474353 on OpenAlexaff
Sarang Gupta, Sam Seleq, Nikko Gimpaya, Rishad Khan, Michael A. Scaffidi, Rishi Bansal, Samir C. Grover

Bibliographic record

VenueJournal of the Canadian Association of Gastroenterology · 2023
Typereview
Languageen
FieldMedicine
TopicColorectal Cancer Screening and Detection
Canadian institutionsSt. Michael's HospitalUniversity of Toronto
Fundersnot available
KeywordsMedicineSystematic reviewObservational studyReliability (semiconductor)Classification schemeRadiologyPathologyMEDLINEInformation retrieval

Abstract

fetched live from OpenAlex

Background and study aims: The Paris classification characterizes the morphology of superficial gastrointestinal tract neoplasms. This system has been shown to predict the risk of submucosal invasion in certain subtypes of lesions. There is limited data that assesses its agreement amongst endoscopists. We performed a systematic review to summarize the available literature on the interobserver reliability (IOR) of the Paris classification. Methods: We conducted a search through December 2020 for studies reporting IOR of the Paris classification. Studies were included if they quantitatively evaluated the IOR of the Paris classification with at least five participating endoscopists. Two authors independently screened studies and abstracted data using an a priori-designed data collection form. Evaluation of study quality and risk of bias was performed using an adapted version of the Guidelines for Reporting Reliability and Agreement Studies. Results: = 0.54) and substantial in one study that evaluated gastric neoplasms (κw = 0.65). An educational intervention was conducted by three studies with variable methodology and no significant change in IOR. Conclusions: IOR of the Paris classification is moderate for superficial colonic neoplasms. Further study is needed to determine the reliability of this system for superficial gastric lesions. Standardized training programs are required to investigate the impact of educational intervention on the Paris classification amongst endoscopists.

Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.

How this classification was reachedexpand

Full frame machine prediction

Teacher imitation

Not calibrated prevalence, not ground truth. Human validation pending. The Gemma side is a direct model label for every work in the frame, read from the title-only record. The Codex side is a classifier learned from the 10,348 direct Codex labels and calibrated to design-weighted sample rates; fields without enough sample support carry no Codex call. Candidate is the union of the two sides; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels.

metaresearch head score (Codex)0.042
metaresearch head score (Gemma)0.190
Version: metacan-v3-hybrid-931329e0061cValidation status: machine_predicted_unvalidated
Candidate categoriesnone
Consensus categoriesnone
DomainCandidate signal: none · Consensus signal: none
Study designCandidate signal: Systematic review · Consensus signal: Systematic review
GenreCandidate signal: Review · Consensus signal: Review
Teacher disagreement score0.042
Threshold uncertainty score0.224

Distilled classifier scores by category (both heads)

CategoryCodexGemma
Metaresearch0.0420.190
Meta-epidemiology (narrow)0.0020.001
Meta-epidemiology (broad)0.0090.009
Bibliometrics0.0120.011
Science and technology studies0.0010.002
Scholarly communication0.0030.003
Open science0.0030.002
Research integrity0.0020.001
Insufficient payload (model declined to judge)0.0030.000

Machine scores (provisional)

The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.

Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.

Opus teacher head0.048
GPT teacher head0.310
Teacher spread0.262 · how far apart the two teachers sit on this one work
Validation statusscore_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from it

Classification

machine, unvalidated

Machine predicted; a candidate call from one source (direct Gemma or distilled Codex), not a consensus.

The models applied no category: nothing in the taxonomy fit this work.
Study designSystematic review
Domainnot available
GenreReview

How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".

Quick stats

Citations4
Published2023
Admission routes1
Has abstractyes

Explore more

Same venueJournal of the Canadian Association of GastroenterologySame topicColorectal Cancer Screening and DetectionFrench-language works237,207