Differences in diagnostic criteria for esophageal squamous cell carcinoma between Japanese and Western pathologists
Bibliographic record
Abstract
BACKGROUND: Large discrepancies have been found between Western and Japanese pathologists in the diagnosis of adenoma/dysplasia versus carcinoma for gastric and colorectal glandular lesions. It is important to determine whether similar differences exist in the diagnosis of esophageal squamous lesions. METHODS: Eleven expert gastrointestinal pathologists from Japan, North America, and Europe individually reviewed a set of microscopic slides containing 21 sections of biopsies and corresponding endoscopic mucosal resection specimens from Japanese patients with superficial esophageal squamous neoplastic lesions. The pathologists indicated the pathologic findings on which they based each diagnosis. RESULTS: Invasion was the most important diagnostic criterion of carcinoma for the Western pathologists whereas nuclear and structural features were more important for the Japanese pathologists. For two sections showing low grade dysplasia according to most Western pathologists, the Japanese pathologists diagnosed suspected carcinoma in one case and definite carcinoma in the other. For nine sections with high grade dysplasia according to the Western pathologists, the Japanese pathologists diagnosed suspected carcinoma in two cases and definite carcinoma in seven cases. For six sections with suspected carcinoma according to most Western pathologists, the Japanese pathologists diagnosed suspected carcinoma in one case and definite carcinoma in five cases. Four sections showed definite carcinoma according to both the Western and Japanese pathologists. Thus, there was agreement among the Western and Japanese pathologists for only 5 of the 21 sections (kappa value, 0.04). However, when high grade dysplasia, noninvasive carcinoma, and suspected carcinoma were grouped together, the agreement was excellent (19 of the 21 sections; kappa value, 0.75). CONCLUSIONS: In Japan, esophageal squamous cell carcinoma is diagnosed mainly based on nuclear criteria, even in cases judged to be noninvasive low grade dysplasia in the West. This difference in diagnostic practice may contribute to the relatively high incidence rate and good prognosis of superficial esophageal carcinoma in Japan. To improve the comparability of research data, the authors recommend that high grade dysplasia, noninvasive carcinoma, and suspected carcinoma be grouped together into one category of "noninvasive high grade neoplasia." [See editorial on pages 969-70, this issue.]
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame distilled prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.
Codex and Gemma teacher scores by category
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.000 | 0.000 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.000 | 0.000 |
| Bibliometrics | 0.000 | 0.000 |
| Science and technology studies | 0.000 | 0.000 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.000 | 0.000 |
| Research integrity | 0.000 | 0.000 |
| Insufficient payload (model declined to judge) | 0.001 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one teacher head, not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".