Demonstrating the undermining of science and health policy after the Fukushima nuclear accident by applying the Toolkit for detecting misused epidemiological methods
Bibliographic record
Abstract
It is well known that science can be misused to hinder the resolution (i.e., the elimination and/or control) of a health problem. To recognize distorted and misapplied epidemiological science, a 33-item "Toolkit for detecting misused epidemiological methods" (hereinafter, the Toolkit) was published in 2021. Applying the Toolkit, we critically evaluated a review paper entitled, "Lessons learned from Chernobyl and Fukushima on thyroid cancer screening and recommendations in the case of a future nuclear accident" in Environment International in 2021, published by the SHAMISEN (Nuclear Emergency Situations - Improvement of Medical and Health Surveillance) international expert consortium. The article highlighted the claim that overdiagnosis of childhood thyroid cancers greatly increased the number of cases detected in ultrasound thyroid screening following the 2011 Fukushima nuclear accident. However, the reasons cited in the SHAMISEN review paper for overdiagnosis in mass screening lacked important information about the high incidence of thyroid cancers after the accident. The SHAMISEN review paper ignored published studies of screening results in unexposed areas, and included an invalid comparison of screenings among children with screenings among adults. The review omitted the actual state of screening in Fukushima after the nuclear accident, in which only nodules > 5 mm in diameter were examined. The growth rate of thyroid cancers was not slow, as emphasized in the SHAMISEN review paper; evidence shows that cancers detected in second-round screening grew to more than 5 mm in diameter over a 2-year period. The SHAMISEN consortium used an unfounded overdiagnosis hypothesis and misguided evidence to refute that the excess incidence of thyroid cancer was attributable to the nuclear accident, despite the findings of ongoing ultrasound screening for thyroid cancer in Fukushima and around Chernobyl. By our evaluation, the SHAMISEN review paper includes 20 of the 33 items in the Toolkit that demonstrate the misuse of epidemiology. The International Agency for Research on Cancer meeting in 2017 and its publication cited in the SHAMISEN review paper includes 12 of the 33 items in the Toolkit. Finally, we recommend a few enhancements to the Toolkit to increase its utility.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame distilled prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.
Codex and Gemma teacher scores by category
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.013 | 0.001 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.001 | 0.000 |
| Bibliometrics | 0.000 | 0.000 |
| Science and technology studies | 0.002 | 0.001 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.000 | 0.000 |
| Research integrity | 0.000 | 0.001 |
| Insufficient payload (model declined to judge) | 0.000 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one teacher head, not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".