A critique of the European Commission Document, “State of the Art Assessment of Endocrine Disrupters”
Bibliographic record
Abstract
In this commentary, we critique a recently finalized document titled "State of the Art Assessment of Endocrine Disrupters" (SOA Assessment). The SOA Assessment was commissioned by the European Union Directorate-General for the Environment to provide a basis for developing scientific criteria for identifying endocrine disruptors and reviewing and possibly revising the European Community Strategy on Endocrine Disrupters. In our view, the SOA Assessment takes an anecdotal approach rather than attempting a comprehensive assessment of the state of the art or synthesis of current knowledge. To do the latter, the document would have had to (i) distinguish between apparent associations of outcomes with exposure and the inference of an endocrine-disruption (ED) basis for those outcomes; (ii) constitute a complete and unbiased survey of new literature since 2002 (when the WHO/IPCS document, "Global Assessment of the State-of-the-Science of Endocrine Disruptors" was published); (iii) consider strengths and weaknesses and issues in interpretation of the cited literature; (iv) follow a weight-of-evidence methodology to evaluate evidence of ED; (v) document the evidence for its conclusions or the reasoning behind them; and (vi) present the evidence for or reasoning behind why conclusions that differ from those drawn in the 2002 WHO/IPCS document need to be changed. In its present form, the SOA Assessment fails to provide a balanced and critical assessment or synthesis of literature relevant to ED. We urge further evidence-based evaluations to develop the needed scientific basis to support future policy decisions.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame distilled prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.
Codex and Gemma teacher scores by category
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.002 | 0.002 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.002 | 0.001 |
| Bibliometrics | 0.000 | 0.000 |
| Science and technology studies | 0.000 | 0.001 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.001 | 0.001 |
| Research integrity | 0.000 | 0.001 |
| Insufficient payload (model declined to judge) | 0.001 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one teacher head, not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".