The OECD Program to Validate the Rat Hershberger Bioassay to Screen Compounds for <i>in Vivo</i> Androgen and Antiandrogen Responses: Phase 2 Dose–Response Studies
Bibliographic record
Abstract
OBJECTIVE: The Organisation for Economic Co-operation and Development (OECD) has completed phase 2 of an international program to validate the rodent Hershberger bioassay. DESIGN: The Hershberger bioassay is designed to identify suspected androgens and antiandrogens based on changes in the weights of five androgen-responsive tissues (ventral prostate, paired seminal vesicles and coagulating glands, the levator ani and bulbocavernosus muscles, the glans penis, and paired Cowper's or bulbourethral glands). Protocol sensitivity and reproducibility were tested using two androgen agonists (17alpha-methyl testosterone and 17beta-trenbolone), four antagonists [procymi-done, vinclozolin, linuron, and 1,1-dichoro-2,2-bis-(p-chlorophenyl)ethylene (p,p'-DDE)], and a 5alpha-reductase inhibitor (finasteride). Sixteen laboratories from seven countries participated in phase 2. RESULTS: In 40 of 41 studies, the laboratories successfully detected substance-related weight changes in one or more tissues. The one exception was with the weakest antiandrogen, linuron, in a laboratory with reduced sensitivity because of high coefficients of variation in all tissue weights. The protocols performed well under different experimental conditions (e.g., strain, diet, housing protocol, bedding, vehicle). There was good agreement and reproducibility among laboratories with regard to the lowest dose inducing significant effects on tissue weights. CONCLUSIONS: The results show that the OECD Hershberger bioassay protocol is reproducible and transferable across laboratories with androgen agonists, weak androgen antagonists, and a 5alpha-reductase inhibitor. The next validation phase will employ coded test substances, including positive substances and negative substances having no androgenic or antiandrogenic activity.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame distilled prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.
Codex and Gemma teacher scores by category
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.003 | 0.000 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.000 | 0.000 |
| Bibliometrics | 0.000 | 0.000 |
| Science and technology studies | 0.001 | 0.001 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.000 | 0.000 |
| Research integrity | 0.000 | 0.000 |
| Insufficient payload (model declined to judge) | 0.000 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one teacher head, not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".