Improving the design and conduct of aquatic toxicity studies with oils based on 20 years of CROSERF experience
Bibliographic record
Abstract
Laboratory toxicity testing is a key tool used in oil spill science, spill effects assessment, and mitigation strategy decisions to minimize environmental impacts. A major consideration in oil toxicity testing is how to replicate real-world spill conditions, oil types, weathering states, receptor organisms, and modifying environmental factors under laboratory conditions. Oils and petroleum-derived products are comprised of thousands of compounds with different physicochemical and toxicological properties, and this leads to challenges in conducting and interpreting oil toxicity studies. Experimental methods used to mix oils with aqueous test media have been shown to influence the aqueous-phase hydrocarbon composition and concentrations, hydrocarbon phase distribution (i.e., dissolved phase versus in oil droplets), and the stability of oil:water solutions which, in turn, influence the bioavailability and toxicity of the oil containing media. Studies have shown that differences in experimental methods can lead to divergent test results. Therefore, it is imperative to standardize the methods used to prepare oil:water solutions in order to improve the realism and comparability of laboratory tests. The CROSERF methodology, originally published in 2005, was developed as a standardized method to prepare oil:water solutions for testing and evaluating dispersants and dispersed oil. However, it was found equally applicable for use in testing oil-derived petroleum substances. The goals of the current effort were to: (1) build upon two decades of experience to update existing CROSERF guidance for conducting aquatic toxicity tests and (2) to improve the design of laboratory toxicity studies for use in hazard evaluation and development of quantitative effects models that can then be applied in spill assessment. Key experimental design considerations discussed include species selection (standard vs field collected), test substance (single compound vs whole oil), exposure regime (static vs flow-through) and duration, exposure metrics, toxicity endpoints, and quality assurance and control.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame machine prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. The Gemma side is a direct model label for every work in the frame, read from the title-only record. The Codex side is a classifier learned from the 10,348 direct Codex labels and calibrated to design-weighted sample rates; fields without enough sample support carry no Codex call. Candidate is the union of the two sides; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels.
Distilled classifier scores by category (both heads)
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.023 | 0.010 |
| Meta-epidemiology (narrow) | 0.001 | 0.000 |
| Meta-epidemiology (broad) | 0.001 | 0.001 |
| Bibliometrics | 0.002 | 0.000 |
| Science and technology studies | 0.001 | 0.001 |
| Scholarly communication | 0.002 | 0.003 |
| Open science | 0.002 | 0.002 |
| Research integrity | 0.001 | 0.001 |
| Insufficient payload (model declined to judge) | 0.003 | 0.001 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one source (direct Gemma or distilled Codex), not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".