Advantages of model averaging of species sensitivity distributions used for regulating produced water discharges
Bibliographic record
Abstract
Produced water (PW) generated by Australian offshore oil and gas activities is typically discharged to the ocean after treatment. These complex mixtures of organic and inorganic compounds can pose significant environmental risk to receiving waters, if not managed appropriately. Oil and gas operators in Australia are required to demonstrate that environmental impacts of their activity are managed to levels that are as low as reasonably practicable, for example, through risk assessments comparing predicted no-effect concentrations (PNECs) with predicted environmental concentrations of PW. Probabilistic species sensitivity distribution (SSD) approaches are increasingly being used to derive PW PNECs and subsequently calculating dilutions of PW (termed "safe" dilutions) required to protect a nominated percentage of species in the receiving environment (e.g., 95% and 99% or PC95 and PC99, respectively). Limitations associated with SSDs include fitting a single model to small (six to eight species) data sets, resulting in large uncertainty (very wide 95% confidence limits) in the region associated with PC99 and PC95 results. Recent advances in SSD methodology, in the form of model averaging, claim to overcome some of these limitations by applying the average model fit of multiple models to a data set. We assessed the advantages and limitations of four different SSD software packages for determining PNECs for five PWs from a gas and condensate platform off the North West Shelf of Australia. Model averaging reduced occurrences of extreme uncertainty around PC95 and PC99 values compared with single model fitting and was less prone to the derivation of overly conservative PC99 and PC95 values that resulted from lack of fit to single models. Our results support the use of model averaging for improved robustness in derived PNEC and subsequent "safe" dilution values for PW discharge management and risk assessment. In addition, we present and discuss the toxicity of PW considering the paucity of such information in peer-reviewed literature. Integr Environ Assess Manag 2024;20:498-517. © 2023 Commonwealth Scientific and Industrial Research Organisation. Integrated Environmental Assessment and Management published by Wiley Periodicals LLC on behalf of Society of Environmental Toxicology & Chemistry (SETAC).
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame distilled prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.
Codex and Gemma teacher scores by category
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.001 | 0.000 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.000 | 0.000 |
| Bibliometrics | 0.000 | 0.000 |
| Science and technology studies | 0.000 | 0.000 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.000 | 0.000 |
| Research integrity | 0.000 | 0.000 |
| Insufficient payload (model declined to judge) | 0.000 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one teacher head, not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".