MétaCan
Menu
Back to cohort
Record W2056117213 · doi:10.1002/ieam.278

It is time for changes in the analysis of whole effluent toxicity data

2011· article· en· W2056117213 on OpenAlexaboutno aff
Jerry Diamond, Debra L. Denton, Brian S. Anderson, Bryn M. Phillips

Bibliographic record

VenueIntegrated Environmental Assessment and Management · 2011
Typearticle
Languageen
FieldAgricultural and Biological Sciences
TopicPesticide Residue Analysis and Safety
Canadian institutionsnot available
Fundersnot available
KeywordsEffluentTest (biology)Environmental scienceControl (management)Null hypothesisToxicityReliability engineeringToxicologyComputer scienceEnvironmental engineeringStatisticsEngineeringMathematicsEcologyChemistryBiology

Abstract

fetched live from OpenAlex

The whole effluent toxicity (WET) program in the United States, Canada, and other countries typically requires multi concentration testing of effluents. While multiconcentration testing of chemicals is desirable for regulatory and scientific reasons, we believe this requirement is not as efficient for evaluating effluent compliance in a WET program. The key regulatory question of concern is whether an effluent is toxic or not, which is best answered statistically using a hypothesis approach, not a point estimate approach. However, the traditional hypothesis approach currently recommended does not reward high within-test precision. This report describes the need for 3 specific changes in the analysis of WET compliance data that we believe would yield a more robust WET regulatory program: (1) restate the null hypothesis so that test power is associated with demonstrating that the effluent is not toxic, (2) use USEPA's Test of Significant Toxicity (based on the noninferiority approach) to identify unacceptable toxicity as well as acceptable effects with a high probability, and (3) evaluate only the test control and the critical concentration of concern (e.g., instream waste concentration). We demonstrate that instituting these 3 changes would provide: Positive incentives for permittees to produce high-quality WET data, a transparent analysis approach in which the permittee could have greater control over regulatory decisions based on test results, and potentially a less expensive testing program because fewer effluent concentrations need to be examined within a test. As a result, WET test frequency could be increased for the same cost as current testing programs while providing greater representativeness of effluent quality.

Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.

How this classification was reachedexpand

Full frame machine prediction

Teacher imitation

Not calibrated prevalence, not ground truth. Human validation pending. The Gemma side is a direct model label for every work in the frame, read from the title-only record. The Codex side is a classifier learned from the 10,348 direct Codex labels and calibrated to design-weighted sample rates; fields without enough sample support carry no Codex call. Candidate is the union of the two sides; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels.

metaresearch head score (Codex)0.089
metaresearch head score (Gemma)0.136
Version: metacan-v3-hybrid-931329e0061cValidation status: machine_predicted_unvalidated
Candidate categoriesMetaresearch
Consensus categoriesnone
DomainCandidate signal: Methods · Consensus signal: none
Study designCandidate signal: Theoretical or conceptual · Consensus signal: none
GenreCandidate signal: Commentary · Consensus signal: Commentary
Teacher disagreement score0.911
Threshold uncertainty score0.469

Distilled classifier scores by category (both heads)

CategoryCodexGemma
Metaresearch0.0890.136
Meta-epidemiology (narrow)0.0010.001
Meta-epidemiology (broad)0.0010.003
Bibliometrics0.0040.003
Science and technology studies0.0030.005
Scholarly communication0.0090.009
Open science0.0070.004
Research integrity0.0080.017
Insufficient payload (model declined to judge)0.0070.006

Machine scores (provisional)

The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.

Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.

Opus teacher head0.042
GPT teacher head0.266
Teacher spread0.224 · how far apart the two teachers sit on this one work
Validation statusscore_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from it

Classification

machine, unvalidated

Machine predicted; a candidate call from one source (direct Gemma or distilled Codex), not a consensus.

Study designTheoretical or conceptual
DomainMethods
GenreCommentary

How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".

Quick stats

Citations4
Published2011
Admission routes1
Has abstractyes

Explore more

Same venueIntegrated Environmental Assessment and ManagementSame topicPesticide Residue Analysis and SafetyFrench-language works237,207