Taxonomic Blind Spots: A Limitation of Environmental DNA Metabarcoding‐Based Detection for Canadian Freshwater Fishes
Bibliographic record
Abstract
ABSTRACT With increasing utilization of eDNA metabarcoding for fish community assessment, it is critical to identify, address, and communicate its capabilities and limitations. One limitation of great concern is the reliability of taxonomic coverage. Taxonomic blind spots, defined as consistent false negatives for specific taxa despite known presence, reduce corroboration with conventional surveys and can limit the uptake of eDNA metabarcoding for biomonitoring. These blind spots result from gaps in reference sequence libraries, issues with taxonomic resolution, inefficient binding of universal primers to the DNA of certain species, and ineffective collection during the sampling of eDNA. To explore this, a multiproject empirical dataset was compiled and analyzed to evaluate the taxonomic coverage of eDNA metabarcoding for a subset of Canadian freshwater fishes using a standardized workflow for two genetic markers: 12S MiFish‐U and Vertebrate COI. The compiled dataset consists of species lists generated by eDNA surveys, paired conventional surveys, and historical records. In total, 59 fish species across 15 families were evaluated of which approximately 40% were unable to be consistently detected by either marker because of a blind spot. The 12S and COI markers also differed in which kinds of blind spots were most frequently observed, with 12S markers exhibiting more reference and resolution blind spots and the COI marker exhibiting more unclassified blind spots. Additionally, in silico primer testing exhibited inconsistent predictions for amplification when using multiple software packages, suggesting the need for further in vitro analysis to troubleshoot primer‐related blind spots. This study highlights the impact of these blind spots in taxonomic coverage on eDNA metabarcoding studies of Canadian freshwater fishes. The limitations imposed by taxonomic blind spots should be addressed in future optimization efforts as eDNA metabarcoding sees broader acceptance as an applied method for fish biomonitoring.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame distilled prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.
Codex and Gemma teacher scores by category
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.000 | 0.000 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.000 | 0.000 |
| Bibliometrics | 0.000 | 0.000 |
| Science and technology studies | 0.000 | 0.001 |
| Scholarly communication | 0.000 | 0.001 |
| Open science | 0.000 | 0.000 |
| Research integrity | 0.000 | 0.000 |
| Insufficient payload (model declined to judge) | 0.004 | 0.002 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; both teacher heads agree on what is shown here.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".