MétaCan
Menu
Back to cohort
Record W4284670634 · doi:10.3897/zookeys.1110.85491

More discussion of minimalist species descriptions and clarifying some misconceptions contained in Meier et al. 2021

2022· article· en· W4284670634 on OpenAlexaff
Michael J. Sharkey, Erika M. Tucker, Austin Baker, M. Alex Smith, Sujeevan Ratnasingham, Ramya Manjunath, Paul D. N. Hebert, Winnie Hallwachs, Daniel H. Janzen

Bibliographic record

VenueZooKeys · 2022
Typearticle
Languageen
FieldEnvironmental Science
TopicEnvironmental DNA in Biodiversity Studies
Canadian institutionsUniversity of Guelph
Fundersnot available
KeywordsIdentification (biology)BarcodeSpecies nameTaxonCode (set theory)Computer scienceBiologyInformation retrievalEvolutionary biologyGenealogyHistoryTaxonomy (biology)PaleontologyEcology

Abstract

fetched live from OpenAlex

This is a response to a preprint version of “A re-analysis of the data in Sharkey et al.’s (2021) minimalist revision reveals that BINs do not deserve names, but BOLD Systems needs a stronger commitment to open science”, https://www.biorxiv.org/content/10.1101/2021.04.28.441626v2. Meier et al. strongly criticized Sharkey et al.’s publication in which 403 new species were deliberately minimally described, based primarily on COI barcode sequence data. Here we respond to these criticisms. The following points are made: 1) Sharkey et al. did not equate BINs with species, as demonstrated in several examples in which multiple species were found to be in single BINs. 2) We reiterate that BINs were used as a preliminary sorting tool, just as preliminary morphological identification commonly sorts specimens based on color and size into unit trays; despite BINs and species concepts matching well over 90% of species, this matching does not equate to equality. 3) Consensus barcodes were used only to provide a diagnosis to conform to the rules of the International Code of Zoological Nomenclature just as consensus morphological diagnoses are. The barcode of a holotype is definitive and simply part of its cellular morphology. 4) Minimalist revisions will facilitate and accelerate future taxonomic research, not hinder it. 5) We refute the claim that the BOLD sequences of Plesiocoelus vanachterbergi are pseudogenes and demonstrate that they simply represent a frameshift mutation. 6) We reassert our observation that morphological evidence alone is insufficient to recognize species within species-rich higher taxa and that its usefulness lies in character states that are congruent with molecular data. 7) We show that in the cases in which COI barcodes code for the same amino acids in different putative species, data from morphology, host specificity, and other ecological traits reaffirm their utility as indicators of genetically distinct lineages.

Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.

How this classification was reachedexpand

Full frame distilled prediction

Teacher imitation

Not calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.

metaresearch head score (Codex)0.000
metaresearch head score (Gemma)0.000
Version: codex-gemma-dda1882f352aValidation status: machine_predicted_unvalidated
Candidate categoriesInsufficient payload (model declined to judge)
Consensus categoriesnone
DomainCandidate signal: none · Consensus signal: none
Study designCandidate signal: Observational · Consensus signal: Observational
GenreCandidate signal: Empirical · Consensus signal: Empirical
Teacher disagreement score0.024
Threshold uncertainty score0.996

Codex and Gemma teacher scores by category

CategoryCodexGemma
Metaresearch0.0000.000
Meta-epidemiology (narrow)0.0000.000
Meta-epidemiology (broad)0.0000.000
Bibliometrics0.0000.000
Science and technology studies0.0000.000
Scholarly communication0.0000.000
Open science0.0000.001
Research integrity0.0000.000
Insufficient payload (model declined to judge)0.0050.000

Machine scores (provisional)

The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.

Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.

Opus teacher head0.021
GPT teacher head0.236
Teacher spread0.215 · how far apart the two teachers sit on this one work
Validation statusscore_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from it

Classification

machine, unvalidated

Machine predicted; a candidate call from one teacher head, not a consensus.

Study designObservational
Domainnot available
GenreEmpirical

How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".

Quick stats

Citations5
Published2022
Admission routes1
Has abstractyes

Explore more

Same venueZooKeysSame topicEnvironmental DNA in Biodiversity StudiesFrench-language works237,207