DNA barcoding as a screening tool for cryptic diversity: an example from Caryocolum, with description of a new species (Lepidoptera, Gelechiidae)
Bibliographic record
Abstract
We explore the potential value of DNA barcode divergence for species delimitation in the genus Caryocolum Gregor & Povolný, 1954 (Lepidoptera, Gelechiidae), based on data from 44 European species (including 4 subspecies). Low intraspecific divergence of the DNA barcodes of the mtCOI (cytochrome c oxidase 1) gene and/or distinct barcode gaps to the nearest neighbor support species status for all examined nominal taxa. However, in 8 taxa we observed deep splits with a maximum intraspecific barcode divergence beyond a threshold of 3%, thus indicating possible cryptic diversity. The taxonomy of these taxa has to be re-assessed in the future. We investigated one such deep split in Caryocolum amaurella (Hering, 1924) and found it in congruence with yet unrecognized diagnostic morphological characters and specific host-plants. The integrative species delineation leads to the description of Caryocolum crypticum sp. n. from northern Italy, Switzerland and Greece. The new species and the hitherto intermixed closest relative C. amaurella are described in detail and adults and genitalia of both species are illustrated and a lectotype of C. amaurella is designated; a diagnostic comparison of the closely related C. iranicum Huemer, 1989, is added.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame machine prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. The Gemma side is a direct model label for every work in the frame, read from the title-only record. The Codex side is a classifier learned from the 10,348 direct Codex labels and calibrated to design-weighted sample rates; fields without enough sample support carry no Codex call. Candidate is the union of the two sides; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels.
Distilled classifier scores by category (both heads)
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.000 | 0.001 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.000 | 0.000 |
| Bibliometrics | 0.002 | 0.002 |
| Science and technology studies | 0.001 | 0.001 |
| Scholarly communication | 0.001 | 0.000 |
| Open science | 0.000 | 0.001 |
| Research integrity | 0.001 | 0.001 |
| Insufficient payload (model declined to judge) | 0.001 | 0.001 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one source (direct Gemma or distilled Codex), not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".