DNA Barcoding Reveals Cryptic Diversity in Lumbricus terrestris L., 1758 (Clitellata): Resurrection of L. herculeus (Savigny, 1826)
Bibliographic record
Abstract
The widely studied and invasive earthworm, Lumbricus terrestris L., 1758 has been the subject of nomenclatural debate for many years. However these disputes were not based on suspicions of heterogeneity, but rather on the descriptions and nomenclatural acts associated with the species name. Large numbers of DNA barcode sequences of the cytochrome oxidase I obtained for nominal L. terrestris and six congeneric species reveal that there are two distinct lineages within nominal L. terrestris. One of those lineages contains the Swedish population from which the name-bearing specimen of L. terrestris was obtained. The other contains the population from which the syntype series of Enterion herculeum Savigny, 1826 was collected. In both cases modern and old representatives yielded barcode sequences allowing us to clearly establish that these are two distinct species, as different from one another as any other pair of congeners in our data set. The two are morphologically indistinguishable, except by overlapping size-related characters. We have designated a new neotype for L. terrestris. The newly designated neotype and a syntype of L. herculeus yielded DNA adequate for sequencing part of the cytochrome oxidase I gene (COI). The sequence data make possible the objective determination of the identities of earthworms morphologically identical to L. terrestris and L. herculeus, regardless of body size and segment number. Past work on nominal L. terrestris could have been on either or both species, although L. herculeus has yet to be found outside of Europe.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame machine prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. The Gemma side is a direct model label for every work in the frame, read from the title-only record. The Codex side is a classifier learned from the 10,348 direct Codex labels and calibrated to design-weighted sample rates; fields without enough sample support carry no Codex call. Candidate is the union of the two sides; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels.
Distilled classifier scores by category (both heads)
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.000 | 0.000 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.000 | 0.000 |
| Bibliometrics | 0.001 | 0.001 |
| Science and technology studies | 0.000 | 0.000 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.000 | 0.000 |
| Research integrity | 0.000 | 0.000 |
| Insufficient payload (model declined to judge) | 0.001 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one source (direct Gemma or distilled Codex), not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".