Minimum sample size, population genomics and morphological variation in the terrestrial gastropod Webbhelix multilineata (Mollusca, Polygyridae)
Bibliographic record
Abstract
Patterns in evolutionary and ecological biology help shape our view of the physical world both in the present and its past. Biogeographical studies in particular afford understanding of how biological systems have been shaped by processes over geological time, both genetically and morphologically. Some systems, like terrestrial gastropod mollusks, are well suited for these types of studies owing to their near ubiquitous occupation of terrestrial habitats, while also individually being very locally restricted and having poor reproductive dispersal. A notable species is the striped white-lip snail Webbhelix multilineata, a morphologically charming and biogeographically interesting study system because of its somewhat unusually broad geographic range spanning large regions of North America previously glaciated and unglaciated during the Pleistocene. It is also quite restricted to moist, shaded riparian woodlands, a historically profuse habitat that has been largely lost and become highly fragmented due to human development. These features and others make Webbhelix an enticing system for biogeographical and population study, and the chosen focus of the studies presented here. While there are many different molecular and morphological approaches that can be taken to pursue such studies, this dissertation focused primarily on the generation and use of genomic data for a non-model system, and secondarily on generating and utilizing morphological data, to accomplish three major aims: (1) test for a minimum sample size to accurately estimate basic population genomic parameters in Webbhelix multilineata; (2) generate the first genomic dataset for Webbhelix and utilize it for population genomic assessment, as well as coalescent simulations to assess models of post-glacial refugial expansion; and (3) test variation in Webbhelix shell size and brightness within the context of established ecogeographic patterns. Addressing the first aim using SNPs generated through a ddRADseq approach showed that as few as 6 individuals per sampled locality was sufficient to accurately estimate observed and expected heterozygosity, inbreeding coefficient, and pairwise differentiation in Webbhelix. Pursuit of the second aim netted an 8,480 SNP dataset composed of 18 sample localities from across the species range, and its assessment for population genomic structure revealed 4 population clusters: two Midwestern clusters, a southern Mississippi River cluster, and an eastern cluster on the shores of Lake Erie in Ontario. Estimates of genetic diversity found it was highest in the southern-most sampled localities and generally decreased further north and away from the Mississippi River, indicative of serial founder effects resulting from range expansion. Analysis of coalescent simulations suggested that post-glacial range expansion in Webbhelix occurred along major river corridors rather than openly in all directions, and that expansion either began as early as some 20,000 years ago around the LGM or a more northerly refugia existed for this species than initially considered. Lastly, work toward the third aim resulted in a moderately strong negative correlation being found between shell size and latitude in Webbhelix, indicating a pattern described by a reverse Bergmann cline. Additionally, a weak positive correlation was found between shell brightness and latitude, suggesting (with heavy caveats) that Webbhelix may exhibit the pattern described by Gloger’s ecogeographic hypothesis.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame machine prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. The Gemma side is a direct model label for every work in the frame, read from the title-only record. The Codex side is a classifier learned from the 10,348 direct Codex labels and calibrated to design-weighted sample rates; fields without enough sample support carry no Codex call. Candidate is the union of the two sides; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels.
Distilled classifier scores by category (both heads)
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.002 | 0.006 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.000 | 0.000 |
| Bibliometrics | 0.000 | 0.000 |
| Science and technology studies | 0.000 | 0.000 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.000 | 0.000 |
| Research integrity | 0.000 | 0.000 |
| Insufficient payload (model declined to judge) | 0.001 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one source (direct Gemma or distilled Codex), not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".