Coastal Transient Niches Shape the Microdiversity Pattern of a Bacterioplankton Population with Reduced Genomes
Bibliographic record
Abstract
Prokaryotic species, defined with operational thresholds, such as 95% of the whole-genome average nucleotide identity (ANI) or 98.7% similarity of the 16S rRNA gene sequences, commonly contain extensive fine-grained diversity in both the core genome and the accessory genome. However, the ways in which this genomic microdiversity and its associated phenotypic microdiversity are organized and structured is poorly understood, which disconnects microbial diversity and ecosystem functioning. Population genomic approaches that allow this question to be addressed are commonly applied to cultured species because linkages between different loci are necessary but are missing from metagenome-assembled genomes. In the past, these approaches were only applied to easily cultivable bacteria and archaea, which, nevertheless, are often not representative of natural communities. Here, we focus on the recently discovered cluster, CHUG, which are representative in marine bacterioplankton communities and possess some of the smallest genomes in the globally dominant marine Roseobacter group. Despite being over 95% ANI and identical in the 16S rRNA gene, the 33 CHUG genomes we analyzed have undergone multiple speciation events, with the first split event predominantly structuring the genomic diversity. The observed pattern of genomic microdiversity correlates with CHUG members' differential utilization of carbon sources and differential ability to explore low-oxygen niches. The available data are consistent with the idea that brown algae may be home to CHUG, though other habitats, such as fresh organic aggregates, are also possible.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame distilled prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.
Codex and Gemma teacher scores by category
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.000 | 0.000 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.000 | 0.000 |
| Bibliometrics | 0.000 | 0.000 |
| Science and technology studies | 0.000 | 0.000 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.000 | 0.000 |
| Research integrity | 0.000 | 0.000 |
| Insufficient payload (model declined to judge) | 0.006 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one teacher head, not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".