Genome-based taxonomic framework for the class Negativicutes: division of the class Negativicutes into the orders Selenomonadales emend., Acidaminococcales ord. nov. and Veillonellales ord. nov.
Bibliographic record
Abstract
The class Negativicutes is currently divided into one order and two families on the basis of 16S rRNA gene sequence phylogenies. We report here comprehensive comparative genomic analyses of the sequenced members of the class Negativicutes to demarcate its different evolutionary groups in molecular terms, independently of phylogenetic trees. Our comparative genomic analyses have identified 14 conserved signature indels (CSIs) and 48 conserved signature proteins (CSPs) that either are specific for the entire class or differentiate four main groups within the class. Two CSIs and nine CSPs are shared uniquely by all or most members of the class Negativicutes, distinguishing this class from all other sequenced members of the phylum Firmicutes. Four other CSIs and six CSPs were specific characteristics of the family Acidaminococcaceae, two CSIs and four CSPs were uniquely present in the family Veillonellaceae, six CSIs and eight CSPs were found only in Selenomonas and related genera, and 17 CSPs were identified uniquely in Sporomusa and related genera. Four additional CSPs support a pairing of the groups containing the genera Selenomonas and Sporomusa. We also report detailed phylogenetic analyses for the Negativicutes based on core protein sequences and 16S rRNA gene sequences, which strongly support the four main groups identified by CSIs and by CSPs. Based on the results from different lines of investigation, we propose a division of the class Negativicutes into an emended order Selenomonadales containing the new families Selenomonadaceae fam. nov. and Sporomusaceae fam. nov. and two new orders, Acidaminococcales ord. nov. and Veillonellales ord. nov., respectively containing the families Acidaminococcaceae and Veillonellaceae.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame distilled prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.
Codex and Gemma teacher scores by category
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.001 | 0.001 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.000 | 0.000 |
| Bibliometrics | 0.000 | 0.000 |
| Science and technology studies | 0.000 | 0.000 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.001 | 0.000 |
| Research integrity | 0.000 | 0.000 |
| Insufficient payload (model declined to judge) | 0.000 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one teacher head, not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".