High quality <i>Bathyarchaeia</i> MAGs from lignocellulose-impacted environments elucidate metabolism and evolutionary mechanisms
Bibliographic record
Abstract
Abstract The archaeal class Bathyarchaeia is widely and abundantly distributed in anoxic habitats. Metagenomic studies have suggested that they are mixotrophic, capable of CO2 fixation and heterotrophic growth, and involved in acetogenesis and lignin degradation. We analyzed 35 Bathyarchaeia metagenome-assembled genomes (MAGs), including the first complete circularized MAG (cMAG) of the Bathy-6 subgroup, from the metagenomes of three full-scale pulp and paper mill anaerobic digesters and three laboratory methanogenic enrichment cultures maintained on pre-treated poplar. Thirty-three MAGs belong to the Bathy-6, lineage while two are from the Bathy-8 lineage. In our previous analysis of the microbial community in the pulp mill digesters, Bathyarchaeia were abundant and positively correlated to hydrogenotrophic and methylotrophic methanogenesis. Several factors likely contribute to the success of the Bathy-6 lineage compared to Bathy-8 in the reactors. The Bathy-6 genomes are larger than those of Bathy-8 and have more genes involved in lignocellulose degradation, including carbohydrate-active enzymes not present in the Bathy-8. Bathy-6 also shares the Bathyarchaeal O-demethylase system recently identified in Bathy-8. All the Bathy-6 MAGs had numerous membrane-associated pyrroloquinoline quinone-domain proteins that we suggest are involved in lignin modification or degradation, together with Radical-S-adenosylmethionine (SAM) and Rieske domain proteins, and AA2, AA3, and AA6-family oxidoreductases. We also identified a complete B12 synthesis pathway and a complete nitrogenase gene locus. Finally, comparative genomic analyses revealed that Bathyarchaeia genomes are dynamic and have interacted with other organisms in their environments through gene transfer to expand their gene repertoire.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame machine prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. The Gemma side is a direct model label for every work in the frame, read from the title-only record. The Codex side is a classifier learned from the 10,348 direct Codex labels and calibrated to design-weighted sample rates; fields without enough sample support carry no Codex call. Candidate is the union of the two sides; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels.
Distilled classifier scores by category (both heads)
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.000 | 0.001 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.000 | 0.001 |
| Bibliometrics | 0.001 | 0.001 |
| Science and technology studies | 0.000 | 0.000 |
| Scholarly communication | 0.001 | 0.000 |
| Open science | 0.000 | 0.001 |
| Research integrity | 0.000 | 0.000 |
| Insufficient payload (model declined to judge) | 0.001 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one source (direct Gemma or distilled Codex), not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".