The molecular basis of cereal mixed-linkage β-glucan utilization by the human gut bacterium Segatella copri
Bibliographic record
Abstract
Mixed-linkage β(1,3)/β(1,4)-glucan (MLG) is abundant in the human diet through the ingestion of cereal grains and is widely associated with healthful effects on metabolism and cholesterol levels. MLG is also a major source of fermentable glucose for the human gut microbiota (HGM). Bacteria from the family Prevotellaceae are highly represented in the HGM of individuals who eat plant-rich diets, including certain indigenous people and vegetarians in postindustrial societies. Here, we have defined and functionally characterized an exemplar Prevotellaceae MLG polysaccharide utilization locus (MLG-PUL) in the type-strain Segatella copri (syn. Prevotella copri) DSM 18205 through transcriptomic, biochemical, and structural biological approaches. In particular, structure-function analysis of the cell-surface glycan-binding proteins and glycoside hydrolases of the S. copri MLG-PUL revealed the molecular basis for glycan capture and saccharification. Notably, syntenic MLG-PULs from human gut, human oral, and ruminant gut Prevotellaceae are distinguished from their counterparts in Bacteroidaceae by the presence of a β(1,3)-specific endo-glucanase from glycoside hydrolase family 5, subfamily 4 (GH5_4) that initiates MLG backbone cleavage. The definition of a family of homologous MLG-PULs in individual species enabled a survey of nearly 2000 human fecal microbiomes using these genes as molecular markers, which revealed global population-specific distributions of Bacteroidaceae- and Prevotellaceae-mediated MLG utilization. Altogether, the data presented here provide new insight into the molecular basis of β-glucan metabolism in the HGM, as a basis for informing the development of approaches to improve the nutrition and health of humans and other animals.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame distilled prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.
Codex and Gemma teacher scores by category
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.001 | 0.000 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.000 | 0.000 |
| Bibliometrics | 0.000 | 0.000 |
| Science and technology studies | 0.000 | 0.001 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.001 | 0.000 |
| Research integrity | 0.000 | 0.001 |
| Insufficient payload (model declined to judge) | 0.000 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one teacher head, not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".