Metagenomic Discovery and Characterization of Multi-Functional and Monomodular Processive Endoglucanases as Biocatalysts
Bibliographic record
Abstract
Biomass includes cellulose, hemicelluloses, pectin and lignin; constitutes the components of dietary fibre of plant and alge origins in animals and humans; and can potentially provide inexhaustible basic monomer compounds for developing sustainable biofuels and biomaterials for the world. Development of efficacious cellulases is the key to unlock the biomass polymer and unleash its potential applications in society. Upon reviewing the current literature of cellulase research, two characterized and/or engineered glycosyl hydrolase family-5 (GH5) cellulases have displayed unique properties of processive endoglucanases, including GH5-tCel5A1 that was engineered and was originally identified via targeted genome sequencing of the extremely thermophilic Thermotoga maritima and GH5-p4818Cel5_2A that was screened out of the porcine hindgut microbial metagenomic expression library. Both GH5-tCel5A1 and GH5-p4818Cel5_2A have been characterized as having small molecular weights with an estimated spherical diameter at or < 4.6 nm; being monomodular without a required carbohydrate-binding domain; and acting as processive β-1,4-endoglucanases. These two unique GH5-tCel5A1 and GH5-p4818Cel5_2A processive endocellulases are active in hydrolyzing natural crystalline and pre-treated cellulosic substrates and have multi-functionality towards several hemicelluloses including β-glucans, xylan, xylogulcans, mannans, galactomannans and glucomannans. Therefore, these two multifunctional and monomodular GH5-tCel5A1 and GH5-p4818Cel5_2A endocellulases already have promising structural and functional properties for further optimization and industrial applications.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame distilled prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.
Codex and Gemma teacher scores by category
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.000 | 0.000 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.000 | 0.000 |
| Bibliometrics | 0.000 | 0.000 |
| Science and technology studies | 0.000 | 0.000 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.000 | 0.000 |
| Research integrity | 0.000 | 0.000 |
| Insufficient payload (model declined to judge) | 0.000 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one teacher head, not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".