Insights into richness of PKS and NRPS gene clusters and genome guided bioprospection for bioactive natural products
Bibliographic record
Abstract
Advent of next-generation sequencing and genome-mining tools witnessed a rejuvenation of research on Actinobacteria to meet the growing drug-resistance of pathogenic microbes. Therefore, the Actinobacteria present in underexplored environments are being largely studied in recent days. Intertidal areas, which endure regular periods of immersion and emersion, are important in the coastal or estuarine environment and represent an underexplored biological niche that could be of interest for the discovery. In this study, we have evaluated biosynthetic heterogeneity and richness of intertidal Actinobacteria isolated from Diu Island (India) and demonstrated genome mining in selected potential strain to facilitate the discovery of novel bioactive compounds relevant to antibiotics development. A total of 62 strains affiliated with seven different genera, Streptomyces, Micromonospora, Saccharomonospora, Nocardia, Nocardiopsis, Actinomadura, and Glycomyces were studied. The amplified fragment restriction fingerprinting was done by targeting specific domains of polyketide synthase type II (PKS-II) and non-ribosomal peptide synthetase (NRPS) to reveal the biosynthetic potential and functional heterogeneity of Actinobacteria. The restriction profiles were scored as binary data and visualized in UPGMA dendrograms. Notably, strains affiliated with Streptomyces and Nocardiopsis showed relatively high biosynthetic richness and heterogeneity among the actinobacterial strains. Indeed, those that had a close relation in the 16S rRNA gene-based phylogeny also showed significant heterogeneity, which suggested a quite diverse biosynthetic potential even among closely related isolates. Sequence analysis of randomly selected PKS-II and NRPS fragments revealed their relative similarity to naphthoquinone and anthracycline group compound producing strains. Based on the phylogenetic novelty and biosynthetic richness, three streptomycete strains, JJ36, JJ38, and JJ66 were selected, and their genomes were sequenced using Illumina platform. Resulted draft genome sizes were 6.45, 5.89 and 4.83 Mb, respectively. Genome mining of the assembled draft genomes for secondary metabolite biosynthetic gene clusters relevant to bioactive compounds production was performed using antiSMASH. Interestingly, results revealed the presence of a total 183 secondary metabolite biosynthetic gene clusters including 109 putative gene clusters in the three genomes. This genomic data was further mapped with secondary metabolites profile of particular strains and resulted in the identification of novel compounds affiliated with aromatic ketones.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame machine prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. The Gemma side is a direct model label for every work in the frame, read from the title-only record. The Codex side is a classifier learned from the 10,348 direct Codex labels and calibrated to design-weighted sample rates; fields without enough sample support carry no Codex call. Candidate is the union of the two sides; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels.
Distilled classifier scores by category (both heads)
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.000 | 0.001 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.000 | 0.001 |
| Bibliometrics | 0.002 | 0.002 |
| Science and technology studies | 0.000 | 0.000 |
| Scholarly communication | 0.001 | 0.001 |
| Open science | 0.000 | 0.000 |
| Research integrity | 0.000 | 0.000 |
| Insufficient payload (model declined to judge) | 0.001 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one source (direct Gemma or distilled Codex), not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".