Fusobacterium nucleatum subsp. polymorphum recovered from malignant and potentially malignant oral disease exhibit heterogeneity in adhesion phenotypes and adhesin gene copy number, shaped by inter-subspecies horizontal gene transfer and recombination-derived mosaicism
Bibliographic record
Abstract
Fusobacterium nucleatum is an anaerobic commensal of the oral cavity associated with periodontitis and extra-oral diseases, including colorectal cancer. Previous studies have shown an increased relative abundance of this bacterium associated with oral dysplasia or within oral tumours. Using direct culture, we found that 75 % of Fusobacterium species isolated from malignant or potentially malignant oral mucosa were F. nucleatum subsp. polymorphum. Whole genome sequencing and pangenome analysis with Panaroo was carried out on 76 F. nucleatum subsp. polymorphum genomes. F. nucleatum subsp. polymorphum was shown to possesses a relatively small core genome of 1604 genes in a pangenome of 7363 genes. Phylogenetic analysis based on the core genome shows the isolates can be separated into three main clades with no obvious genotypic associations with disease. Isolates recovered from healthy and diseased sites in the same patient are generally highly related. A large repertoire of adhesins belonging to the type V secretion system (TVSS) could be identified with major variation in repertoire and copy number between strains. Analysis of intergenic recombination using fastGEAR showed that adhesin complement is shaped by horizontal gene transfer and recombination. Recombination events at TVSS adhesin genes were not only common between lineages of subspecies polymorphum, but also between different subspecies of F. nucleatum. Strains of subspecies polymorphum with low copy numbers of TVSS adhesin encoding genes tended to have the weakest adhesion to oral keratinocytes. This study highlights the genetic heterogeneity of F. nucleatum subsp. polymorphum and provides a new framework for defining virulence in this organism.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame machine prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. The Gemma side is a direct model label for every work in the frame, read from the title-only record. The Codex side is a classifier learned from the 10,348 direct Codex labels and calibrated to design-weighted sample rates; fields without enough sample support carry no Codex call. Candidate is the union of the two sides; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels.
Distilled classifier scores by category (both heads)
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.000 | 0.000 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.000 | 0.000 |
| Bibliometrics | 0.001 | 0.001 |
| Science and technology studies | 0.000 | 0.000 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.000 | 0.000 |
| Research integrity | 0.000 | 0.000 |
| Insufficient payload (model declined to judge) | 0.001 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one source (direct Gemma or distilled Codex), not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".