Burkholderia cepacia Complex Taxon K: Where to Split?
Bibliographic record
Abstract
The objective of the present study was to provide an updated classification for Burkholderia cepacia complex (Bcc) taxon K isolates. A representative set of 39 taxon K isolates were analyzed through multilocus sequence typing (MLST) and phylogenomic analyses. MLST analysis revealed the presence of at least six clusters of sequence types (STs) within taxon K, two of which contain the type strains of Burkholderia contaminans (ST-102) and Burkholderia lata (ST-101), and four corresponding to the previously defined taxa Other Bcc groups C, G, H and M. This clustering was largely supported by a phylogenomic tree which revealed three main clades. Isolates of B. contaminans and of Other Bcc groups C, G and H represented a first clade which generally shared average nucleotide identity (ANI) and average digital DNA-DNA hybridization (dDDH) values at or above the 95-96% ANI and 70% dDDH thresholds for species delineation. A second clade consisted of Other Bcc group M bacteria and of four B. lata isolates and was supported by average ANI and dDDH values of 97.2% and 76.1% within this clade and average ANI and dDDH values of 94.5% and 57.2% towards the remaining B. lata isolates (including the type strain), which represented a third clade. We therefore concluded that isolates known as Other Bcc groups C, G and H should be classified as B. contaminans, and propose a novel species, Burkholderia aenigmatica sp. nov., to accommodate Other Bcc M and B. lata ST-98, ST-103 and ST-119 isolates. Optimized MALDI-TOF MS databases for the identification of clinical Burkholderia isolates may provide correct species-level identification for some of these bacteria but would identify most of them as B. cepacia complex. MLST facilitates species-level identification of many taxon K strains but some may require comparative genomics for accurate species-level assignment. Finally, the inclusion of Other Bcc groups C, G and H into B. contaminans affects the phenotype of this species minimally and the proposal to classify Other Bcc group M and B. lata ST-98, ST-103 and ST-119 strains as a novel Burkholderia species is supported by a distinctive phenotype, i.e. growth at 42°C and lysine decarboxylase activity.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame machine prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. The Gemma side is a direct model label for every work in the frame, read from the title-only record. The Codex side is a classifier learned from the 10,348 direct Codex labels and calibrated to design-weighted sample rates; fields without enough sample support carry no Codex call. Candidate is the union of the two sides; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels.
Distilled classifier scores by category (both heads)
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.001 | 0.002 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.001 | 0.000 |
| Bibliometrics | 0.001 | 0.001 |
| Science and technology studies | 0.001 | 0.001 |
| Scholarly communication | 0.003 | 0.003 |
| Open science | 0.001 | 0.001 |
| Research integrity | 0.001 | 0.001 |
| Insufficient payload (model declined to judge) | 0.002 | 0.001 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one source (direct Gemma or distilled Codex), not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".