Genomic Analysis Distinguishes <i>Mycobacterium africanum</i>
Bibliographic record
Abstract
Mycobacterium africanum is thought to comprise a unique species within the Mycobacterium tuberculosis complex. M. africanum has traditionally been identified by phenotypic criteria, occupying an intermediate position between M. tuberculosis and M. bovis according to biochemical characteristics. Although M. africanum isolates present near-identical sequence homology to other species of the M. tuberculosis complex, several studies have uncovered large genomic regions variably deleted from certain M. africanum isolates. To further investigate the genomic characteristics of organisms characterized as M. africanum, the DNA content of 12 isolates was interrogated by using Affymetrix GeneChip. Analysis revealed genomic regions of M. tuberculosis deleted from all isolates of putative diagnostic and biological consequence. The distribution of deleted sequences suggests that M. africanum subtype II isolates are situated among strains of "modern" M. tuberculosis. In contrast, other M. africanum isolates (subtype I) constitute two distinct evolutionary branches within the M. tuberculosis complex. To test for an association between deleted sequences and biochemical attributes used for speciation, a phenotypically diverse panel of "M. africanum-like" isolates from Guinea-Bissau was tested for these deletions. These isolates clustered together within one of the M. africanum subtype I branches, irrespective of phenotype. These results indicate that convergent biochemical profiles can be independently obtained for M. tuberculosis complex members, challenging the traditional approach to M. tuberculosis complex speciation. Furthermore, the genomic results suggest a rational framework for defining M. africanum and provide tools to accurately assess its prevalence in clinical specimens.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame distilled prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.
Codex and Gemma teacher scores by category
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.003 | 0.009 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.002 | 0.001 |
| Bibliometrics | 0.000 | 0.000 |
| Science and technology studies | 0.000 | 0.000 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.000 | 0.000 |
| Research integrity | 0.000 | 0.001 |
| Insufficient payload (model declined to judge) | 0.000 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one teacher head, not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".