Multilocus Sequence Typing of Historical <i>Burkholderia pseudomallei</i> Isolates Collected in Southeast Asia from 1964 to 1967 Provides Insight into the Epidemiology of Melioidosis
Bibliographic record
Abstract
A collection of 207 historically relevant Burkholderia pseudomallei isolates was analyzed by multilocus sequence typing (MLST). The strain collection contains environmental isolates obtained from a geographical distribution survey of B. pseudomallei isolates in Thailand (1964 to 1967), as well as stock cultures and colony variants from the U.S. Army Medical Research Unit (Malaysia), the Walter Reed Army Institute for Research, and the Pasteur Institute (Vietnam). The 207 isolates of the collection were resolved into 80 sequence types (STs); 56 of these were novel. eBURST diagrams predict that the historical-collection STs segregate into three complexes when analyzed separately. When added to the 760 isolates and 365 STs of the B. pseudomallei MLST database, the historical-collection STs cluster significantly within the main complex of the eBURST diagram in an ancestral pattern and alter the B. pseudomallei "population snapshot." Differences in colony morphology among reference isolates were found not to affect the STs assigned, which were consistent with the original isolates. Australian ST84 is likely characteristic of B. pseudomallei isolates of Southeast Asia rather than Australia, since multiple environmental isolates from Thailand and Malaysia share this ST with the single Australian clinical isolate in the MLST database. Phylogenetic evidence is also provided suggesting that Australian isolates may not be distinct from those of Thailand, since ST60 is common to environmental isolates from both countries. MLST and eBURST are useful tools for the study of population biology and epidemiology, since they provide methods to elucidate new genetic relationships among bacterial isolates.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame distilled prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.
Codex and Gemma teacher scores by category
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.003 | 0.006 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.002 | 0.001 |
| Bibliometrics | 0.000 | 0.001 |
| Science and technology studies | 0.000 | 0.000 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.000 | 0.000 |
| Research integrity | 0.001 | 0.001 |
| Insufficient payload (model declined to judge) | 0.000 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one teacher head, not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".