Molecular Epidemiology Reveals Genetic Diversity amongst Isolates of the Cryptococcus neoformans/C. gattii Species Complex in Thailand
Bibliographic record
Abstract
To gain a more detailed picture of cryptococcosis in Thailand, a retrospective study of 498 C. neoformans and C. gattii isolates has been conducted. Among these, 386, 83 and 29 strains were from clinical, environmental and veterinary sources, respectively. A total of 485 C. neoformans and 13 C. gattii strains were studied. The majority of the strains (68.9%) were isolated from males (mean age of 37.97 years), 88.5% of C. neoformans and only 37.5% of C. gattii strains were from HIV patients. URA5-RFLP and/or M13 PCR-fingerprinting analysis revealed that the majority of the isolates were C. neoformans molecular type VNI regardless of their sources (94.8%; 94.6% of the clinical, 98.8% of the environmental and 86.2% of the veterinary isolates). In addition, the molecular types VNII (2.4%; 66.7% of the clinical and 33.3% of the veterinary isolates), VNIV (0.2%; 100% environmental isolate), VGI (0.2%; 100% clinical isolate) and VGII (2.4%; 100% clinical isolates) were found less frequently. Multilocus Sequence Type (MLST) analysis using the ISHAM consensus MLST scheme for the C. neoformans/C. gattii species complex identified a total of 20 sequence types (ST) in Thailand combining current and previous data. The Thai isolates are an integrated part of the global cryptococcal population genetic structure, with ST30 for C. gattii and ST82, ST83, ST137, ST141, ST172 and ST173 for C. neoformans being unique to Thailand. Most of the C. gattii isolates were ST7 = VGIIb, which is identical to the less virulent minor Vancouver island outbreak genotype, indicating Thailand as a stepping stone in the global spread of this outbreak strain. The current study revealed a greater genetic diversity and a wider range of major molecular types being present amongst Thai cryptococcal isolates than previously reported.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame machine prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. The Gemma side is a direct model label for every work in the frame, read from the title-only record. The Codex side is a classifier learned from the 10,348 direct Codex labels and calibrated to design-weighted sample rates; fields without enough sample support carry no Codex call. Candidate is the union of the two sides; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels.
Distilled classifier scores by category (both heads)
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.000 | 0.001 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.000 | 0.000 |
| Bibliometrics | 0.002 | 0.001 |
| Science and technology studies | 0.000 | 0.000 |
| Scholarly communication | 0.001 | 0.000 |
| Open science | 0.000 | 0.000 |
| Research integrity | 0.000 | 0.000 |
| Insufficient payload (model declined to judge) | 0.001 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one source (direct Gemma or distilled Codex), not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".