Analyses of the Global Multilocus Genotypes of the Human Pathogenic Yeast Candida tropicalis
Bibliographic record
Abstract
Candida tropicalis is a globally distributed human pathogenic yeast, especially prevalent in tropical and sub-tropical regions. Over the last several decades, a large number of studies have been published on the genetic diversity and molecular epidemiology of C. tropicalis from different parts of the world. However, the global pattern of genetic variation remains largely unknown. Here we analyzed the published multilocus sequence data at six loci for 876 isolates from 16 countries representing five continents. Our results showed that 280 of the 2677 (10.5%) analyzed nucleotides were polymorphic, resulting in a mean of 82 (a range of 38 to 150) genotypes per locus and a total of 633 combined diploid sequence types (DSTs). Among these, 93 combined DSTs were shared by 336 strains, including 10 by strains from different continents. Analysis of Molecular Variance (AMOVA) showed that 89% of the observed genetic variations were found within regional and national populations while <10% was due to among-country separations. Pairwise geographic population analyses showed overall low but statistically significant genetic differentiation between most geographic populations, with the Singaporean and Indian populations being the most distinct from other populations. However, the Mantel test showed no significant correlation between genetic distance and geographic distance among the geographic populations. Consistent with high genetic variation within and limited variations among geographic populations, results from STRUCTURE analyses showed that the 876 isolates could be grouped into 15 genetic clusters, with each cluster having a broad geographic distribution. Together, our results suggest frequent gene flows among certain regional, national, and continental populations of C. tropicalis, resulting in abundant regional and national genetic diversities of this important human fungal pathogen.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame machine prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. The Gemma side is a direct model label for every work in the frame, read from the title-only record. The Codex side is a classifier learned from the 10,348 direct Codex labels and calibrated to design-weighted sample rates; fields without enough sample support carry no Codex call. Candidate is the union of the two sides; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels.
Distilled classifier scores by category (both heads)
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.000 | 0.001 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.000 | 0.000 |
| Bibliometrics | 0.001 | 0.001 |
| Science and technology studies | 0.000 | 0.000 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.000 | 0.000 |
| Research integrity | 0.000 | 0.000 |
| Insufficient payload (model declined to judge) | 0.001 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one source (direct Gemma or distilled Codex), not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".