Emergence and Pathogenicity of Highly Virulent Cryptococcus gattii Genotypes in the Northwest United States
Bibliographic record
Abstract
Cryptococcus gattii causes life-threatening disease in otherwise healthy hosts and to a lesser extent in immunocompromised hosts. The highest incidence for this disease is on Vancouver Island, Canada, where an outbreak is expanding into neighboring regions including mainland British Columbia and the United States. This outbreak is caused predominantly by C. gattii molecular type VGII, specifically VGIIa/major. In addition, a novel genotype, VGIIc, has emerged in Oregon and is now a major source of illness in the region. Through molecular epidemiology and population analysis of MLST and VNTR markers, we show that the VGIIc group is clonal and hypothesize it arose recently. The VGIIa/IIc outbreak lineages are sexually fertile and studies support ongoing recombination in the global VGII population. This illustrates two hallmarks of emerging outbreaks: high clonality and the emergence of novel genotypes via recombination. In macrophage and murine infections, the novel VGIIc genotype and VGIIa/major isolates from the United States are highly virulent compared to similar non-outbreak VGIIa/major-related isolates. Combined MLST-VNTR analysis distinguishes clonal expansion of the VGIIa/major outbreak genotype from related but distinguishable less-virulent genotypes isolated from other geographic regions. Our evidence documents emerging hypervirulent genotypes in the United States that may expand further and provides insight into the possible molecular and geographic origins of the outbreak.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame machine prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. The Gemma side is a direct model label for every work in the frame, read from the title-only record. The Codex side is a classifier learned from the 10,348 direct Codex labels and calibrated to design-weighted sample rates; fields without enough sample support carry no Codex call. Candidate is the union of the two sides; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels.
Distilled classifier scores by category (both heads)
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.000 | 0.001 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.000 | 0.000 |
| Bibliometrics | 0.000 | 0.000 |
| Science and technology studies | 0.000 | 0.000 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.000 | 0.000 |
| Research integrity | 0.000 | 0.000 |
| Insufficient payload (model declined to judge) | 0.000 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one source (direct Gemma or distilled Codex), not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".