Analysis of <i>rbc</i>L sequences reveals the global biodiversity, community structure, and biogeographical pattern of thermoacidophilic red algae (Cyanidiales)
Bibliographic record
Abstract
Thermoacidophilic cyanidia (Cyanidiales) are the primary photosynthetic eukaryotes in volcanic areas. These red algae also serve as important model organisms for studying life in extreme habitats. The global biodiversity and community structure of Cyanidiales remain unclear despite previous sampling efforts. Here, we surveyed the Cyanidiales biodiversity in the Tatun Volcano Group (TVG) area in Taiwan using environmental DNA sequencing. We generated 174 rbcL sequences from eight samples from four regions in the TVG area, and combined them with 239 publicly available rbcL sequences collected worldwide. Species delimita-tion using this large rbcL data set suggested at least 20 Cyanidiales OTUs (operational taxono-mic units) worldwide, almost three times the presently recognized seven species. Results from environmental DNA showed that OTUs in the TVG area were divided into three groups: (i) dominant in hot springs with 92%-99% sequence identity to Galdieria maxima; (ii) largely distributed in drier and more acidic microhabitats with 99% identity to G. partita; and (iii) primarily distributed in cooler microhabitats and lacking identity to known cyanidia species (a novel Cyanidiales lineage). In both global and individual area analyses, we observed greater species diversity in non-aquatic than aquatic habitats. Community structure analysis showed high similarity between the TVG community and West Pacific-Iceland communities, reflecting their geographic proximity to each other. Our study is the first examination of the global species diversity and biogeographic affinity of cyanidia. Additionally, our data illuminate the influence of microhabitat type on Cyanidiales diversity and highlight intriguing questions for future ecological research.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame distilled prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.
Codex and Gemma teacher scores by category
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.001 | 0.000 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.001 | 0.000 |
| Bibliometrics | 0.000 | 0.001 |
| Science and technology studies | 0.000 | 0.002 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.001 | 0.000 |
| Research integrity | 0.000 | 0.000 |
| Insufficient payload (model declined to judge) | 0.001 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one teacher head, not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".