A phylogenetic approach for identifying new sources of economically important fatty acids in plants and algae
Bibliographic record
Abstract
Societal Impact Statement New sources of essential and other nutritionally and/or pharmacologically important fatty acids for human use can be discovered through bioprospecting. Demonstrating how fatty acids in plants and algae are distributed across a phylogeny is a critical first step in this process. Here, new sources of essential omega‐3 fatty acids that are critically important to human health were identified, such as green and Chromista algae, bryophytes and some angiosperm families. The identification of certain taxa that are high in critical fatty acids is important for developing new sources of these healthful compounds. Summary Essential fatty acids (EFA) including long‐chain omega‐3 and omega‐6 EFA are key nutrients that broadly support the health of individuals and the ecosystems they inhabit. Production of these EFA is critical to global nutritional security, and therefore, it is imperative to assess the distribution patterns of FA in plants to bioprospect new resources that do not threaten biodiversity. We used a meta‐analytic approach (phylogenetic signal analysis), which incorporated a wide variety of taxa to map FA profiles including both presence and abundance of specific FA and EFA onto a plant phylogeny including algae, bryophytes, pteridophytes, gymnosperms and angiosperms. Our phylogenetic signal analysis revealed that presence/absence of most FA of interest were randomly or ubiquitously distributed, while FA content was significantly clustered in specific clades for most key FA (including eicosapentaenoic acid (EPA), docosahexaenoic acid (DHA) and nervonic acid). Both green and Chromista algae displayed relatively high proportions of EPA and DHA, while green algae contained higher proportional amounts of stearidonic acid (SDA). Further, signal analysis revealed high proportional content of gamma‐linoleic acid (GLA) in bryophytes and some angiosperm families, suggesting potential sources for bioprospecting GLA. This analysis provides a crucial foundation for understanding the influence of phylogenetic relationships in the relative proportions of key FA in plants and algae across a diverse phylogeny. Furthermore, in demonstrating how fatty acids in plants and algae are distributed across a phylogeny, we provide a critical first step in bioprospecting for new sources of essential and other nutritionally‐ and/or pharmacologically important fatty acids for human use.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame machine prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. The Gemma side is a direct model label for every work in the frame, read from the title-only record. The Codex side is a classifier learned from the 10,348 direct Codex labels and calibrated to design-weighted sample rates; fields without enough sample support carry no Codex call. Candidate is the union of the two sides; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels.
Distilled classifier scores by category (both heads)
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.006 | 0.010 |
| Meta-epidemiology (narrow) | 0.001 | 0.001 |
| Meta-epidemiology (broad) | 0.002 | 0.005 |
| Bibliometrics | 0.015 | 0.011 |
| Science and technology studies | 0.003 | 0.001 |
| Scholarly communication | 0.002 | 0.002 |
| Open science | 0.001 | 0.002 |
| Research integrity | 0.001 | 0.002 |
| Insufficient payload (model declined to judge) | 0.003 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one source (direct Gemma or distilled Codex), not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".