Investigation of NLR Genes Reveals Divergent Evolution on NLRome in Diploid and Polyploid Species in Genus Trifolium
Bibliographic record
Abstract
Crop wild relatives contain a greater variety of phenotypic and genotypic diversity compared to their domesticated counterparts. Trifolium crop species have limited genetic diversity to cope with biotic and abiotic stresses due to artificial selection for consumer preferences. Here, we investigated the distribution and evolution of nucleotide-binding site leucine-rich repeat receptor (NLR) genes in the genus of Trifolium with the objective to identify reference NLR genes. We identified 412, 350, 306, 389 and 241 NLR genes were identified from Trifolium. subterraneum, T. pratense, T. occidentale, subgenome-A of T. repens and subgenome-B of T. repens, respectively. Phylogenetic and clustering analysis reveals seven sub-groups in genus Trifolium. Specific subgroups such as G4-CNL, CCG10-CNL and TIR-CNL show distinct duplication patterns in specific species, which suggests subgroup duplications that are the hallmarks of their divergent evolution. Furthermore, our results strongly suggest the overall expansion of NLR repertoire in T. subterraneum is due to gene duplication events and birth of gene families after speciation. Moreover, the NLRome of the allopolyploid species T. repens has evolved asymmetrically, with the subgenome -A showing expansion, while the subgenome-B underwent contraction. These findings provide crucial background data for comprehending NLR evolution in the Fabaceae family and offer a more comprehensive analysis of NLR genes as disease resistance genes.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame distilled prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.
Codex and Gemma teacher scores by category
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.000 | 0.000 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.000 | 0.000 |
| Bibliometrics | 0.000 | 0.001 |
| Science and technology studies | 0.000 | 0.000 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.000 | 0.000 |
| Research integrity | 0.000 | 0.000 |
| Insufficient payload (model declined to judge) | 0.000 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one teacher head, not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".