Genome-wide analysis of MAPKKKs shows expansion and evolution of a new MEKK class involved in solanaceous species sexual reproduction
Bibliographic record
Abstract
BACKGROUND: Members of the plant MAP Kinases superfamily have been mostly studied in Arabidopsis thaliana and little is known in most other species. In Solanum chacoense, a wild species close to the common potato, it had been reported that members of a specific group in the MEKK subfamily, namely ScFRK1 and ScFRK2, are involved in male and female reproductive development. Apart from these two kinases, almost nothing is known about the roles of this peculiar family. METHODS: MEKKs were identified using BLAST and hidden Markov model (HMM) to build profiles using the 21 MEKKs from A. thaliana. Following protein sequence alignments, the neighbor-joining method was used to reconstruct phylogenetic trees of the MEKK subfamily. Kinase subdomains sequence logos were generated with WebLogo in order to pinpoint FRK distinct motifs. Codon alignments of the FRKs kinase subdomains and maximum-likelihood phylogenetic trees were used in the codon substitution models of the codeml program in the PAML package to detect selective pressure between FRK groups. RESULTS: With the recent progress in Next-Generation Sequencing technologies, the genomes and transcriptomes of numerous plant species have been recently sequenced, giving access to a vast amount of data. With the aim of finding all members of the MEKK subfamily members in plants, we screened the genomes of 15 species from different clades of the plant kingdom. Interestingly, the whole MEKK subfamily has significantly expanded throughout evolution, especially in solanaceous species. This holds true for members of the FRK class, which have also strongly expanded and diverged. CONCLUSIONS: Expansion and rapid evolution of the FRK class members in solanaceous species support the hypothesis that they have acquired new roles, mainly in male and female reproductive development.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame machine prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. The Gemma side is a direct model label for every work in the frame, read from the title-only record. The Codex side is a classifier learned from the 10,348 direct Codex labels and calibrated to design-weighted sample rates; fields without enough sample support carry no Codex call. Candidate is the union of the two sides; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels.
Distilled classifier scores by category (both heads)
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.000 | 0.000 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.000 | 0.001 |
| Bibliometrics | 0.001 | 0.001 |
| Science and technology studies | 0.000 | 0.000 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.000 | 0.000 |
| Research integrity | 0.000 | 0.000 |
| Insufficient payload (model declined to judge) | 0.002 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one source (direct Gemma or distilled Codex), not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".