The sole LSm complex in <i>Cyanidioschyzon merolae</i> associates with pre-mRNA splicing and mRNA degradation factors
Bibliographic record
Abstract
Proteins of the Sm and Sm-like (LSm) families, referred to collectively as (L)Sm proteins, are found in all three domains of life and are known to promote a variety of RNA processes such as base-pair formation, unwinding, RNA degradation, and RNA stabilization. In eukaryotes, (L)Sm proteins have been studied, inter alia, for their role in pre-mRNA splicing. In many organisms, the LSm proteins form two distinct complexes, one consisting of LSm1–7 that is involved in mRNA degradation in the cytoplasm, and the other consisting of LSm2–8 that binds spliceosomal U6 snRNA in the nucleus. We recently characterized the splicing proteins from the red alga Cyanidioschyzon merolae and found that it has only seven LSm proteins. The identities of CmLSm2–CmLSm7 were unambiguous, but the seventh protein was similar to LSm1 and LSm8. Here, we use in vitro binding measurements, microscopy, and affinity purification-mass spectrometry to demonstrate a canonical splicing function for the C. merolae LSm complex and experimentally validate our bioinformatic predictions of a reduced spliceosome in this organism. Copurification of Pat1 and its associated mRNA degradation proteins with the LSm proteins, along with evidence of a cytoplasmic fraction of CmLSm complexes, argues that this complex is involved in both splicing and cytoplasmic mRNA degradation. Intriguingly, the Pat1 complex also copurifies with all four snRNAs, suggesting the possibility of a spliceosome-associated pre-mRNA degradation complex in the nucleus.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame distilled prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.
Codex and Gemma teacher scores by category
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.000 | 0.000 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.000 | 0.000 |
| Bibliometrics | 0.000 | 0.000 |
| Science and technology studies | 0.001 | 0.000 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.000 | 0.000 |
| Research integrity | 0.000 | 0.000 |
| Insufficient payload (model declined to judge) | 0.000 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one teacher head, not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".