The combinatorial control of alternative splicing in C. elegans
Bibliographic record
Abstract
Normal development requires the right splice variants to be made in the right tissues at the right time. The core splicing machinery is engaged in all splicing events, but which precise splice variant is made requires the choice between alternative splice sites-for this to occur, a set of splicing factors (SFs) must recognize and bind to short RNA motifs in the pre-mRNA. In C. elegans, there is known to be extensive variation in splicing patterns across development, but little is known about the targets of each SF or how multiple SFs combine to regulate splicing. Here we combine RNA-seq with in vitro binding assays to study how 4 different C. elegans SFs, ASD-1, FOX-1, MEC-8, and EXC-7, regulate splicing. The 4 SFs chosen all have well-characterised biology and well-studied loss-of-function genetic alleles, and all contain RRM domains. Intriguingly, while the SFs we examined have varied roles in C. elegans development, they show an unexpectedly high overlap in their targets. We also find that binding sites for these SFs occur on the same pre-mRNAs more frequently than expected suggesting extensive combinatorial control of splicing. We confirm that regulation of splicing by multiple SFs is often combinatorial and show that this is functionally significant. We also find that SFs appear to combine to affect splicing in two modes-they either bind in close proximity within the same intron or they appear to bind to separate regions of the intron in a conserved order. Finally, we find that the genes whose splicing are regulated by multiple SFs are highly enriched for genes involved in the cytoskeleton and in ion channels that are key for neurotransmission. Together, this shows that specific classes of genes have complex combinatorial regulation of splicing and that this combinatorial regulation is critical for normal development to occur.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame distilled prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.
Codex and Gemma teacher scores by category
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.000 | 0.000 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.000 | 0.000 |
| Bibliometrics | 0.000 | 0.000 |
| Science and technology studies | 0.000 | 0.000 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.000 | 0.000 |
| Research integrity | 0.000 | 0.000 |
| Insufficient payload (model declined to judge) | 0.000 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one teacher head, not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".