A neuron-specific microexon ablates the novel DNA-binding function of a histone H3K4me0 reader PHF21A
Bibliographic record
Abstract
Abstract How cell-type-specific chromatin landscapes emerge and progress during metazoan ontogenesis remains an important question. Transcription factors are expressed in a cell-type-specific manner and recruit chromatin-regulatory machinery to specific genomic loci. In contrast, chromatin-regulatory proteins are expressed broadly and are assumed to exert the same intrinsic function across cell types. However, human genetics studies have revealed an unexpected vulnerability of neurodevelopment to chromatin factor mutations with unknown mechanisms. Here, we report that 14 chromatin regulators undergo evolutionary-conserved neuron-specific splicing events involving microexons. Of the 14 chromatin regulators, two are integral components of a histone H3K4 demethylase complex; the catalytic subunit LSD1 and an H3K4me0-reader protein PHF21A adopt neuron-specific forms. We found that canonical PHF21A (PHF21A-c) binds to DNA by AT-hook motif, and the neuronal counterpart PHF21A-n lacks this DNA-binding function yet maintains H3K4me0 recognition intact. In-vitro reconstitution of the canonical and neuronal PHF21A-LSD1 complexes identified the neuronal complex as a hypomorphic H3K4 demethylating machinery with reduced nucleosome engagement. Furthermore, an autism-associated PHF21A missense mutation, 1285 G>A, at the last nucleotide of the common exon immediately upstream of the neuronal microexon led to impaired splicing of PHF21A -n. Thus, ubiquitous chromatin regulatory complexes exert unique intrinsic functions in neurons via alternative splicing of their subunits and potentially contribute to faithful human brain development.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame machine prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. The Gemma side is a direct model label for every work in the frame, read from the title-only record. The Codex side is a classifier learned from the 10,348 direct Codex labels and calibrated to design-weighted sample rates; fields without enough sample support carry no Codex call. Candidate is the union of the two sides; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels.
Distilled classifier scores by category (both heads)
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.000 | 0.000 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.000 | 0.000 |
| Bibliometrics | 0.000 | 0.000 |
| Science and technology studies | 0.000 | 0.000 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.000 | 0.000 |
| Research integrity | 0.000 | 0.000 |
| Insufficient payload (model declined to judge) | 0.002 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one source (direct Gemma or distilled Codex), not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".