Missense variants in <i>TAF1</i> and developmental phenotypes: Challenges of determining pathogenicity
Bibliographic record
Abstract
We recently described a new neurodevelopmental syndrome (TAF1/MRXS33 intellectual disability syndrome) (MIM# 300966) caused by pathogenic variants involving the X-linked gene TAF1, which participates in RNA polymerase II transcription. The initial study reported eleven families, and the syndrome was defined as presenting early in life with hypotonia, facial dysmorphia, and developmental delay that evolved into intellectual disability (ID) and/or autism spectrum disorder (ASD). We have now identified an additional 27 families through a genotype-first approach. Familial segregation analysis, clinical phenotyping, and bioinformatics were capitalized on to assess potential variant pathogenicity, and molecular modelling was performed for those variants falling within structurally characterized domains of TAF1. A novel phenotypic clustering approach was also applied, in which the phenotypes of affected individuals were classified using 51 standardized Human Phenotype Ontology (HPO) terms. Phenotypes associated with TAF1 variants show considerable pleiotropy and clinical variability, but prominent among previously unreported effects were brain morphological abnormalities, seizures, hearing loss, and heart malformations. Our allelic series broadens the phenotypic spectrum of TAF1/MRXS33 intellectual disability syndrome and the range of TAF1 molecular defects in humans. It also illustrates the challenges for determining the pathogenicity of inherited missense variants, particularly for genes mapping to chromosome X. This article is protected by copyright. All rights reserved.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame distilled prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.
Codex and Gemma teacher scores by category
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.000 | 0.000 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.000 | 0.000 |
| Bibliometrics | 0.000 | 0.000 |
| Science and technology studies | 0.000 | 0.000 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.000 | 0.000 |
| Research integrity | 0.000 | 0.000 |
| Insufficient payload (model declined to judge) | 0.000 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one teacher head, not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".