The PAX3 and 7 homeodomains have evolved unique determinants that influence DNA-binding, structure and communication with the paired domain
Bibliographic record
Abstract
ABSTRACT The PAX ( pa ired bo x ) family is a collection of metazoan transcription factors defined by the paired domain, which confers sequence-specific DNA-binding. Ancestral PAX proteins also contained a homeodomain, which can communicate with the paired domain to modulate DNA-binding. In the present study, we sought to identify determinants of this functional interaction using the paralogous PAX3 and 7 proteins. First, we evaluated a group of heterologous paired domains and homeodomains for the ability to bind DNA cooperatively through formation of a ternary complex (paired domain:homeodomain:DNA). This revealed that capacity for ternary complex formation was unique to the PAX3 and 7 homeodomains and therefore not simply a consequence of DNA-binding. We also found PAX3 and 7 were distinguished by an extended region of conservation N-terminal to the homeodomain (NTE). Phylogenetic analyses established the NTE was restricted to PAX3/7 orthologs of segmented metazoans, indicating it arose in a bilaterian precursor prior to separation of deuterostomes and protostomes. In DNA-binding assays, presence of the NTE caused a decrease in monomeric binding by the PAX3 homeodomain that reflected a lack of secondary structure in 1D- 1 H-NMR. Nevertheless, this inhibitory effect could be overcome by homeodomain dimerization or cooperative binding with the paired domain, establishing that protein interactions could induce homeodomain folding in the presence of the NTE. Strikingly, the PAX7 counterpart did not impair homeodomain binding, revealing inherent differences that could account for its distinct target profile in vivo. Collectively, these findings identify critical determinants of PAX3 and 7 activity, which contribute to their functional diversification.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame machine prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. The Gemma side is a direct model label for every work in the frame, read from the title-only record. The Codex side is a classifier learned from the 10,348 direct Codex labels and calibrated to design-weighted sample rates; fields without enough sample support carry no Codex call. Candidate is the union of the two sides; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels.
Distilled classifier scores by category (both heads)
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.000 | 0.000 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.000 | 0.000 |
| Bibliometrics | 0.000 | 0.000 |
| Science and technology studies | 0.000 | 0.000 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.000 | 0.000 |
| Research integrity | 0.000 | 0.000 |
| Insufficient payload (model declined to judge) | 0.001 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one source (direct Gemma or distilled Codex), not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".