Functional Domains and Evolutionary History of the PMEL and GPNMB Family Proteins
Bibliographic record
Abstract
The ancient paralogs premelanosome protein (PMEL) and glycoprotein nonmetastatic melanoma protein B (GPNMB) have independently emerged as intriguing disease loci in recent years. Both proteins possess common functional domains and variants that cause a shared spectrum of overlapping phenotypes and disease associations: melanin-based pigmentation, cancer, neurodegenerative disease and glaucoma. Surprisingly, these proteins have yet to be shown to physically or genetically interact within the same cellular pathway. This juxtaposition inspired us to compare and contrast this family across a breadth of species to better understand the divergent evolutionary trajectories of two related, but distinct, genes. In this study, we investigated the evolutionary history of PMEL and GPNMB in clade-representative species and identified TMEM130 as the most ancient paralog of the family. By curating the functional domains in each paralog, we identified many commonalities dating back to the emergence of the gene family in basal metazoans. PMEL and GPNMB have gained functional domains since their divergence from TMEM130, including the core amyloid fragment (CAF) that is critical for the amyloid potential of PMEL. Additionally, the PMEL gene has acquired the enigmatic repeat domain (RPT), composed of a variable number of imperfect tandem repeats; this domain acts in an accessory role to control amyloid formation. Our analyses revealed the vast variability in sequence, length and repeat number in homologous RPT domains between craniates, even within the same taxonomic class. We hope that these analyses inspire further investigation into a gene family that is remarkable from the evolutionary, pathological and cell biology perspectives.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame distilled prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.
Codex and Gemma teacher scores by category
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.000 | 0.000 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.000 | 0.000 |
| Bibliometrics | 0.000 | 0.000 |
| Science and technology studies | 0.000 | 0.000 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.000 | 0.000 |
| Research integrity | 0.000 | 0.000 |
| Insufficient payload (model declined to judge) | 0.000 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one teacher head, not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".