BRCA1 Circos: a visualisation resource for functional analysis of missense variants
Bibliographic record
Abstract
BACKGROUND: Inactivating germline mutations in the tumour suppressor gene BRCA1 are associated with a significantly increased risk of developing breast and ovarian cancer. A large number (>1500) of unique BRCA1 variants have been identified in the population and can be classified as pathogenic, non-pathogenic or as variants of unknown significance (VUS). Many VUS are rare missense variants leading to single amino acid changes. Their impact on protein function cannot be directly inferred from sequence information, precluding assessment of their pathogenicity. Thus, functional assays are critical to assess the impact of these VUS on protein activity. BRCA1 is a multifunctional protein and different assays have been used to assess the impact of variants on different biochemical activities and biological processes. METHODS AND RESULTS: To facilitate VUS analysis, we have developed a visualisation resource that compiles and displays functional data on all documented BRCA1 missense variants. BRCA1 Circos is a web-based visualisation tool based on the freely available Circos software package. The BRCA1 Circos web tool (http://research.nhgri.nih.gov/bic/circos/) aggregates data from all published BRCA1 missense variants for functional studies, harmonises their results and presents various functionalities to search and interpret individual-level functional information for each BRCA1 missense variant. CONCLUSIONS: This research visualisation tool will serve as a quick one-stop publically available reference for all the BRCA1 missense variants that have been functionally assessed. It will facilitate meta-analysis of functional data and improve assessment of pathogenicity of VUS.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame distilled prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.
Codex and Gemma teacher scores by category
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.001 | 0.001 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.000 | 0.000 |
| Bibliometrics | 0.000 | 0.000 |
| Science and technology studies | 0.000 | 0.000 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.000 | 0.000 |
| Research integrity | 0.000 | 0.000 |
| Insufficient payload (model declined to judge) | 0.000 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one teacher head, not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".