<i>RadiationGeneSigDB</i> : a database of oxic and hypoxic radiation response gene signatures and their utility in pre-clinical research
Bibliographic record
Abstract
Objective: Radiation therapy is among the most effective and widely used modalities of cancer therapy in current clinical practice. In this era of personalized radiation medicine, high-throughput data now provide the means to investigate novel biomarkers of radiation response. Large-scale efforts have identified several radiation response signatures, which poses two challenges, namely, their analytical validity and redundancy of gene signatures. Methods: To address these fundamental radiogenomics questions, we curated a database of gene expression signatures predictive of radiation response under oxic and hypoxic conditions. RadiationGeneSigDB has a collection of 11 oxic and 24 hypoxic signatures with the standardized gene list as a gene symbol, Entrez gene ID, and its function. We present the utility of this database by gaining an understanding of hypoxia-associated miRNA by applying a penalized multivariate model; by comparing breast cancer oxic signatures in cell line data vs patient data; and by comparing the similarity of head and neck cancer hypoxia signatures at the pathway level in clinical tumour data. Results: We obtained a set of miRNA highly associated both positively and negatively to the hypoxia gene signatures, across pan-cancer. In addition, we identified moderate correlations between breast cancer oxic signatures in patient data, and significant differences across molecular subtypes. Moreover, we also found that different set of pathways to be enriched using the head and neck hypoxia signatures, although, they are found to be concordant when applied on the patient data. Conclusion: This valuable, curated repertoire of published gene expression signatures provides motivating case studies for how to search for similarities in radiation response for tumours arising from different tissues across model systems under oxic and hypoxic conditions, and how a well-curated set of gene signatures can be used to generate novel biological hypotheses about the functions of non-coding RNA. Advances in knowledge: We envision that RadiationSigDB database will help accelerate preclinical radiotherapeutic discovery pipelines in terms of analytical validity of novel biomarkers of radiation response and the need for ensemble approaches to clinical genomic biomarkers.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame distilled prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.
Codex and Gemma teacher scores by category
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.004 | 0.001 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.000 | 0.000 |
| Bibliometrics | 0.000 | 0.000 |
| Science and technology studies | 0.000 | 0.000 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.000 | 0.000 |
| Research integrity | 0.000 | 0.000 |
| Insufficient payload (model declined to judge) | 0.000 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one teacher head, not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".