Influence of C-5 substituted cytosine and related nucleoside analogs on the formation of benzo[a]pyrene diol epoxide-dG adducts at CG base pairs of DNA
Bibliographic record
Abstract
Endogenous 5-methylcytosine ((Me)C) residues are found at all CG dinucleotides of the p53 tumor suppressor gene, including the mutational 'hotspots' for smoking induced lung cancer. (Me)C enhances the reactivity of its base paired guanine towards carcinogenic diolepoxide metabolites of polycyclic aromatic hydrocarbons (PAH) present in cigarette smoke. In the present study, the structural basis for these effects was investigated using a series of unnatural nucleoside analogs and a representative PAH diolepoxide, benzo[a]pyrene diolepoxide (BPDE). Synthetic DNA duplexes derived from a frequently mutated region of the p53 gene (5'-CCCGGCACCC GC[(15)N(3),(13)C(1)-G]TCCGCG-3', + strand) were prepared containing [(15)N(3), (13)C(1)]-guanine opposite unsubstituted cytosine, (Me)C, abasic site, or unnatural nucleobase analogs. Following BPDE treatment and hydrolysis of the modified DNA to 2'-deoxynucleosides, N(2)-BPDE-dG adducts formed at the [(15)N(3), (13)C(1)]-labeled guanine and elsewhere in the sequence were quantified by mass spectrometry. We found that C-5 alkylcytosines and related structural analogs specifically enhance the reactivity of the base paired guanine towards BPDE and modify the diastereomeric composition of N(2)-BPDE-dG adducts. Fluorescence and molecular docking studies revealed that 5-alkylcytosines and unnatural nucleobase analogs with extended aromatic systems facilitate the formation of intercalative BPDE-DNA complexes, placing BPDE in a favorable orientation for nucleophilic attack by the N(2) position of guanine.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame machine prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. The Gemma side is a direct model label for every work in the frame, read from the title-only record. The Codex side is a classifier learned from the 10,348 direct Codex labels and calibrated to design-weighted sample rates; fields without enough sample support carry no Codex call. Candidate is the union of the two sides; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels.
Distilled classifier scores by category (both heads)
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.000 | 0.000 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.000 | 0.000 |
| Bibliometrics | 0.000 | 0.000 |
| Science and technology studies | 0.000 | 0.000 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.000 | 0.000 |
| Research integrity | 0.000 | 0.000 |
| Insufficient payload (model declined to judge) | 0.001 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one source (direct Gemma or distilled Codex), not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".