Structural basis for error-free bypass of bulky adduct by DNA polymerase Polκ
Bibliographic record
Abstract
Humans are frequently exposed to the environmentally ubiquitous and potentially carcinigenic polycyclic aromatic hydrocarbon, benzo[a]pyrene (BP). BP is metabolized to highly reactive benzo[a]pyrene diol epoxides (BPDEs) in the cells. BPDEs react with DNA predominantly at the N2 position of guanine and form bulky adducts. The major BP adduct is (+)-trans-anti-[BP]-N2-dG (BP-N2-dG) that is carcinogenic. The bulky adduct block DNA synthesis by replicative or high-fidelity DNA polymerases. Some of the specialized lesion bypass polymerases (mostly belonging to Y-family) can replicate through this bulky adduct but often in an error prone manner, resulting in mutagenesis. Among the four human Y-family polymerases Polη, Polκ, Polι and Rev1, Polκ is unique in its ability for efficient and error-free replication through BP induced BP-N2-dG adduct. In this study, we determined the crystal structures of human Polκ (hPolκ) in ternary complex with DNA and an incoming nucleotide dCTP analogue. The crystals contain DNA with either G base or (+)-trans-anti-[BP]-N2-dG adduct at a template-primer junction and diffract to 2.5 Å and 2.8 Å, respectively. The structures reveal that hPolκ is able to accommodate the bulky adducted DNA in its minor groove without base flipping and nucleotide looping out. The bulky adduct has the polycyclic BP moiety in the minor groove in the regular helical conformation. Polκ has a unique active site that is more open at the minor groove side than other Y-family polymerases. The damaged guanine is in the anti-conformation, the dCMPNPP incoming nucleotide maintains normal Watson-Crick pairing with the G* base. This is the first structure of eukaryotic Y-family polymerase carrying the minor groove BP adduct. The structure and biochemical analysis provides a basis for understanding how hPolκ can correctly bypass and tolerate BP induced BP-N2-dG adduct in human cells.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame machine prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. The Gemma side is a direct model label for every work in the frame, read from the title-only record. The Codex side is a classifier learned from the 10,348 direct Codex labels and calibrated to design-weighted sample rates; fields without enough sample support carry no Codex call. Candidate is the union of the two sides; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels.
Distilled classifier scores by category (both heads)
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.000 | 0.001 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.000 | 0.000 |
| Bibliometrics | 0.000 | 0.000 |
| Science and technology studies | 0.000 | 0.001 |
| Scholarly communication | 0.001 | 0.001 |
| Open science | 0.001 | 0.000 |
| Research integrity | 0.001 | 0.001 |
| Insufficient payload (model declined to judge) | 0.002 | 0.001 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one source (direct Gemma or distilled Codex), not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".