Highly Efficient Site-Specific and Cassette Mutagenesis of Plasmids Harboring GC-Rich Sequences
Bibliographic record
Abstract
GC-rich sequences affect DNA replication, recombination and repair, as well as RNA transcription in vivo. Such sequences may also impede site-directed mutagenesis in vitro. P3a site-directed mutagenesis is a highly efficient method, but it has not been tested with plasmids possessing GC-rich sequences. Here we report that it is very efficient with a BRPF3 expression vector but unsuccessful with that for KAT2B. Because two GC-rich regions located within the synthetic CAG promoter and the KAT2B coding region may form guanine (G)-quadruplexes and hinder plasmid denaturation during PCR, we developed P3b site-specific mutagenesis, achieving an average efficiency of 97.5% in engineering ten KAT2B mutants. Importantly, deletion mutagenesis revealed that either of the two GC-rich regions is sufficient for rendering the plasmid incompatible with P3a mutagenesis. Consistent with this, only P3b mutagenesis worked efficiently with several widely used sgRNA/Cas9 expression vectors, which contain the CAG promoter, and with an expression vector for CDK13, which possesses an intrinsically disordered domain encoded by a GC-rich DNA fragment. Thus, this study highlights serious challenges posed by GC-rich sequences to site-directed mutagenesis and provides an effective remedy to address such challenges. The findings support that G-quadruplex formation is one mechanism whereby such sequences impede regular PCR-based mutagenesis methods.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame distilled prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.
Codex and Gemma teacher scores by category
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.000 | 0.000 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.000 | 0.000 |
| Bibliometrics | 0.000 | 0.000 |
| Science and technology studies | 0.000 | 0.000 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.000 | 0.000 |
| Research integrity | 0.000 | 0.000 |
| Insufficient payload (model declined to judge) | 0.000 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one teacher head, not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".