Quantitative Analysis of the Binding of Simian Virus 40 Large T Antigen to DNA
Bibliographic record
Abstract
SV40 large T antigen (T-ag) is a multifunctional protein that successively binds to 5'-GAGGC-3' sequences in the viral origin of replication, melts the origin, unwinds DNA ahead of the replication fork, and interacts with host DNA replication factors to promote replication of the simian virus 40 genome. The transition of T-ag from a sequence-specific binding protein to a nonspecific helicase involves its assembly into a double hexamer whose formation is likely dictated by the propensity of T-ag to oligomerize and its relative affinities for the origin as well as for nonspecific double- and single-stranded DNA. In this study, we used a sensitive assay based on fluorescence anisotropy to measure the affinities of wild-type and mutant forms of the T-ag origin-binding domain (OBD), and of a larger fragment containing the N-terminal domain (N260), for different DNA substrates. We report that the N-terminal domain does not contribute to binding affinity but reduces the propensity of the OBD to self-associate. We found that the OBD binds with different affinities to its four sites in the origin and determined a consensus binding site by systematic mutagenesis of the 5'-GAGGC-3' sequence and of the residue downstream of it, which also contributes to affinity. Interestingly, the OBD also binds to single-stranded DNA with an approximately 10-fold higher affinity than to nonspecific duplex DNA and in a mutually exclusive manner. Finally, we provide evidence that the sequence specificity of full-length T-ag is lower than that of the OBD. These results provide a quantitative basis onto which to anchor our understanding of the interaction of T-ag with the origin and its assembly into a double hexamer.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame distilled prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.
Codex and Gemma teacher scores by category
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.001 | 0.000 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.001 | 0.000 |
| Bibliometrics | 0.001 | 0.001 |
| Science and technology studies | 0.000 | 0.000 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.000 | 0.000 |
| Research integrity | 0.000 | 0.000 |
| Insufficient payload (model declined to judge) | 0.000 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one teacher head, not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".