The Initial Hepatitis B Virus-Hepatocyte Genomic Integrations and Their Role in Hepatocellular Oncogenesis
Bibliographic record
Abstract
Hepatitis B virus (HBV) remains a dominant cause of hepatocellular carcinoma (HCC). Recently it was shown that HBV and woodchuck hepatitis virus (WHV) integrate into hepatocyte genome minutes after invasion. Retrotransposons and transposable sequences were frequent sites of the initial insertions suggesting a mechanism for spontaneous HBV DNA disperse throughout hepatocyte genome. Several somatic genes were also identified as early insertional targets in infected hepatocytes and woodchuck livers. Head-to-tail joints (HTJs) dominated amongst fusions indicating their creation by non-homologous-end-joining (NHEJ). Their formation coincided with robust oxidative damage of hepatocyte DNA. This was associated with activation of the poly(ADP-ribose) polymerase 1 (PARP1)-mediated dsDNA repair as reflected by augmented transcription of PARP1 and XRCC1, the PARP1 binding partner, OGG1, a responder to oxidative DNA damage, and by increased activity of NAD+, a marker of PARP1 activation, and HO1, an indicator of cell oxidative stress. The engagement of the PARP1-mediated NHEJ repair pathway explains HTJ format of the initial merges. The findings showed that HBV and WHV are immediate inducers of oxidative DNA damage, hijack dsDNA repair to integrate into hepatocyte genome and by this may initiate pro-oncogenic process. Tracking initial integrations may uncover early markers of HCC and help to explain HBV-associated oncogenesis.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame machine prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. The Gemma side is a direct model label for every work in the frame, read from the title-only record. The Codex side is a classifier learned from the 10,348 direct Codex labels and calibrated to design-weighted sample rates; fields without enough sample support carry no Codex call. Candidate is the union of the two sides; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels.
Distilled classifier scores by category (both heads)
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.000 | 0.000 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.000 | 0.000 |
| Bibliometrics | 0.000 | 0.000 |
| Science and technology studies | 0.000 | 0.000 |
| Scholarly communication | 0.001 | 0.000 |
| Open science | 0.000 | 0.000 |
| Research integrity | 0.000 | 0.000 |
| Insufficient payload (model declined to judge) | 0.002 | 0.001 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one source (direct Gemma or distilled Codex), not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".