Comprehensive Analysis of Hepatitis B Virus Promoter Region Mutations
Bibliographic record
Abstract
Over 250 million people are infected chronically with hepatitis B virus (HBV), the leading cause of liver cancer worldwide. HBV persists, due, in part, to its compact, stable minichromosome, the covalently-closed, circular DNA (cccDNA), which resides in the hepatocytes' nuclei. Current therapies target downstream replication products, however, a true virological cure will require targeting the cccDNA. Finding targets on such a small, compact genome is challenging. For HBV, to remain replication-competent, it needs to maintain nucleotide fidelity in key regions, such as the promoter regions, to ensure that it can continue to utilize the necessary host proteins. HBVdb (HBV database) is a repository of HBV sequences spanning all genotypes (A⁻H) amplified from clinical samples, and hence implying an extensive collection of replication-competent viruses. Here, we analyzed the HBV sequences from HBVdb using bioinformatics tools to comprehensively assess the HBV core and X promoter regions amongst the nearly 70,000 HBV sequences for highly-conserved nucleotides and variant frequencies. Notably, there is a high degree of nucleotide conservation within specific segments of these promoter regions highlighting their importance in potential host protein-viral interactions and thus the virus' viability. Such findings may have key implications for designing antivirals to target these areas.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame distilled prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.
Codex and Gemma teacher scores by category
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.000 | 0.000 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.000 | 0.000 |
| Bibliometrics | 0.000 | 0.001 |
| Science and technology studies | 0.000 | 0.000 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.000 | 0.000 |
| Research integrity | 0.000 | 0.000 |
| Insufficient payload (model declined to judge) | 0.000 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one teacher head, not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".