Tracing hepatitis B virus (HBV) genotype B5 (formerly B6) evolutionary history in the circumpolar Arctic through phylogeographic modelling
Bibliographic record
Abstract
Background Indigenous populations of the circumpolar Arctic are considered to be endemically infected (>2% prevalence) with hepatitis B virus (HBV), with subgenotype B5 (formerly B6) unique to these populations. The distinctive properties of HBV/B5, including high nucleotide diversity yet no significant liver disease, suggest virus adaptation through long-term host-pathogen association. Methods To investigate the origin and evolutionary spread of HBV/B5 into the circumpolar Arctic, fifty-seven partial and full genome sequences from Alaska, Canada and Greenland, having known location and sampling dates spanning 40 years, were phylogeographically investigated by Bayesian analysis (BEAST 2) using a reversible-jump-based substitution model and a clock rate estimated at 4.1 × 10 −5 substitutions/site/year. Results Following an initial divergence from an Asian viral ancestor approximately 1954 years before present (YBP; 95% highest probability density interval [1188, 2901]), HBV/B5 coalescence occurred almost 1000 years later. Surprisingly, the HBV/B5 ancestor appears to locate first to Greenland in a rapid coastal route progression based on the landscape aware geographic model, with subsequent B5 evolution and spread westward. Bayesian skyline plot analysis demonstrated an HBV/B5 population expansion occurring approximately 400 YBP, coinciding with the disruption of the Neo-Eskimo Thule culture into more heterogeneous and regionally distinct Inuit populations throughout the North American Arctic. Discussion HBV/B5 origin and spread appears to occur coincident with the movement of Neo-Eskimo (Inuit) populations within the past 1000 years, further supporting the hypothesis of HBV/host co-expansion, and illustrating the concept of host-pathogen adaptation and balance.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame distilled prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.
Codex and Gemma teacher scores by category
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.000 | 0.000 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.000 | 0.000 |
| Bibliometrics | 0.000 | 0.000 |
| Science and technology studies | 0.001 | 0.000 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.000 | 0.000 |
| Research integrity | 0.000 | 0.001 |
| Insufficient payload (model declined to judge) | 0.000 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one teacher head, not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".