PoxiPred: An Artificial-Intelligence-Based Method for the Prediction of Potential Antigens and Epitopes to Accelerate Vaccine Development Efforts against Poxviruses
Bibliographic record
Abstract
Poxviridae is a family of large, complex, enveloped, and double-stranded DNA viruses. The members of this family are ubiquitous and well known to cause contagious diseases in humans and other types of animals as well. Taxonomically, the poxviridae family is classified into two subfamilies, namely Chordopoxvirinae (affecting vertebrates) and Entomopoxvirinae (affecting insects). The members of the Chordopoxvirinae subfamily are further divided into 18 genera based on the genome architecture and evolutionary relationship. Of these 18 genera, four genera, namely Molluscipoxvirus, Orthopoxvirus, Parapoxvirus, and Yatapoxvirus, are known for infecting humans. Some of the popular members of poxviridae are variola virus, vaccine virus, Mpox (formerly known as monkeypox), cowpox, etc. There is still a pressing demand for the development of effective vaccines against poxviruses. Integrated immunoinformatics and artificial-intelligence (AI)-based methods have emerged as important approaches to design multi-epitope vaccines against contagious emerging infectious diseases. Despite significant progress in immunoinformatics and AI-based techniques, limited methods are available to predict the epitopes. In this study, we have proposed a unique method to predict the potential antigens and T-cell epitopes for multiple poxviruses. With PoxiPred, we developed an AI-based tool that was trained and tested with the antigens and epitopes of poxviruses. Our tool was able to locate 3191 antigen proteins from 25 distinct poxviruses. From these antigenic proteins, PoxiPred redundantly located up to five epitopes per protein, resulting in 16,817 potential T-cell epitopes which were mostly (i.e., 92%) predicted as being reactive to CD8+ T-cells. PoxiPred is able to, on a single run, identify antigens and T-cell epitopes for poxviruses with one single input, i.e., the proteome file of any poxvirus.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame distilled prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.
Codex and Gemma teacher scores by category
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.000 | 0.000 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.000 | 0.000 |
| Bibliometrics | 0.000 | 0.000 |
| Science and technology studies | 0.000 | 0.000 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.000 | 0.000 |
| Research integrity | 0.000 | 0.000 |
| Insufficient payload (model declined to judge) | 0.000 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one teacher head, not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".