Prediction of KIR3DL1/Human Leukocyte Antigen binding
Notice bibliographique
Résumé
Abstract KIR3DL1 is a polymorphic inhibitory Natural Killer (NK) cell receptor that recognizes Human Leukocyte Antigen (HLA) class I allotypes that contain the Bw4 motif. Structural analyses have shown that in addition to residues 77-83 that span the Bw4 motif, polymorphism at other sites throughout the HLA molecule can influence the interaction with KIR3DL1. Given the extensive polymorphism of both KIR3DL1 and HLA class I, we built a machine learning prediction model to describe the influence of allotypic variation on the binding of KIR3DL1 to HLA class I. Nine KIR3DL1 tetramers were screened for reactivity against a panel of HLA class I molecules which revealed different patterns of specificity for each KIR3DL1 allotype. Separate models were trained for each of KIR3DL1 allotypes based on the full amino sequence of exons 2 and 3 encoding the α 1 and α 2 domains of the class I HLA allotypes, the set of polymorphic positions that span the Bw4 motif, or the positions that encode α 1 and α 2 but exclude the connecting loops. The Multi-Label-Vector-Optimization (MLVO) model trained on all alpha helix positions performed best with AUC scores ranging from 0.74 to 0.974 for the 9 KIR3DL1 allotype models. We show that a binary division into binder and non-binder is not precise, and that intermediate levels exist. Using the same models, within the binder group, high- and low-binder categories can also be predicted, the regions in HLA affecting the high vs low binder being completely distinct from the classical Bw4 motif. We further show that these positions affect binding affinity in a nonadditive way and induce deviations from linear models used to predict interaction strength. We propose that this approach should be used in lieu of simpler binding models based on a single HLA motif.
Récupéré en direct depuis OpenAlex et désinversé. Les résumés ne sont pas conservés dans cette base de données : les index inversés représentent 8,6 Go des 9,3 Go de texte de la base, et le serveur dispose de 13 Go libres.
Comment cette classification a été obtenuedéplier
Prédiction machine sur la base complète
Imitation des enseignantsNi prévalence calibrée, ni vérité terrain. Validation humaine à venir. Le volet Gemma est une étiquette directe du modèle pour chaque travail de la base, lue sur la notice réduite au titre. Le volet Codex est un classifieur appris des 10 348 étiquettes directes de Codex et calibré sur les taux pondérés de l'échantillon; les champs sans appui suffisant ne portent aucun appel Codex. Le mode candidate est l'union des deux volets; le consensus est leur intersection. Ces sorties portent le statut machine_predicted_unvalidated et ne sont pas des étiquettes humaines.
Scores du classifieur distillé par catégorie (deux têtes)
| Catégorie | Codex | Gemma |
|---|---|---|
| Métarecherche | 0,001 | 0,002 |
| Méta-épidémiologie (sens strict) | 0,001 | 0,000 |
| Méta-épidémiologie (sens large) | 0,001 | 0,001 |
| Bibliométrie | 0,001 | 0,000 |
| Études des sciences et des technologies | 0,000 | 0,000 |
| Communication savante | 0,001 | 0,000 |
| Science ouverte | 0,000 | 0,000 |
| Intégrité de la recherche | 0,001 | 0,001 |
| Charge utile insuffisante (le modèle a refusé de juger) | 0,002 | 0,001 |
Scores machine (provisoires)
Les deux têtes enseignantes du modèle étudiant, lues sur ce travail. Un score ordonne la base pour la relecture; il n'affirme jamais une catégorie, et le statut de validation accompagne chaque rangée tel quel.
Scores de référence d'un modèle non mature (critères de maturité non atteints, 7 itérations). Un score ordonne; il n'affirme jamais une catégorie.
score_only:v0-immature-baseline · tel quel depuis la passe de notation : score_only signifie que le nombre peut ordonner les travaux, et qu'aucune étiquette de catégorie n'en découleClassification
machine, non validéePrédiction automatique; un appel candidat d’une seule source (Gemma direct ou Codex distillé), pas un consensus.
Le détail, modèle par modèle et score par score, se trouve en fin de page sous « Comment cette classification a été obtenue ».