Identification of the Carbohydrate Moieties and Glycosylation Motifs in Campylobacter jejuni Flagellin
Bibliographic record
Abstract
Flagellins from three strains of Campylobacter jejuni and one strain of Campylobacter coli were shown to be extensively modified by glycosyl residues, imparting an approximate 6000-Da shift from the molecular mass of the protein predicted from the DNA sequence. Tryptic peptides from C. jejuni 81-176 flagellin were subjected to capillary liquid chromatography-electrospray mass spectrometry with a high/low orifice stepping to identify peptide segments of aberrant masses together with their corresponding glycosyl appendages. These modified peptides were further characterized by tandem mass spectrometry and preparative high performance liquid chromatography followed by nano-NMR spectroscopy to identify the nature and precise site of glycosylation. These analyses have shown that there are 19 modified Ser/Thr residues in C. jejuni 81-176 flagellin. The predominant modification found on C. jejuni flagellin was O-linked 5,7-diacetamido-3,5,7,9-tetradeoxy-l-glycero-l-manno-nonulosonic acid (pseudaminic acid, Pse5Ac7Ac) with additional heterogeneity conferred by substitution of the acetamido groups with acetamidino and hydroxyproprionyl groups. In C. jejuni 81-176, the gene Cj1316c, encoding a protein of unknown function, was shown to be involved in the biosynthesis and/or the addition of the acetamidino group on Pse5Ac7Ac. Glycosylation is not random, since 19 of the total 107 Ser/Thr residues are modified, and all but one of these are restricted to the central, surface-exposed domain of flagellin when folded in the filament. The mechanism of attachment appears unrelated to a consensus peptide sequence but is rather based on surface accessibility of Ser/Thr residues in the folded protein.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame machine prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. The Gemma side is a direct model label for every work in the frame, read from the title-only record. The Codex side is a classifier learned from the 10,348 direct Codex labels and calibrated to design-weighted sample rates; fields without enough sample support carry no Codex call. Candidate is the union of the two sides; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels.
Distilled classifier scores by category (both heads)
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.000 | 0.000 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.000 | 0.000 |
| Bibliometrics | 0.000 | 0.000 |
| Science and technology studies | 0.000 | 0.000 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.000 | 0.000 |
| Research integrity | 0.000 | 0.000 |
| Insufficient payload (model declined to judge) | 0.000 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one source (direct Gemma or distilled Codex), not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".