Cell surface glycoproteins from Thermoplasma acidophilum are modified with an N-linked glycan containing 6-C-sulfofucose†
Bibliographic record
Abstract
Thermoplasma acidophilum is a thermoacidophilic archaeon that grows optimally at pH 2 and 59°C. This extremophile is remarkable by the absence of a cell wall or an S-layer. Treating the cells with Triton X-100 at pH 3 allowed the extraction of all of the cell surface glycoproteins while keeping cells intact. The extracted glycoproteins were partially purified by cation-exchange chromatography, and we identified five glycoproteins by N-terminal sequencing and mass spectrometry of in-gel tryptic digests. These glycoproteins are positive for periodic acid-Schiff staining, have a high content of Asn including a large number in the Asn-X-Ser/Thr sequon and have apparent masses that are 34-48% larger than the masses deduced from their amino acid sequences. The pooled glycoproteins were digested with proteinase K and the purified glycopeptides were analyzed by NMR. Structural determination showed that the carbohydrate part was represented by two structures in nearly equal amounts, differing by the presence of one terminal mannose residue. The larger glycan chain consists of eight residues: six hexoses, one heptose and one sugar with an unusual residue mass of 226 Da which was identified as 6-deoxy-6-C-sulfo-D-galactose (6-C-sulfo-D-fucose). Mass spectrometry analyses of the peptides obtained by trypsin and chymotrypsin digestion confirmed the principal structures to be those determined by NMR and identified 14 glycopeptides derived from the main glycoprotein, Ta0280, all containing the Asn-X-Ser/Thr sequons. Thermoplasma acidophilum appears to have a "general" protein N-glycosylation system that targets a number of cell surface proteins.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame machine prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. The Gemma side is a direct model label for every work in the frame, read from the title-only record. The Codex side is a classifier learned from the 10,348 direct Codex labels and calibrated to design-weighted sample rates; fields without enough sample support carry no Codex call. Candidate is the union of the two sides; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels.
Distilled classifier scores by category (both heads)
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.000 | 0.000 |
| Meta-epidemiology (narrow) | 0.001 | 0.000 |
| Meta-epidemiology (broad) | 0.000 | 0.000 |
| Bibliometrics | 0.000 | 0.000 |
| Science and technology studies | 0.000 | 0.000 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.000 | 0.000 |
| Research integrity | 0.000 | 0.000 |
| Insufficient payload (model declined to judge) | 0.000 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one source (direct Gemma or distilled Codex), not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".