Genetic diversity of Salmonella enterica isolated over 13 years from raw California almonds and from an almond orchard
Bibliographic record
Abstract
A comparative genomic analysis was conducted for 171 Salmonella isolates recovered from raw inshell almonds and raw almond kernels between 2001 and 2013 and for 30 Salmonella Enteritidis phage type (PT) 30 isolates recovered between 2001 and 2006 from a 2001 salmonellosis outbreak-associated almond orchard. Whole genome sequencing was used to measure the genetic distance among isolates by single nucleotide polymorphism (SNP) analyses and to predict the presence of plasmid DNA and of antimicrobial resistance (AMR) and virulence genes. Isolates were classified by serovars with Parsnp, a fast core-genome multi aligner, before being analyzed with the CFSAN SNP Pipeline (U.S. Food and Drug Administration Center for Food Safety and Applied Nutrition). Genetically similar (≤18 SNPs) Salmonella isolates were identified among several serovars isolated years apart. Almond isolates of Salmonella Montevideo (2001 to 2013) and Salmonella Newport (2003 to 2010) differed by ≤9 SNPs. Salmonella Enteritidis PT 30 isolated between 2001 and 2013 from survey, orchard, outbreak, and clinical samples differed by ≤18 SNPs. One to seven plasmids were found in 106 (62%) of the Salmonella isolates. Of the 27 plasmid families that were identified, IncFII and IncFIB plasmids were the most predominant. AMR genes were identified in 16 (9%) of the survey isolates and were plasmid encoded in 11 of 16 cases; 12 isolates (7%) had putative resistance to at least one antibiotic in three or more drug classes. A total of 303 virulence genes were detected among the assembled genomes; a plasmid that harbored a combination of pef, rck, and spv virulence genes was identified in 23% of the isolates. These data provide evidence of long-term survival (years) of Salmonella in agricultural environments.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame machine prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. The Gemma side is a direct model label for every work in the frame, read from the title-only record. The Codex side is a classifier learned from the 10,348 direct Codex labels and calibrated to design-weighted sample rates; fields without enough sample support carry no Codex call. Candidate is the union of the two sides; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels.
Distilled classifier scores by category (both heads)
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.000 | 0.000 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.000 | 0.000 |
| Bibliometrics | 0.002 | 0.001 |
| Science and technology studies | 0.001 | 0.000 |
| Scholarly communication | 0.001 | 0.000 |
| Open science | 0.000 | 0.000 |
| Research integrity | 0.000 | 0.000 |
| Insufficient payload (model declined to judge) | 0.001 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one source (direct Gemma or distilled Codex), not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".