Comparative analysis of a<i>Brassica</i>BAC clone containing several major aliphatic glucosinolate genes with its corresponding<i>Arabidopsis</i>sequence
Bibliographic record
Abstract
We compared the sequence of a 101-kb-long bacterial artificial chromosome (BAC) clone (B21H13) from Brassica oleracea with its homologous region in Arabidopsis thaliana. This clone contains a gene family involved in the synthesis of aliphatic glucosinolates. The A. thaliana homologs for this gene family are located on chromosome IV and correspond to three 2-oxoglutarate-dependent dioxygenase (AOP) genes. We found that B21H13 harbors 23 genes, whereas the equivalent region in Arabidopsis contains 37 genes. All 23 common genes have the same order and orientation in both Brassica and Arabidopsis. The 16 missing genes in the broccoli BAC clone were arranged in two major blocks of 5 and 7 contiguous genes, two singletons, and a twosome. The 118 exons comprising these 23 genes have high conservation between the two species. The arrangement of the AOP gene family in A. thaliana is as follows: AOP3 (GS-OHP) - AOP2 (GS-ALK) - pseudogene - AOP1. In contrast, in B. oleracea (broccoli and collard), two of the genes are duplicated and the third, AOP3, is missing. The remaining genes are arranged as follows: Bo-AOP2.1 (BoGSL-ALKa) - pseudogene - AOP2.2 (BoGSL-ALKb) - AOP1.1 - AOP1.2. When the survey was expanded to other Brassica accessions, we found variation in copy number and sequence for the Brassica AOP2 homologs. This study confirms that extensive rearrangements have taken place during the evolution of the Brassicacea at both gene and chromosomal levels.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame machine prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. The Gemma side is a direct model label for every work in the frame, read from the title-only record. The Codex side is a classifier learned from the 10,348 direct Codex labels and calibrated to design-weighted sample rates; fields without enough sample support carry no Codex call. Candidate is the union of the two sides; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels.
Distilled classifier scores by category (both heads)
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.000 | 0.000 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.000 | 0.000 |
| Bibliometrics | 0.001 | 0.001 |
| Science and technology studies | 0.000 | 0.000 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.000 | 0.000 |
| Research integrity | 0.000 | 0.000 |
| Insufficient payload (model declined to judge) | 0.003 | 0.001 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one source (direct Gemma or distilled Codex), not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".