MétaCan
Menu
Back to cohort
Record W2893532394 · doi:10.1007/s00425-018-3012-9

De novo sequencing of the Lavandula angustifolia genome reveals highly duplicated and optimized features for essential oil production

2018· article· en· W2893532394 on OpenAlexafffund
Radesh P. N. Malli, Ayelign M. Adal, Lukman S. Sarker, Ping Liang, Soheil S. Mahmoud

Bibliographic record

VenuePlanta · 2018
Typearticle
Languageen
FieldBiochemistry, Genetics and Molecular Biology
TopicPlant biochemistry and biosynthesis
Canadian institutionsOkanagan University CollegeUniversity of British Columbia, Okanagan CampusUniversity of British ColumbiaBrock University
FundersAgriculture and Agri-Food CanadaNatural Sciences and Engineering Research Council of CanadaCompute Canada
KeywordsGenomeBiologyGeneGenome sizeGeneticsLavandula angustifoliaGene familySecondary metabolismSequence assemblyWhole genome sequencingComputational biologyBotanyEssential oilLavenderBiosynthesis

Abstract

fetched live from OpenAlex

MAIN CONCLUSION: The first draft genome for a member of the genus Lavandula is described. This 870 Mbp genome assembly is composed of over 688 Mbp of non-gap sequences comprising 62,141 protein-coding genes. Lavenders (Lavandula: Lamiaceae) are economically important plants widely grown around the world for their essential oils (EOs), which contribute to the cosmetic, personal hygiene, and pharmaceutical industries. To better understand the genetic mechanisms involved in EO production, identify genes involved in important biological processes, and find genetic markers for plant breeding, we generated the first de novo draft genome assembly for L. angustifolia (Maillette). This high-quality draft reveals a moderately repeated (> 48% repeated elements) 870 Mbp genome, composed of over 688 Mbp of non-gap sequences in 84,291 scaffolds with an N50 value of 96,735 bp. The genome contains 62,141 protein-coding genes and 2003 RNA-coding genes, with a large proportion of genes showing duplications, possibly reflecting past genome polyploidization. The draft genome contains full-length coding sequences for all genes involved in both cytosolic and plastidial pathways of isoprenoid metabolism, and all terpene synthase genes previously described from lavenders. Of particular interest is the observation that the genome contains a high copy number (14 and 7, respectively) of DXS (1-deoxyxylulose-5-phosphate synthase) and HDR (4-hydroxy-3-methylbut-2-enyl diphosphate reductase) genes, encoding the two known regulatory steps in the plastidial isoprenoid biosynthetic pathway. The latter generates precursors for the production of monoterpenes, the most abundant essential oil constituents in lavender. Furthermore, the draft genome contains a variety of monoterpene synthase genes, underlining the production of several monoterpene essential oil constituents in lavender. Taken together, these findings indicate that the genome of L. angustifolia is highly duplicated and optimized for essential oil production.

Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.

How this classification was reachedexpand

Full frame distilled prediction

Teacher imitation

Not calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.

metaresearch head score (Codex)0.000
metaresearch head score (Gemma)0.000
Version: codex-gemma-dda1882f352aValidation status: machine_predicted_unvalidated
Candidate categoriesnone
Consensus categoriesnone
DomainCandidate signal: none · Consensus signal: none
Study designCandidate signal: Bench or experimental · Consensus signal: Bench or experimental
GenreCandidate signal: Empirical · Consensus signal: Empirical
Teacher disagreement score0.002
Threshold uncertainty score0.296

Codex and Gemma teacher scores by category

CategoryCodexGemma
Metaresearch0.0000.000
Meta-epidemiology (narrow)0.0000.000
Meta-epidemiology (broad)0.0000.000
Bibliometrics0.0000.000
Science and technology studies0.0000.000
Scholarly communication0.0000.000
Open science0.0000.000
Research integrity0.0000.000
Insufficient payload (model declined to judge)0.0000.000

Machine scores (provisional)

The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.

Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.

Opus teacher head0.009
GPT teacher head0.227
Teacher spread0.218 · how far apart the two teachers sit on this one work
Validation statusscore_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from it

Classification

machine, unvalidated

Machine predicted; a candidate call from one teacher head, not a consensus.

The models applied no category: nothing in the taxonomy fit this work.
Study designBench or experimental
Domainnot available
GenreEmpirical

How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".

Quick stats

Citations33
Published2018
Admission routes2
Has abstractyes

Explore more

Same venuePlantaSame topicPlant biochemistry and biosynthesisFrench-language works237,207