Single nucleotide polymorphism discovery through Illumina-based transcriptome sequencing and mapping in lentil
Bibliographic record
Abstract
Lentil, which belongs to the family Leguminosae (Fabaceae), is a diploid (2n = 2x = 14 chromosomes) self-pollinating crop with a genome size of 4063 Mbp. Because of its nutritional importance and role in the fixation of nitrogen from the atmosphere, lentil is a widely used crop species in molecular genetic studies. By using DNA markers, to date, a limited number of polymorphic bands have been generated. Therefore, it is necessary to develop additional markers to saturate the genome at high density. Single nucleotide polymorphism (SNP) markers are promising for this purpose because of their abundance, stability, and heredity; they can be used to generate a large number of markers over a short distance that are distributed in both intragenic and intergenic regions. Transcriptome sequencing technology was applied to 2 lentil genotypes, and cDNAs were sequenced using the Illumina platform. A total of 111,105,153 sequence reads were generated after trimming. The high-quality reads were assembled, producing 97,528 contigs with an N50 of 1996 bp. The Genome Analysis Tool Kit Unified Genotyper algorithm detected 50,960 putative SNP primers. A genetic linkage map was constructed by using JoinMap4.0 and the map consists of 7 major linkage groups that could be represented as 7 chromosomes of lentil. The extensive sequence information and large number of SNPs obtained in this study could potentially be used for future high-density linkage map construction and association mapping. The large number of contigs obtained in this study could be used for the identification of orthologous transcripts from cDNA data on other organisms.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame distilled prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.
Codex and Gemma teacher scores by category
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.000 | 0.000 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.000 | 0.000 |
| Bibliometrics | 0.000 | 0.000 |
| Science and technology studies | 0.000 | 0.000 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.000 | 0.000 |
| Research integrity | 0.000 | 0.000 |
| Insufficient payload (model declined to judge) | 0.000 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one teacher head, not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".