Classification of archaic rice grains excavated at the Mojiaoshan site within the Liangzhu site complex reveals an Indica and Japonica chloroplast complex
Bibliographic record
Abstract
Abstract To understand rice types that were utilized during postdomestication and in the modern age and the potential of genetic research in aged rice materials, archaeogenetic analysis was conducted for two populations of archaic rice grains from the Mojiaoshan site during the Liangzhu Period in China (2940 to 2840 BC). Sequencing after the PCR amplification of three regions of the chloroplast genome and one region of the nuclear genome showed recovery rates that were comparable to those in previous studies except for one chloroplast genome region, suggesting that the materials used in this work were appropriate for recovering genetic information related to domestication traits by using advanced technology. Classification after sequencing in these regions proved the existence ofJaponicaandIndicachloroplasts in archaic grains from the west trench, which were subsequently classified into eight plastid groups (type I–VIII), and indicated that these rice grains derived from different maternal lineages were stored together in storage houses at the Mojiaohsan site. Among these plastid groups, type V exhibited the same sequences as two modernIndicaaccessions that are utilized in basic studies and rice breeding. It was inferred that part of the chloroplast genome of archaic rice has been preserved in modern genetic resources in these two modernIndicaaccessions, and the results indicated that rice related to their maternal ancestor was present at the Mojiaoshan site during the Liangzhu Period in China. The usefulness of archaeogenetic analysis can be demonstrated by our research data as well as previous studies, providing encouragement for the possibility that archaeogenetic analysis can be applied to older rice materials that were utilized in the rice-domesticated period. Graphical abstract
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame machine prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. The Gemma side is a direct model label for every work in the frame, read from the title-only record. The Codex side is a classifier learned from the 10,348 direct Codex labels and calibrated to design-weighted sample rates; fields without enough sample support carry no Codex call. Candidate is the union of the two sides; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels.
Distilled classifier scores by category (both heads)
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.000 | 0.000 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.000 | 0.000 |
| Bibliometrics | 0.001 | 0.001 |
| Science and technology studies | 0.001 | 0.001 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.000 | 0.000 |
| Research integrity | 0.000 | 0.000 |
| Insufficient payload (model declined to judge) | 0.000 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one source (direct Gemma or distilled Codex), not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".