Do NIR spectra collected from laboratory-reared mosquitoes differ from those collected from wild mosquitoes?
Bibliographic record
Abstract
BACKGROUND: Near infrared spectroscopy (NIRS) is a high throughput technique that measures absorbance of specific wavelengths of light by biological samples and uses this information to classify the age of lab-reared mosquitoes as younger or older than seven days with an average accuracy greater than 80%. For NIRS to estimate ages of wild mosquitoes, a sample of wild mosquitoes with known age in days would be required to train and test the model. Mark-release-recapture is the most reliable method to produce wild-caught mosquitoes of known age in days. However, it is logistically demanding, time inefficient, subject to low recapture rates, and raises ethical issues due to the release of mosquitoes. Using labels from Detinova dissection results in a mathematical model with poor accuracy. Alternatively, a model trained on spectra from laboratory-reared mosquitoes where age in days is known can be applied to estimate the age of wild mosquitoes, but this would be appropriate only if spectra collected from laboratory-reared and wild mosquitoes are similar. METHODS AND FINDINGS: We performed k-means (k = 2) cluster analysis on a mixture of spectra collected from lab-reared and wild Anopheles arabiensis to determine if there is any significant difference between these two groups. While controlling the numbers of mosquitoes included in the model at each age, we found two clusters with no significant difference in distribution of spectra collected from lab-reared and wild mosquitoes (p = 0.25). We repeated the analysis using hierarchical clustering, and similarly, no significant difference was observed (p = 0.13). CONCLUSION: We find no difference between spectra collected from laboratory-reared and wild mosquitoes of the same age and species. The results strengthen and support the on-going practice of applying the model trained on spectra collected from laboratory-reared mosquitoes, especially first-generation laboratory-reared mosquitoes.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame distilled prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.
Codex and Gemma teacher scores by category
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.000 | 0.001 |
| Meta-epidemiology (narrow) | 0.001 | 0.001 |
| Meta-epidemiology (broad) | 0.001 | 0.000 |
| Bibliometrics | 0.000 | 0.001 |
| Science and technology studies | 0.000 | 0.000 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.000 | 0.000 |
| Research integrity | 0.000 | 0.001 |
| Insufficient payload (model declined to judge) | 0.004 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one teacher head, not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".