Survey and DNA barcoding of flat bugs (Hemiptera: Aradidae) in the Tanzanian Forest Archipelago reveal a phylogeographically structured fauna largely unknown at the species level
Bibliographic record
Abstract
We report results of a faunal survey of Aradidae flat bugs sampled by sifting litter in 14 wet and discrete Tanzanian primary forests (= Tanzanian Forest Archipelago, TFA) of different geological origins and ages. Images, locality data and, when available, DNA barcoding sequences of 300 Aradidae adults and nymphs forming the core of the herein analyzed data are publicly available online at dx.doi.org/10.5883/DS-ARADTZ. Three Aradidae subfamilies and seven genera were recorded: Aneurinae (Paraneurus), Carventinae (Dundocoris) and Mezirinae (Afropictinus, Embuana, Linnavuoriessa, Neochelonoderus, Usumbaraia); the two latter subfamilies were also represented by specimens not assignable to nominal genera. Barring the six nominal species of Neochelonoderus and Afropictinus described earlier by us from these samples and representing 11 of the herein defined Operational Taxonomic Units (OTU), only one of the remaining 52 OTUs could be assigned to a named species; the remaining 51 OTUs (81%) represent unnamed species. Average diversity of Aradidae is 4.64 species per locality; diversity on the three geologically young volcanoes (Mts Hanang, Meru, Kilimanjaro) is significantly lower (1.33) than on the nine Eastern Arc Mountains (5.67) and in two lowland forests (5). Observed phylogeographic structure of Aradidae in TFA can be attributed to vicariance, while the depauperate fauna of Aradidae on geologically young Tanzanian volcanoes was likely formed anew by colonisation from nearby and geologically older forests.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame machine prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. The Gemma side is a direct model label for every work in the frame, read from the title-only record. The Codex side is a classifier learned from the 10,348 direct Codex labels and calibrated to design-weighted sample rates; fields without enough sample support carry no Codex call. Candidate is the union of the two sides; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels.
Distilled classifier scores by category (both heads)
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.000 | 0.000 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.000 | 0.000 |
| Bibliometrics | 0.001 | 0.001 |
| Science and technology studies | 0.000 | 0.000 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.000 | 0.000 |
| Research integrity | 0.000 | 0.000 |
| Insufficient payload (model declined to judge) | 0.001 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one source (direct Gemma or distilled Codex), not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".