Ichneumonid subfamily abundances from 38 datasets collected between 25°S and 81°N
Bibliographic record
Abstract
Data from a study of latitudinal diversity gradients in the parasitoid wasp family Ichneumonidae, using subfamily abundance from 38 collections made between 25°S and 81°N. All collections in this study were made using Malaise traps. Data for most of the collections were found in published papers (see below); data from Alberta was collected by Marla Schwarzfeld in association with the Ecoystem Management Emulating Natural Disturbance (EMEND) project; data from northern Canada was collected by Laura Timms in association with the Northern Biodiversity Program (NBP); data from Peru was collected by Ilari Sääksjärvi. These data are associated with manuscript in review: Timms, L.L., Schwarzfeld, M., and Sääksjärvi, I.E. Extending understanding of latitudinal patterns in parasitoid wasp diversity. This dataset has 44 columns and 38 rows of data, where each row is the information associated with one collection. Column titles: Collection = Abbreviation for the collection Location = Geographic location (country, state, or province) where the collection took place Year(s) = Year(s) during which the collection took place Acae-Xori = these 28 columns contain the abundances 28 ichneumonid subfamilies; the four letter code title of each column is the first four letters of the subfamily name (except for Ortc = Orthocentrinae and Ortp = Orthopelmatinae) TOT = the total abundance of all ichneumonids in each collection Lat = latitude of the collection site (where negative values indicate southern hemisphere) H1 = northern/southern hemisphere Long = longitude of the collection site H2 = eastern/western hemisphere MAP = mean annual precipitation at the collection site LGP.mean = mean length of growing period (LGP) at each collection site, based on data from the FAO (http://www.fao.org/geonetwork/srv/en/main.home), where LGP is calculated based on the number of days with a temperature of at least 5°C and with enough moisture for plant growth. Note that LGP values are usually provided by the FAO in the form of ranges, the values given here are the midpoints of those ranges. Also note, the values for NBP.BKS and NBP.HAZ are equal to the mean number of days above 5°C according to the Environment Canada national climate archives; the FAO's LGP values for these regions were erroneously provided as 0. TD.tot = total number of Malaise trap days in the collection Traps = total number of Malaise traps used in the collection dpt = days per trap; mean number of days each Malaise trap was deployed dpt.lgp = days per trap divided by length of growing period; a measure of the proportion of the active season trapped Ind.TD.tot = total number of individuals per trap day for the collection iabd = abundance of idiobionts in the collection kabd = abundance of koinobionts in the collection Related reference(s) = associated literature reference(s) for the collection
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame machine prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. The Gemma side is a direct model label for every work in the frame, read from the title-only record. The Codex side is a classifier learned from the 10,348 direct Codex labels and calibrated to design-weighted sample rates; fields without enough sample support carry no Codex call. Candidate is the union of the two sides; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels.
Distilled classifier scores by category (both heads)
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.001 | 0.002 |
| Meta-epidemiology (narrow) | 0.001 | 0.000 |
| Meta-epidemiology (broad) | 0.001 | 0.001 |
| Bibliometrics | 0.003 | 0.005 |
| Science and technology studies | 0.001 | 0.000 |
| Scholarly communication | 0.001 | 0.001 |
| Open science | 0.001 | 0.001 |
| Research integrity | 0.000 | 0.001 |
| Insufficient payload (model declined to judge) | 0.007 | 0.005 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one source (direct Gemma or distilled Codex), not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".