MSB Mammal Observations (Arctos)
Bibliographic record
Abstract
The Museum of Southwestern Biology (MSB) occasionally records non-vouchered mammal observations in Arctos. Some of these are accompanied by photographs linked to the observational record. These records are curated with the same minimal data standards as vouchered specimens. The Division of Mammals contains over 327,000 catalogued specimens and is among the 3 largest in the world. Specimens date back to the 1880's with the majority documenting the rapid environmental change that has occurred since the 1950's. The collections are taxonomically broad, representing 25 orders, 106 families, 543 genera, ~1,750 species. The majority from the Orders Rodentia (249,000), Chiroptera (27,000), Carnivora (20,000), Eulipotyphla (16,000) and Artiodactyla (7,800). The collections are world-wide in scope (78 countries and all 50 US states) with particularly strong holdings from Western North America (225,000 specimens), Beringia and high latitudes (39,000 from Alaska, Russia and Canada), Mongolia (6,500 specimens and parasites), and Latin America (10,200 specimens from Bolivia, 10,000 from Panama, 7,000 from Chile, and 5,400 from Argentina). The collection contains 89 holotypes or paratypes, 185 parasite symbiotypes and 22 virus symbiotypes. Important collections integrated into the MSB include the USGS Biological Surveys collection (30,000), the UIMNH (Hoffmeister) Collection (33,000), and the Rausch Collection (4,000 specimens and parasites). Specimens range from traditional skin/skull and fluid vouchers to "holistic vouchers" containing skin, skull, post-cranial skeleton, up to seven tissue types (heart, kidney, liver, lung, spleen, muscle, blood), cell suspensions, and ecto and endo parasites. Additionally, frozen tissue samples are available for about 200,000 individual mammals and date back to the late 1970's. 37,000 specimens have serology data associated with hantavirus surveillance programs in the Americas. The MSB houses an extensive archive of field journals and catalogues that date to the 1900's and are associated with specimens held in the collection. The collections are growing rapidly through the active research programs of curators, staff and students and ongoing collaborations with other institutions and governmental agencies.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame machine prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. The Gemma side is a direct model label for every work in the frame, read from the title-only record. The Codex side is a classifier learned from the 10,348 direct Codex labels and calibrated to design-weighted sample rates; fields without enough sample support carry no Codex call. Candidate is the union of the two sides; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels.
Distilled classifier scores by category (both heads)
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.000 | 0.001 |
| Meta-epidemiology (narrow) | 0.001 | 0.000 |
| Meta-epidemiology (broad) | 0.000 | 0.000 |
| Bibliometrics | 0.005 | 0.004 |
| Science and technology studies | 0.001 | 0.000 |
| Scholarly communication | 0.001 | 0.001 |
| Open science | 0.001 | 0.001 |
| Research integrity | 0.000 | 0.000 |
| Insufficient payload (model declined to judge) | 0.157 | 0.048 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one source (direct Gemma or distilled Codex), not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".