The extinct marine megafauna of the Phanerozoic
Bibliographic record
Abstract
Abstract The modern marine megafauna is known to play important ecological roles and includes many charismatic species that have drawn the attention of both the scientific community and the public. However, the extinct marine megafauna has never been assessed as a whole, nor has it been defined in deep time. Here, we review the literature to define and list the species that constitute the extinct marine megafauna, and to explore biological and ecological patterns throughout the Phanerozoic. We propose a size cut-off of 1 m of length to define the extinct marine megafauna. Based on this definition, we list 706 taxa belonging to eight main groups. We found that the extinct marine megafauna was conspicuous over the Phanerozoic and ubiquitous across all geological eras and periods, with the Mesozoic, especially the Cretaceous, having the greatest number of taxa. Marine reptiles include the largest size recorded (21 m; Shonisaurus sikanniensis ) and contain the highest number of extinct marine megafaunal taxa. This contrasts with today’s assemblage, where marine animals achieve sizes of >30 m. The extinct marine megafaunal taxa were found to be well-represented in the Paleobiology Database, but not better sampled than their smaller counterparts. Among the extinct marine megafauna, there appears to be an overall increase in body size through time. Most extinct megafaunal taxa were inferred to be macropredators preferentially living in coastal environments. Across the Phanerozoic, megafaunal species had similar extinction risks as smaller species, in stark contrast to modern oceans where the large species are most affected by human perturbations. Our work represents a first step towards a better understanding of the marine megafauna that lived in the geological past. However, more work is required to expand our list of taxa and their traits so that we can obtain a more complete picture of their ecology and evolution.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame machine prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. The Gemma side is a direct model label for every work in the frame, read from the title-only record. The Codex side is a classifier learned from the 10,348 direct Codex labels and calibrated to design-weighted sample rates; fields without enough sample support carry no Codex call. Candidate is the union of the two sides; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels.
Distilled classifier scores by category (both heads)
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.000 | 0.001 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.000 | 0.000 |
| Bibliometrics | 0.002 | 0.002 |
| Science and technology studies | 0.000 | 0.001 |
| Scholarly communication | 0.001 | 0.001 |
| Open science | 0.000 | 0.001 |
| Research integrity | 0.000 | 0.000 |
| Insufficient payload (model declined to judge) | 0.002 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one source (direct Gemma or distilled Codex), not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".