Global Biogeography of Reef Fishes: A Hierarchical Quantitative Delineation of Regions
Bibliographic record
Abstract
Delineating regions is an important first step in understanding the evolution and biogeography of faunas. However, quantitative approaches are often limited at a global scale, particularly in the marine realm. Reef fishes are the most diversified group of marine fishes, and compared to most other phyla, their taxonomy and geographical distributions are relatively well known. Based on 169 checklists spread across all tropical oceans, the present work aims to quantitatively delineate biogeographical entities for reef fishes at a global scale. Four different classifications were used to account for uncertainty related to species identification and the quality of checklists. The four classifications delivered converging results, with biogeographical entities that can be hierarchically delineated into realms, regions and provinces. All classifications indicated that the Indo-Pacific has a weak internal structure, with a high similarity from east to west. In contrast, the Atlantic and the Eastern Tropical Pacific were more strongly structured, which may be related to the higher levels of endemism in these two realms. The "Coral Triangle", an area of the Indo-Pacific which contains the highest species diversity for reef fishes, was not clearly delineated by its species composition. Our results show a global concordance with recent works based upon endemism, environmental factors, expert knowledge, or their combination. Our quantitative delineation of biogeographical entities, however, tests the robustness of the results and yields easily replicated patterns. The similarity between our results and those from other phyla, such as corals, suggests that our approach may be of broad utility in describing and understanding global marine biodiversity patterns.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame distilled prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.
Codex and Gemma teacher scores by category
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.000 | 0.000 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.000 | 0.000 |
| Bibliometrics | 0.000 | 0.000 |
| Science and technology studies | 0.000 | 0.000 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.000 | 0.000 |
| Research integrity | 0.000 | 0.000 |
| Insufficient payload (model declined to judge) | 0.000 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one teacher head, not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".