A high-resolution database of historical and future climate for Africa developed with deep neural networks
Bibliographic record
Abstract
Abstract This study contributes an accessible, comprehensive database of interpolated climate data for Africa that includes monthly, annual, decadal, and 30-year normal climate data for the last 120 years (1901 to present) as well as multi-model CMIP6 climate change projections for the 21 st century. The database includes variables relevant for ecological research and infrastructure planning, and it comprises more than 25,000 climate grids that can be queried with a provided ClimateAF software package. In addition, 30 arcsecond (~1 km) resolution gridded data are available for download. The climate grids were developed with a three-step approach, using thin-plate spline interpolations of weather station data as a first approximation. Subsequently, a novel deep learning approach is used to model orographic precipitation, rain shadows, lake and coastal effects at moderate resolution. Lastly, lapse-rate based downscaling is applied to generate high-resolution grids. The climate estimates were optimized and cross-validated with a checkerboard approach to ensure that training data was spatially distanced from validation data. We conclude with a discussion of applications and limitations of this database.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame distilled prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.
Codex and Gemma teacher scores by category
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.001 | 0.000 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.000 | 0.000 |
| Bibliometrics | 0.000 | 0.000 |
| Science and technology studies | 0.000 | 0.000 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.001 | 0.001 |
| Research integrity | 0.000 | 0.000 |
| Insufficient payload (model declined to judge) | 0.000 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one teacher head, not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".