A database of Holocene temperature records for north‐eastern North America and the north‐western Atlantic
Bibliographic record
Abstract
Abstract Centennial‐to‐millennial temperature records of the past provide a context for the interpretation of current and future changes in climate. Quaternary climates have been relatively well studied in north‐east North America and the adjacent Atlantic Ocean over the last decades, and new research methods have been developed to improve reconstructions. We present newly inferred reconstructions of sea surface temperature for the north‐western Atlantic region, together with a compilation of published temperature records. The database thus comprises a total of 86 records from both marine and terrestrial sites, including lakes, peatlands, ice and tree rings, each covering at least part of the Holocene. For each record, we present details on seasons covered, chronologies and information on radiocarbon dates and analytical time steps. The 86 records contain a total of 154 reconstructions of temperature and temperature‐related variables. Main proxies include pollen and dinocysts, while summer was the season for which the highest number of reconstructions were available. Many records covered most of the Holocene, but many dinocyst records did not extend to the surface, due to sediment mixing, and dendroclimate records were limited to the last millennium. The database allows for the exploration of linkages between sea ice and climate and may be used in conjunction with other palaeoclimate and palaeoenvironmental records, such as wildfire records and peatland dynamics. This inventory may also aid the identification of gaps in the geographic distribution of past temperature records thus guiding future research efforts.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame machine prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. The Gemma side is a direct model label for every work in the frame, read from the title-only record. The Codex side is a classifier learned from the 10,348 direct Codex labels and calibrated to design-weighted sample rates; fields without enough sample support carry no Codex call. Candidate is the union of the two sides; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels.
Distilled classifier scores by category (both heads)
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.000 | 0.002 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.000 | 0.000 |
| Bibliometrics | 0.003 | 0.005 |
| Science and technology studies | 0.000 | 0.000 |
| Scholarly communication | 0.001 | 0.001 |
| Open science | 0.000 | 0.000 |
| Research integrity | 0.000 | 0.000 |
| Insufficient payload (model declined to judge) | 0.003 | 0.001 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one source (direct Gemma or distilled Codex), not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".