The plural of anecdote is not data: Rigorously testing a boreal forest chronosequence
Bibliographic record
Abstract
Abstract "Forests appear stable because the ecologists who study them die." The chronosequence, with its space-for-time substitution, is a widely-employed workaround for a difficult problem: many interesting ecosystem processes occur at much longer time scales than researchers can afford to spend studying them. But their use is problematic, particularly for vegetation succession but also for biogeochemical cycling: having sites of different ages is only a necessary, and not a sufficient, condition. How do we test the validity and representativeness of a chronosequence, particularly given ecosystem variability in time and space? Here we use data from an intensively-studied group of stands in northern Manitoba, Canada, to assess the spatial and temporal variability of carbon fluxes in this boreal forest, and examine the suitability of a chronosequence study design for making larger-scale (in space and time) generalizations. This ecosystem is well suited for examining this question, being floristically simple, frequently disturbed by wildfire, and thus generally composed of even-aged forests of known origin. A number of techniques can be brought to bear: re-visiting chronosequence sites after a significant period converts single-point measurements into data vectors; measurements that integrate fluxes over longer time periods (e.g., tree ring cores) provide a similar capability, extending our observation window; replication of the chronosequence stands extends the spatial domain; process modeling may indicate site selection errors. We can also examine data at local, regional and global scales to ask, for any particular pool or flux, if replication in time or space is more useful. For example, global soil respiration studies indicate that spatial and interannual variability are of roughly equal magnitude; in contrast, boreal tree ring data suggest that carbon sequestration is more variable year-to-year than it is site-to-site, after controlling for forest age and soil drainage. Different sampling strategies may thus be appropriate for each flux; historical records such as databases of wildfire occurrence also help constrain this problem. We conclude that while unreplicated chronosequences provide anecdotes, not data, extending measurements in space and time lets us quantifiably assess their performance.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame distilled prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.
Codex and Gemma teacher scores by category
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.002 | 0.002 |
| Meta-epidemiology (narrow) | 0.001 | 0.000 |
| Meta-epidemiology (broad) | 0.001 | 0.000 |
| Bibliometrics | 0.000 | 0.000 |
| Science and technology studies | 0.000 | 0.000 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.004 | 0.004 |
| Research integrity | 0.001 | 0.003 |
| Insufficient payload (model declined to judge) | 0.000 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one teacher head, not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".