Intercomparison of homogenization techniques for precipitation data continued: Comparison of two recent Bayesian change point models
Bibliographic record
Abstract
In this paper, two new Bayesian change point techniques are described and compared to eight other techniques presented in previous work to detect inhomogeneities in climatic series. An inhomogeneity can be defined as a change point (a time point in a series such that the observations have a different distribution before and after this time) in the data series induced from changes in measurement conditions at a given station. It is important to be able to detect and correct an inhomogeneity, as it can interfere with the real climate change signal. The first technique is a Bayesian method of multiple change point detection in a multiple linear regression. The second one allows the detection of a single change point in a multiple linear regression. These two techniques have never been used for homogenization purposes. The ability of the two techniques to discriminate homogeneous and inhomogeneous series was evaluated using simulated data series. Various sets of synthetic series (homogeneous, with a single shift, and with multiple shifts) representing the typical total annual precipitation observed in the southern and central parts of the province of Quebec, Canada, and nearby areas were generated for the purpose of this study. The two techniques gave small false detection rates on the homogeneous series. Furthermore, the two techniques proved to be efficient for the detection of a single shift in a series. For the series with multiple shifts, the Bayesian method of multiple change point detection performed better. An application to a real data set is also provided and validated with the available metadata.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame distilled prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.
Codex and Gemma teacher scores by category
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.002 | 0.000 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.000 | 0.000 |
| Bibliometrics | 0.000 | 0.000 |
| Science and technology studies | 0.000 | 0.000 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.001 | 0.000 |
| Research integrity | 0.000 | 0.000 |
| Insufficient payload (model declined to judge) | 0.000 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one teacher head, not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".