Multi-Model Ensemble Sub-Seasonal Forecasting of Precipitation over the Maritime Continent in Boreal Summer
Bibliographic record
Abstract
The Maritime Continent (MC) is a critical region with unique geographical conditions and significant monsoon activities that plays a vital role in global climate variation. In this study, the weekly prediction of precipitation over the MC during boreal summer (from May to September) was analyzed using the 12-year reforecasts data from five Sub-seasonal to Seasonal (S2S) models, including the China Meteorological Administration (CMA), the European Centre for Medium-Range Weather Forecasts (ECMWF), Environment and Climate Change Canada (ECCC), the National Centers for Environmental Prediction (NCEP), and the Met Office (UKMO). The result shows that, compared with the individual models, our newly derived median multi-model ensemble (MME) can significantly improve the prediction skill of sub-seasonal precipitation in the MC. Both the Temporal Correlation Coefficient (TCC) skill and the Pattern Correlation Coefficient (PCC) skill reached 0.6 in lead week 1, dropped the following week, did not exceed 0.2 in lead week 3, and then lost their significance. The results show higher prediction skill near the Equator than in the north at 10° N. It is difficult to make effective predictions with the models beyond three weeks. The prediction ability of the median MME improves significantly as the total number of model members increases. The prediction performance of the median MME depends not only on the diversity of models but also on the number of model members. Moreover, the prediction skill is particularly sensitive to the intensity and phase of Boreal Summer Intraseasonal Oscillation 1 (BSISO1) with the highest skills appearing at initial phases 1 and 5.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame distilled prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.
Codex and Gemma teacher scores by category
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.000 | 0.000 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.000 | 0.000 |
| Bibliometrics | 0.000 | 0.000 |
| Science and technology studies | 0.000 | 0.000 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.000 | 0.000 |
| Research integrity | 0.000 | 0.000 |
| Insufficient payload (model declined to judge) | 0.000 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one teacher head, not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".