Proceedings of the Survey Methods Section A Review of the Weighting Strategy for the Canadian Community Health Survey
Bibliographic record
Abstract
The regional component of the Canadian Community Health Survey (CCHS) is a cross-sectional survey with a complex, multi-stage, multi-frame design. It collects general health-related information from a sample large enough to provide estimates for more than 120 health regions across Canada. To date, there have been three regional component surveys conducted in the years 2001, 2003, and 2005. The year 2007 marks a turning point for the survey, as it has been redesigned to incorporate a continuous collection process. In the past, data was collected over a period of one year biennially. Starting in January of 2007, data is collected continually with no breaks in the collection schedule. As part of the CCHS redesign, the methodology of the weighting process is being reviewed. This revision involves some improvements to the weighting strategy, including the methodology of the nonresponse adjustments and the integration of the different frames. As well, the overall process will be simplified to reduce the number of adjustments required.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame distilled prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.
Codex and Gemma teacher scores by category
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.434 | 0.103 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.000 | 0.000 |
| Bibliometrics | 0.000 | 0.001 |
| Science and technology studies | 0.002 | 0.001 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.001 | 0.000 |
| Research integrity | 0.000 | 0.000 |
| Insufficient payload (model declined to judge) | 0.000 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; both teacher heads agree on what is shown here.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".