At Odds: School Achievement -- Pan-Canadian Assessments, the Setting the Record Straight
Bibliographic record
Abstract
Ms. Robertson's In Canada column last May demonstrated a wish to discredit an organization and its processes for ideological and personal reasons, Mr. Fournier charges. THE TITLE Bogus Points is indeed appropriate for Heather-jane Robertson's May 1999 article on the Council of Ministers of Education Canada (CMEC) and its School Achievement Indicators Program (SAIP). The article contains many misconceptions and inaccuracies and demonstrates a lack of knowledge or understanding about a process that the author should have researched adequately before making a judgment. It becomes quite apparent when reading the article that Robertson wishes to discredit an organization and its processes for ideological and personal reasons. Corrections and clarifications are in order. The initial reference to testing as a game is quite surprising, coming from a director of professional development services with a federation that wishes to represent the views of Canadian teachers. Testing is not a game, and it is one of the primary activities teachers engage in with their students. A pan-Canadian assessment program provides insight into the factors affecting student performance. Why does Robertson not find value in informing and training teachers to assess their pupils properly? It is sincerely hoped that teachers are not playing with our children's lives when assessing their efforts. If so, students participating in assessments should be forewarned about these games, so that they can make the necessary effort to amuse their keepers. Robertson's article goes on to imply that, because of the introduction of a pan-Canadian assessment program by CMEC, Canadian provinces and territories have ceded their autonomy in the evaluation of student performance. This is not the case, nor has any recent activity on the part of CMEC implied the relinquishing of provincial or territorial authority concerning matters related to education. A clarification as to the nature and role of CMEC might be instructive. Since education in Canada falls within the jurisdiction of individual provinces and territories, no federal agency exists to coordinate and report on educational activities from a pan-Canadian perspective. Founded more than 30 years ago as a means by which Canadian provinces and territories can work together on a variety of projects, CMEC is the body that speaks for education as it relates to pan-Canadian interests. Therefore, the preparation of a set of indicators of Canadian student performance falls under CMEC's mandate, as originally intended by the ministers of education; no abdication of jurisdictional rights has taken place. Contrary to what Robertson asserts, the need for accountability in education is not a U.S.-led initiative. Many countries - including France, the Netherlands, Norway, Scotland, and England, to name a few - have adopted some form of testing on a national level. In response to the growing need for a comparative measure of student performance, the ministers of education established a testing program, developed with Canadian students and priorities in mind, to assess achievement in various academic subjects. The SAIP was established in 1989. The first assessment, in mathematics, was administered in 1993 to a sample of 13- and 16-year-old Canadian students. An assessment of reading and writing skills followed in 1994, and another, in science literacy and practical task skills, took place in 1996. While the assessments did not reflect the curriculum of any one jurisdiction, SAIP was designed as a measure of the general skills expected of students in those particular age groups. SAIP was not designed to be used as a national exit exam. To imply that this activity was an initiative of the business community, partially funded by it, is entirely erroneous. When SAIP was first established, outside funding sources were sought. Private enterprise provided less than 0.1% of the total funds for the project, and this for only three years. …
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame distilled prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.
Codex and Gemma teacher scores by category
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.001 | 0.000 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.000 | 0.000 |
| Bibliometrics | 0.000 | 0.000 |
| Science and technology studies | 0.003 | 0.000 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.001 | 0.000 |
| Research integrity | 0.000 | 0.000 |
| Insufficient payload (model declined to judge) | 0.012 | 0.001 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; both teacher heads agree on what is shown here.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".