MétaCan
Menu
Back to cohort
Record W285260377

At Odds: School Achievement -- Pan-Canadian Assessments, the Setting the Record Straight

2000· article· en· W285260377 on OpenAlexaboutno aff
Gilles Fournier

Bibliographic record

VenuePhi Delta Kappan · 2000
Typearticle
Languageen
FieldSocial Sciences
TopicEducational Practices and Policies
Canadian institutionsnot available
Fundersnot available
KeywordsReading (process)PsychologyIdeologyValue (mathematics)OddsSociologyPublic relationsLawPedagogySocial psychologyPolitical sciencePoliticsComputer science
DOInot available

Abstract

fetched live from OpenAlex

Ms. Robertson's In Canada column last May demonstrated a wish to discredit an organization and its processes for ideological and personal reasons, Mr. Fournier charges. THE TITLE Bogus Points is indeed appropriate for Heather-jane Robertson's May 1999 article on the Council of Ministers of Education Canada (CMEC) and its School Achievement Indicators Program (SAIP). The article contains many misconceptions and inaccuracies and demonstrates a lack of knowledge or understanding about a process that the author should have researched adequately before making a judgment. It becomes quite apparent when reading the article that Robertson wishes to discredit an organization and its processes for ideological and personal reasons. Corrections and clarifications are in order. The initial reference to testing as a game is quite surprising, coming from a director of professional development services with a federation that wishes to represent the views of Canadian teachers. Testing is not a game, and it is one of the primary activities teachers engage in with their students. A pan-Canadian assessment program provides insight into the factors affecting student performance. Why does Robertson not find value in informing and training teachers to assess their pupils properly? It is sincerely hoped that teachers are not playing with our children's lives when assessing their efforts. If so, students participating in assessments should be forewarned about these games, so that they can make the necessary effort to amuse their keepers. Robertson's article goes on to imply that, because of the introduction of a pan-Canadian assessment program by CMEC, Canadian provinces and territories have ceded their autonomy in the evaluation of student performance. This is not the case, nor has any recent activity on the part of CMEC implied the relinquishing of provincial or territorial authority concerning matters related to education. A clarification as to the nature and role of CMEC might be instructive. Since education in Canada falls within the jurisdiction of individual provinces and territories, no federal agency exists to coordinate and report on educational activities from a pan-Canadian perspective. Founded more than 30 years ago as a means by which Canadian provinces and territories can work together on a variety of projects, CMEC is the body that speaks for education as it relates to pan-Canadian interests. Therefore, the preparation of a set of indicators of Canadian student performance falls under CMEC's mandate, as originally intended by the ministers of education; no abdication of jurisdictional rights has taken place. Contrary to what Robertson asserts, the need for accountability in education is not a U.S.-led initiative. Many countries - including France, the Netherlands, Norway, Scotland, and England, to name a few - have adopted some form of testing on a national level. In response to the growing need for a comparative measure of student performance, the ministers of education established a testing program, developed with Canadian students and priorities in mind, to assess achievement in various academic subjects. The SAIP was established in 1989. The first assessment, in mathematics, was administered in 1993 to a sample of 13- and 16-year-old Canadian students. An assessment of reading and writing skills followed in 1994, and another, in science literacy and practical task skills, took place in 1996. While the assessments did not reflect the curriculum of any one jurisdiction, SAIP was designed as a measure of the general skills expected of students in those particular age groups. SAIP was not designed to be used as a national exit exam. To imply that this activity was an initiative of the business community, partially funded by it, is entirely erroneous. When SAIP was first established, outside funding sources were sought. Private enterprise provided less than 0.1% of the total funds for the project, and this for only three years. …

Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.

How this classification was reachedexpand

Full frame distilled prediction

Teacher imitation

Not calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.

metaresearch head score (Codex)0.001
metaresearch head score (Gemma)0.000
Version: codex-gemma-dda1882f352aValidation status: machine_predicted_unvalidated
Candidate categoriesScience and technology studies, Insufficient payload (model declined to judge)
Consensus categoriesInsufficient payload (model declined to judge)
DomainCandidate signal: none · Consensus signal: none
Study designCandidate signal: Not applicable · Consensus signal: Not applicable
GenreCandidate signal: Empirical · Consensus signal: none
Teacher disagreement score0.629
Threshold uncertainty score1.000

Codex and Gemma teacher scores by category

CategoryCodexGemma
Metaresearch0.0010.000
Meta-epidemiology (narrow)0.0000.000
Meta-epidemiology (broad)0.0000.000
Bibliometrics0.0000.000
Science and technology studies0.0030.000
Scholarly communication0.0000.000
Open science0.0010.000
Research integrity0.0000.000
Insufficient payload (model declined to judge)0.0120.001

Machine scores (provisional)

The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.

Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.

Opus teacher head0.043
GPT teacher head0.371
Teacher spread0.328 · how far apart the two teachers sit on this one work
Validation statusscore_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from it

Classification

machine, unvalidated

Machine predicted; both teacher heads agree on what is shown here.

Study designNot applicable
Domainnot available
GenreEmpirical

How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".

Quick stats

Citations0
Published2000
Admission routes1
Has abstractyes

Explore more

Same venuePhi Delta KappanSame topicEducational Practices and PoliciesFrench-language works237,207