MétaCan
Menu
Back to cohort
Record W163729223

Grades That Mean Something: Kentucky Develops Standards-Based Report Cards a Group of Teachers, School Leaders, and Education Researchers Create Report Cards That Link Course Grades to Student Progress on Mastering State Standards

2011· article· en· W163729223 on OpenAlexaboutno aff
Thomas R. Guskey, Gerry Swan, Lee Ann Jung

Bibliographic record

VenuePhi Delta Kappan · 2011
Typearticle
Languageen
FieldDecision Sciences
TopicEducational Assessment and Improvement
Canadian institutionsnot available
Fundersnot available
KeywordsReport cardGrading (engineering)AccountabilityLearning standardsAcademic standardsMathematics educationStandardized testPsychologyWork (physics)Medical educationPublic relationsPedagogyPolitical scienceCurriculumHigher educationMedicineEngineering
DOInot available

Abstract

fetched live from OpenAlex

Nearly all states today have standards for student learning that describe what students should learn and be able to do. Nearly all states also have large-scale accountability assessment programs designed to measure students' proficiency on those standards. Despite these commonalities, schools in each state are left to develop their own standards-based student report cards as the primary means of communicating information about students' performance on state standards. Although school leaders would undoubtedly like to align their reporting procedures with the same standards and assessments that guide instructional programs, most lack the time and resources to do so. Those few leaders who take up the challenge rarely have expertise in developing effective standards-based reporting forms and inevitably encounter significant design and implementation problems (Guskey & Bailey, 2010). To help Kentucky educators address this challenge, we worked with a group of teachers and school leaders to develop a common, statewide, standards-based student report card for all grade levels. While some Canadian provinces have used standards-based report cards for many years, Kentucky educators are the first in the U.S. to attempt such a statewide reform. Data from the early implementation demonstrate that schools can implement more effective ways of communicating student learning with little additional work by teachers and that parents and community members can be strong supporters of such reforms. This shows great promise for revolutionizing reporting systems in Kentucky and elsewhere. STANDARDS-BASED GRADING Grades have long been identified by those in the measurement community as prime examples of unreliable measurement. Huge differences exist among teachers in the criteria they use when assigning grades. Even in schools where established policies offer guidelines for grading, significant variation remains in individual teachers' grading practices. The unique adaptations teachers use in assigning grades to students with disabilities and English learners make that variation wider still. These varying grading practices result in part from the lack of formal training teachers receive on grading and reporting. Most teachers have scant knowledge of various grading methods, the advantages and shortcomings of each, or the effects of different grading policies on students. As a result, most simply replicate what they experienced as students. Because the nature of these experiences widely vary, so do the grading practices and policies teachers employ. Rarely do these policies and practices reflect those recommended by researchers and aligned with a standards-based approach. Standards-based approaches to grading and reporting address these grading dilemmas in two important ways. First, they require teachers to base grades on explicit criteria derived from the articulated learning standards. To assign grades, teachers must analyze the meaning of each standard and decide what evidence best reflects achievement of that specific standard. Second, they compel teachers to distinguish product, process, and progress criteria in assigning grades (Guskey, 2006, 2009). THE KENTUCKY INITIATIVE We began our standards-based grading initiative in Kentucky by bringing together educators from three diverse school districts who had been working to develop standards-based report cards, unaware of each other's efforts. District and school leaders, along with teacher leaders from each district were invited to a three-day, summer workshop on standards-based report cards led by researchers with expertise in grading and reporting policies and practices. The first part of the workshop focused on the unique challenges of standards-based grading, recommended practices in grading and reporting, and methods of applying these practices to students with disabilities and English learners. The second part featured school leaders and teachers working to create two standards-based reporting forms: one for grades K-5, and another for grades 6-12. …

Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.

How this classification was reachedexpand

Full frame machine prediction

Teacher imitation

Not calibrated prevalence, not ground truth. Human validation pending. The Gemma side is a direct model label for every work in the frame, read from the title-only record. The Codex side is a classifier learned from the 10,348 direct Codex labels and calibrated to design-weighted sample rates; fields without enough sample support carry no Codex call. Candidate is the union of the two sides; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels.

metaresearch head score (Codex)0.025
metaresearch head score (Gemma)0.035
Version: metacan-v3-hybrid-931329e0061cValidation status: machine_predicted_unvalidated
Candidate categoriesnone
Consensus categoriesnone
DomainCandidate signal: none · Consensus signal: none
Study designCandidate signal: Observational · Consensus signal: none
GenreCandidate signal: Empirical · Consensus signal: Empirical
Teacher disagreement score0.498
Threshold uncertainty score0.989

Distilled classifier scores by category (both heads)

CategoryCodexGemma
Metaresearch0.0250.035
Meta-epidemiology (narrow)0.0010.001
Meta-epidemiology (broad)0.0000.000
Bibliometrics0.0030.003
Science and technology studies0.0080.002
Scholarly communication0.0070.007
Open science0.0030.007
Research integrity0.0010.002
Insufficient payload (model declined to judge)0.0100.004

Machine scores (provisional)

The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.

Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.

Opus teacher head0.267
GPT teacher head0.480
Teacher spread0.213 · how far apart the two teachers sit on this one work
Validation statusscore_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from it

Classification

machine, unvalidated

Machine predicted; a candidate call from one source (direct Gemma or distilled Codex), not a consensus.

The models applied no category: nothing in the taxonomy fit this work.
Study designObservational
Domainnot available
GenreEmpirical

How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".

Quick stats

Citations0
Published2011
Admission routes1
Has abstractyes

Explore more

Same venuePhi Delta KappanSame topicEducational Assessment and ImprovementFrench-language works237,207