Department Grading Calibration: Norming Scores on a Shared Eugénie Grandet Essay
Published on October 5th, 2026 by the GraideMind team
When several teachers assign the same essay on Eugénie Grandet, students in different sections often receive noticeably different scores for similar work. These inconsistencies frustrate students and parents and make it hard for department heads to trust grade data. A norming session, in which teachers score the same sample essays and compare their reasoning, is one of the most effective ways to address the problem. It does not require extensive resources, but it does require planning and a willingness to discuss disagreements openly.

The session begins with the selection of sample essays. A department head or lead teacher should choose four to six anonymous essays that represent a range of quality, including at least one that is difficult to score. Essays from previous years work well, provided identifying information is removed. The samples should respond to the same prompt that the current assignment uses so that the discussion is directly relevant.
Each teacher scores the essays independently using the shared rubric before the meeting. This step is essential because it prevents the most confident voice from shaping everyone else's judgment. When the group convenes, the scores are collected and compared, and the discussion focuses on the essays with the widest disagreement. Those are the papers where the rubric language is most open to interpretation.
Turning Disagreements Into Clearer Criteria
Disagreements are the point of the exercise, not a problem to be avoided. If one teacher gives a high score to an essay with a vivid but loosely supported argument while another gives a low score, the group can discuss what the rubric says about evidence and decide how to apply it. Often the resolution is a clarification, such as specifying how many distinct scenes constitute sufficient support. Documenting these decisions creates a shared reference for future grading.
- Score each sample essay independently before discussing results
- Focus the conversation on the essays with the largest score gaps
- Identify which rubric phrases caused different interpretations
- Record agreed definitions and examples for each performance level
- Revisit the anchor papers at the start of each grading cycle
A rubric only becomes reliable when the people using it agree about what its words mean.
Stop spending your evenings grading essays
Let AI generate rubric-based feedback instantly, so you can focus on teaching instead.
Try it free in secondsBuilding a Library of Anchor Papers
Anchor papers are annotated sample essays that illustrate each score level. After a norming session, the department can compile the agreed examples into a reference set, with brief notes explaining why each essay earned its score. New teachers and substitutes can use the set to understand departmental expectations quickly. Over time, the library grows to cover different prompts, texts, and levels of difficulty.
Anchor papers should be updated periodically to reflect changes in the curriculum and in student writing. If the department introduces a new emphasis, such as integrating historical context, the anchors should show how that emphasis is evaluated. Keeping the library current prevents the standards from drifting. It also gives students a transparent view of what success looks like.
Measuring Whether Calibration Worked
Departments can check the effect of norming by comparing score distributions across sections after grading. If one section's average is significantly higher or lower without an obvious explanation, it may indicate that calibration needs another round. Exact agreement is not necessary, but wide gaps deserve attention. A brief conversation among the teachers involved can often resolve the discrepancy.
Another useful measure is the number of grade disputes. When standards are shared and documented, teachers can explain scores more confidently, and students are less likely to feel they were treated unfairly. Fewer disputes also save time for everyone involved. This is an underappreciated benefit of calibration.
Supporting Calibration With Technology
AI-assisted grading tools can add a useful layer of consistency by applying the same rubric to every essay regardless of the section. When teachers compare the tool's scores with their own, they can see where their interpretations differ and decide whether to adjust. The tool does not replace the norming conversation but gives it a concrete data point. It can also serve as a second reader that flags essays where the score might be questioned.
Administrators interested in grading consistency can use these tools to monitor patterns across a department without reading every essay. The data may reveal that certain criteria are scored differently across classrooms, prompting targeted professional development. Combined with regular norming sessions, this creates a feedback loop that steadily improves reliability. The goal is a system students and teachers can trust.
See how fast your grading workflow can be
Most teachers go from hours per batch to minutes.
Create free account


