Grade Norming for English Departments Using The Good Earth Essays

Published on September 18th, 2026 by the GraideMind team

Any English department that shares a novel eventually faces the same problem: two teachers score the same kind of essay and give very different grades. Students compare notes, parents ask questions, and confidence in the process suffers. Grade norming is the practice of calibrating scoring so that similar work earns similar marks. The Good Earth, a common shared text, is a good place to start.

A stack of exam papers waiting to be graded

Begin by choosing a small set of anchor essays. Pick four to six responses to the same prompt that represent a range of quality, and remove the names. These become the shared reference for the department.

Have every teacher score the essays independently using the department rubric. Then compare results without arguing about who is right. The goal is to find where the descriptors are being read differently.

Discuss the biggest gaps first. If one teacher scored an essay a level higher on analysis than another, look at the specific sentences that drove the difference. Often the disagreement comes down to whether a paragraph counts as analysis or as a well-written summary.

Running a Norming Session

A norming session takes about an hour and pays off all year. Keep it structured so it does not turn into a debate club. Here is a format that works.

  • Distribute three to five anonymous essays and the shared rubric at least a day ahead
  • Ask each teacher to score independently and record scores before the meeting
  • Compare scores criterion by criterion and identify where they diverge
  • Discuss the language of the rubric and revise any descriptors that are unclear
  • Agree on final anchor scores and save the essays as a reference set for new teachers

Consistency is not about making teachers identical; it is about making expectations visible.

Stop spending your evenings grading essays

Let AI generate rubric-based feedback instantly, so you can focus on teaching instead.

Try it free in seconds

What Disagreements Usually Reveal

Most scoring differences trace back to vague descriptors. Terms like insightful or thorough mean different things to different readers. Rewriting them with observable behaviors, such as explains how a scene supports the claim, removes much of the guesswork.

Other differences reflect personal preferences about style. One teacher may value polished prose while another cares more about argument. Norming sessions help departments decide how much weight each element deserves.

Bringing New Teachers Into the Process

New and student teachers benefit greatly from anchor essays. Seeing how experienced colleagues score real examples is faster than reading a rubric alone. It also gives them permission to ask questions about scoring in a low-stakes setting.

Keep the anchor set fresh by adding a new example each year. Over time, you build a library that covers common issues in Good Earth essays and other units. That library also makes it easier to onboard substitutes and long-term guest teachers.

Checking Consistency With Data

After a grading cycle, compare score distributions across sections. If one teacher's average on analysis is a full level above another's, it may be worth another conversation. AI grading tools such as GraideMind can provide rubric-level breakdowns that make these comparisons easier, and they can score the same anchor essays so you have an additional data point.

Treat the data as a starting point for discussion, not a verdict. Differences may reflect real differences in classes as well as scoring. What matters is that the department can explain and defend the grades it gives.

See how fast your grading workflow can be

Most teachers go from hours per batch to minutes.

Create free account