Building a Department-Wide Common Assessment for a Till We Have Faces Unit

Published on September 30th, 2026 by the GraideMind team

When four teachers in a department assign the same essay on Till We Have Faces, a student's grade should not depend on which room they happened to sit in. Yet it often does, because each teacher brings different expectations about what counts as strong analysis. A common assessment, built and scored together, brings those expectations closer and gives administrators data they can actually trust.

The first step is agreeing on what the assessment is meant to measure. Departments sometimes try to measure everything at once, including reading comprehension, literary analysis, argument writing, and mechanics, and end up with a muddy result. A more useful design chooses two or three skills, such as constructing a claim about a narrator and supporting it with textual evidence, and builds the rubric around those.

The prompt itself should be broad enough for different teaching styles but specific enough to score. A prompt asking how Orual's narration shapes the reader's judgment of her fits most classrooms, since every teacher will have covered the narration in some way. Prompts that depend on one teacher's particular emphasis, such as a long unit on Greek myth, will disadvantage students whose classes did not spend that time.

Aligning Scoring Across Teachers

Common assessments fail most often at the scoring stage. Teachers read the same rubric and interpret it differently, and nobody finds out until grades are compared months later. A calibration meeting, where everyone scores the same set of anonymous papers and discusses the differences, catches most of these problems before they affect real students. Skipping that meeting is the single most common reason shared assessments end up producing numbers nobody trusts.

  • Agree on two or three skills the assessment will measure
  • Write performance descriptors in observable terms, not vague adjectives
  • Score a shared set of anonymous papers and compare results
  • Choose model papers for each score level and keep them on file
  • Plan a second calibration after the first set of papers is graded

A shared rubric only helps if everyone reads it the same way.

Stop spending your evenings grading essays

Let AI generate rubric-based feedback instantly, so you can focus on teaching instead.

Try it free in seconds

Using the Data Without Misusing It

Once the papers are scored, the department has a rare thing: comparable data on how students across sections performed on the same task. The temptation is to rank teachers by their class averages, but that is both unfair and unhelpful, since sections differ in size, preparation, and student needs. A better use is to look for patterns, such as whether students struggled with Part Two or whether most claims lacked a clear explanation of the evidence.

Those patterns can guide instruction in the following unit. If students across sections consistently confuse plot summary with analysis, the department can build a shared mini-lesson and practice set. Because the data came from a common task, the response can be common as well. Teachers also gain a reason to share what worked in their own classrooms, which turns the assessment into a conversation about teaching and not only about scores.

Practical Logistics for Busy Departments

The biggest obstacle to common assessments is time. Teachers already carry full grading loads, and a department meeting to score papers feels like an additional burden. Building the calibration into an existing professional learning block, and limiting it to ten or twelve sample papers, keeps the cost manageable. Departments that protect that time on the calendar early in the year are far more likely to follow through.

Sharing the grading workload is another option. Some departments have each teacher score a mix of papers from other sections, which reduces bias and spreads the effort evenly. This also exposes teachers to the way colleagues think about student writing, which is a form of professional development on its own. Teachers who try it often report that the discussion afterward is the most useful part of the whole process.

Where Technology Supports the Process

AI grading tools can serve as an additional, consistent reader in a common assessment. When every paper is scored against the same rubric by the same system, the department gets a reference point that does not tire or drift, and teachers can compare their own scores against it. Large differences are worth discussing, since they may point to a rubric that needs clearer language.

The final grades should still belong to teachers, and the tool is best treated as a second opinion rather than a verdict. Departments that use it this way tend to find that conversations about scoring become more concrete, because there is a shared set of results to discuss. That focus on evidence helps teachers align their standards without anyone feeling their judgment has been replaced.

See how fast your grading workflow can be

Most teachers go from hours per batch to minutes.

Create free account