Designing a Department-Wide Common Assessment Around Harrison Bergeron
Published on September 29th, 2026 by the GraideMind team
Common assessments give English departments a way to compare student performance across classrooms, evaluate curriculum, and identify areas for growth. Harrison Bergeron is a strong candidate for a shared assessment because it is short, widely available, and rich enough to support analysis at multiple skill levels. The challenge is making the assessment reliable, meaning that a paper scored in one classroom would receive the same score in another. Achieving that reliability requires deliberate design and calibration.

Start by agreeing on the standards the assessment is meant to measure, such as constructing an argument based on textual evidence or analyzing how an author's choices shape meaning. Every element of the assessment, from the prompt to the rubric, should trace back to those standards. Without this alignment, teachers may interpret the goals differently and produce inconsistent results. A short planning meeting to define the target skills is time well spent.
Next, write a single prompt that all teachers will use, and test it by drafting sample responses at different levels. If teachers cannot agree on what a strong response would include, the prompt may be ambiguous. Adjust the wording until the expectations are clear. Also decide the conditions of administration, such as whether students write in class or at home, how much time they have, and what materials they may consult.
Building the Shared Rubric
A common rubric should be specific enough to guide scoring but flexible enough to accommodate different valid approaches. Define each performance level with observable descriptors, and include examples of the kinds of evidence and analysis that would earn each score. Test the rubric by having each teacher score the same set of sample essays and comparing results. Discrepancies point to descriptors that need refinement.
- Agree on standards and translate them into rubric criteria
- Write descriptors with observable features for every performance level
- Select anchor papers that illustrate each level for each criterion
- Hold a calibration session where teachers score the same papers independently
- Discuss differences and revise the rubric before the live assessment
Calibration is the difference between a shared rubric and a shared understanding of the rubric.
Stop spending your evenings grading essays
Let AI generate rubric-based feedback instantly, so you can focus on teaching instead.
Try it free in secondsRunning a Calibration Session
A calibration session begins with teachers independently scoring three to five anchor essays and then comparing their scores. The conversation about why scores differ is the most valuable part, as it exposes assumptions and clarifies the meaning of the rubric. Keep notes on the decisions made so that new teachers can learn from them. Plan to repeat the process periodically, since scorers drift over time.
Some departments add a second scorer for a random sample of essays to measure agreement. If agreement falls below a target, additional discussion or training may be needed. This step may sound burdensome, but it provides data on the quality of the assessment itself. Reliable scores make the results far more useful for instructional decisions.
Using the Results
Once the assessment is scored, analyze the results by criterion rather than just by total score. If the department finds that students struggle with analysis but perform well on evidence, that finding suggests where to focus instruction. Sharing results in a supportive, non-evaluative way helps teachers learn from one another. Departments can also compare performance across sections to identify effective practices.
Be cautious about using common assessment data to evaluate individual teachers, which can discourage honesty and collaboration. The purpose is to improve instruction and curriculum, not to rank colleagues. Framing the data as a shared tool builds trust. When teachers feel safe, they are more willing to experiment and share what works.
Reducing the Scoring Burden
Scoring hundreds of essays for a department-wide assessment is a heavy load, particularly when each requires careful reading and comments. Departments may set aside a scoring day, share the workload, or limit the number of comments required. These strategies help but still demand substantial time. Efficiency matters because delayed results reduce the usefulness of the data.
AI grading tools that apply a shared rubric can help departments achieve both speed and consistency by scoring every essay against the same criteria. Teachers can review the results, focus on borderline cases, and use the tool's output to inform calibration discussions. Because the same standards are applied across sections, the data becomes easier to compare. This support lets departments run common assessments more often and act on the results sooner.
See how fast your grading workflow can be
Most teachers go from hours per batch to minutes.
Create free account