A Rubric for Grading "Hills Like White Elephants" Subtext Essays

Published on October 9th, 2026 by the GraideMind team

"Hills Like White Elephants" is almost entirely dialogue between a man and a woman waiting for a train, and the central issue is never named outright. That design is exactly why the story produces such uneven student essays. Some students read the subtext with real precision, while others invent backstory the text never supports, and a rubric needs to tell those two groups apart.

The most important decision is whether your rubric rewards a particular interpretation or a particular method. Most readers agree the conversation concerns a medical procedure and a strained relationship, but students vary in how much weight they give each of those threads. Grading the method, meaning how well the student builds an argument from the dialogue, lets you reward thoughtful readings that differ from yours.

Because the story is short, students should be able to support every claim with a precise reference to the page. Ask them to treat repeated words, deflected questions, and the setting of the station between two landscapes as pieces of evidence rather than as decoration. A rubric that names those features explicitly gives students a map of where strong analysis tends to come from.

Criteria Worth Including

Start with a criterion for textual evidence, scored on whether quotations are specific, brief, and tied to a claim. Add a criterion for interpretation that separates retelling the conversation from explaining what each speaker wants and avoids saying. A third criterion for structure can look at whether the essay builds toward a considered conclusion instead of listing observations about symbolism in the order they appear in the story.

  • Quotes dialogue accurately and explains how the phrasing reveals intent
  • Analyzes the two speakers separately instead of merging them into one viewpoint
  • Uses the setting, including the contrasting hills and the station, as supporting evidence
  • Acknowledges what the text leaves unresolved without dodging the question
  • Avoids importing facts about the characters that the story does not establish

A confident reading is only as strong as the lines it can point to.

Stop spending your evenings grading essays

Let AI generate rubric-based feedback instantly, so you can focus on teaching instead.

Try it free in seconds

Handling Sensitive Content With Care

The story's subject matter is mature, and teachers sometimes worry about how to frame the assignment for younger students. A short note in the assignment sheet explaining that the story addresses a difficult personal decision, and that essays should stay focused on literary craft, helps keep the discussion analytical. Students can discuss the characters' conflict thoroughly without the essay needing graphic detail that the text itself never provides.

Grading criteria can reinforce this focus by rewarding attention to how Hemingway conveys the tension rather than to what a student personally believes about the decision. Essays that slide into a policy argument usually stop analyzing the text, and feedback can redirect them with a prompt such as asking which line shows the woman's changing attitude. That redirection keeps the paper on the literary task and makes the grading more defensible if a parent asks about it.

Writing Feedback That Moves the Reader Forward

Useful comments on these essays are usually questions or concrete instructions, not verdicts. Telling a student that an argument "needs more evidence" leaves them guessing, while asking them to find a second line where the woman changes the subject gives them a clear next step. The difference is especially large for students who are capable but unsure how to turn a hunch into a supported claim.

It also helps to name what is working so students repeat it. If a student notices that the man keeps calling the situation "simple" while the woman keeps looking at the landscape, saying that this contrast is a strong piece of evidence teaches the habit of close observation. An AI grading tool configured with your rubric can draft this kind of targeted comment at scale, leaving you free to adjust tone and priorities.

Calibrating Scores Across Sections

If several teachers assign this story, calibrate before grading by scoring the same three sample essays independently and comparing results. Disagreements usually appear around the interpretation criterion, where one teacher rewards boldness and another rewards caution. Settling that question together, and writing down the decision, prevents students in one section from earning notably higher marks for the same quality of work.

A shared set of annotated sample responses is more valuable than a longer rubric. Short notes explaining why a particular paragraph earned a three instead of a four make the descriptors concrete, and new teachers can use them to learn the department's standard quickly. Reusing the same samples each year also lets you notice when the assignment itself, not the students, needs revision.

See how fast your grading workflow can be

Most teachers go from hours per batch to minutes.

Create free account