Skip to main content
Evaluator last updated June 24, 2026.

Overview

The Student Response Anchor evaluator assesses whether a piece of feedback is clearly based on the student’s specific response — i.e., referencing, building on, or responding to the student’s own idea, wording, or use of evidence. In short, it assesses whether the feedback demonstrates an accurate understanding of what the student wrote. The evaluator considers:
  • Specific reference to the student’s work (the feedback names or builds on the student’s particular idea, wording, or evidence)
  • Whether the feedback is generic or template-based (a stock phrase that could apply to any response)
  • Whether feedback is based on an accurate understanding and interpretation of the student’s response
  • Whether the feedback addresses something the student did not actually say or attempt

At a glance

The evaluator was built and validated using the model and temperature below (other configurations will produce different results and may have lower accuracy):

Getting started

Follow the Quickstart to start using this evaluator:

Inputs

Inputs must be de-identified. Do not submit student PII or any regulated or sensitive personal information.
Example input

Output

Example output

Interpreting results

Accuracy and validation

This evaluator is provided as Early access. Reported metrics come from a small held-out test split (24 examples) with wide confidence intervals and should be read as directional. Validation testing is ongoing.
We assessed performance against Quill.org ↗ classroom writing data (64 labeled pairs; 20 train / 20 validation / 24 test) — expert-annotated student-response and teacher-feedback pairs labeled by Leanlab Education ↗ using the Productive Coaching rubric.
On this dimension, GEPA optimization improved GPT-5.4 over its naive baseline on the held-out test set (+4 points accuracy, +6 points macro-F1).

Evaluator release history