Overview
The Actionability evaluator assesses whether a piece of feedback gives the student a clear, usable next step they can reasonably act on without additional clarification. The evaluator considers:- Presence of a directive verb (e.g., add, replace, clarify, explain, revise) or focused question that points somewhere specific
- Whether feedback goes beyond evaluative comments (e.g., “good job,” “this is incomplete”)
- Clarity of the target (is it clear what should be revised?)
- Specificity of the next move (a concrete directive rather than a vague instruction)
- Whether the student could reasonably act without further clarification or more information
- Whether a next step is warranted at all (a response that didn’t require revision should not draw a forced directive)
At a glance
The evaluator was built and validated using the model and temperature below (other configurations will produce different results and may have lower accuracy):
Getting started
Follow the Quickstart to start using this evaluator:Inputs
Inputs must be de-identified. Do not submit student PII or any regulated or
sensitive personal information.
Example input
Output
Example output
Interpreting results
Accuracy and validation
This evaluator is provided as Early access. Reported metrics come from a small
held-out test split (19 examples) with wide confidence intervals and should be
read as directional. Validation testing is ongoing.
On this dimension, GEPA optimization did not improve GPT-5.4 over its naive
baseline on the held-out test set — both reached 95% accuracy. GPT-5.4 was
selected for its strong mean performance across the full suite rather than its
margin on this single dimension.