Skip to main content
v0.4.0

What you’ll do

Evaluate a batch of text from a CSV file using all literacy evaluators. Results are output in both CSV and HTML format.

What you’ll need

  • Install the SDK globally
  • Create a CSV file with the text you want to evaluate
    • Must be 50 or fewer input rows (unless using the --bypass-row-limit option)
    • Must have text and grade columns
    • May include additional columns (will be preserved as-is in the output)
example.csv

Running the batch evaluator

Run the batch evaluator using npx from any directory:
You will be prompted for the following information:
  • CSV file path
  • Google and OpenAI API keys
    • Copy and paste directly in terminal window
    • Alternatively, provide as environment variables (GOOGLE_API_KEY and OPENAI_API_KEY, by default)
  • Output directory
    • Defaults to a folder in the current directory with a human-readable timestamp (e.g. batch-results-2024-02-07_14-30-22/)

Options

Pass in options to override the batch evaluator’s defaults:

Results

You’ll see a real-time display of the batch evaluator’s progress:
The batch evaluator will generate 2 files in your output directory:
results.csv
  • Spreadsheet-compatible format
  • Original CSV columns preserved
  • New CSV columns for each evaluator
    • {evaluator}_score
    • {evaluator}_reasoning
    • {evaluator}_status
results.html
  • Summary dashboard with grade-level distribution and text complexity charts
  • Scores and reasoning for each evaluator
If any evaluations fail (even after retries), only those rows will error out. The batch evaluator will skip those rows and then ultimately surface those failures in the results with an error status.

Graceful shutdown

If you press Ctrl+C during evaluation:
  • In-flight evaluations finish processing
  • Pending tasks are cancelled
  • Completed results are saved to results-partial.* files to preserve progress
If you press Ctrl+C twice to force quit immediately, you may lose in-flight results.