Blog / Questions / FIG. 131
Can Jev Score Essays?
Can Jev do AI essay scoring? It can check rubric criteria fast, but essays stay human-graded. The short-form vs essay line and the formative rule.
Partly, and the "partly" is the whole answer. For AI essay scoring, Jev, TypeSafe AI's decision model, can check an essay against rubric criteria as a battery of closed questions (does the thesis appear in the first paragraph? is each claim supported by a cited source?). It should not assign the grade. On essays, verdicts are support for a human grader, never the mark itself.
That line isn't timidity. Essays reward qualities a checklist can't fully capture, and the stakes (grades, placement, transcripts) are exactly where a human has to own the decision.
The short-form vs essay line
Short answers: strong fit. A question with a defined correct concept decomposes cleanly into rubric verdicts with expert boundary clauses. Calibrated against teacher marks, that's reliable enough for practice feedback and pre-grading, per the AI grading page, which owns the policy side.
Essays: criteria yes, grade no. You can run a dozen rubric questions over an essay cheaply. Decomposition at that scale is normal in the catalog: one build asks 61 questions of a draft post in about a second for $0.0004, as reported (build). What you get is a map of which criteria are present, not a judgment of how good the essay is.
Jev also can't write the feedback comments. It picks from answers you define; if students need prose feedback, either map each rubric verdict to a pre-written note or hand the flagged criteria to a chat model to draft, with a teacher reviewing.
The formative rule
Use verdicts where mistakes are cheap and fast feedback is valuable: formative, not summative. Practice drafts checked on submission, students revising immediately, teachers seeing class-wide patterns. A speech-practice build shows the loop in another medium: it scores 30 seconds of speaking for pauses, filler words, and clarity, with feedback to improve (build).
For anything that counts, low-confidence verdicts go to the teacher rather than being forced, the same escalation shape as confidence routing, and the teacher decides every grade. Log question versions and verdicts so any mark can be explained if a student or parent asks.
Frequently asked questions
Can Jev grade essays automatically?
It shouldn't assign essay grades. It can check rubric criteria to support a human grader, who makes the final decision.
What can Jev grade well?
Rubric-decomposed short answers with clear boundary clauses, calibrated against teacher marks, especially for formative practice. The AI grading page covers deployment.
Can Jev give students written feedback?
Not directly, since it doesn't generate text. Map verdicts to pre-written feedback, or have a chat model draft comments that a teacher reviews.
How accurate is AI essay scoring with Jev?
No official benchmarks exist. Measure agreement with your own teachers' marks on a labeled set before using verdicts, per the getting-started guide.
Numbers throughout are as reported by the build authors, not verified by shipwithjev. Code-shaped examples are pseudocode; the official docs live at docs.typesafe.ai.