Small decisions. Interesting possibilities.Submit contentSubmit

How it uses Jev

An Instant Eval job selects traces or spans and passes their text to Jev as state. Configured classifier questions produce Boolean, integer Score, or Category Choice judgments, and the run service stores those verdicts with the evaluated traces.

What Jev decides

Illustrative classifier requestExample answers · not a recorded Jev response · Source ↗
Question 1 · configured boolean question
YOUR APP
INSTRUCTION

Judge whether the configured criterion holds for the supplied trace text.

STATE

Text extracted from a selected trace or span, plus configured evaluator questions.

JEV · BOOLEAN
YesNo

App workflow

  1. Start Instant Eval

    A user or API requests an evaluation job over selected traces or spans.

  2. Page through traces

    The run service loads a bounded page and extracts text for the configured evaluator.

  3. Ask Jev

    Build typed classifier questions and send them with the trace text.

  4. Store verdicts

    Validate and persist results against the evaluated traces or spans.

Why it is interesting

LangWatch connects typed evaluation criteria to observability data: a user can define an evaluator once, run it over selected traces, and inspect stored verdicts beside the original activity.