Get Shadow Eval Job
One job with derived counts, judge spend, latest error, and stratified results.
Path parameters
Response
Successful Response
Every auto-router this job runs as a shadow arm. Multi-router jobs sample one slice of traffic and judge every arm against the same real responses
Model groups the sampled traffic is narrowed to; empty means every model the targets use
The operator who stopped the job early, recorded by the stop endpoint; 'unknown' backfilled by migration for jobs that displayed stopped when the column arrived; None when the job ended on its own. Its presence is what makes a job read stopped rather than completed
Verdicts recorded; detail endpoint only
Sampled attempts that errored; detail endpoint only
Judge cost so far; detail endpoint only
Most recent attempt error; detail endpoint only
The first router, kept for callers that predate router_names; derived so the two fields can never disagree.
Three recorded facts, no history-guessing: a stop is stopped_by (the migration backfills it for every job that displayed stopped when the column arrived, so the pre-column population is closed), completion is the window passing or every target spending its budget, and anything else is running. The all-targets-stamped fallback covers only stops written by pre-column pods during a rolling deploy.