Run a batch of cases through an agent or swarm and grade them

The whole measure loop in one call: attach evals, run the cases, return the per-case, per-agent, per-step graded results.

Request

This endpoint expects an object.
targetobjectRequired
caseslist of objectsRequired
evalslist of objectsOptional
run_typeenumOptional
Allowed values:
scoring_modeenumOptional
Allowed values:
labelstringOptional

Response

Simulation complete

Errors

401
Unauthorized Error