Run a batch of cases through an agent or swarm and grade them
The whole measure loop in one call: attach evals, run the cases, return the per-case, per-agent, per-step graded results.
Request
This endpoint expects an object.
target
cases
evals
run_type
Allowed values:
scoring_mode
Allowed values:
label
Response
Simulation complete
Errors
401
Unauthorized Error