Skip to main content
POST
Create an agent evaluation

Authorizations

Authorization
string
header
required

Use this header with a Bearer token to authenticate requests.

Path Parameters

tenant_id
string
required

Tenant ID

workspace_id
string
required

Workspace ID

Body

application/json

Agent evaluation definition

agent_id
string
config
object
evaluation_id
string
name
string
schedule
enum<string>
Available options:
daily,
weekly,
monthly
weight
number

Response

Created agent evaluation

agent_id
string
config
object
created_at
string
evaluation_id
string
evaluation_name
string

Joined from eval_evaluations.

id
string
latest_run_cancel_requested
boolean
latest_run_finished_at
string
latest_run_id
string

Latest run, populated via lateral join on eval_evaluation_runs. LatestRunID is what a row-level "stop run" action posts to, and LatestRunCancelRequested lets the row show the pending state during the window between the 202 and the worker writing 'cancelled'. Both are nil when the evaluation has never run.

latest_run_progress
number
latest_run_started_at
string
latest_run_status
string
latest_run_summary
object
latest_scored_run_finished_at
string
latest_scored_run_id
string

Latest run carrying a numeric score, via its own lateral. A failed, queued, or running run has no summary, so LatestRunSummary is null for exactly the cases where a score is still worth showing. Consumers read the state from LatestRun* and the numbers from here: folding the two together would hide the failure. Equal to LatestRunID when the newest run is scored, and all nil when no run has ever scored.

latest_scored_run_summary
object
name
string
run_requested
boolean
schedule
enum<string>
Available options:
daily,
weekly,
monthly
score_trend
number[]

Score trend: last 5 run scores in chronological order (oldest first).

updated_at
string
weight
number