Skip to main content
GET
Get an agent evaluation by ID

Authorizations

Authorization
string
header
required

Use this header with a Bearer token to authenticate requests.

Path Parameters

tenant_id
string
required

Tenant ID

workspace_id
string
required

Workspace ID

agent_evaluation_id
string
required

Agent evaluation ID (UUID)

Response

Agent evaluation details

agent_id
string
config
object
created_at
string
evaluation_id
string
evaluation_name
string

Joined from eval_evaluations.

id
string
latest_run_cancel_requested
boolean
latest_run_finished_at
string
latest_run_id
string

Latest run, populated via lateral join on eval_evaluation_runs. LatestRunID is what a row-level "stop run" action posts to, and LatestRunCancelRequested lets the row show the pending state during the window between the 202 and the worker writing 'cancelled'. Both are nil when the evaluation has never run.

latest_run_progress
number
latest_run_started_at
string
latest_run_status
string
latest_run_summary
object
latest_scored_run_finished_at
string
latest_scored_run_id
string

Latest run carrying a numeric score, via its own lateral. A failed, queued, or running run has no summary, so LatestRunSummary is null for exactly the cases where a score is still worth showing. Consumers read the state from LatestRun* and the numbers from here: folding the two together would hide the failure. Equal to LatestRunID when the newest run is scored, and all nil when no run has ever scored.

latest_scored_run_summary
object
name
string
run_requested
boolean
schedule
enum<string>
Available options:
daily,
weekly,
monthly
score_trend
number[]

Score trend: last 5 run scores in chronological order (oldest first).

updated_at
string
weight
number