Bulk update scores for all results in an eval run.
const url = 'https://app.everruns.com/api/v1/evals/example/runs/example/scores';const options = { method: 'PATCH', headers: {'Content-Type': 'application/json'}, body: '{"metadata":"example","results":[{"result_id":"evalresult_01933b5a000070008000000000000001","scores":[{"pass":true,"reason":"Output contains expected text","value":1}],"status":"passed"}]}'};
try { const response = await fetch(url, options); const data = await response.json(); console.log(data);} catch (error) { console.error(error);}curl --request PATCH \ --url https://app.everruns.com/api/v1/evals/example/runs/example/scores \ --header 'Content-Type: application/json' \ --data '{ "metadata": "example", "results": [ { "result_id": "evalresult_01933b5a000070008000000000000001", "scores": [ { "pass": true, "reason": "Output contains expected text", "value": 1 } ], "status": "passed" } ] }'Parameters
Section titled “ Parameters ”Path Parameters
Section titled “Path Parameters”Prefixed public identifier
Prefixed public identifier
Request Bodyrequired
Section titled “Request Bodyrequired”Request to write external scores to several results of an eval run.
object
Free-form metadata attached to this resource.
Per-result score updates applied together.
Score update for one eval case result.
object
Eval case result to update.
Example
evalresult_01933b5a000070008000000000000001Externally computed scores to store on the result.
Result from a single scorer evaluation.
object
Whether this scorer passed.
Example
trueHuman-readable explanation.
Example
Output contains expected textScore value 0.0–1.0.
Example
1Responses
Section titled “ Responses ”Success
Response wrapper for list endpoints.
All list endpoints return responses wrapped in a data field.
object
Array of items returned by the list operation.
Result of a single case within a run.
object
Collected session file contents keyed by artifact name.
Case name (denormalized for display).
When the result was created.
Error message if errored.
The case this result is for.
External identifier (evalresult_<32-hex>).
Token usage.
Execution time in milliseconds.
External scorer metadata captured during deferred write-back.
Output tokens used.
Per-scorer results.
Session created for this case (browsable in UI).
Execution status of the case.
Session creation parameters (mirrors CreateSessionRequest).
object
Agent to work in this session.
Harness for the session. If omitted, org default harness is used.
Addressable harness name (alternative to harness_id).
Max LLM iterations per turn.
LLM model override.
System prompt override (prepended to agent prompt).
Reference to a deployed app.
object
Label-only target for externally-executed runs (e.g. imported from Mira).
Carries provider/model labels and opaque params instead of session setup:
external runs are ingested already-complete, so everruns never builds a
session from this. Mirrors a provider-agnostic (provider, model) pair.
object
Session creation parameters (mirrors CreateSessionRequest).
object
Agent to work in this session.
Harness for the session. If omitted, org default harness is used.
Addressable harness name (alternative to harness_id).
Max LLM iterations per turn.
LLM model override.
System prompt override (prepended to agent prompt).
Reference to a deployed app.
object
Label-only target for externally-executed runs (e.g. imported from Mira).
Carries provider/model labels and opaque params instead of session setup:
external runs are ingested already-complete, so everruns never builds a
session from this. Mirrors a provider-agnostic (provider, model) pair.
object
Turn count.
When the result was last updated.
Example
{ "data": [ { "artifacts": { "patch": "diff --git a/src/lib.rs b/src/lib.rs" }, "case_name": "fix-failing-test", "created_at": "2026-01-15T10:30:00Z", "error_message": "Session timed out", "eval_case_id": "evalcase_01933b5a000070008000000000000001", "id": "evalresult_01933b5a000070008000000000000001", "input_tokens": 1200, "latency_ms": 8450, "output_tokens": 850, "session_id": "session_01933b5a00007000800000000000001", "status": "pending", "target": { "type": "session" }, "target_snapshot": { "type": "session" }, "turns": 3, "updated_at": "2026-01-15T10:30:00Z" } ]}