Update scores for one eval result.
const url = 'https://app.everruns.com/api/v1/evals/example/runs/example/results/example/scores';const options = { method: 'PATCH', headers: {'Content-Type': 'application/json'}, body: '{"metadata":"example","scores":[{"pass":true,"reason":"Output contains expected text","value":1}],"status":"passed"}'};
try { const response = await fetch(url, options); const data = await response.json(); console.log(data);} catch (error) { console.error(error);}curl --request PATCH \ --url https://app.everruns.com/api/v1/evals/example/runs/example/results/example/scores \ --header 'Content-Type: application/json' \ --data '{ "metadata": "example", "scores": [ { "pass": true, "reason": "Output contains expected text", "value": 1 } ], "status": "passed" }'Parameters
Section titled “ Parameters ”Path Parameters
Section titled “Path Parameters”Prefixed public identifier
Prefixed public identifier
Prefixed public identifier
Request Bodyrequired
Section titled “Request Bodyrequired”Request to write external scores to one eval case result.
object
Free-form metadata attached to this resource.
Externally computed scores to store on the result.
Result from a single scorer evaluation.
object
Whether this scorer passed.
Example
trueHuman-readable explanation.
Example
Output contains expected textScore value 0.0–1.0.
Example
1Responses
Section titled “ Responses ”Success
Result of a single case within a run.
object
Collected session file contents keyed by artifact name.
Case name (denormalized for display).
When the result was created.
Error message if errored.
The case this result is for.
External identifier (evalresult_<32-hex>).
Token usage.
Execution time in milliseconds.
External scorer metadata captured during deferred write-back.
Output tokens used.
Per-scorer results.
Session created for this case (browsable in UI).
Execution status of the case.
Session creation parameters (mirrors CreateSessionRequest).
object
Agent to work in this session.
Harness for the session. If omitted, org default harness is used.
Addressable harness name (alternative to harness_id).
Max LLM iterations per turn.
LLM model override.
System prompt override (prepended to agent prompt).
Reference to a deployed app.
object
Label-only target for externally-executed runs (e.g. imported from Mira).
Carries provider/model labels and opaque params instead of session setup:
external runs are ingested already-complete, so everruns never builds a
session from this. Mirrors a provider-agnostic (provider, model) pair.
object
Session creation parameters (mirrors CreateSessionRequest).
object
Agent to work in this session.
Harness for the session. If omitted, org default harness is used.
Addressable harness name (alternative to harness_id).
Max LLM iterations per turn.
LLM model override.
System prompt override (prepended to agent prompt).
Reference to a deployed app.
object
Label-only target for externally-executed runs (e.g. imported from Mira).
Carries provider/model labels and opaque params instead of session setup:
external runs are ingested already-complete, so everruns never builds a
session from this. Mirrors a provider-agnostic (provider, model) pair.
object
Turn count.
When the result was last updated.
Example
{ "artifacts": { "patch": "diff --git a/src/lib.rs b/src/lib.rs" }, "case_name": "fix-failing-test", "created_at": "2026-01-15T10:30:00Z", "error_message": "Session timed out", "eval_case_id": "evalcase_01933b5a000070008000000000000001", "id": "evalresult_01933b5a000070008000000000000001", "input_tokens": 1200, "latency_ms": 8450, "output_tokens": 850, "session_id": "session_01933b5a00007000800000000000001", "status": "pending", "target": { "type": "session" }, "target_snapshot": { "type": "session" }, "turns": 3, "updated_at": "2026-01-15T10:30:00Z"}Eval result not found