Skip to content
Everruns Cloud is open in early access. Run agents without operating the platform.

list_model_profiles

GET
/v1/model-profiles
curl --request GET \
--url https://app.everruns.com/api/v1/model-profiles

List curated and account-scoped profiles visible to the caller, optionally filtered by service or provider.

service
One of:

A typed service a provider driver can offer (see knowledge/foundations/providers.md).

Drivers declare supported services in code; catalog models persist their selected service. One provider account composes typed chat, decision and embedding drivers over shared authentication.

string
Allowed values: chat decisions embeddings realtime images rerank

Return only profiles for this service.

provider_id
string | null

Prefixed provider account ID used to restrict available profiles.

Authorized model profiles

Media typeapplication/json

Response wrapper for list endpoints. All list endpoints return responses wrapped in a data field.

object
data
required

Array of items returned by the list operation.

Array<object>

Read-only model behavior, independent from the credentials of a serving account.

object
key
required

Stable profile identity, including a vendor or custom-account namespace.

string
profile
required

Capabilities, decision semantics, limits and pricing for this model.

object
attachment
required

Whether the model supports file/image attachments

boolean
cost
One of:

Cost per million tokens

object
cache_read

Cached read cost per million tokens (USD), if supported

number | null format: double
cache_write

Cache write cost per million tokens (USD); absent falls back to input.

number | null format: double
cost_tiers

Tiered pricing that applies when prompt tokens exceed context thresholds. When present, the highest matching tier replaces the base rates for the whole request.

Array<object>

A pricing tier that activates above a context token threshold. For example, OpenAI charges higher rates for prompts exceeding 200K tokens.

object
above_tokens
required

Context token threshold above which this tier applies

integer format: int32
cache_read

Cached read cost per million tokens (USD) for this tier, if supported

number | null format: double
cache_write

Cache write cost per million tokens (USD); absent falls back to input.

number | null format: double
input
required

Input cost per million tokens (USD) for this tier

number format: double
output
required

Output cost per million tokens (USD) for this tier

number format: double
input
required

Input cost per million tokens (USD)

number format: double
output
required

Output cost per million tokens (USD)

number format: double
decisions
One of:

Typed semantics when this profile describes a decision model.

object
calibrated
required

Whether returned probabilities are calibrated native decision measurements.

boolean
max_choice_options

Maximum options in one choice question, when known.

integer | null
max_score_levels

Maximum levels in one score question, when known.

integer | null
primitives
required

Native primitive names supported by this profile (noul, choice or score).

Array<string>
request_tokens

Maximum token count for the complete request, when known.

integer | null
state_tokens

Maximum token count for the evaluated state, when known.

integer | null
description

Short human-readable description of the model’s strengths and intended use

string | null
family
required

Model family (e.g., “gpt-5.6-sol”, “claude-sonnet-5”)

string
knowledge

Knowledge cutoff date (YYYY-MM-DD format)

string | null
last_updated

Last updated date (YYYY-MM-DD format)

string | null
limits
One of:

Token limits

object
context
required

Maximum context window size in tokens

integer format: int32
input

Maximum input tokens (if different from context - output)

integer | null format: int32
max_media

Maximum images or PDF pages per request

integer | null format: int32
output
required

Maximum output tokens

integer format: int32
modalities
One of:

Supported modalities

object
input
required

Supported input modalities

Array<string>
Allowed values: text image audio video pdf
output
required

Supported output modalities

Array<string>
Allowed values: text image audio video pdf
name
required

Display name of the model

string
open_weights
required

Whether the model has open weights

boolean
reasoning
required

Whether the model has reasoning/chain-of-thought capabilities

boolean
reasoning_effort
One of:

Reasoning effort configuration (for reasoning models)

object
default
required

Default reasoning effort for this model

string
Allowed values: none minimal low medium high xhigh max
values
required

Available reasoning effort values for this model

Array<object>

Named reasoning effort value for UI display

object
name
required

Display name (e.g., “Low”, “Medium”)

string
value
required

The API value (e.g., “low”, “medium”)

string
Allowed values: none minimal low medium high xhigh max
release_date

Release date (YYYY-MM-DD format)

string | null
speed
One of:

Speed (service tier) configuration, for models served with selectable latency/price tiers (OpenAI service_tier).

object
default
required

Default speed for this model

string
Allowed values: flex default priority fast ultrafast
values
required

Available speed values for this model

Array<object>

Named speed value for UI display

object
cost_multiplier

Price of this tier relative to the standard rate (e.g. 2.0 for Fast, 0.5 for Flex), applied to every token bucket. None when the tier’s rate is not recorded; cost estimates then use the standard rate.

number | null format: double
name
required

Display name (e.g., “Flex”, “Fast”)

string
value
required

The API value (e.g., “flex”, “priority”)

string
Allowed values: flex default priority fast ultrafast
structured_output
required

Whether the model supports structured output (JSON mode)

boolean
supported_parameters

Provider-advertised request parameters supported by this model.

Array<string>
supports_phases

Whether the model supports native execution phases (“commentary” / “final_answer”). When true, the driver sends the phase field on assistant messages in the wire format. Currently supported by GPT-5.4 and newer via OpenAI Responses API.

boolean
supports_server_compaction

Whether the direct provider API supports threshold server-side compaction.

This is an explicit rollout control. Unknown models and provider surfaces that do not opt in remain disabled.

boolean
temperature
required

Whether temperature control is supported

boolean
tool_call
required

Whether the model supports tool/function calling

boolean
tool_search

Whether the model supports tool_search (deferred tool loading). When true, the driver can use namespaces and defer_loading to reduce token usage for large tool sets. Currently supported by GPT-5.4 and newer.

boolean
verbosity
One of:

Verbosity configuration, for models that expose output-length control (OpenAI verbosity).

object
default
required

Default verbosity for this model

string
Allowed values: low medium high
values
required

Available verbosity values for this model

Array<object>

Named verbosity value for UI display

object
name
required

Display name (e.g., “Low”, “High”)

string
value
required

The API value (e.g., “low”, “high”)

string
Allowed values: low medium high
service
required

Model service described by the profile.

string
Allowed values: chat decisions embeddings realtime images rerank
source
required

Profile origin: curated, discovered, predefined or manual.

string
vendor
One of:

Model developer, when known; this may differ from the serving provider.

string
Allowed values: openai anthropic google nvidia qwen microsoft meta minimax moonshot typesafe xai llmsim
Example
{
"data": [
{
"profile": {
"modalities": {
"input": [
"text"
],
"output": [
"text"
]
},
"reasoning_effort": {
"default": "none",
"values": [
{
"value": "none"
}
]
},
"speed": {
"default": "flex",
"values": [
{
"value": "flex"
}
]
},
"verbosity": {
"default": "low",
"values": [
{
"value": "low"
}
]
}
},
"service": "chat",
"vendor": "openai"
}
]
}