get_model_profile
const url = 'https://app.everruns.com/api/v1/model-profiles/example';const options = {method: 'GET'};
try { const response = await fetch(url, options); const data = await response.json(); console.log(data);} catch (error) { console.error(error);}curl --request GET \ --url https://app.everruns.com/api/v1/model-profiles/exampleRead a stable model profile visible to the caller. Profiles describe behavior independently from provider authentication.
Parameters
Section titled “ Parameters ”Path Parameters
Section titled “Path Parameters”Responses
Section titled “ Responses ”Model profile
Read-only model behavior, independent from the credentials of a serving account.
object
Stable profile identity, including a vendor or custom-account namespace.
Capabilities, decision semantics, limits and pricing for this model.
object
Whether the model supports file/image attachments
Cost per million tokens
object
Cached read cost per million tokens (USD), if supported
Cache write cost per million tokens (USD); absent falls back to input.
Tiered pricing that applies when prompt tokens exceed context thresholds. When present, the highest matching tier replaces the base rates for the whole request.
A pricing tier that activates above a context token threshold. For example, OpenAI charges higher rates for prompts exceeding 200K tokens.
object
Context token threshold above which this tier applies
Cached read cost per million tokens (USD) for this tier, if supported
Cache write cost per million tokens (USD); absent falls back to input.
Input cost per million tokens (USD) for this tier
Output cost per million tokens (USD) for this tier
Input cost per million tokens (USD)
Output cost per million tokens (USD)
Typed semantics when this profile describes a decision model.
object
Whether returned probabilities are calibrated native decision measurements.
Maximum options in one choice question, when known.
Maximum levels in one score question, when known.
Native primitive names supported by this profile (noul, choice or score).
Maximum token count for the complete request, when known.
Maximum token count for the evaluated state, when known.
Short human-readable description of the model’s strengths and intended use
Model family (e.g., “gpt-5.6-sol”, “claude-sonnet-5”)
Knowledge cutoff date (YYYY-MM-DD format)
Last updated date (YYYY-MM-DD format)
Token limits
object
Maximum context window size in tokens
Maximum input tokens (if different from context - output)
Maximum images or PDF pages per request
Maximum output tokens
Display name of the model
Whether the model has open weights
Whether the model has reasoning/chain-of-thought capabilities
Reasoning effort configuration (for reasoning models)
object
Default reasoning effort for this model
Available reasoning effort values for this model
Named reasoning effort value for UI display
object
Display name (e.g., “Low”, “Medium”)
The API value (e.g., “low”, “medium”)
Release date (YYYY-MM-DD format)
Speed (service tier) configuration, for models served with
selectable latency/price tiers (OpenAI service_tier).
object
Default speed for this model
Available speed values for this model
Named speed value for UI display
object
Price of this tier relative to the standard rate (e.g. 2.0 for Fast,
0.5 for Flex), applied to every token bucket. None when the tier’s
rate is not recorded; cost estimates then use the standard rate.
Display name (e.g., “Flex”, “Fast”)
The API value (e.g., “flex”, “priority”)
Whether the model supports structured output (JSON mode)
Provider-advertised request parameters supported by this model.
Whether the model supports native execution phases (“commentary” / “final_answer”).
When true, the driver sends the phase field on assistant messages in the wire format.
Currently supported by GPT-5.4 and newer via OpenAI Responses API.
Whether the direct provider API supports threshold server-side compaction.
This is an explicit rollout control. Unknown models and provider surfaces that do not opt in remain disabled.
Whether temperature control is supported
Whether the model supports tool/function calling
Whether the model supports tool_search (deferred tool loading). When true, the driver can use namespaces and defer_loading to reduce token usage for large tool sets. Currently supported by GPT-5.4 and newer.
Verbosity configuration, for models that expose output-length
control (OpenAI verbosity).
object
Default verbosity for this model
Available verbosity values for this model
Named verbosity value for UI display
object
Display name (e.g., “Low”, “High”)
The API value (e.g., “low”, “high”)
Model service described by the profile.
Profile origin: curated, discovered, predefined or manual.
Example
{ "profile": { "modalities": { "input": [ "text" ], "output": [ "text" ] }, "reasoning_effort": { "default": "none", "values": [ { "value": "none" } ] }, "speed": { "default": "flex", "values": [ { "value": "flex" } ] }, "verbosity": { "default": "low", "values": [ { "value": "low" } ] } }, "service": "chat", "vendor": "openai"}Unknown profile