ModelService
List models
List all available models along with their capabilities, pricing, and specifications.
/llm/v1alpha/modelsBrowse inworld.ai/models without an API key. To retrieve model metadata programmatically from this endpoint, authenticate with Authorization: Basic $INWORLD_API_KEY; requests without credentials return 401.
This is the endpoint for listing models. The OpenAI-style GET https://api.inworld.ai/v1/models (client.models.list() in the OpenAI SDK) is not available. Use the optional provider and model query parameters to filter by exact match, for example ?provider=openai. For an Inworld-hosted model, combine model with provider=inworld. By default the list holds the models you can call with chat completions; add ?output_modalities=video for video models or ?output_modalities=decisions for decision models.
Authorizationstringrequired
Your authentication credentials. For Basic authentication, please populate Basic $INWORLD_API_KEY. You can create a key in one command with the Inworld CLI: inworld workspace add-key.
providerstring
Return only models from this provider. Must match the model's provider value exactly (for example, openai).
modelstring
Return only the model with this exact model value (for example, gpt-5.4). Combine with provider to select a single provider's listing. For an Inworld-hosted model, set provider=inworld.
output_modalitiesstring
Return only models whose spec.outputModalities includes one of these values: text, image, video or decisions. Separate several values with commas, or pass all to list every model. Without it, the response lists the models you can call with chat completions; video and decision models are listed only when you ask for them (for example, ?output_modalities=video).
modelsobject[]
The list of available models.
Show child attributes
modelstringrequired
The model identifier. Combine it with provider as provider/model in the model field of Router requests. Models hosted by Inworld are listed as models/<name> (use inworld/models/<name>).
providerstringrequired
The service provider hosting the model, such as openai, anthropic, or google.
modelCreatorstring
The organization that created the model. May differ from the provider (e.g., a Meta model hosted on Groq).
pricingobject
Per-token pricing information for the model (in USD). Fields without a price are omitted.
Show child attributes
promptTokennumber
Cost per input (prompt) token in USD.
completionTokennumber
Cost per output (completion) token in USD.
promptCacheReadTokennumber
Cost per input token read from the prompt cache, in USD.
promptCacheWriteTokennumber
Cost per input token written to the prompt cache, in USD.
specobject
Technical specifications and capabilities of the model.
Show child attributes
inputModalitiesenum<string>[]
The input modalities supported by the model.
outputModalitiesenum<string>[]
The output modalities supported by the model.
contextLengthinteger
The maximum number of tokens the model can process as input (context window size).
maxCompletionTokensinteger
The maximum number of tokens the model can generate in a single response.
supportedParametersstring[]
The generation parameters supported by this model. Common values include: max_tokens, temperature, top_p, stop, seed, response_format, structured_outputs, tools, tool_choice, reasoning, include_reasoning, frequency_penalty, presence_penalty.
capabilitiesobject
High-level capability flags for the model. Only capabilities that are true are included in the response.
Show child attributes
functionCallingboolean
Whether the model supports function/tool calling.
webSearchboolean
Whether the model supports web search (grounding).
reasoningboolean
Whether the model supports chain-of-thought reasoning.
promptCachingboolean
Whether prompt caching is available for this model.
responseSchemaboolean
Whether the model supports structured output with a response schema (JSON mode).
visionboolean
Whether the model supports image/vision input.
reasoningCapabilityobject
Reasoning support in detail. Present when the model supports reasoning.
Show child attributes
supportedboolean
Whether the model supports reasoning.
supportedLevelsenum<string>[]
The reasoning effort levels the model accepts.
responseFormatCapabilityobject
Structured output support in detail. Present when the model supports JSON output.
Show child attributes
jsonboolean
Whether the model supports JSON mode (response_format: {"type": "json_object"}).
jsonSchemaboolean
Whether the model supports strict structured output with a JSON schema.
deprecationDatestring
The date the model is scheduled to be deprecated, when announced.
isSupportedboolean
Indicates whether this model is currently supported and recommended for use. Omitted when false; such models may be experimental, deprecated, or not yet generally available.
healthobject
Recent reliability and latency statistics for the model. Omitted when no data is available.
Show child attributes
successRatenumber
Fraction of recent requests that succeeded, from 0 to 1.
avgLatencyMsstring
Average request latency in milliseconds. Returned as a string (64-bit integer JSON encoding).
p95LatencyMsstring
95th percentile request latency in milliseconds. Returned as a string (64-bit integer JSON encoding).