List Available Models
List the language models currently available to your workspace.
Returns every language model your workspace can use, so you can pass a valid id as llmModel when creating or updating a call agent.
The available models change over time. Read this endpoint rather than hardcoding a list.
Headers
x-workspace-idstringrequiredThe unique identifier of the workspace being accessed.
Response Fields
idstringThe model identifier. This is the value to pass as llmModel on agent create and update.
namestringDisplay name for the model.
descriptionstringShort description, usually the context window. May be empty.
speedstringRelative speed score, as a percentage. May be empty.
qualitystringRelative quality score, as a percentage. May be empty.
tierstringBroad performance band, for example fast or balanced.
isRecommendedbooleanWhether the model is recommended for general use.
isTopPickbooleanWhether the model is the current default choice. Exactly one model carries this.
averageLatencynumberMean time to first token in seconds, measured across recent calls. Present only once a model has enough traffic to report on.
Response
[
{
"id": "platform-llm",
"name": "Platform LLM",
"description": "",
"speed": "",
"quality": "",
"tier": "fast",
"tags": [],
"isRecommended": true,
"isTopPick": true,
"allowedFlowIds": [],
"averageLatency": 0.32
},
{
"id": "gpt-4.1",
"name": "GPT-4.1",
"description": "1M context window",
"speed": "93%",
"quality": "94%",
"tier": "balanced",
"tags": [],
"isRecommended": true,
"isTopPick": false,
"allowedFlowIds": [],
"averageLatency": 0.88
},
{
"id": "gemini-2.5-flash",
"name": "Gemini 2.5 Flash",
"description": "1M context window",
"speed": "99%",
"quality": "98%",
"tier": "balanced",
"tags": [],
"isRecommended": false,
"isTopPick": false,
"allowedFlowIds": [],
"averageLatency": 0.83
}
]