curl --request GET \
--url https://app.thinnest.ai/api/v1/models \
--header 'Authorization: Bearer <token>'const options = {method: 'GET', headers: {Authorization: 'Bearer <token>'}};
fetch('https://app.thinnest.ai/api/v1/models', options)
.then(res => res.json())
.then(res => console.log(res))
.catch(err => console.error(err));import requests
url = "https://app.thinnest.ai/api/v1/models"
headers = {"Authorization": "Bearer <token>"}
response = requests.get(url, headers=headers)
print(response.text){
"items": [
{
"id": "prana",
"name": "Prana",
"voice": true,
"minPlan": "free",
"available": true
},
{
"id": "prana-voice",
"name": "Prana [Voice]",
"voice": true,
"minPlan": "free",
"available": true
},
{
"id": "gpt-5-mini",
"name": "GPT-5 Mini",
"voice": true,
"minPlan": "free",
"available": true
},
{
"id": "gpt-5.6-luna",
"name": "GPT-5.6 Luna",
"voice": true,
"minPlan": "payg",
"available": false
},
{
"id": "gpt-4.1",
"name": "GPT-4.1",
"voice": false,
"minPlan": "payg",
"available": false
}
],
"byok": {
"provider": {
"id": "openai",
"label": "OpenAI"
},
"workspaceModel": "gpt-5.4-nano",
"models": [
"gpt-5.4",
"gpt-5.4-mini",
"gpt-5.4-nano",
"gpt-4.1-mini"
]
}
}{
"error": "Send a valid API key as `Authorization: Bearer <key>`."
}{
"error": "Over 240 requests a minute. Slow down and retry."
}List Models
Every model an agent may answer with, by the name the console shows. available says whether your plan may pick it — the rest are listed so you can see what an upgrade adds — and voice whether it is quick enough to answer a call (an agent’s model must be). On a workspace that brings its own keys there is also byok: your LLM provider, your key’s own model and every model the key reaches, which is what agents run on while your own keys are on.
curl --request GET \
--url https://app.thinnest.ai/api/v1/models \
--header 'Authorization: Bearer <token>'const options = {method: 'GET', headers: {Authorization: 'Bearer <token>'}};
fetch('https://app.thinnest.ai/api/v1/models', options)
.then(res => res.json())
.then(res => console.log(res))
.catch(err => console.error(err));import requests
url = "https://app.thinnest.ai/api/v1/models"
headers = {"Authorization": "Bearer <token>"}
response = requests.get(url, headers=headers)
print(response.text){
"items": [
{
"id": "prana",
"name": "Prana",
"voice": true,
"minPlan": "free",
"available": true
},
{
"id": "prana-voice",
"name": "Prana [Voice]",
"voice": true,
"minPlan": "free",
"available": true
},
{
"id": "gpt-5-mini",
"name": "GPT-5 Mini",
"voice": true,
"minPlan": "free",
"available": true
},
{
"id": "gpt-5.6-luna",
"name": "GPT-5.6 Luna",
"voice": true,
"minPlan": "payg",
"available": false
},
{
"id": "gpt-4.1",
"name": "GPT-4.1",
"voice": false,
"minPlan": "payg",
"available": false
}
],
"byok": {
"provider": {
"id": "openai",
"label": "OpenAI"
},
"workspaceModel": "gpt-5.4-nano",
"models": [
"gpt-5.4",
"gpt-5.4-mini",
"gpt-5.4-nano",
"gpt-4.1-mini"
]
}
}{
"error": "Send a valid API key as `Authorization: Bearer <key>`."
}{
"error": "Over 240 requests a minute. Slow down and retry."
}Authorizations
Your API key (ta_live_…) from Settings → API keys, sent as Authorization: Bearer <key>. Keep it on a server: it can message every customer you have. A key is full, build or read-only; a request its level does not allow is refused with 403.
Headers
Developers only: the customer workspace this request acts in — its org_… id from POST /customers. Leave it out to act in your own workspace.
"org_3fKq9TzQ1mN8vB2xR7cLpA"