One endpoint.
A live catalog.
Search the models listed by the live routing.run API. Prices, context limits, and reported availability come from the backend catalog—not a separate marketing table.
GPT-5.6 Sol$5input / 1M
OpenAI's flagship GPT-5.6 reasoning model.
gpt-5.6-sol- Context
- 1.05M
- Max output
- 128K
- Input / 1M
- $5
- Output / 1M
- $30
GPT-5.6 Terra$2input / 1M
OpenAI's balanced GPT-5.6 reasoning model.
gpt-5.6-terra- Context
- 1.05M
- Max output
- 128K
- Input / 1M
- $2
- Output / 1M
- $12
GPT-5.6 Luna$0.2input / 1M
OpenAI's efficient GPT-5.6 reasoning model.
gpt-5.6-luna- Context
- 1.05M
- Max output
- 128K
- Input / 1M
- $0.2
- Output / 1M
- $1.2
DeepSeek V4 Flash$0.15input / 1M
DeepSeek's speed-focused model for chat, coding, and tool use.
deepseek-v4-flash- Context
- 1M
- Max output
- 64K
- Input / 1M
- $0.15
- Output / 1M
- $0.4
DeepSeek V4 Pro$0.4input / 1M
DeepSeek's reasoning model for complex coding and agent workflows.
deepseek-v4-pro- Context
- 1M
- Max output
- 64K
- Input / 1M
- $0.4
- Output / 1M
- $1.4
Glm 5.3$0.107input / 1M
Routing chat model listed in the routing.run catalog.
glm-5.3- Context
- 1M
- Max output
- 128K
- Input / 1M
- $0.107
- Output / 1M
- $0.335
Kimi K3$3input / 1M
Moonshot's Kimi model for multimodal reasoning and agentic work.
kimi-k3- Context
- 1.05M
- Max output
- 16K
- Input / 1M
- $3
- Output / 1M
- $15
Kimi K2.7 Code$0.11input / 1M
Moonshot's coding-focused model for tool-heavy agent workflows.
kimi-k2.7-code- Context
- 262K
- Max output
- 66K
- Input / 1M
- $0.11
- Output / 1M
- $0.542
Kimi K2.6$0.157input / 1M
Moonshot's model for agentic coding and long-context tool use.
kimi-k2.6- Context
- 262K
- Max output
- 66K
- Input / 1M
- $0.157
- Output / 1M
- $0.731
Glm 5.2$0.425input / 1M
Routing chat model listed in the routing.run catalog.
glm-5.2- Context
- 1M
- Max output
- 128K
- Input / 1M
- $0.425
- Output / 1M
- $1.34
Glm 5.3 Flash$0.0082input / 1M
Routing chat model listed in the routing.run catalog.
glm-5.3-flash- Context
- 1.05M
- Max output
- 128K
- Input / 1M
- $0.0082
- Output / 1M
- $0.0272
Deepseek V4 Flash 0731$0.0308input / 1M
DeepSeek chat model listed in the routing.run catalog.
deepseek-v4-flash-0731- Context
- 1.05M
- Max output
- 66K
- Input / 1M
- $0.0308
- Output / 1M
- $0.0616
Deepseek V4 Pro 0813$0.277input / 1M
DeepSeek chat model listed in the routing.run catalog.
deepseek-v4-pro-0813- Context
- 1.05M
- Max output
- 128K
- Input / 1M
- $0.277
- Output / 1M
- $0.831
Deepseek V4 Flash 0731 Fast$0.192input / 1M
DeepSeek chat model listed in the routing.run catalog.
deepseek-v4-flash-0731-fast- Context
- 1M
- Max output
- 33K
- Input / 1M
- $0.192
- Output / 1M
- $0.385
Claude Fable 5.1$4.95input / 1M
Routing chat model listed in the routing.run catalog.
claude-fable-5.1- Context
- 1M
- Max output
- 128K
- Input / 1M
- $4.95
- Output / 1M
- $24.8
Claude Opus 5$1.84input / 1M
Routing chat model listed in the routing.run catalog.
claude-opus-5- Context
- 1M
- Max output
- 128K
- Input / 1M
- $1.84
- Output / 1M
- $9.21
Claude Opus 4.6$1.59input / 1M
Routing chat model listed in the routing.run catalog.
claude-opus-4.6- Context
- 1M
- Max output
- 128K
- Input / 1M
- $1.59
- Output / 1M
- $7.98
Claude Opus 4.7$1.9input / 1M
Routing chat model listed in the routing.run catalog.
claude-opus-4.7- Context
- 1M
- Max output
- 128K
- Input / 1M
- $1.9
- Output / 1M
- $9.49
Claude Opus 4.8$0.93input / 1M
Routing chat model listed in the routing.run catalog.
claude-opus-4.8- Context
- 1M
- Max output
- 128K
- Input / 1M
- $0.93
- Output / 1M
- $4.65
Claude Sonnet 5$0.627input / 1M
Routing chat model listed in the routing.run catalog.
claude-sonnet-5- Context
- 1M
- Max output
- 64K
- Input / 1M
- $0.627
- Output / 1M
- $3.14
Claude Sonnet 4.6$0.974input / 1M
Routing chat model listed in the routing.run catalog.
claude-sonnet-4.6- Context
- 1M
- Max output
- 64K
- Input / 1M
- $0.974
- Output / 1M
- $4.87
Claude Haiku 4.5$0.185input / 1M
Routing chat model listed in the routing.run catalog.
claude-haiku-4.5- Context
- 200K
- Max output
- 64K
- Input / 1M
- $0.185
- Output / 1M
- $0.926
Gpt 6 Astra$3.3input / 1M
OpenAI chat model listed in the routing.run catalog.
gpt-6-astra- Context
- 1.05M
- Max output
- 128K
- Input / 1M
- $3.3
- Output / 1M
- $16.5
Gpt 5.5$1.31input / 1M
OpenAI chat model listed in the routing.run catalog.
gpt-5.5- Context
- 1M
- Max output
- 128K
- Input / 1M
- $1.31
- Output / 1M
- $7.84
Gpt 5.4$0.522input / 1M
OpenAI chat model listed in the routing.run catalog.
gpt-5.4- Context
- 1M
- Max output
- 128K
- Input / 1M
- $0.522
- Output / 1M
- $3.14
Gpt 5.4 Mini$0.157input / 1M
OpenAI chat model listed in the routing.run catalog.
gpt-5.4-mini- Context
- 400K
- Max output
- 128K
- Input / 1M
- $0.157
- Output / 1M
- $0.941
Gpt 5.4 Nano$0.055input / 1M
OpenAI chat model listed in the routing.run catalog.
gpt-5.4-nano- Context
- 400K
- Max output
- 128K
- Input / 1M
- $0.055
- Output / 1M
- $0.344
Gpt 5.3 Codex$0.366input / 1M
OpenAI chat model listed in the routing.run catalog.
gpt-5.3-codex- Context
- 400K
- Max output
- 128K
- Input / 1M
- $0.366
- Output / 1M
- $2.93
Gpt Oss 120b$0.001input / 1M
OpenAI chat model listed in the routing.run catalog.
gpt-oss-120b- Context
- 128K
- Max output
- 16K
- Input / 1M
- $0.001
- Output / 1M
- $0.0033
Gemini 3.8 Flash$0.0412input / 1M
Routing chat model listed in the routing.run catalog.
gemini-3.8-flash- Context
- 1.05M
- Max output
- 66K
- Input / 1M
- $0.0412
- Output / 1M
- $0.206
Gemini 3.5 Flash Lite$0.0825input / 1M
Routing chat model listed in the routing.run catalog.
gemini-3.5-flash-lite- Context
- 1M
- Max output
- 66K
- Input / 1M
- $0.0825
- Output / 1M
- $0.688
Grok 4.5$0.396input / 1M
Routing chat model listed in the routing.run catalog.
grok-4.5- Context
- 500K
- Max output
- 128K
- Input / 1M
- $0.396
- Output / 1M
- $1.19
Grok 4.3$0.0312input / 1M
Routing chat model listed in the routing.run catalog.
grok-4.3- Context
- 1M
- Max output
- 128K
- Input / 1M
- $0.0312
- Output / 1M
- $0.0623
Qwen3.8 Max$0.688input / 1M
Qwen chat model listed in the routing.run catalog.
qwen3.8-max- Context
- 1M
- Max output
- 128K
- Input / 1M
- $0.688
- Output / 1M
- $2.06
Qwen3.8 Flash$0.0314input / 1M
Qwen chat model listed in the routing.run catalog.
qwen3.8-flash- Context
- 1M
- Max output
- 128K
- Input / 1M
- $0.0314
- Output / 1M
- $0.0982
Qwen3 Coder$0.0066input / 1M
Qwen chat model listed in the routing.run catalog.
qwen3-coder- Context
- 256K
- Max output
- 66K
- Input / 1M
- $0.0066
- Output / 1M
- $0.022
Deepseek V4.1 Flash$0.149input / 1M
DeepSeek chat model listed in the routing.run catalog.
deepseek-v4.1-flash- Context
- 1.05M
- Max output
- 128K
- Input / 1M
- $0.149
- Output / 1M
- $0.594
Minimax M3$0.33input / 1M
Routing chat model listed in the routing.run catalog.
minimax-m3- Context
- 200K
- Max output
- 33K
- Input / 1M
- $0.33
- Output / 1M
- $1.32
Mistral Small 4$0.103input / 1M
Routing chat model listed in the routing.run catalog.
mistral-small-4- Context
- 256K
- Max output
- 33K
- Input / 1M
- $0.103
- Output / 1M
- $0.412
No models match that search.
Prices are USD per one million tokens and refresh from thelive model catalog. Each card shows the API's availability value; when the API does not report one, the status is explicitly unknown.