Model catalog
35 models. 13 providers. One key.
Every model below is addressable through the same endpoint, on your own provider keys, at the provider's exact price. TokenRouter adds routing — never a markup.
Models
35
Providers
13
Widest context
1.05M tokens
Cheapest input
$0.02 / 1M
Can't decide? Pass model="auto" and the router picks from the providers you have keys for.
- auto
- balanced default
- auto:cost
- cheapest capable model
- auto:latency
- fastest first token
- auto:quality
- strongest available
OpenAI 8 models
GPT-5
openai/gpt-5
cached $0.125
GPT-5 mini
openai/gpt-5-mini
cached $0.025
GPT-5 nano
openai/gpt-5-nano
cached $0.005
GPT-4.1
openai/gpt-4.1
cached $0.5
GPT-4.1 mini
openai/gpt-4.1-mini
cached $0.1
GPT-4o
openai/gpt-4o
cached $1.25
OpenAI o3
openai/o3
cached $0.5
Embedding 3 small
openai/text-embedding-3-small
Anthropic 3 models
Claude Opus 4.1
anthropic/claude-opus-4-1
cached $1.5
Claude Sonnet 4.5
anthropic/claude-sonnet-4-5
cached $0.3
Claude Haiku 4.5
anthropic/claude-haiku-4-5
cached $0.1
Google 3 models
Gemini 2.5 Pro
google/gemini-2.5-pro
cached $0.31
Gemini 2.5 Flash
google/gemini-2.5-flash
cached $0.075
Gemini 2.5 Flash-Lite
google/gemini-2.5-flash-lite
cached $0.025
xxAI 2 models
Grok 4
xai/grok-4
cached $0.75
Grok 3 mini
xai/grok-3-mini
DeepSeek 2 models
DeepSeek V3.1
deepseek/deepseek-chat
cached $0.07
DeepSeek R1
deepseek/deepseek-reasoner
cached $0.14
Mistral 3 models
Mistral Large
mistral/mistral-large-latest
Mistral Small
mistral/mistral-small-latest
Codestral
mistral/codestral-latest
AAlibaba Qwen 3 models
Qwen Max
qwen/qwen-max
Qwen Plus
qwen/qwen-plus
Qwen Turbo
qwen/qwen-turbo
Zhipu GLM 2 models
GLM-4.5
glm/glm-4.5
GLM-4.5 Air
glm/glm-4.5-air
Moonshot Kimi 1 model
Kimi K2
kimi/kimi-k2-0711-preview
cached $0.15
MiniMax 1 model
MiniMax M1
minimax/MiniMax-M1
GGroq 2 models
Llama 3.3 70B
groq/llama-3.3-70b-versatile
Llama 4 Maverick
groq/meta-llama/llama-4-maverick-17b-128e-instruct
TTogether AI 3 models
Llama 4 Scout
together/meta-llama/Llama-4-Scout-17B-16E-Instruct
DeepSeek V3
together/deepseek-ai/DeepSeek-V3
Qwen3 235B
together/Qwen/Qwen3-235B-A22B-fp8
FFireworks AI 2 models
Llama 4 Maverick
fireworks/accounts/fireworks/models/llama4-maverick-instruct-basic
DeepSeek R1
fireworks/accounts/fireworks/models/deepseek-r1
Prices are provider list rates per 1M tokens from the July 2026 catalog, billed to you directly by each provider on your own keys — TokenRouter adds no markup. New and unlisted models pass straight through the gateway; check each provider for current rates.
Every model above, one base URL
Add your provider keys once, and switch between any of these models by changing a string — or let the router choose.