Supported models and pricing
Compare every model family, context window, capability, traffic share, and input/output price in one live catalog.
- Pricing syncs automatically
- Active and recently unused models
- One API key across every provider
Transparent provider pricing.
Tokenly price = upstream list price × the provider service rate. Rates are shown before you make a request.
Latest models
New providers appear here first, then move into their family directory below.
How to read this catalog. Traffic share is usage analytics, not a whitelist or quota. Every Tokenly account can call every model listed here. Cached-input prices appear when available upstream.
Claude models
15 active models · all plans can call these models
| Model | Input / M | Output / M | Context | Capabilities | Traffic |
|---|---|---|---|---|---|
Claude Opus 5.5claude-opus-5-5 | $4.00cached · $4.00 | $20.00 | 1M / 128K | VisionToolsCache | <0.1% |
Claude Opus 5claude-opus-5 | $5.00cached · $5.00 | $25.00 | 1M / 128K | VisionToolsCache | 4.4% |
Claude Sonnet 5claude-sonnet-5 | $2.00cached · $2.00 | $10.00 | 1M / 128K | VisionToolsCache | 16.8% |
Claude Haiku 4.5claude-haiku-4-5 | $1.00cached · $1.00 | $5.00 | 200K / 64K | VisionToolsCache | 12.0% |
Claude Fable 5claude-fable-5 | $10.00cached · $10.00 | $50.00 | 1M / 128K | VisionToolsCache | 1.1% |
Claude Mythos 5claude-mythos-5 | $3.00cached · $3.00 | $15.00 | 1M / 128K | ReasoningTools | — |
Claude Opus 4.6claude-opus-4-6 | $5.00cached · $5.00 | $25.00 | 1M / 128K | VisionTools | — |
Claude Sonnet 4.6claude-sonnet-4-6 | $2.00cached · $2.00 | $10.00 | 1M / 128K | VisionTools | — |
Claude Opus 4.5claude-opus-4-5 | $5.00cached · $5.00 | $25.00 | 200K / 64K | VisionTools | — |
Claude Haiku 4.5 20251001claude-haiku-4-5-20251001 | $1.00cached · $1.00 | $5.00 | 200K / 64K | VisionTools | — |
OpenAI / Codex models
12 active models · all plans can call these models
| Model | Input / M | Output / M | Context | Capabilities | Traffic |
|---|---|---|---|---|---|
GPT 6 solgpt-6-sol | $2.00cached · $2.00 | $10.00 | 922K / 128K | VisionToolsCache | <0.1% |
GPT 5.6 solgpt-5-6-sol | $5.00cached · $5.00 | $30.00 | 922K / 128K | VisionToolsCache | 8.7% |
GPT 5.6 terragpt-5-6-terra | $2.00cached · $2.00 | $12.00 | 922K / 128K | VisionTools | 6.4% |
GPT 6 lunagpt-6-luna | $0.10cached · $0.10 | $0.50 | 922K / 128K | VisionTools | 1.2% |
GPT Image 2gpt-image-2 | Per imagecached · Per image | Per image | 128K | ImageTools | 0.6% |
GPT Image 2.5gpt-image-2-5 | Per imagecached · Per image | Per image | 128K | Image | — |
GPT 6 astragpt-6-astra | $3.00cached · $3.00 | $15.00 | 922K / 128K | ReasoningTools | — |
GPT 6 terragpt-6-terra | $1.00cached · $1.00 | $5.00 | 922K / 128K | ReasoningTools | — |
Google Gemini models
16 active models · all plans can call these models
| Model | Input / M | Output / M | Context | Capabilities | Traffic |
|---|---|---|---|---|---|
Gemini 3.8 Flashgemini-3-8-flash | $0.40cached · $0.40 | $1.60 | 1M / 128K | VisionToolsCache | <0.1% |
Gemini 3.5 Flashgemini-3-5-flash | $3.00cached · $3.00 | $18.00 | 1M / 65K | VisionToolsCache | — |
Gemini 3.5 Progemini-3-5-pro | $5.00cached · $5.00 | $20.00 | 1M / 65K | VisionToolsCache | <0.1% |
Gemini 3.1 Progemini-3-1-pro | $2.50cached · $2.50 | $12.00 | 1M / 128K | VisionTools | 3.4% |
Gemini 2.5 Flashgemini-2-5-flash | $0.30cached · $0.30 | $2.50 | 1M / 64K | VisionAudio | 2.8% |
Gemini 2.5 Progemini-2-5-pro | $1.25cached · $1.25 | $10.00 | 1M / 64K | VisionTools | — |
Gemini Embedding 001gemini-embedding-001 | $0.15cached · $0.15 | — | 32K | Embedding | — |
Gemini 3.1 Flash Imagegemini-3-1-flash-image | $0.40cached · $0.40 | $1.60 | 1M / 64K | ImageVision | — |
GLM models
5 active models · all plans can call these models
| Model | Input / M | Output / M | Context | Capabilities | Traffic |
|---|---|---|---|---|---|
GLM 5.3glm-5-3 | $1.20cached · $1.20 | $4.20 | 200K / 128K | ReasoningTools | 1.1% |
GLM 5.3 Flashglm-5-3-flash | $0.14cached · $0.14 | $0.49 | 1M / 128K | VisionToolsCache | 2.1% |
GLM 5.2glm-5-2 | $2.38cached · $2.38 | $8.32 | 200K / 131K | VisionToolsCache | 0.7% |
GLM 4.5glm-4-5 | $1.00cached · $1.00 | $4.00 | 128K | ReasoningTools | — |
GLM 4.5 Airglm-4-5-air | $0.20cached · $0.20 | $0.80 | 128K | FastTools | — |
Kimi models
4 active models · all plans can call these models
| Model | Input / M | Output / M | Context | Capabilities | Traffic |
|---|---|---|---|---|---|
Kimi K3kimi-k3 | $21.00cached · $21.00 | $105.00 | 1M / 131K | VisionToolsCache | 0.1% |
Kimi K2kimi-k2 | $4.00cached · $4.00 | $16.00 | 1M / 128K | ReasoningTools | 0.2% |
Kimi K2 Thinkingkimi-k2-thinking | $8.00cached · $8.00 | $32.00 | 1M / 128K | ReasoningTools | — |
Kimi K1.5kimi-k1-5 | $3.00cached · $3.00 | $12.00 | 200K / 64K | VisionReasoning | — |
DeepSeek models
6 active models · all plans can call these models
| Model | Input / M | Output / M | Context | Capabilities | Traffic |
|---|---|---|---|---|---|
DeepSeek V4 Prodeepseek-v4-pro | $0.66cached · $0.66 | $1.98 | 1M / 384K | ReasoningTools | 1.7% |
DeepSeek V4.1 Flashdeepseek-v4-1-flash | $0.57cached · $0.57 | $1.70 | 1M / 384K | ReasoningFast | 11.3% |
DeepSeek V4 Flashdeepseek-v4-flash | $0.20cached · $0.20 | $0.80 | 1M / 128K | ReasoningFast | 2.2% |
DeepSeek R1deepseek-r1 | $0.80cached · $0.80 | $2.40 | 128K | Reasoning | — |
DeepSeek V3deepseek-v3 | $0.27cached · $0.27 | $1.10 | 128K | ReasoningTools | — |
Qwen models
8 active models · all plans can call these models
| Model | Input / M | Output / M | Context | Capabilities | Traffic |
|---|---|---|---|---|---|
Qwen 3.8 Maxqwen3-8-max | $4.20cached · $4.20 | $12.59 | 1M / 128K | VisionToolsCache | 0.2% |
Qwen 3.8 Flashqwen3-8-flash | $0.35cached · $0.35 | $1.05 | 1M / 128K | VisionToolsCache | 0.5% |
Qwen 3.7 Plusqwen3-7-plus | $1.80cached · $1.80 | $7.20 | 1M / 128K | ReasoningTools | 1.1% |
Qwen 3.7qwen3-7 | $0.80cached · $0.80 | $3.20 | 256K | ReasoningTools | — |
Qwen 3 Coderqwen3-coder | $0.50cached · $0.50 | $2.00 | 256K | CodeTools | — |
Use any model with one Tokenly key.
Create a token, choose a provider, and start with the same OpenAI-compatible endpoint.
Common questions.
How are model prices calculated?
We display upstream list price and apply the provider service rate shown above. Your usage ledger records input, output, and cached input separately.
Can every plan call every model?
Yes. Plans control credits, quotas, and support. They do not hide models from the catalog.
How often does pricing update?
Pricing is synchronized regularly and this catalog shows the latest available rate.