LIVE PRICING DATA

Supported models and pricing

Compare every model family, context window, capability, traffic share, and input/output price in one live catalog.

  • Pricing syncs automatically
  • Active and recently unused models
  • One API key across every provider
SERVICE RATES

Transparent provider pricing.

Tokenly price = upstream list price × the provider service rate. Rates are shown before you make a request.

Anthropic Claude ×1.0Vision, tool calling, prompt caching
OpenAI / Codex ×1.0Reasoning, coding, image generation
Google Gemini ×2.0Multimodal and long-context workflows

How to read this catalog. Traffic share is usage analytics, not a whitelist or quota. Every Tokenly account can call every model listed here. Cached-input prices appear when available upstream.

CLAUDE

Claude models

15 active models · all plans can call these models

Provider online
Show 10 recently unused Claude models
ModelInput / MOutput / MContextCapabilitiesTraffic
C
Claude Opus 5.5claude-opus-5-5
$4.00$20.001M / 128K
VisionToolsCache
<0.1%
C
Claude Opus 5claude-opus-5
$5.00$25.001M / 128K
VisionToolsCache
4.4%
C
Claude Sonnet 5claude-sonnet-5
$2.00$10.001M / 128K
VisionToolsCache
16.8%
C
Claude Haiku 4.5claude-haiku-4-5
$1.00$5.00200K / 64K
VisionToolsCache
12.0%
C
Claude Fable 5claude-fable-5
$10.00$50.001M / 128K
VisionToolsCache
1.1%
C
Claude Mythos 5claude-mythos-5
$3.00$15.001M / 128K
ReasoningTools
C
Claude Opus 4.6claude-opus-4-6
$5.00$25.001M / 128K
VisionTools
C
Claude Sonnet 4.6claude-sonnet-4-6
$2.00$10.001M / 128K
VisionTools
C
Claude Opus 4.5claude-opus-4-5
$5.00$25.00200K / 64K
VisionTools
C
Claude Haiku 4.5 20251001claude-haiku-4-5-20251001
$1.00$5.00200K / 64K
VisionTools
OPENAI

OpenAI / Codex models

12 active models · all plans can call these models

Provider online
Show 7 recently unused OpenAI models
ModelInput / MOutput / MContextCapabilitiesTraffic
G
GPT 6 solgpt-6-sol
$2.00$10.00922K / 128K
VisionToolsCache
<0.1%
G
GPT 5.6 solgpt-5-6-sol
$5.00$30.00922K / 128K
VisionToolsCache
8.7%
G
GPT 5.6 terragpt-5-6-terra
$2.00$12.00922K / 128K
VisionTools
6.4%
G
GPT 6 lunagpt-6-luna
$0.10$0.50922K / 128K
VisionTools
1.2%
G
GPT Image 2gpt-image-2
Per imagePer image128K
ImageTools
0.6%
G
GPT Image 2.5gpt-image-2-5
Per imagePer image128K
Image
G
GPT 6 astragpt-6-astra
$3.00$15.00922K / 128K
ReasoningTools
G
GPT 6 terragpt-6-terra
$1.00$5.00922K / 128K
ReasoningTools
GOOGLE

Google Gemini models

16 active models · all plans can call these models

Provider online
Show 11 recently unused Gemini models
ModelInput / MOutput / MContextCapabilitiesTraffic
G
Gemini 3.8 Flashgemini-3-8-flash
$0.40$1.601M / 128K
VisionToolsCache
<0.1%
G
Gemini 3.5 Flashgemini-3-5-flash
$3.00$18.001M / 65K
VisionToolsCache
G
Gemini 3.5 Progemini-3-5-pro
$5.00$20.001M / 65K
VisionToolsCache
<0.1%
G
Gemini 3.1 Progemini-3-1-pro
$2.50$12.001M / 128K
VisionTools
3.4%
G
Gemini 2.5 Flashgemini-2-5-flash
$0.30$2.501M / 64K
VisionAudio
2.8%
G
Gemini 2.5 Progemini-2-5-pro
$1.25$10.001M / 64K
VisionTools
G
Gemini Embedding 001gemini-embedding-001
$0.1532K
Embedding
G
Gemini 3.1 Flash Imagegemini-3-1-flash-image
$0.40$1.601M / 64K
ImageVision
GLM

GLM models

5 active models · all plans can call these models

Provider online
Show 2 recently unused GLM models
ModelInput / MOutput / MContextCapabilitiesTraffic
G
GLM 5.3glm-5-3
$1.20$4.20200K / 128K
ReasoningTools
1.1%
G
GLM 5.3 Flashglm-5-3-flash
$0.14$0.491M / 128K
VisionToolsCache
2.1%
G
GLM 5.2glm-5-2
$2.38$8.32200K / 131K
VisionToolsCache
0.7%
G
GLM 4.5glm-4-5
$1.00$4.00128K
ReasoningTools
G
GLM 4.5 Airglm-4-5-air
$0.20$0.80128K
FastTools
KIMI

Kimi models

4 active models · all plans can call these models

Provider online
Show 2 recently unused Kimi models
ModelInput / MOutput / MContextCapabilitiesTraffic
K
Kimi K3kimi-k3
$21.00$105.001M / 131K
VisionToolsCache
0.1%
K
Kimi K2kimi-k2
$4.00$16.001M / 128K
ReasoningTools
0.2%
K
Kimi K2 Thinkingkimi-k2-thinking
$8.00$32.001M / 128K
ReasoningTools
K
Kimi K1.5kimi-k1-5
$3.00$12.00200K / 64K
VisionReasoning
DEEPSEEK

DeepSeek models

6 active models · all plans can call these models

Provider online
Show 3 recently unused DeepSeek models
ModelInput / MOutput / MContextCapabilitiesTraffic
D
DeepSeek V4 Prodeepseek-v4-pro
$0.66$1.981M / 384K
ReasoningTools
1.7%
D
DeepSeek V4.1 Flashdeepseek-v4-1-flash
$0.57$1.701M / 384K
ReasoningFast
11.3%
D
DeepSeek V4 Flashdeepseek-v4-flash
$0.20$0.801M / 128K
ReasoningFast
2.2%
D
DeepSeek R1deepseek-r1
$0.80$2.40128K
Reasoning
D
DeepSeek V3deepseek-v3
$0.27$1.10128K
ReasoningTools
QWEN

Qwen models

8 active models · all plans can call these models

Provider online
Show 5 recently unused Qwen models
ModelInput / MOutput / MContextCapabilitiesTraffic
Q
Qwen 3.8 Maxqwen3-8-max
$4.20$12.591M / 128K
VisionToolsCache
0.2%
Q
Qwen 3.8 Flashqwen3-8-flash
$0.35$1.051M / 128K
VisionToolsCache
0.5%
Q
Qwen 3.7 Plusqwen3-7-plus
$1.80$7.201M / 128K
ReasoningTools
1.1%
Q
Qwen 3.7qwen3-7
$0.80$3.20256K
ReasoningTools
Q
Qwen 3 Coderqwen3-coder
$0.50$2.00256K
CodeTools
READY TO BUILD?

Use any model with one Tokenly key.

Create a token, choose a provider, and start with the same OpenAI-compatible endpoint.

Create your account ↗
MODEL FAQ

Common questions.

How are model prices calculated?

We display upstream list price and apply the provider service rate shown above. Your usage ledger records input, output, and cached input separately.

Can every plan call every model?

Yes. Plans control credits, quotas, and support. They do not hide models from the catalog.

How often does pricing update?

Pricing is synchronized regularly and this catalog shows the latest available rate.