LIVE PRICING DATA

Supported models and pricing

Compare model families, context windows, capabilities, traffic share, and input/output pricing in one live catalog.

  • Pricing syncs automatically
  • Seven model families in one catalog
  • One API key across every provider
SERVICE RATES

Transparent provider pricing.

Tokenly price = upstream list price × the provider service rate. Rates are shown before you make a request.

Anthropic Claude ×1.0Vision, tool calling, prompt caching
OpenAI / Codex ×1.0Reasoning, coding, image generation
Google Gemini ×2.0Multimodal and long-context workflows

How to read this catalog. Traffic share is usage analytics, not a whitelist or quota. Every Tokenly account can call every model listed here. Cached-input prices are shown only when available upstream.

MODEL FAMILY

Claude models

3 active models · all plans can call these models

Provider online
ModelInput / MOutput / MContextCapabilitiesTraffic
C
Claude Opus 5claude-opus-5
$5.00$25.001M / 128K
VisionToolsCache
4.4%
C
Claude Sonnet 5claude-sonnet-5
$2.00$10.001M / 128K
VisionToolsCache
16.8%
C
Claude Haiku 4.5claude-haiku-4-5
$1.00$5.00200K / 64K
VisionToolsCache
12.0%
MODEL FAMILY

OpenAI / Codex models

3 active models · all plans can call these models

Provider online
ModelInput / MOutput / MContextCapabilitiesTraffic
G
GPT 5.6 solgpt-5.6-sol
$5.00$30.00922K / 128K
VisionToolsCache
8.7%
G
GPT 5.6 terragpt-5.6-terra
$2.00$12.00922K / 128K
VisionToolsCache
6.4%
G
GPT 6 lunagpt-6-luna
$0.10$0.50922K / 128K
VisionTools
<0.1%
MODEL FAMILY

Google Gemini models

2 active models · all plans can call these models

Provider online
ModelInput / MOutput / MContextCapabilitiesTraffic
G
Gemini 3.5 Flashgemini-3.5-flash
$3.00$18.001M / 65K
VisionToolsCache
G
Gemini 3.5 Progemini-3.5-pro
$5.00$20.001M / 65K
VisionToolsCache
<0.1%
MODEL FAMILY

GLM models

2 active models · all plans can call these models

Provider online
ModelInput / MOutput / MContextCapabilitiesTraffic
G
GLM 5.2glm-5.2
$2.38$8.32200K / 131K
VisionToolsCache
0.7%
G
GLM 5.3 Flashglm-5.3-flash
$0.14$0.491M / 128K
VisionToolsCache
2.1%
MODEL FAMILY

Kimi models

1 active model · all plans can call these models

Provider online
ModelInput / MOutput / MContextCapabilitiesTraffic
K
Kimi K3kimi-k3
$21.00$105.001M / 131K
VisionToolsCache
0.1%
MODEL FAMILY

DeepSeek models

2 active models · all plans can call these models

Provider online
ModelInput / MOutput / MContextCapabilitiesTraffic
D
DeepSeek V4 Prodeepseek-v4-pro
$0.66$1.981M / 384K
ReasoningTools
1.7%
D
DeepSeek V4.1 Flashdeepseek-v4.1-flash
$0.57$1.701M / 384K
ReasoningFast
11.3%
MODEL FAMILY

Qwen models

2 active models · all plans can call these models

Provider online
ModelInput / MOutput / MContextCapabilitiesTraffic
Q
Qwen 3.8 Flashqwen3.8-flash
$0.35$1.051M / 128K
VisionToolsCache
0.5%
Q
Qwen 3.8 Maxqwen3.8-max
$4.20$12.591M / 128K
VisionToolsCache
0.2%
READY TO BUILD?

Use any model with one Tokenly key.

Create a token, choose a provider, and start with the same OpenAI-compatible endpoint.

Create your account ↗
MODEL FAQ

Common questions.

How are model prices calculated?

We display the upstream list price and apply the provider service rate shown above. Your usage ledger records input, output, and cached input separately.

Can every plan call every model?

Yes. Plans control credits, quotas, and support. They do not hide models from the catalog.

How often does pricing update?

Pricing is synchronized regularly and the catalog shows the latest available rate for each provider.