PRICING

Pay per token. Nothing else.

No subscription, no seat count, no minimum. You are charged for the tokens a call actually used — at the same rates below whichever provider ends up serving it.

Rates per 1M tokens

Generated from the live catalog — the same numbers the API bills against.

ModelProviderInputOutput1k in → 500 out
deepseek/deepseek-r1-distill-llama-8bchatDeepSeek$0.04$0.04$0.00006
openai/gpt-5.4-minichatOpenAI$0.05$0.4$0.00025
openai/gpt-5-nanochatOpenAI$0.05$0.4$0.00025
qwen/qwen3-coder-30b-a3b-instructchatAlibaba$0.07$0.27$0.000205
qwen/qwen3-235b-a22b-instruct-2507chatAlibaba$0.09$0.58$0.00038
google/gemini-2.5-flash-litechatGoogle$0.1$0.4$0.0003
openai/gpt-4.1-nanochatOpenAI$0.1$0.4$0.0003
openai/text-embedding-3-largeembeddingOpenAI$0.13$0$0.00013
deepseek/deepseek-v4-flashchatDeepSeek$0.14$0.28$0.00028
openai/gpt-4o-minichatOpenAI$0.15$0.6$0.00045
zai/glm-4.5-airchatZAI$0.2$1.1$0.00075
qwen/qwen3-coder-nextchatAlibaba$0.2$1.5$0.00095
qwen/qwen3-vl-30b-a3b-instructchatAlibaba$0.2$0.7$0.00055
google/gemini-3.1-flash-lite-previewchatGoogle$0.25$1.5$0.001
qwen/qwen3.5-35b-a3bchatAlibaba$0.25$2$0.00125
deepseek/deepseek-v3-0324chatDeepSeek$0.27$1.1$0.00082
deepseek/deepseek-v3.1chatDeepSeek$0.27$1.1$0.00082
deepseek/deepseek-v3.1-terminuschatDeepSeek$0.27$1.1$0.00082
deepseek/deepseek-v3.2chatDeepSeek$0.27$0.4$0.00047
deepseek/deepseek-v3.2-expchatDeepSeek$0.27$0.41$0.000475
google/gemini-2.5-flashchatGoogle$0.3$2.5$0.00155
zai/glm-4.6vchatZAI$0.3$0.9$0.00075
minimax/minimax-m2chatMiniMax$0.3$1.2$0.0009
minimax/minimax-m2.1chatMiniMax$0.3$1.2$0.0009
minimax/minimax-m2.5chatMiniMax$0.3$1.2$0.0009
minimax/minimax-m2.7chatMiniMax$0.3$1.2$0.0009
qwen/qwen3.5-27bchatAlibaba$0.3$2.4$0.0015
qwen/qwen3-coder-480b-a35b-instructchatAlibaba$0.3$1.3$0.00095
qwen/qwen-2.5-72b-instructchatAlibaba$0.38$0.4$0.00058
deepseek/deepseek-v3-turbochatDeepSeek$0.4$1.3$0.00105
minimaxai/minimax-m1-80kchatMiniMax$0.4$2.2$0.0015
qwen/qwen3.5-122b-a10bchatAlibaba$0.4$3.2$0.002
qwen/qwen3.5-pluschatAlibaba$0.4$2.4$0.0016
openai/codex-auto-reviewchatOpenAI$0.5$4$0.0025
google/gemini-3-flash-previewchatGoogle$0.5$3$0.002
openai/gpt-5.3-codex-sparkchatOpenAI$0.5$4$0.0025
openai/gpt-5.4chatOpenAI$0.5$4$0.0025
openai/gpt-5.5chatOpenAI$0.5$4$0.0025
deepseek/deepseek-r1-0528chatDeepSeek$0.55$2.19$0.001645
moonshotai/kimi-k2-instructchatMoonshotAI$0.57$2.3$0.00172
zai/glm-4.5chatZAI$0.6$2.2$0.0017
zai/glm-4.5vchatZAI$0.6$1.8$0.0015
zai/glm-4.6chatZAI$0.6$2.2$0.0017
zai/glm-4.7chatZAI$0.6$2.2$0.0017
moonshotai/kimi-k2-0905chatMoonshotAI$0.6$2.5$0.00185
moonshotai/kimi-k2.5chatMoonshotAI$0.6$3$0.0021
moonshotai/kimi-k2-thinkingchatMoonshotAI$0.6$2.5$0.00185
qwen/qwen3.5-397b-a17bchatAlibaba$0.6$3.6$0.0024
deepseek/deepseek-prover-v2-671bchatDeepSeek$0.7$2.5$0.00195
deepseek/deepseek-r1-turbochatDeepSeek$0.7$2.5$0.00195
deepseek/deepseek-r1-distill-llama-70bchatDeepSeek$0.8$0.8$0.0012
moonshotai/kimi-k2.6chatMoonshotAI$0.8$3.4$0.0025
anthropic/claude-haiku-4.5chatAnthropic$1$5$0.0035
zai/glm-5chatZAI$1$3.2$0.0026
openai/o4-minichatOpenAI$1.1$4.4$0.0033
google/gemini-2.5-prochatGoogle$1.25$10$0.00625
openai/gpt-5chatOpenAI$1.25$10$0.00625
openai/gpt-5.6-lunachatOpenAI$1.25$10$0.00625
openai/gpt-5.6-solchatOpenAI$1.25$10$0.00625
openai/gpt-5.6-terrachatOpenAI$1.25$10$0.00625
zai/glm-5.1chatZAI$1.4$4.4$0.0036
zai/glm-5.2chatZAI$1.4$4.4$0.0036
google/gemini-3.5-flashchatGoogle$1.5$9$0.006
deepseek/deepseek-v4-prochatDeepSeek$1.74$3.48$0.00348
anthropic/claude-sonnet-5chatAnthropic$2$10$0.007
google/gemini-3.1-pro-previewchatGoogle$2$12$0.008
openai/gpt-4.1chatOpenAI$2$8$0.006
openai/o3chatOpenAI$2$8$0.006
openai/gpt-4ochatOpenAI$2.5$10$0.0075
anthropic/claude-sonnet-4chatAnthropic$3$15$0.0105
anthropic/claude-sonnet-4.5chatAnthropic$3$15$0.0105
anthropic/claude-sonnet-4.6chatAnthropic$3$15$0.0105
moonshotai/kimi-k3chatMoonshotAI$3$15$0.0105
anthropic/claude-opus-4.5chatAnthropic$5$25$0.0175
anthropic/claude-opus-4.6chatAnthropic$5$25$0.0175
anthropic/claude-opus-4.7chatAnthropic$5$25$0.0175
anthropic/claude-opus-4.8chatAnthropic$5$25$0.0175

per 1M tokens

How a call is billed

  1. 1

    Estimated before the call

    Your prompt plus max_tokens gives an upper bound on the cost. If your available balance cannot cover it, the request is rejected with 402 and never reaches a provider — a runaway script cannot push the account negative.

  2. 2

    Held, not charged

    That estimate is held against your balance while the call runs. Available balance is credits minus everything currently held. Holds live in the database with an expiry, so a crash releases them instead of stranding your credit.

  3. 3

    Settled on actual usage

    When the response comes back you are charged for the tokens actually used, and the rest of the hold is returned. Since max_tokens is an upper bound, the settled amount is usually lower than the hold.

  4. 4

    Failed attempts cost nothing

    If a provider errors, the request is retried on the next channel carrying that model. Only the attempt that produced a response is billed. Failed attempts are logged at zero cost so you can still see what happened.

Pricing questions

The rates above are what you pay. There is no separate platform fee and no monthly minimum.

Ready to send the first call?

Create a key, change one base URL, and keep the OpenAI client you already have.