Pricing
Simple plans. Monthly or yearly.
Pick a monthly allowance of weighted tokens and spend it on any model in your plan, from one key. Each model has its own multiplier; the table further down turns your allowance into the real tokens it buys on every one of them.
Billed monthly. Cancel any time.
Starter
For side projects and trying things out.
$4.99/month
30M weighted tokens / month
What's included
- 30M weighted tokens every month
- No daily or weekly caps. You set the pace.
- Every Starter-tier model
- 2 API keys
Pro
Most popularFor daily coding with CLI agents.
$14.99/month
120M weighted tokens / month
What's included
- 120M weighted tokens every month
- No daily or weekly caps. You set the pace.
- Everything in Starter
- 10 API keys
Max
For heavy, all-day agent workloads.
$29.99/month
300M weighted tokens / month
What's included
- 300M weighted tokens every month
- No daily or weekly caps. You set the pace.
- Everything in Pro
- Unlimited API keys and priority routing
FreeTry before you pay.
Every new account starts with 2M free tokens, shared across 5 free models. No daily limits, no expiry, no card needed. One offer per person.
- glm-5.3-flash-free
- deepseek-v4.1-flash-free
- gpt-5.6-luna-free
- muse-spark-1.3-free
- kimi-k3-free
Prices in USD. Taxes may apply depending on your country.
How it works
One allowance. One multiplier per model.
Your plan is a monthly budget of weighted tokens. Every model converts the tokens you actually send and receive into weighted tokens at its own rate, so a fast model stretches your allowance further than a frontier one.
- 01
Choose an allowance
Starter, Pro or Max: 30M, 120M or 300M weighted tokens a month, shared across every model in the plan and every key on the account.
- 02
Every model has a multiplier
1× on fast models, higher on frontier and reasoning models. Prompt and completion tokens are both counted, then multiplied by the model's rate.
- 03
Divide to get real tokens
Allowance ÷ multiplier is what you can actually send to that model in a month. On a 1× model the whole allowance is yours; at 3×, a third of it.
Pro allowance: 120M weighted tokens a month ÷ Model multiplier 1.5× = 80M Real tokens a month on that model.
Model rates
42 models. One multiplier each.
The multiplier is how many weighted tokens one real token costs on a model. The plan columns show the real tokens a full monthly allowance buys there — the number you can actually send.
Included from Starter11 models
grok-4.6xAI0.5×1K tokens → 500 weightedContext 500K- Free tier
- Not included
- Starter
- 60M
- Pro
- 240M
- Max
- 600M
grok-4.5xAI0.5×1K tokens → 500 weightedContext 500K- Free tier
- Not included
- Starter
- 60M
- Pro
- 240M
- Max
- 600M
deepseek-v4-proDeepSeek0.5×1K tokens → 500 weightedContext 1M- Free tier
- Not included
- Starter
- 60M
- Pro
- 240M
- Max
- 600M
deepseek-v4-flashDeepSeek0.1×1K tokens → 100 weightedContext 1M- Free tier
- Not included
- Starter
- 300M
- Pro
- 1.2B
- Max
- 3B
glm-5.3-flashZ.ai (GLM)0.3×1K tokens → 300 weightedContext 1M- Free tier
- Not included
- Starter
- 100M
- Pro
- 400M
- Max
- 1B
minimax-m3MiniMax0.3×1K tokens → 300 weightedContext 512K- Free tier
- Not included
- Starter
- 100M
- Pro
- 400M
- Max
- 1B
mimo-v2.5-proXiaomi MiMo0.3×1K tokens → 300 weightedContext 1M- Free tier
- Not included
- Starter
- 100M
- Pro
- 400M
- Max
- 1B
mimo-v2.5Xiaomi MiMo0.05×1K tokens → 50 weightedContext 1M- Free tier
- Not included
- Starter
- 600M
- Pro
- 2.4B
- Max
- 6B
deepseek-v4.1-flashDeepSeek0.3×1K tokens → 300 weightedContext 1M- Free tier
- Not included
- Starter
- 100M
- Pro
- 400M
- Max
- 1B
muse-spark-1.3Meta1×1K tokens → 1K weighted- Free tier
- Not included
- Starter
- 30M
- Pro
- 120M
- Max
- 300M
gpt-5.6-lunaOpenAI1×1K tokens → 1K weighted- Free tier
- Not included
- Starter
- 30M
- Pro
- 120M
- Max
- 300M
Included from Pro15 models
gemini-3.8-flashGoogle2×1K tokens → 2K weightedContext 1M- Free tier
- Not included
- Starter
- Not included
- Pro
- 60M
- Max
- 150M
gemini-3.7-flashGoogle2×1K tokens → 2K weightedContext 1M- Free tier
- Not included
- Starter
- Not included
- Pro
- 60M
- Max
- 150M
gemini-3.6-flashGoogle2×1K tokens → 2K weightedContext 1M- Free tier
- Not included
- Starter
- Not included
- Pro
- 60M
- Max
- 150M
gemini-3.5-flashGoogle2×1K tokens → 2K weightedContext 1M- Free tier
- Not included
- Starter
- Not included
- Pro
- 60M
- Max
- 150M
qwen3.8-flashAlibaba Qwen0.7×1K tokens → 700 weightedContext 1M- Free tier
- Not included
- Starter
- Not included
- Pro
- 171.4M
- Max
- 428.6M
hy4-previewTencent Hunyuan1×1K tokens → 1K weightedContext 1M- Free tier
- Not included
- Starter
- Not included
- Pro
- 120M
- Max
- 300M
claude-haiku-4-5Anthropic0.9×1K tokens → 900 weightedContext 200K- Free tier
- Not included
- Starter
- Not included
- Pro
- 133.3M
- Max
- 333.3M
gpt-5.6-terraOpenAI1.5×1K tokens → 1.5K weightedContext 1.1M- Free tier
- Not included
- Starter
- Not included
- Pro
- 80M
- Max
- 200M
claude-sonnet-5Anthropic2×1K tokens → 2K weightedContext 1M- Free tier
- Not included
- Starter
- Not included
- Pro
- 60M
- Max
- 150M
gemini-3.1-proGoogle2×1K tokens → 2K weightedContext 1M- Free tier
- Not included
- Starter
- Not included
- Pro
- 60M
- Max
- 150M
glm-5.3Z.ai (GLM)1.5×1K tokens → 1.5K weightedContext 1M- Free tier
- Not included
- Starter
- Not included
- Pro
- 80M
- Max
- 200M
glm-5.2Z.ai (GLM)1.5×1K tokens → 1.5K weightedContext 1M- Free tier
- Not included
- Starter
- Not included
- Pro
- 80M
- Max
- 200M
glm-5-turboZ.ai (GLM)1×1K tokens → 1K weightedContext 202.8K- Free tier
- Not included
- Starter
- Not included
- Pro
- 120M
- Max
- 300M
kimi-k2.7-codeMoonshot Kimi0.9×1K tokens → 900 weightedContext 262.1K- Free tier
- Not included
- Starter
- Not included
- Pro
- 133.3M
- Max
- 333.3M
claude-sonnet-4-6Anthropic2×1K tokens → 2K weightedContext 200K- Free tier
- Not included
- Starter
- Not included
- Pro
- 60M
- Max
- 150M
Included from Max11 models
gpt-5.6-solOpenAI3×1K tokens → 3K weightedContext 1.1M- Free tier
- Not included
- Starter
- Not included
- Pro
- Not included
- Max
- 100M
gpt-5.5OpenAI3×1K tokens → 3K weightedContext 258K- Free tier
- Not included
- Starter
- Not included
- Pro
- Not included
- Max
- 100M
- Qwen3.8 MaxNew
qwen3.8-maxAlibaba Qwen2.5×1K tokens → 2.5K weightedContext 1M- Free tier
- Not included
- Starter
- Not included
- Pro
- Not included
- Max
- 120M
claude-opus-5Anthropic4×1K tokens → 4K weightedContext 1M- Free tier
- Not included
- Starter
- Not included
- Pro
- Not included
- Max
- 75M
claude-opus-4-8Anthropic4×1K tokens → 4K weightedContext 1M- Free tier
- Not included
- Starter
- Not included
- Pro
- Not included
- Max
- 75M
claude-opus-4-7Anthropic4×1K tokens → 4K weightedContext 1M- Free tier
- Not included
- Starter
- Not included
- Pro
- Not included
- Max
- 75M
claude-opus-4-6Anthropic4×1K tokens → 4K weightedContext 200K- Free tier
- Not included
- Starter
- Not included
- Pro
- Not included
- Max
- 75M
- GPT-6 AstraNew
gpt-6-astraOpenAI7.5×1K tokens → 7.5K weightedContext 1.1M- Free tier
- Not included
- Starter
- Not included
- Pro
- Not included
- Max
- 40M
claude-fable-5Anthropic8×1K tokens → 8K weightedContext 1M- Free tier
- Not included
- Starter
- Not included
- Pro
- Not included
- Max
- 37.5M
claude-fable-5-1Anthropic8×1K tokens → 8K weightedContext 1M- Free tier
- Not included
- Starter
- Not included
- Pro
- Not included
- Max
- 37.5M
kimi-k3Moonshot Kimi2.5×1K tokens → 2.5K weighted- Free tier
- Not included
- Starter
- Not included
- Pro
- Not included
- Max
- 120M
Free tier5 models
glm-5.3-flash-freeZ.ai (GLM)1×1K tokens → 1K weightedContext 1M- Free tier
- 2M
- Starter
- 30M
- Pro
- 120M
- Max
- 300M
deepseek-v4.1-flash-freeDeepSeek1×1K tokens → 1K weightedContext 1M- Free tier
- 2M
- Starter
- 30M
- Pro
- 120M
- Max
- 300M
gpt-5.6-luna-freeOpenAI1×1K tokens → 1K weighted- Free tier
- 2M
- Starter
- 30M
- Pro
- 120M
- Max
- 300M
muse-spark-1.3-freeMeta1×1K tokens → 1K weighted- Free tier
- 2M
- Starter
- 30M
- Pro
- 120M
- Max
- 300M
- Kimi K3 FreeFree
kimi-k3-freeMoonshot Kimi1×1K tokens → 1K weighted- Free tier
- 2M
- Starter
- 30M
- Pro
- 120M
- Max
- 300M
Multipliers are set per model and can change when a provider changes its prices. This page, each model's page and your dashboard always show the current value; requests already made keep the rate they were billed at.
FAQ
Pricing questions, answered.
Does every plan include every model?
Plans are tiered. Each plan includes the models in its own tier plus everything in the tiers below, so Max covers the whole catalog. The table above shows exactly which plans include a model, and the model's own page says the same. Every plan uses the same key and the same endpoint.
What are weighted tokens?
One unit across every model. Every prompt and completion token — reasoning tokens included — is multiplied by the model's rate and deducted from your monthly allowance. A 1,000-token exchange on a 1× model costs 1,000 weighted tokens; on a 3× model, 3,000.
Do yearly plans have a bigger allowance?
No. The allowance is monthly on both. A yearly plan is the same monthly allowance for twelve months at a lower price, paid once.
What happens when I use up the allowance?
Requests return a clear error until the monthly reset, or until you move to a bigger plan. An upgrade applies right away and the difference is charged pro rata. You can also set your own daily or weekly limits so nothing surprises you.
Can I change or cancel my plan?
Any time, from the dashboard. Upgrades take effect immediately. Downgrades and cancellations take effect at the end of the paid period, and your keys keep working until then.
Are there refunds?
Monthly plans are refundable in full while less than 5% of that month's allowance has been used. Yearly plans are not refundable. The refund policy has the exact numbers per plan.
Start with any plan. Change it whenever you like.
Choose your allowance, create a key, and point your tools at apmix.

