Skip to content
APMIX.AI

Pricing

Simple plans. Monthly or yearly.

Pick a monthly allowance of weighted tokens and spend it on any model in your plan, from one key. Each model has its own multiplier; the table further down turns your allowance into the real tokens it buys on every one of them.

Billing period

Billed monthly. Cancel any time.

FreeTry before you pay.

Every new account starts with 2M free tokens, shared across 5 free models. No daily limits, no expiry, no card needed. One offer per person.

  • glm-5.3-flash-free
  • deepseek-v4.1-flash-free
  • gpt-5.6-luna-free
  • muse-spark-1.3-free
  • kimi-k3-free
Create an account

Prices in USD. Taxes may apply depending on your country.

How it works

One allowance. One multiplier per model.

Your plan is a monthly budget of weighted tokens. Every model converts the tokens you actually send and receive into weighted tokens at its own rate, so a fast model stretches your allowance further than a frontier one.

  1. 01

    Choose an allowance

    Starter, Pro or Max: 30M, 120M or 300M weighted tokens a month, shared across every model in the plan and every key on the account.

  2. 02

    Every model has a multiplier

    1× on fast models, higher on frontier and reasoning models. Prompt and completion tokens are both counted, then multiplied by the model's rate.

  3. 03

    Divide to get real tokens

    Allowance ÷ multiplier is what you can actually send to that model in a month. On a 1× model the whole allowance is yours; at 3×, a third of it.

Pro allowance: 120M weighted tokens a month ÷ Model multiplier 1.5× = 80M Real tokens a month on that model.

Model rates

42 models. One multiplier each.

The multiplier is how many weighted tokens one real token costs on a model. The plan columns show the real tokens a full monthly allowance buys there — the number you can actually send.

Plan
Provider
Showing 42 of 42

Included from Starter11 models

  • grok-4.6xAI

    0.5×1K tokens → 500 weightedContext 500K
    Free tier
    Not included
    Starter
    60M
    Pro
    240M
    Max
    600M
  • grok-4.5xAI

    0.5×1K tokens → 500 weightedContext 500K
    Free tier
    Not included
    Starter
    60M
    Pro
    240M
    Max
    600M
  • deepseek-v4-proDeepSeek

    0.5×1K tokens → 500 weightedContext 1M
    Free tier
    Not included
    Starter
    60M
    Pro
    240M
    Max
    600M
  • deepseek-v4-flashDeepSeek

    0.1×1K tokens → 100 weightedContext 1M
    Free tier
    Not included
    Starter
    300M
    Pro
    1.2B
    Max
    3B
  • glm-5.3-flashZ.ai (GLM)

    0.3×1K tokens → 300 weightedContext 1M
    Free tier
    Not included
    Starter
    100M
    Pro
    400M
    Max
    1B
  • minimax-m3MiniMax

    0.3×1K tokens → 300 weightedContext 512K
    Free tier
    Not included
    Starter
    100M
    Pro
    400M
    Max
    1B
  • mimo-v2.5-proXiaomi MiMo

    0.3×1K tokens → 300 weightedContext 1M
    Free tier
    Not included
    Starter
    100M
    Pro
    400M
    Max
    1B
  • mimo-v2.5Xiaomi MiMo

    0.05×1K tokens → 50 weightedContext 1M
    Free tier
    Not included
    Starter
    600M
    Pro
    2.4B
    Max
    6B
  • deepseek-v4.1-flashDeepSeek

    0.3×1K tokens → 300 weightedContext 1M
    Free tier
    Not included
    Starter
    100M
    Pro
    400M
    Max
    1B
  • muse-spark-1.3Meta

    1K tokens → 1K weighted
    Free tier
    Not included
    Starter
    30M
    Pro
    120M
    Max
    300M
  • gpt-5.6-lunaOpenAI

    1K tokens → 1K weighted
    Free tier
    Not included
    Starter
    30M
    Pro
    120M
    Max
    300M

Included from Pro15 models

  • gemini-3.8-flashGoogle

    1K tokens → 2K weightedContext 1M
    Free tier
    Not included
    Starter
    Not included
    Pro
    60M
    Max
    150M
  • gemini-3.7-flashGoogle

    1K tokens → 2K weightedContext 1M
    Free tier
    Not included
    Starter
    Not included
    Pro
    60M
    Max
    150M
  • gemini-3.6-flashGoogle

    1K tokens → 2K weightedContext 1M
    Free tier
    Not included
    Starter
    Not included
    Pro
    60M
    Max
    150M
  • gemini-3.5-flashGoogle

    1K tokens → 2K weightedContext 1M
    Free tier
    Not included
    Starter
    Not included
    Pro
    60M
    Max
    150M
  • qwen3.8-flashAlibaba Qwen

    0.7×1K tokens → 700 weightedContext 1M
    Free tier
    Not included
    Starter
    Not included
    Pro
    171.4M
    Max
    428.6M
  • hy4-previewTencent Hunyuan

    1K tokens → 1K weightedContext 1M
    Free tier
    Not included
    Starter
    Not included
    Pro
    120M
    Max
    300M
  • claude-haiku-4-5Anthropic

    0.9×1K tokens → 900 weightedContext 200K
    Free tier
    Not included
    Starter
    Not included
    Pro
    133.3M
    Max
    333.3M
  • gpt-5.6-terraOpenAI

    1.5×1K tokens → 1.5K weightedContext 1.1M
    Free tier
    Not included
    Starter
    Not included
    Pro
    80M
    Max
    200M
  • claude-sonnet-5Anthropic

    1K tokens → 2K weightedContext 1M
    Free tier
    Not included
    Starter
    Not included
    Pro
    60M
    Max
    150M
  • gemini-3.1-proGoogle

    1K tokens → 2K weightedContext 1M
    Free tier
    Not included
    Starter
    Not included
    Pro
    60M
    Max
    150M
  • glm-5.3Z.ai (GLM)

    1.5×1K tokens → 1.5K weightedContext 1M
    Free tier
    Not included
    Starter
    Not included
    Pro
    80M
    Max
    200M
  • glm-5.2Z.ai (GLM)

    1.5×1K tokens → 1.5K weightedContext 1M
    Free tier
    Not included
    Starter
    Not included
    Pro
    80M
    Max
    200M
  • glm-5-turboZ.ai (GLM)

    1K tokens → 1K weightedContext 202.8K
    Free tier
    Not included
    Starter
    Not included
    Pro
    120M
    Max
    300M
  • kimi-k2.7-codeMoonshot Kimi

    0.9×1K tokens → 900 weightedContext 262.1K
    Free tier
    Not included
    Starter
    Not included
    Pro
    133.3M
    Max
    333.3M
  • claude-sonnet-4-6Anthropic

    1K tokens → 2K weightedContext 200K
    Free tier
    Not included
    Starter
    Not included
    Pro
    60M
    Max
    150M

Included from Max11 models

  • gpt-5.6-solOpenAI

    1K tokens → 3K weightedContext 1.1M
    Free tier
    Not included
    Starter
    Not included
    Pro
    Not included
    Max
    100M
  • gpt-5.5OpenAI

    1K tokens → 3K weightedContext 258K
    Free tier
    Not included
    Starter
    Not included
    Pro
    Not included
    Max
    100M
  • qwen3.8-maxAlibaba Qwen

    2.5×1K tokens → 2.5K weightedContext 1M
    Free tier
    Not included
    Starter
    Not included
    Pro
    Not included
    Max
    120M
  • claude-opus-5Anthropic

    1K tokens → 4K weightedContext 1M
    Free tier
    Not included
    Starter
    Not included
    Pro
    Not included
    Max
    75M
  • claude-opus-4-8Anthropic

    1K tokens → 4K weightedContext 1M
    Free tier
    Not included
    Starter
    Not included
    Pro
    Not included
    Max
    75M
  • claude-opus-4-7Anthropic

    1K tokens → 4K weightedContext 1M
    Free tier
    Not included
    Starter
    Not included
    Pro
    Not included
    Max
    75M
  • claude-opus-4-6Anthropic

    1K tokens → 4K weightedContext 200K
    Free tier
    Not included
    Starter
    Not included
    Pro
    Not included
    Max
    75M
  • gpt-6-astraOpenAI

    7.5×1K tokens → 7.5K weightedContext 1.1M
    Free tier
    Not included
    Starter
    Not included
    Pro
    Not included
    Max
    40M
  • claude-fable-5Anthropic

    1K tokens → 8K weightedContext 1M
    Free tier
    Not included
    Starter
    Not included
    Pro
    Not included
    Max
    37.5M
  • claude-fable-5-1Anthropic

    1K tokens → 8K weightedContext 1M
    Free tier
    Not included
    Starter
    Not included
    Pro
    Not included
    Max
    37.5M
  • kimi-k3Moonshot Kimi

    2.5×1K tokens → 2.5K weighted
    Free tier
    Not included
    Starter
    Not included
    Pro
    Not included
    Max
    120M

Free tier5 models

  • glm-5.3-flash-freeZ.ai (GLM)

    1K tokens → 1K weightedContext 1M
    Free tier
    2M
    Starter
    30M
    Pro
    120M
    Max
    300M
  • deepseek-v4.1-flash-freeDeepSeek

    1K tokens → 1K weightedContext 1M
    Free tier
    2M
    Starter
    30M
    Pro
    120M
    Max
    300M
  • gpt-5.6-luna-freeOpenAI

    1K tokens → 1K weighted
    Free tier
    2M
    Starter
    30M
    Pro
    120M
    Max
    300M
  • muse-spark-1.3-freeMeta

    1K tokens → 1K weighted
    Free tier
    2M
    Starter
    30M
    Pro
    120M
    Max
    300M
  • kimi-k3-freeMoonshot Kimi

    1K tokens → 1K weighted
    Free tier
    2M
    Starter
    30M
    Pro
    120M
    Max
    300M

Multipliers are set per model and can change when a provider changes its prices. This page, each model's page and your dashboard always show the current value; requests already made keep the rate they were billed at.

FAQ

Pricing questions, answered.

Does every plan include every model?

Plans are tiered. Each plan includes the models in its own tier plus everything in the tiers below, so Max covers the whole catalog. The table above shows exactly which plans include a model, and the model's own page says the same. Every plan uses the same key and the same endpoint.

What are weighted tokens?

One unit across every model. Every prompt and completion token — reasoning tokens included — is multiplied by the model's rate and deducted from your monthly allowance. A 1,000-token exchange on a 1× model costs 1,000 weighted tokens; on a 3× model, 3,000.

Do yearly plans have a bigger allowance?

No. The allowance is monthly on both. A yearly plan is the same monthly allowance for twelve months at a lower price, paid once.

What happens when I use up the allowance?

Requests return a clear error until the monthly reset, or until you move to a bigger plan. An upgrade applies right away and the difference is charged pro rata. You can also set your own daily or weekly limits so nothing surprises you.

Can I change or cancel my plan?

Any time, from the dashboard. Upgrades take effect immediately. Downgrades and cancellations take effect at the end of the paid period, and your keys keep working until then.

Are there refunds?

Monthly plans are refundable in full while less than 5% of that month's allowance has been used. Yearly plans are not refundable. The refund policy has the exact numbers per plan.

Start with any plan. Change it whenever you like.

Choose your allowance, create a key, and point your tools at apmix.