tokker.dev

Early index, not launched yet.

What AI tokens really cost, per million.

About 60 API providers price tokens differently, and subscriptions hide their limits in 5-hour windows and weekly caps. Tokker puts every model and every plan in dollars per 1M tokens, with the source beside each number.

As of 5 Oct 2026

Open-weight, 18 offers from 16 providers

Cheapest way to buy GLM-5.3, $ per 1M tokens

Input $0.90, output $4.00 per 1M tokens.

DeepInfra, fp4.

GLM-5.3: cheapest offers by provider, US dollars per 1M tokens
ProviderInOutChecked
DeepInfrafp4$0.90$4.005 Oct 2026 source: api.deepinfra.com
Hugging Face Inference ProvidersStandard$0.90$4.005 Oct 2026 source: router.huggingface.co
Runwaresecondary source$1.20$4.005 Oct 2026 source: router.requesty.ai
Z.AI (Zhipu)China site, from CNY$1.19$4.185 Oct 2026 source: docs.bigmodel.cn
$ per 1M
  • Claude Opus 5.5, cheapest at Anthropic: input $4.00, output $20.00 per 1M tokensINOUT
  • GPT-6 Sol, cheapest at OpenAI: input $2.00, output $10.00 per 1M tokensINOUT
  • Claude Max 20x subscription: about $0.077 per 1M tokens at full useFULL USE
  • Gemini 3.1 Pro Preview, cheapest at Google (Gemini API): input $2.00, output $12.00 per 1M tokensINOUT
  • Claude Sonnet 5.5, cheapest at Anthropic: input $2.00, output $10.00 per 1M tokensINOUT
  • ChatGPT Pro $200 subscription: about $0.044 per 1M tokens at full useFULL USE
  • Gemini 3.8 Flash, cheapest at Google (Gemini API): input $0.75, output $3.75 per 1M tokensINOUT
  • Mistral Large 3, cheapest at Mistral AI: input $0.50, output $1.50 per 1M tokensINOUT
  • Google AI Pro subscription: about $0.021 per 1M tokens at full useFULL USE
  • MiniMax-M3, cheapest at DeepInfra: input $0.28, output $1.10 per 1M tokensINOUT
  • Kimi K3, cheapest at Hugging Face Inference Providers: input $2.70, output $13.50 per 1M tokensINOUT
  • Copilot Pro+ subscription: about $0.59 per 1M tokens at full useFULL USE
  • GLM-5.3, cheapest at DeepInfra: input $0.90, output $4.00 per 1M tokensINOUT
  • DeepSeek V4.1 Flash, cheapest at SiliconFlow: input $0.15, output $0.60 per 1M tokensINOUT
  • GLM Coding Max subscription: about $0.10 per 1M tokens at full useFULL USE
  • gpt-oss-120b, cheapest at Runware: input $0.032, output $0.14 per 1M tokensINOUT
  • gpt-oss-20b, cheapest at DeepInfra: input $0.03, output $0.14 per 1M tokensINOUT
  • Coding Plan Pro subscription: about $0.026 per 1M tokens at full useFULL USE
  • Qwen3.7-Flash, cheapest at Alibaba Cloud Model Studio: input $0.030, output $0.119 per 1M tokensINOUT
  • ModelArk Coding Lite subscription: about $0.039 per 1M tokens at full useFULL USE
  • GLM-4.7-Flash on Z.AI (Zhipu): freePRICE

62 API providers, 480 offers, 103 plans from 32 vendors, 158 sources. Prices as of 5 Oct 2026.

First snapshot. Once checks repeat, moves will show as ▼ cheaper or ▲ pricier.

Cheapest right now

US dollars per 1M tokens, cheapest provider per model. Every number links to the page it came from.

Frontier models

Frontier models: the cheapest provider for each, US dollars per 1M tokens
ModelCheapest providerInputOutputCache readOffersSource
Claude Opus 5.5Anthropic$4.00$20.00$0.207checked 5 Oct 2026 (source: platform.claude.com)
GPT-6 SolOpenAI$2.00$10.00$0.205checked 5 Oct 2026 (source: developers.openai.com)
Gemini 3.1 Pro PreviewGoogle (Gemini API)$2.00$12.00$0.203checked 5 Oct 2026 (source: ai.google.dev)

Cheapest means the lowest blended price, three parts input to one part output. OpenRouter is listed but never ranked cheapest: its headline price can pair one endpoint's input price with another's output price. Prices in other currencies are converted at the ECB rate of 2 Oct 2026.

Subscriptions

Ranked by estimated $ per 1M tokens at full use, for agentic coding: 20k input tokens (90% cached) and 1.5k output per call, four 5-hour windows a day.

Subscription plans ranked by estimated US dollars per 1M tokens at full use
PlanPrice / month$ / 1M at full useRuns out firstLimits fromSource
Google AI Pro AI ProGoogle$19.99$0.021Daily capVendor pagechecked 5 Oct 2026 (source: geminicli.com)
Synthetic subscription Subscription Pack (1)Synthetic$30$0.023Not publishedVendor pagechecked 5 Oct 2026 (source: synthetic.new)
Coding Plan ProAlibaba Cloud (Model Studio, intl)$50$0.026Monthly capVendor pagechecked 5 Oct 2026 (source: alibabacloud.com)
ModelArk Coding Plan LiteBytePlus (ByteDance intl)$20$0.039Monthly capSecondary sourcechecked 5 Oct 2026 (source: byteplus.com)
ChatGPT / Codex Pro $200OpenAI$200$0.044Not publishedSecondary sourcechecked 5 Oct 2026 (source: learn.chatgpt.com)
Cerebras Code Code MaxCerebras$200$0.056Daily capVendor pagechecked 5 Oct 2026 (source: cerebras.ai)
Claude ProAnthropic$20$0.065Not publishedSecondary sourcechecked 5 Oct 2026 (source: claude.com)
Claude Max 5xAnthropic$100$0.067Not publishedSecondary sourcechecked 5 Oct 2026 (source: support.claude.com)
Cerebras Code Code ProCerebras$50$0.069Daily capVendor pagechecked 5 Oct 2026 (source: cerebras.ai)
Claude Max 20xAnthropic$200$0.077Weekly capSecondary sourcechecked 5 Oct 2026 (source: support.claude.com)
Google AI Ultra AI Ultra $100 (5x Pro)Google$100$0.077Daily capVendor pagechecked 5 Oct 2026 (source: geminicli.com)
ChatGPT / Codex PlusOpenAI$20$0.089Not publishedVendor pagechecked 5 Oct 2026 (source: learn.chatgpt.com)
ChatGPT / Codex Pro $100OpenAI$100$0.089Not publishedSecondary sourcechecked 5 Oct 2026 (source: learn.chatgpt.com)
ChatGPT / Codex Pro $500OpenAI$500$0.089Not publishedSecondary sourcechecked 5 Oct 2026 (source: learn.chatgpt.com)
GLM Coding Plan MaxZ.AI (Zhipu)$168$0.10Weekly capSecondary sourcechecked 5 Oct 2026 (source: docs.z.ai)
ChatGPT / Codex BusinessOpenAI$25$0.11Not publishedVendor pagechecked 5 Oct 2026 (source: learn.chatgpt.com)
GLM Coding Plan ProZ.AI (Zhipu)$80$0.12Weekly capSecondary sourcechecked 5 Oct 2026 (source: docs.z.ai)
Google AI Ultra AI Ultra $200 (20x Pro)Google$200$0.15Daily capVendor pagechecked 5 Oct 2026 (source: geminicli.com)
GLM Coding Plan LiteZ.AI (Zhipu)$18$0.16Weekly capVendor pagechecked 5 Oct 2026 (source: docs.z.ai)
T3 Chat ProT3 Chat$8$0.16Monthly creditSecondary sourcechecked 5 Oct 2026 (source: t3.chat)
GitHub Copilot MaxGitHub$100$0.53Monthly creditVendor pagechecked 5 Oct 2026 (source: docs.github.com)
GitHub Copilot Pro+GitHub$39$0.59Monthly creditVendor pagechecked 5 Oct 2026 (source: docs.github.com)
GitHub Copilot ProGitHub$10$0.70Monthly creditVendor pagechecked 5 Oct 2026 (source: docs.github.com)
Warp MaxWarp$200$0.88Monthly creditVendor pagechecked 5 Oct 2026 (source: warp.dev)
JetBrains AI AI UltimateJetBrains$30$0.90Monthly creditVendor pagechecked 5 Oct 2026 (source: jetbrains.com)
JetBrains AI AI ProJetBrains$10$1.05Monthly creditVendor pagechecked 5 Oct 2026 (source: jetbrains.com)
GitHub Copilot BusinessGitHub$19$1.05Monthly creditVendor pagechecked 5 Oct 2026 (source: docs.github.com)
Warp BuildWarp$20$1.05Monthly creditVendor pagechecked 5 Oct 2026 (source: warp.dev)
Augment StandardAugment Code$20$1.05Monthly creditVendor pagechecked 5 Oct 2026 (source: augmentcode.com)
GitHub Copilot EnterpriseGitHub$39$1.05Monthly creditVendor pagechecked 5 Oct 2026 (source: docs.github.com)
Augment BusinessAugment Code$100$1.05Monthly creditVendor pagechecked 5 Oct 2026 (source: augmentcode.com)
Zed ProZed Industries$10$2.31Monthly creditVendor pagechecked 5 Oct 2026 (source: zed.dev)
Bolt.new ProStackBlitz$25$2.50Not publishedVendor pagechecked 5 Oct 2026 (source: bolt.new)
Warp BusinessWarp$50$2.63Monthly creditVendor pagechecked 5 Oct 2026 (source: warp.dev)
Bolt.new TeamsStackBlitz$30$3.00Not publishedVendor pagechecked 5 Oct 2026 (source: bolt.new)

No absolute limits published, so no estimate: AWS (Kiro), Alibaba Cloud (Model Studio, intl), Alibaba Cloud Bailian (China), Amp (Sourcegraph), Anthropic, BytePlus (ByteDance intl), Chutes, Cognition (Devin / Windsurf), Cursor (Anysphere), Factory, Featherless AI, GitHub, Google, MiniMax, Mistral AI, Moonshot AI, OpenAI, Perplexity, Replit, Tencent Cloud, Volcengine (ByteDance China), Z.AI (Zhipu) and xAI. Those plans are in the index with their verbatim wording.

Same model, international site and China site

Converted to dollars at the ECB rate, input / output per 1M tokens:

  • GLM-5.3: api.z.ai $1.40 / $4.40, bigmodel.cn $1.19 / $4.18; the China site is cheaper.
  • Kimi K3: platform.kimi.ai $3.00 / $15.00, platform.moonshot.cn $2.98 / $14.92; about the same price.
  • MiniMax-M3: platform.minimax.io $0.30 / $1.20, platform.minimaxi.com $0.313 / $1.25; the international site is cheaper.
  • DeepSeek V4.1 Flash: BytePlus $0.30 / $1.20, Volcengine $0.298 / $1.19; about the same price.

Subscriptions, decoded

5-hour windows, weekly caps and credit pools, turned into real tokens and a price per 1M at full use. At full use, the coding plans below cost $0.021 to $0.10 per 1M tokens. The same agentic workload through the Claude Sonnet 5.5 API costs $1.05: 10 to 51 times more.

Workload
20k in (90% cached) / 1.5k out per call

Claude Max 20x, effective price at full use

$0.077 per 1M tokens

Claude Max 20x ≈ $0.077 per 1M tokens at full use; the weekly cap runs out before the 5-hour windows do.

Assumes 20k in (90% cached) / 1.5k out per call, four 5-hour windows a day, every cap used. The same workload on the Claude Sonnet 5.5 API costs $1.05 per 1M, 14× more. Based on secondary measurements, not published limits.

$ per 1M at full use, log scale. Pick a plan.

  • Claude Sonnet 5.5 API$1.05
Tokens each cap allows at full use for this workload.
PlanPrice / monthPer 5-hour windowPer weekPer month$ / 1M at full useSource
$200not published609M binds first2.61B$0.077checked 5 Oct 2026 (source: support.claude.com)
$200nonenot published4.51B$0.044checked 5 Oct 2026 (source: learn.chatgpt.com)
$19.99not published226M968M binds first$0.021checked 5 Oct 2026 (source: geminicli.com)
$39not publishednot published66.6M binds first$0.59checked 5 Oct 2026 (source: docs.github.com)
$16874.9M374M binds first1.60B$0.10checked 5 Oct 2026 (source: docs.z.ai)
$50129M968M1.94B binds first$0.026checked 5 Oct 2026 (source: alibabacloud.com)
$2040.9M258M516M binds first$0.039checked 5 Oct 2026 (source: byteplus.com)

How it stays true

Four steps for every number. Step 1 is done for the seed index; the daily re-check and the review are being built.

  1. Sourced

    Every price links to the provider's own page. If a provider doesn't publish a number, the index says “unknown”. Nothing is guessed.

    GLM-5.3 on api.z.ai: $1.40 in, $4.40 out, from docs.z.ai

  2. Re-checked daily

    Each of the 158 sources has a written fetch recipe. Scheduled agents will re-read every provider once a day, and a missed check will mark the row stale.

    Planned. Not running yet.

  3. Reviewed change

    When a page says something new, the agent will open a change. A person reviews it before the new price goes live.

    Planned, as pull requests on the open dataset.

  4. Checked stamp

    Each number shows the date it was last checked against its source, everywhere it appears.

    $1.40 per 1M input, checked 5 Oct 2026

For developers and agents

Today: the whole dataset as JSON and CSV on GitHub, and llms.txt. Planned: a free JSON API and an MCP server. The responses below are built from today's data.

Prices for one model (planned)

$ curl https://tokker.dev/api/v1/prices?model=claude-opus-5.5

{
  "model": "claude-opus-5.5",
  "unit": "usd_per_1m_tokens",
  "offers": [
    {
      "provider": "anthropic",
      "input": 4.0,
      "output": 20.0,
      "cache_read": 0.2,
      "batch_discount_pct": 50,
      "source_url": "https://platform.claude.com/docs/en/about-claude/pricing.md",
      "last_verified_at": "2026-10-05"
    },
    {
      "provider": "google-vertex",
      "input": 4.0,
      "output": 20.0,
      "cache_read": 0.2,
      "source_url": "https://cloud.google.com/vertex-ai/generative-ai/pricing",
      "last_verified_at": "2026-10-05"
    }
  ]
}

MCP server (planned)

{
  "mcpServers": {
    "tokker": { "url": "https://tokker.dev/mcp" }
  }
}

// tool call
cheapest_provider({ "model": "glm-5.3" })

// result
{
  "provider": "deepinfra",
  "input": 0.9,
  "output": 4.0,
  "source_url": "https://api.deepinfra.com/models/list",
  "last_verified_at": "2026-10-05"
}

Read the data schema

Price-drop alerts

Planned. Pick a model or plan, and get one email when it gets cheaper anywhere.

Placeholder. The waitlist isn't collecting emails yet. Until it opens, watch Tokker-dev/tokker on GitHub.