AI model pricing

Compare verified input and output token prices across leading AI providers.

37 models priced 15 providers Verified 2026-09-08 USD / 1M tokens How we verify pricing →

Lowest input: $0.035 (Amazon Nova Micro) Lowest output: $0.14 (Amazon Nova Micro) Largest context: 1050K (GPT-5.4 / GPT-5.5)

37 pricing records

ModelProviderInputUSD / 1M tokensOutputUSD / 1M tokensContext
Amazon Nova MicroAmazon$0.035$0.14128K
Command R7BCohere$0.0375$0.15128K
Amazon Nova LiteAmazon$0.06$0.24300K
Granite 4H SmallIBM$0.0636$0.265131K
Muse Spark 1.3 (Contributor)Meta$0.1$0.21M
Mistral Small 4Mistral$0.15$0.6256K
Qwen3.8 FlashQwen$0.15$0.47992K
Command RCohere$0.15$0.6128K
Jamba MiniAI21$0.2$0.4256K
GLM-4.5-AirZ.ai$0.2$1.1128K
DeepSeek V4-FlashDeepSeek$0.22$0.661M
Gemini 2.5 FlashGoogle$0.3$2.51.05M
MiniMax-M3MiniMax$0.3$1.21M
Amazon Nova 2 LiteAmazon$0.33$2.751M
Qwen3.7 PlusQwen$0.4$1.6992K
Mistral Large 3Mistral$0.5$1.5256K
GLM-4.6Z.ai$0.6$2.2200K
DeepSeek V4-ProDeepSeek$0.66$1.981M
GPT-5.4 MiniOpenAI$0.75$4.5400K
Amazon Nova ProAmazon$0.8$3.2300K
Kimi K2.6Kimi$0.95$4262K
Claude Haiku 4.5Anthropic$1$5200K
Grok Build 0.1xAI$1$2256K
Gemini 2.5 ProGoogle$1.25$101.05M
Grok 4.3xAI$1.25$2.51M
Muse Spark 1.3Meta$1.25$4.251M
GLM-5.3Z.ai$1.4$4.41M
Claude Sonnet 5Anthropic$2$101M
Qwen3.8 MaxQwen$2$6992K
Grok 4.6xAI$2$6500K
Jamba LargeAI21$2$8256K
GPT-5.4OpenAI$2.5$151.05M
Amazon Nova PremierAmazon$2.5$12.51M
Kimi K3Kimi$3$151.05M
Claude Opus 5Anthropic$5$251M
GPT-5.5OpenAI$5$301.05M
Claude Fable 5.1Anthropic$10$501M
Amazon Nova MicroAmazon
Input$0.035
Output$0.14
Context: 128K
Command R7BCohere
Input$0.0375
Output$0.15
Context: 128K
Amazon Nova LiteAmazon
Input$0.06
Output$0.24
Context: 300K
Granite 4H SmallIBM
Input$0.0636
Output$0.265
Context: 131K
Muse Spark 1.3 (Contributor)Meta
Input$0.1
Output$0.2
Context: 1M
Mistral Small 4Mistral
Input$0.15
Output$0.6
Context: 256K
Qwen3.8 FlashQwen
Input$0.15
Output$0.47
Context: 992K
Command RCohere
Input$0.15
Output$0.6
Context: 128K
Jamba MiniAI21
Input$0.2
Output$0.4
Context: 256K
GLM-4.5-AirZ.ai
Input$0.2
Output$1.1
Context: 128K
DeepSeek V4-FlashDeepSeek
Input$0.22
Output$0.66
Context: 1M
Gemini 2.5 FlashGoogle
Input$0.3
Output$2.5
Context: 1.05M
MiniMax-M3MiniMax
Input$0.3
Output$1.2
Context: 1M
Amazon Nova 2 LiteAmazon
Input$0.33
Output$2.75
Context: 1M
Qwen3.7 PlusQwen
Input$0.4
Output$1.6
Context: 992K
Mistral Large 3Mistral
Input$0.5
Output$1.5
Context: 256K
GLM-4.6Z.ai
Input$0.6
Output$2.2
Context: 200K
DeepSeek V4-ProDeepSeek
Input$0.66
Output$1.98
Context: 1M
GPT-5.4 MiniOpenAI
Input$0.75
Output$4.5
Context: 400K
Amazon Nova ProAmazon
Input$0.8
Output$3.2
Context: 300K
Kimi K2.6Kimi
Input$0.95
Output$4
Context: 262K
Claude Haiku 4.5Anthropic
Input$1
Output$5
Context: 200K
Grok Build 0.1xAI
Input$1
Output$2
Context: 256K
Gemini 2.5 ProGoogle
Input$1.25
Output$10
Context: 1.05M
Grok 4.3xAI
Input$1.25
Output$2.5
Context: 1M
Muse Spark 1.3Meta
Input$1.25
Output$4.25
Context: 1M
GLM-5.3Z.ai
Input$1.4
Output$4.4
Context: 1M
Claude Sonnet 5Anthropic
Input$2
Output$10
Context: 1M
Qwen3.8 MaxQwen
Input$2
Output$6
Context: 992K
Grok 4.6xAI
Input$2
Output$6
Context: 500K
Jamba LargeAI21
Input$2
Output$8
Context: 256K
GPT-5.4OpenAI
Input$2.5
Output$15
Context: 1.05M
Amazon Nova PremierAmazon
Input$2.5
Output$12.5
Context: 1M
Kimi K3Kimi
Input$3
Output$15
Context: 1.05M
Claude Opus 5Anthropic
Input$5
Output$25
Context: 1M
GPT-5.5OpenAI
Input$5
Output$30
Context: 1.05M
Claude Fable 5.1Anthropic
Input$10
Output$50
Context: 1M
  • ⚠ has an important pricing note (hover or see the model row)
  • † has a long-context surcharge above a token threshold (hover for detail)

Pricing by provider

OpenAI

OpenAI prices its GPT-5 models per million input and output tokens. Some models apply a higher rate once a single request's input tokens exceed a documented threshold.

  • GPT-5.4 Mini: $0.75 / 1M input tokens, $4.50 / 1M output tokens
    Last verified 2026-08-06 · Official pricing source
  • GPT-5.5: $5.00 / 1M input tokens, $30.00 / 1M output tokens — above 272K input tokens per request: $10.00 / $45.00
    Last verified 2026-08-06 · Official pricing source

Try the chatbot calculator →

Anthropic / Claude

Anthropic prices Claude models per million input and output tokens. Claude Sonnet 5's pricing was re-verified directly against Anthropic's official documentation on 2026-09-06.

  • Claude Haiku 4.5: $1.00 / 1M input tokens, $5.00 / 1M output tokens
    Last verified 2026-08-06 · Official pricing source
  • Claude Sonnet 5: $2.00 / 1M input tokens, $10.00 / 1M output tokens
    Last verified 2026-09-06 · Official pricing source
  • Claude Opus 5: $5.00 / 1M input tokens, $25.00 / 1M output tokens
    Last verified 2026-08-06 · Official pricing source

Try the document Q&A calculator →

Google / Gemini

Google prices Gemini models per million input and output tokens. Gemini 2.5 Pro applies a higher rate once a single request's input tokens exceed a documented threshold; Gemini 2.5 Flash does not.

  • Gemini 2.5 Flash: $0.30 / 1M input tokens, $2.50 / 1M output tokens
    Last verified 2026-09-09 · Official pricing source
  • Gemini 2.5 Pro: $1.25 / 1M input tokens, $10.00 / 1M output tokens — above 200K input tokens per request: $2.50 / $15.00
    Last verified 2026-09-09 · Official pricing source

Compare model tiers →

Mistral

Mistral prices its models per million input and output tokens. Mistral Large 3 is priced higher than Mistral Small 4 in T9 Token Budget's current pricing data.

  • Mistral Small 4: $0.15 / 1M input tokens, $0.60 / 1M output tokens
    Last verified 2026-08-06 · Official pricing source
  • Mistral Large 3: $0.50 / 1M input tokens, $1.50 / 1M output tokens
    Last verified 2026-08-06 · Official pricing source

Try the document summary calculator →

DeepSeek

DeepSeek prices its V4 models per million input and output tokens. Provider pricing can change, so verify current rates with DeepSeek before budgeting.

  • DeepSeek V4-Flash: $0.22 / 1M input tokens, $0.66 / 1M output tokens
    Last verified 2026-09-08 · Official pricing source
  • DeepSeek V4-Pro: $0.66 / 1M input tokens, $1.98 / 1M output tokens
    Last verified 2026-09-08 · Official pricing source

Try the meeting summary calculator →

Kimi

Kimi (Moonshot AI) prices its models per million input and output tokens on its international, USD-denominated API. Its mainland-China platform publishes an independent CNY price list, not shown here.

Try the chatbot calculator →

Qwen

Qwen (Alibaba Cloud) prices its models per million input and output tokens on its international API. Qwen3.7 Plus applies a higher rate once a single request's input tokens exceed a documented threshold.

  • Qwen3.8 Flash: $0.15 / 1M input tokens, $0.47 / 1M output tokens
    Last verified 2026-09-07 · Official pricing source
  • Qwen3.7 Plus: $0.40 / 1M input tokens, $1.60 / 1M output tokens — above 256K input tokens per request: $1.20 / $4.80
    Last verified 2026-09-07 · Official pricing source
  • Qwen3.8 Max: $2.00 / 1M input tokens, $6.00 / 1M output tokens
    Last verified 2026-09-09 · Official pricing source

Try the document summary calculator →

Z.ai

Z.ai prices its GLM models per million input and output tokens on its international, USD-denominated API. GLM-5.3 is its current flagship; GLM-4.6 and GLM-4.5-Air remain separately available at their own rates.

Try the document Q&A calculator →

MiniMax

MiniMax prices its models per million input and output tokens on its international, USD-denominated API. MiniMax-M3's listed rate reflects the provider's own stated permanent discount off its list price.

  • MiniMax-M3: $0.30 / 1M input tokens, $1.20 / 1M output tokens — above 512K input tokens per request: $0.60 / $2.40
    Last verified 2026-09-07 · Official pricing source

Try the meeting summary calculator →

xAI

xAI prices its Grok models per million input and output tokens. Grok 4.3 and Grok 4.6 apply a higher rate once a single request's input tokens reach a documented threshold.

  • Grok Build 0.1: $1.00 / 1M input tokens, $2.00 / 1M output tokens — above 200K input tokens per request: $2.00 / $4.00
    Last verified 2026-09-07 · Official pricing source
  • Grok 4.3: $1.25 / 1M input tokens, $2.50 / 1M output tokens — above 200K input tokens per request: $2.50 / $5.00
    Last verified 2026-09-07 · Official pricing source
  • Grok 4.6: $2.00 / 1M input tokens, $6.00 / 1M output tokens — above 200K input tokens per request: $4.00 / $12.00
    Last verified 2026-09-07 · Official pricing source

Try the chatbot calculator →

Cohere

Cohere prices its Command R models per million input and output tokens. Its newest flagship, Command A+, is distributed as free open-weights rather than a metered API and is not included in this pricing data.

  • Command R7B: $0.0375 / 1M input tokens, $0.15 / 1M output tokens
    Last verified 2026-09-07 · Official pricing source
  • Command R: $0.15 / 1M input tokens, $0.60 / 1M output tokens
    Last verified 2026-09-07 · Official pricing source

Compare model tiers →

How to read AI pricing

Input tokens

What you send to the model — your prompt, conversation history, or document text. Priced separately from output, usually at a lower rate.

Output tokens

What the model generates in response. Output is commonly the more expensive side — often 3–10× the input price — because generating text costs more compute than reading it.

Context window

How much text a model can hold in one request, in tokens. A larger window lets you include more history or documents, but filling it means paying for more input on every call.

Whose prices are these?

The prices in the table above belong to the AI providers (OpenAI, Anthropic, Google, Mistral, and DeepSeek — every provider currently priced in the calculator), not to T9. T9 does not sell, resell, or provide access to any model — it only uses published prices to produce an estimate.

Prices change — always re-verify

AI providers update prices frequently. The figures above were manually verified on 2026-09-08, but pricing should always be rechecked on the provider's official pricing page before any budget decision. T9 uses pricing data only for estimation, and never as a guarantee of what you will be charged.

See pricing in context

Beyond comparing raw prices, T9 lets you apply these current rates to your own use case. Open the chatbot, meeting summary, document summary, or document Q&A calculator and switch models to see how the same workload changes cost, or read the models overview and the methodology.