Compare LLM pricing

Flagship models across every provider selected by default. Paste a prompt and compare cost and context usage side by side. Add more from the model search.

0tokens (est.)
0words
0characters
0chars, no spaces

Token count uses the exact GPT (o200k) tokenizer. Counted in your browser, never uploaded.

Models
Qwen3.8 MaxClaude Haiku 4.5 (latest)Claude Opus 4.8Claude Sonnet 4.6DeepSeek ChatGemini 2.5 FlashGemini 3.5 FlashMistral Medium (latest)Kimi K3GPT-4.1 miniGPT-4oGPT-5.5Grok 4.5GLM-5.2
ModelTokensPer requestMonthlyContext
Qwen3.8 Max0 est.$0.003$3.000.1%
Claude Haiku 4.5 (latest)0 est.$0.0025$2.500.3%
Claude Opus 4.80 est.$0.013$12.500.1%
Claude Sonnet 4.60 est.$0.0075$7.500.1%
DeepSeek Chatcheapest0 est.$0.00014$0.140.1%
Gemini 2.5 Flash0 est.$0.0013$1.250.0%
Gemini 3.5 Flash0 est.$0.0045$4.500.0%
Mistral Medium (latest)0 est.$0.0038$3.750.2%
Kimi K30 est.$0.0075$7.500.0%
GPT-4.1 mini0 exact$0.0008$0.80.0%
GPT-4o0 exact$0.005$5.000.4%
GPT-5.50 exact$0.015$15.000.0%
Grok 4.50 est.$0.003$3.000.1%
GLM-5.20 est.$0.0022$2.200.1%

Token counts vary by tokenizer and model; non-OpenAI counts are estimates. Cost is an estimate and provider pricing can change. See each model's source on the pricing page.

How to compare models fairly

A fair comparison fixes the workload and varies only the model. Paste one real prompt, set a realistic output length, and read cost per request and context usage together, since a cheaper model with a smaller context window may not fit your task at all. The cheapest option is highlighted for the exact inputs you enter.

What to look at beyond price

Price is only one dimension. Context window size determines whether your full prompt fits. Speed affects user experience in real-time applications. Quality determines whether you need one attempt or many. A model that costs 3x more but needs no retries and no prompt engineering can be cheaper in total engineering and compute cost than a cheap model that requires significant tuning.

FAQ

Is the cheapest model always best?

No. Balance price against context window, quality and speed. This tool shows cost and context side by side so you can weigh them for your use case.

How do GPT-4 and Claude compare on price?

Claude and GPT-4 class models are in a similar price tier, both significantly more expensive than budget alternatives. The exact difference depends on the specific model variant. Enter your token counts above to see a direct cost comparison for your workload.

Is DeepSeek a reliable cheaper alternative?

DeepSeek models are significantly cheaper and perform well on many benchmarks. The main considerations are data privacy (servers are based in China), occasional availability issues, and slightly lower performance on certain English-language edge cases. Many teams use DeepSeek for cost-sensitive tasks and a flagship model for quality-critical ones.

How often do LLM prices change?

Prices change frequently — sometimes several times per year — as providers compete and infrastructure costs fall. All prices in this calculator link to official sources with verification dates. Re-check before finalizing a budget for a new project.

Current model pricing

ModelProviderInputCachedOutputContextVerified
Qwen3.8 Max Alibaba $2 $0.25 $6 1 000k 2026-08-03
Claude Haiku 4.5 (latest) Anthropic $1 $0.1 $5 200k 2025-10-15
Claude Opus 4.8 Anthropic $5 $0.5 $25 1 000k 2026-05-28
Claude Sonnet 4.6 Anthropic $3 $0.3 $15 1 000k 2026-03-13
DeepSeek Chat DeepSeek $0.14 $0.0028 $0.28 1 000k 2026-02-28
Gemini 2.5 Flash Google $0.3 $0.03 $2.5 1 048,576k 2025-06-17
Gemini 3.5 Flash Google $1.5 $0.15 $9 1 048,576k 2026-05-19
Mistral Medium (latest) Mistral $1.5 n/a $7.5 262,144k 2026-04-29
Kimi K3 Moonshot AI $3 $0.3 $15 1 048,576k 2026-07-16
GPT-4.1 mini OpenAI $0.4 $0.1 $1.6 1 047,576k 2025-04-14
GPT-4o OpenAI $2.5 $1.25 $10 128k 2024-08-06
GPT-5.5 OpenAI $5 $0.5 $30 1 050k 2026-04-23
Grok 4.5 xAI $2 $0.3 $6 500k 2026-07-08
GLM-5.2 Zhipu AI $1.4 $0.26 $4.4 1 000k 2026-06-13

Prices per 1M tokens (USD). Each model links to its official source. Seed values pending re-verification. See the methodology.