Updated 2026 · 13 hosted models

LLM pricing, side by side.

List prices per 1M tokens for the major hosted LLMs in 2026. Use the calculator to see what they would cost for your workload.

ModelProviderTier$ / M input$ / M outputEst. 100k-call month
Claude Opus 4Anthropicflagship$15.00$75.00$15,0005k in / 1k out
GPT-5OpenAIreasoning$5.00$20.00$4,5005k in / 1k out
GPT-4.1OpenAIflagship$2.00$8.00$1,8005k in / 1k out
Gemini 2.5 ProGoogleflagship$1.25$10.00$1,6255k in / 1k out
Claude Sonnet 4Default in calculatorsAnthropicfast$3.00$15.00$3,0005k in / 1k out
Claude Haiku 4Anthropicmini$0.80$4.00$8005k in / 1k out
GPT-4o miniOpenAImini$0.15$0.60$1355k in / 1k out
Gemini 2.5 FlashGooglemini$0.30$2.50$4005k in / 1k out
Mistral Large 2Mistralfast$2.00$6.00$1,6005k in / 1k out
DeepSeek V3DeepSeekopen$0.27$1.10$2455k in / 1k out
Llama 3.1 70BMeta · Togetheropen$0.88$0.88$5285k in / 1k out
Llama 3.1 8BMeta · Togetheropen$0.18$0.18$1085k in / 1k out
Qwen 2.5 72BAlibaba · OpenRouteropen$0.40$0.40$2405k in / 1k out

Estimates assume 100,000 calls/month, 5,000 input tokens, 1,000 output tokens. Your actual bill depends on caching, batching, free tiers and any negotiated enterprise rates — verify with each provider.

Methodology

How to read this table.

Tier is a rough guide, not a guarantee

"Flagship" usually means the strongest model in a provider's line; "mini" usually means the cheapest. The tier label is a navigation aid, not a quality rating — a cheap model can still beat a flagship on your specific task.

Output tokens cost more than input

Almost every provider charges 2–5× more for output tokens. If your workload is generation-heavy (long agent loops, code synthesis, reports), the output column matters more than the input column.

List price is not your price

Enterprise contracts, committed-use discounts, prompt caching and batch APIs can change the math 2–10×. This page is for planning; for actual billing, check your invoice or your provider's pricing dashboard.

Price alone is not the answer

A model that costs 1/10 the price but only succeeds on 60% of your tasks is more expensive per successful task. Cost per successful task models this, and ModelSwitch measures it on your real workload.