LLM Pricing Calculator

Token prices are hard to compare across providers — enter your usage once and see the real monthly bill for every major model.

Live data · updated 2026-09-15 18:18 · sources: LMArena + OpenRouter

Our verdict

For a typical indie workload (50M in / 10M out per month), the spread between the cheapest and most expensive frontier model is over 40× — most apps can run on muse-spark-1.2 (xHigh) for single-digit dollars.

#ModelArena$/1M in$/1M outMonthly cost
1muse-spark-1.2 (xHigh) Meta1500$1.25$4.25$105.00
2qwen3.8-max Alibaba1302$2$6$160.00
3gemini-omni-1.1-flash Google1515$1.5$9$165.00
4gpt-5.6-sol-xhigh OpenAI1257$4$20$400.00
5claude-opus-5-high Anthropic1516$5$25$500.00
6claude-opus-4-6-high Anthropic1505$5$25$500.00
7claude-opus-4-6-search Anthropic1253$5$25$500.00
8claude-opus-4-7-high Anthropic1502$5$25$500.00
9claude-opus-5-max Anthropic1687$5$25$500.00
10claude-opus-4-6 Anthropic1507$5$25$500.00
11gpt-5.5-search OpenAI1242$5$30$550.00
12gpt-image-2 (medium) OpenAI1381$8$30$700.00
13claude-fable-5 Anthropic1506$10$50$1,000
14gpt-6-astra-max OpenAI1800$10$50$1,000
15claude-fable-5.1-max Anthropic1758$10$50$1,000

50M input + 10M output tokens per month. Cheapest: muse-spark-1.2 (xHigh) at $105.00. Prices from OpenRouter, arena ratings from LMArena.

How to estimate your token usage

One English word ≈ 1.3 tokens; a typical chat turn (your prompt + history) runs 500–2,000 input tokens. RAG pipelines multiply input fast — every retrieval re-sends context. Measure one week of real traffic before committing to a model tier.

Input vs output pricing matters

As of 2026-09-15 18:18, output tokens cost 3–5× more than input on most models. Chat-heavy products should weight output price; summarization and RAG products should weight input. Batch APIs (off-peak) cut costs ~50% on several providers if latency is flexible.

Don't optimize price before quality

A model 20% cheaper that fails 5% more often costs more once you count retries and bad user experiences. Use the arena rating column as a quality floor, then optimize price above it.

Ready to build? Access all of these models through one API key. Try OpenRouter →