LLM Cost Calculator — Real Monthly Cost, Not Just Rate
Raw dataset: data.json. Verbosity measured 2026-06-21. Cite this: All AI Ask LLM Verbosity Index and Effective Cost Dataset, retrieved 2026-06-21.
A $/M rate is not your bill. Set your call shape below and see every priced model ranked by effective monthly cost — adjusted for how many output tokens each model actually spends on a job of this size, not its list price alone.
| Model | Provider | List monthly | Effective monthly | Verbosity | Rank Δ |
|---|---|---|---|---|---|
| Amazon Nova Microbudget | Amazon | $26.60 | $24.25 | 0.76× | — |
| Llama 3.1 8Bbudgetlegacy | Groq | $29.60 | $28.31 | 0.77× | — |
| Amazon Nova Litebudget | Amazon | $45.60 | $44.26 | 0.92× | — |
| GPT-5 Nanobudgetlegacy | OpenAI | $52.00 | $52.00 | — | — |
| Gemini 2.5 Flash Litebudgetlegacy | $76.00 | $76.00 | — | ▲1 | |
| Ministral 8Bbudget | Mistral | $82.50 | $81.77 | 0.93× | ▲1 |
| GPT-OSS 20Bbudget | Groq | $57.00 | $97.53 | 2.93× | ▼2 |
| Mistral Small 3.1budget | Mistral | $114 | $108 | 0.85× | ▲4 |
| GPT-4o Minibudgetlegacy | OpenAI | $114 | $114 | — | — |
| Grok-3 Minibudgetlegacy | xAI | $114 | $114 | — | — |
| DeepSeek V4 Flashbudget | DeepSeek | $86.80 | $118 | 2.60× | ▼3 |
| Llama 4 Maverickbudgetlegacy | Groq | $138 | $138 | — | ▲1 |
| GPT-OSS 120Bbudget | Groq | $114 | $149 | 1.83× | ▼2 |
| GPT-5.4 Nanobudgetlegacy | OpenAI | $184 | $165 | 0.79× | ▲1 |
| GPT-5.6 Lunabudget | OpenAI | $180 | $180 | — | ▼1 |
Show all 63 models
| Codestralbudget | Mistral | $207 | $194 | 0.79× | — |
| Gemini 3.1 Flash Litebudgetlegacy | $225 | $212 | 0.88× | ▲2 | |
| Gemini 3.5 Flash Litebudget | $225 | $225 | — | — | |
| GPT-5 Minibudgetlegacy | OpenAI | $260 | $260 | — | ▲1 |
| GPT-OSS 120B (Cerebras)budget | Cerebras | $221 | $290 | 2.33× | ▼3 |
| Mistral Medium 3budget | Mistral | $332 | $310 | 0.84× | ▲2 |
| Gemini 2.5 Flashbudgetlegacy | $319 | $319 | — | — | |
| Llama 3.3 70Bbudgetlegacy | Groq | $339 | $328 | 0.81× | ▲1 |
| DeepSeek V4 Probudget | DeepSeek | $270 | $410 | 3.31× | ▼3 |
| GLM-5.1midlegacy | Z.ai | $442 | $442 | — | — |
| Qwen 3.8 30Bmid | Groq | $498 | $498 | — | — |
| Qwen 3.6 27Bmidlegacy | Groq | $498 | $498 | — | — |
| Qwen 3.7 Plusmid | Qwen | $524 | $524 | — | — |
| Amazon Nova Promid | Amazon | $608 | $608 | — | — |
| GPT-5.4 Minimidlegacy | OpenAI | $675 | $675 | — | — |
| Gemini 3.1 Flashmidlegacy | $675 | $675 | — | — | |
| Claude Haiku 4.5mid | Anthropic | $830 | $830 | — | ▲1 |
| o3-Minimidlegacy | OpenAI | $836 | $836 | — | ▲1 |
| Grok 4.3mid | xAI | $775 | $849 | 1.42× | ▼2 |
| Qwen 3.8 Maxmid | Qwen | $1,216 | $1,216 | — | ▲1 |
| Qwen 3.7 Maxmid | Qwen | $1,216 | $1,216 | — | ▲1 |
| Grok-3midlegacy | xAI | $1,240 | $1,240 | — | ▲1 |
| GPT-5midlegacy | OpenAI | $1,300 | $1,300 | — | ▲2 |
| Gemini 3.6 Flashmid | $1,350 | $1,350 | — | ▲2 | |
| Gemini 3.5 Flashmidlegacy | $1,350 | $1,350 | — | ▲2 | |
| Grok-4.20 Reasoningmid | xAI | $1,380 | $1,380 | — | ▲2 |
| Grok-4.20mid | xAI | $1,380 | $1,380 | — | ▲2 |
| Mistral Large 3mid | Mistral | $1,380 | $1,380 | — | ▲2 |
| Gemini 3.1 Promid | $1,800 | $1,489 | 0.63× | ▲4 | |
| GPT-4.1midlegacy | OpenAI | $1,520 | $1,520 | — | ▲1 |
| GPT-5.6 Terramid | OpenAI | $1,800 | $1,800 | — | ▲1 |
| GLM-5.2mid | Z.ai | $980 | $1,864 | 3.87× | ▼12 |
| GPT-4omidlegacy | OpenAI | $1,900 | $1,900 | — | ▲1 |
| GPT-5.4midlegacy | OpenAI | $2,250 | $2,250 | — | ▲1 |
| Claude Sonnet 4.6mid | Anthropic | $2,490 | $2,490 | — | ▲1 |
| Claude Sonnet 4.5midlegacy | Anthropic | $2,490 | $2,490 | — | ▲1 |
| Claude Sonnet 4midlegacy | Anthropic | $2,490 | $2,490 | — | ▲1 |
| GLM 4.7 (Cerebras)mid | Cerebras | $1,273 | $2,533 | 7.55× | ▼14 |
| Claude Opus 4.8mid | Anthropic | $4,150 | $4,080 | 0.96× | — |
| Claude Opus 4.7midlegacy | Anthropic | $4,150 | $4,150 | — | — |
| Claude Opus 4.6midlegacy | Anthropic | $4,150 | $4,150 | — | — |
| Claude Opus 4.5midlegacy | Anthropic | $4,150 | $4,150 | — | — |
| GPT-5.6 Solfrontier | OpenAI | $4,500 | $4,500 | — | — |
| GPT-4 Turbofrontierlegacy | OpenAI | $6,900 | $6,900 | — | — |
| Claude Fable 5frontier | Anthropic | $8,300 | $8,300 | — | — |
| Claude Opus 4.1frontierlegacy | Anthropic | $12,450 | $12,450 | — | — |
| Claude Opus 4frontierlegacy | Anthropic | $12,450 | $12,450 | — | — |
| GPT-5.4 Profrontierlegacy | OpenAI | $27,000 | $27,504 | 1.04× | — |
List price lied to you
These models move the most once verbosity is priced in — a model that talks more costs more, regardless of its list rate.
- GLM 4.7 (Cerebras) is priced #39 by list rate but #53 once its 7.55× verbosity is billed — $1,273 list vs $2,533 effective.
- GLM-5.2 is priced #35 by list rate but #47 once its 3.87× verbosity is billed — $980 list vs $1,864 effective.
- Mistral Small 3.1 is priced #12 by list rate but #8 once its 0.85× verbosity is billed — $114 list vs $108 effective.
- Gemini 3.1 Pro is priced #48 by list rate but #44 once its 0.63× verbosity is billed — $1,800 list vs $1,489 effective.
- DeepSeek V4 Flash is priced #8 by list rate but #11 once its 2.60× verbosity is billed — $86.80 list vs $118 effective.
- GPT-OSS 120B (Cerebras) is priced #17 by list rate but #20 once its 2.33× verbosity is billed — $221 list vs $290 effective.
The formula, published
outputTokensBilled = outputTokens × (verbosityIndex ?? 1) inputCost = inputPerM × inputTokens / 1e6 outputCost = outputPerM × outputTokensBilled / 1e6 effective = (inputCost + outputCost) × callsPerMonth with caching: inputCost × (1 − cacheableInputPct × 0.9) with batching: (inputCost + outputCost) × (1 − batchDiscountPct/100)
90% cache-read saving is a stated modelling assumption, not a per-provider sourced figure — providers publish cache-read discounts between 75% and 90% off. The free-form estimator above assumes 30% of input is cache-eligible; pick a workload preset below for a shape-specific assumption instead.
Coverage: 21 of 63 priced models have a measured verbosity index (2+ graded runs). The rest render with a verbosity of — and are shown at unadjusted list price.
