Groq API Cost Calculator
How much does the Groq API cost per month?
Across Groq's 7 priced models at 200K calls/month, the cheapest is Llama 3.1 8B at $28.31 per month, verbosity-adjusted rather than list price. The most expensive, Qwen 3.8 30B, runs $498 per month for the identical call shape.
Verified 2026-08-08
Raw dataset: data.json. Data as of August 2026.
Ranked cost — 200K calls/month
| Model | Provider | List monthly | Effective monthly | Verbosity | Rank Δ |
|---|---|---|---|---|---|
| Llama 3.1 8Bbudgetlegacy | Groq | $29.60 | $14.16 | 0.77× | — |
| GPT-OSS 20Bbudget | Groq | $57.00 | $48.76 | 2.93× | — |
| Llama 4 Maverickbudgetlegacy | Groq | $138 | $69.00 | — | ▲1 |
| GPT-OSS 120Bbudget | Groq | $114 | $74.43 | 1.83× | ▼1 |
| Llama 3.3 70Bbudgetlegacy | Groq | $339 | $164 | 0.81× | — |
| Qwen 3.8 30Bmid | Groq | $498 | $249 | — | — |
| Qwen 3.6 27Bmidlegacy | Groq | $498 | $249 | — | — |
Levers live on Groq
The billing gotcha
Groq prices by token like every other provider here, but its selling point — inference speed — does not appear in a monthly-cost projection at all. A model that costs the same per token on Groq and elsewhere will show identical numbers below even though one of them returns the answer several times faster; that tradeoff has to be read on the provider page, not this calculator.
Verbosity movers within Groq's lineup
- GPT-OSS 120B is priced #3 by list rate but #4 once its 1.83× verbosity is billed — $114 list vs $149 effective.
