Groq API Cost Calculator

How much does the Groq API cost per month?

Across Groq's 7 priced models at 200K calls/month, the cheapest is Llama 3.1 8B at $28.31 per month, verbosity-adjusted rather than list price. The most expensive, Qwen 3.8 30B, runs $498 per month for the identical call shape.

Verified 2026-08-08

Raw dataset: data.json. Data as of August 2026.

Ranked cost — 200K calls/month

ModelProviderList monthlyEffective monthlyVerbosityRank Δ
Llama 3.1 8BbudgetlegacyGroq$29.60$14.160.77×
GPT-OSS 20BbudgetGroq$57.00$48.762.93×
Llama 4 MaverickbudgetlegacyGroq$138$69.001
GPT-OSS 120BbudgetGroq$114$74.431.83×1
Llama 3.3 70BbudgetlegacyGroq$339$1640.81×
Qwen 3.8 30BmidGroq$498$249
Qwen 3.6 27BmidlegacyGroq$498$249

Levers live on Groq

Batch API: up to 50%Model verbosity: up to 94%Context trimming: up to 47%

The billing gotcha

Groq prices by token like every other provider here, but its selling point — inference speed — does not appear in a monthly-cost projection at all. A model that costs the same per token on Groq and elsewhere will show identical numbers below even though one of them returns the answer several times faster; that tradeoff has to be read on the provider page, not this calculator.

Verbosity movers within Groq's lineup

  • GPT-OSS 120B is priced #3 by list rate but #4 once its 1.83× verbosity is billed — $114 list vs $149 effective.

Related

Groq provider profile →Compare against another provider →Global cost calculator →How to reduce LLM API costs →