GPT-OSS 120B API Pricing — $0.15/M input, $0.60/M output
OpenAI flagship open-weight MoE model: ultra-fast agentic workflow.
Speed
Tokens / sec
780
TTFT
160 ms
Rank
#4 of 31
$ / M ÷ t/s
$0.0003
Measured with 5 runs on a fixed prompt — see the full methodology.
Cost at scale
| Tokens / month | Est. cost (blended 3:1) |
|---|---|
| 100,000 | $0.03 |
| 1,000,000 | $0.26 |
| 10,000,000 | $2.62 |
| 100,000,000 | $26.25 |
How GPT-OSS 120B compares
Llama 3.1 8B — $0.06/MGPT-OSS 20B — $0.13/MLlama 4 Maverick — $0.30/MLlama 3.3 70B — $0.64/MQwen 3.8 30B — $1.20/MGPT-4o Mini — $0.26/MGrok-3 Mini — $0.26/MMistral Small 3.1 — $0.26/M
Head-to-head comparisons
FAQ
Is GPT-OSS 120B cheaper than GPT-4o Mini?
GPT-OSS 120B costs $0.26/M blended tokens, GPT-4o Mini costs $0.26/M — GPT-4o Mini is cheaper.
How much does 1 million tokens cost with GPT-OSS 120B?
At a 3:1 input:output ratio, 1 million blended tokens costs approximately $0.26. Pure input costs $0.15/M; pure output costs $0.60/M.
Who provides GPT-OSS 120B?
GPT-OSS 120B is served by Groq. Pricing was last verified on 2026-04-06 against https://console.groq.com/docs/models.
