GPT-OSS 20B
Cost-sensitive agentic tasks that don’t need the full 120B model.
What are GPT-OSS 20B's specs and price?
GPT-OSS 20B, built by Groq, ships a 131K-token context window and a 33K-token max output, released 2025-08. It supports text input with a dedicated reasoning mode and costs $0.13 per million blended tokens, the 3rd-cheapest of 32 models we track.
Specs
| Context window | 131K tokens |
| Max output | 33K tokens |
| Modalities | text |
| Extended thinking | Yes |
| Released | 2025-08 |
| Knowledge cutoff | 2025-05 |
| Provider | Groq |
Verified 2026-08-08 — source.
Where it ranks
Strengths
- Compact open-weight model
- Very cheap on Groq
- Still supports reasoning mode
More about GPT-OSS 20B
FAQ
What is GPT-OSS 20B's context window?
GPT-OSS 20B has a 131K-token context window and a 33K-token max output — the 23rd-largest context of the 32 current models we track. Source: https://console.groq.com/docs/models, verified 2026-08-08.
Does GPT-OSS 20B support vision or audio input?
No — GPT-OSS 20B is text-only as of 2026-08-08.
Does GPT-OSS 20B have a reasoning or extended-thinking mode?
Yes — GPT-OSS 20B exposes a dedicated reasoning mode for multi-step problems.
When was GPT-OSS 20B released, and what is its knowledge cutoff?
GPT-OSS 20B was released 2025-08 with a knowledge cutoff of 2025-05.
How much does GPT-OSS 20B cost, and who provides it?
GPT-OSS 20B is served by Groq at $0.13/M blended tokens (3:1 input:output) — the 3rd-cheapest of 32 current models. Full pricing breakdown: /llm-api-pricing/gpt-oss-20b.
