Should you use GPT OSS 20B?
GPT OSS 20B is listed for chat workloads and supports a 131K context window.
Use the comparison and availability sections before choosing this listing, because it is not currently marked as free.
https://api.groq.com/openai/v1 openai/gpt-oss-20b GPT OSS 20B is listed for chat workloads and supports a 131K context window.
Use the comparison and availability sections before choosing this listing, because it is not currently marked as free.
We only found this GPT OSS 20B listing on Groq in the current catalog. Use the related models section for nearby alternatives.
| Provider | Model listing | Access | Context | API | Limits |
|---|---|---|---|---|---|
| | GPT OSS 20B | Paid | 131K | OpenAI-style | Varies |
| Model | Provider | Context | Access |
|---|---|---|---|
| Moonshot Kimi K2 | Groq | 131K | Free tier |
| Moonshot Kimi K2 0905 | Groq | 131K | Free tier |
| groq/compound | Groq | 131K | Free tier |
| groq/compound-mini | Groq | 131K | Free tier |
| llama-4-scout-17b-16e-instruct | Groq | 131K | Check provider |
GPT OSS 20B is tagged for chat in this catalog and works with OpenAI-compatible client libraries.
GPT OSS 20B is not currently marked as free on Groq.
The model ID shown in this catalog is openai/gpt-oss-20b.
The listed context window is 131K tokens with up to 66K output tokens.
GPT-OSS 20B is OpenAI's open-weight 21B-parameter Mixture-of-Experts model with 3.6B active per forward pass, available free on OpenRouter (also on Groq, Cerebras, and Cloudflare). Supports reasoning level configuration, fine-tuning, and agentic capabilities — function calling, tool use, and structured outputs. Trained in OpenAI's Harmony response format. Released under Apache 2.0. The low active parameter count enables lower-latency inference on consumer or single-GPU hardware. 131K context window. Text-only. OpenAI-compatible. Free tier: 200 RPD on OpenRouter.
For API keys, setup steps, and provider-level limits, see the Groq provider page.