Should you use OpenAI: gpt-oss-20b?
OpenAI: gpt-oss-20b is listed for chat, coding workloads and supports a 131K context window.
Use the comparison and availability sections before choosing this listing, because it is not currently marked as free.
https://openrouter.ai/api/v1 openai/gpt-oss-20b OpenAI: gpt-oss-20b is listed for chat, coding workloads and supports a 131K context window.
Use the comparison and availability sections before choosing this listing, because it is not currently marked as free.
We only found this gpt-oss listing on OpenRouter in the current catalog. Use the related models section for nearby alternatives.
| Provider | Model listing | Access | Context | API | Limits |
|---|---|---|---|---|---|
| | OpenAI: gpt-oss-20b | Paid | 131K | OpenAI-style | 200 req/day (free tier) |
| Model | Provider | Context | Access |
|---|---|---|---|
| inclusionAI: Ling 3.0 Flash VL (free) | OpenRouter | 262K | Free tier |
| Nex AGI: Nex-N2.5-Mini (free) | OpenRouter | 262K | Free tier |
| Nex AGI: Nex-N2.5-Pro (free) | OpenRouter | 262K | Free tier |
| inclusionAI: Ling 3.0 Flash Sante (free) | OpenRouter | 262K | Free tier |
| inclusionAI: Ling 3.0 Flash Fin (free) | OpenRouter | 262K | Free tier |
OpenAI: gpt-oss-20b is tagged for chat in this catalog and works with OpenAI-compatible client libraries.
OpenAI: gpt-oss-20b is tagged for coding in this catalog and works with OpenAI-compatible client libraries.
OpenAI: gpt-oss-20b is not currently marked as free on OpenRouter.
The model ID shown in this catalog is openai/gpt-oss-20b.
The listed free tier limit is 200 req/day (free tier). Limits can change per account tier, so confirm against the provider dashboard.
The listed context window is 131K tokens with up to 33K output tokens.
GPT-OSS 20B is OpenAI's open-weight 21B-parameter Mixture-of-Experts model with 3.6B active per forward pass, available free on OpenRouter (also on Groq, Cerebras, and Cloudflare). Supports reasoning level configuration, fine-tuning, and agentic capabilities — function calling, tool use, and structured outputs. Trained in OpenAI's Harmony response format. Released under Apache 2.0. The low active parameter count enables lower-latency inference on consumer or single-GPU hardware. 131K context window. Text-only. OpenAI-compatible. Free tier: 200 RPD on OpenRouter.
For API keys, setup steps, and provider-level limits, see the OpenRouter provider page.