Should you use GPT OSS 120B?
GPT OSS 120B is listed for chat workloads and supports a 131K context window.
Use it when Ollama Cloud's free tier is enough for evaluation, demos, or light production traffic.
https://api.ollama.com gpt-oss:120b GPT OSS 120B is listed for chat workloads and supports a 131K context window.
Use it when Ollama Cloud's free tier is enough for evaluation, demos, or light production traffic.
We only found this GPT OSS 120B listing on Ollama Cloud in the current catalog. Use the related models section for nearby alternatives.
| Provider | Model listing | Access | Context | API | Limits |
|---|---|---|---|---|---|
| | GPT OSS 120B | Free tier | 131K | Native | Varies |
| Model | Provider | Context | Access |
|---|---|---|---|
| minimax-m3 | Ollama Cloud | 1.0M | Free tier |
| gpt-oss:20b | Ollama Cloud | 131K | Free tier |
| nemotron-3-ultra | Ollama Cloud | 262K | Free tier |
| deepseek-v3.1:671b-cloud | Ollama Cloud | 128K | Check provider |
| qwen3-coder:480b-cloud | Ollama Cloud | 128K | Check provider |
GPT OSS 120B is tagged for chat in this catalog.
GPT OSS 120B is listed with free API access on Ollama Cloud, subject to the provider's quota and account policy.
The model ID shown in this catalog is gpt-oss:120b.
The listed context window is 131K tokens with up to 33K output tokens.
Free GPT OSS 120B API.
For API keys, setup steps, and provider-level limits, see the Ollama Cloud provider page.