Should you use GPT OSS 120B?
GPT OSS 120B is listed for chat workloads and supports a 131K context window.
Use it when NVIDIA NIM's free tier is enough for evaluation, demos, or light production traffic.
https://integrate.api.nvidia.com/v1 openai/gpt-oss-120b GPT OSS 120B is listed for chat workloads and supports a 131K context window.
Use it when NVIDIA NIM's free tier is enough for evaluation, demos, or light production traffic.
We only found this GPT OSS 120B listing on NVIDIA NIM in the current catalog. Use the related models section for nearby alternatives.
| Provider | Model listing | Access | Context | API | Limits |
|---|---|---|---|---|---|
| | GPT OSS 120B | Free tier | 131K | Native | Varies |
| Model | Provider | Context | Access |
|---|---|---|---|
| 01-ai/yi-large | NVIDIA NIM | 131K | Free tier |
| adept/fuyu-8b | NVIDIA NIM | 131K | Free tier |
| ai21labs/jamba-1.5-large-instruct | NVIDIA NIM | 131K | Free tier |
| aisingapore/sea-lion-7b-instruct | NVIDIA NIM | 131K | Free tier |
| baai/bge-m3 | NVIDIA NIM | 131K | Free tier |
GPT OSS 120B is tagged for chat in this catalog and works with OpenAI-compatible client libraries.
GPT OSS 120B is listed with free API access on NVIDIA NIM, subject to the provider's quota and account policy.
The model ID shown in this catalog is openai/gpt-oss-120b.
The listed context window is 131K tokens with up to 33K output tokens.
GPT-OSS 120B is OpenAI's open-weight 117B-parameter Mixture-of-Experts model with 5.1B active parameters per forward pass, available free on OpenRouter (also on Groq, Cerebras, and Cloudflare). Supports configurable reasoning depth, full chain-of-thought access, and native tool use — function calling, browsing, and structured outputs. Optimized to run on a single H100 GPU with native MXFP4 quantization. Built for reasoning-heavy and agentic tasks. 131K context window. Text-only. OpenAI-compatible. Free tier: 200 RPD on OpenRouter.
For API keys, setup steps, and provider-level limits, see the NVIDIA NIM provider page.