Should you use nemotron-3-super-120b-a12b?
nemotron-3-super-120b-a12b is listed for chat, reasoning workloads and supports a 262K context window.
Use it when Kilo Code's free tier is enough for evaluation, demos, or light production traffic.
https://api.kilo.ai/api/gateway nvidia/nemotron-3-super-120b-a12b:free nemotron-3-super-120b-a12b is listed for chat, reasoning workloads and supports a 262K context window.
Use it when Kilo Code's free tier is enough for evaluation, demos, or light production traffic.
We found 3 provider listings for nemotron-3-super-120b-a12b. Check model ID, quota, pricing, and API format before switching providers.
| Provider | Model listing | Access | Context | API | Limits |
|---|---|---|---|---|---|
| | nvidia/nemotron-3-super-120b-a12b:free | Free tier | 262K | OpenAI-style | ~200 req/hr |
| | NVIDIA: Nemotron 3 Super (free) | Free tier | 262K | OpenAI-style | 200 req/day (free tier) |
| | Nemotron 3 Super 120B A12B | Free tier | 262K | Native | Varies |
| Model | Provider | Context | Access |
|---|---|---|---|
| NVIDIA: Nemotron 3 Super (free) | OpenRouter | 262K | Free tier |
| Nemotron 3 Super 120B A12B | NVIDIA NIM | 262K | Free tier |
| nvidia/nemotron-3-ultra-550b-a55b:free | Kilo Code | 1.0M | Free tier |
| stepfun/step-3.7-flash:free | Kilo Code | 262K | Free tier |
| nvidia/nemotron-3-nano-omni-30b-a3b-reasoning:free | Kilo Code | 256K | Free tier |
nemotron-3-super-120b-a12b is tagged for chat in this catalog and works with OpenAI-compatible client libraries.
nemotron-3-super-120b-a12b is tagged for reasoning in this catalog and works with OpenAI-compatible client libraries.
nemotron-3-super-120b-a12b is listed with free API access on Kilo Code, subject to the provider's quota and account policy.
The model ID shown in this catalog is nvidia/nemotron-3-super-120b-a12b:free.
The listed free tier limit is ~200 req/hr. Limits can change per account tier, so confirm against the provider dashboard.
The listed context window is 262K tokens with up to 262K output tokens.
NVIDIA Nemotron 3 Super (120B MoE with 12 active experts) is NVIDIA's flagship reasoning model, available free through Kilo Code's OpenAI-compatible API gateway. With 262K context and strong reasoning performance, it is well-suited for complex analytical tasks, multi-step problem-solving, and technical writing. The 32K output cap provides room for detailed responses. As a gateway proxy through Kilo Code, rate limits depend on upstream availability rather than a published quota.
For API keys, setup steps, and provider-level limits, see the Kilo Code provider page.