Kilo Code logo

nemotron-3-super-120b-a12b Free API on Kilo Code

Free API
★★★★★★★★★★ 3.5 Benchmark-backed score

Nemotron middle tier for collaborative agents and high-volume reasoning workloads

Free APIOpenAI compatibleReasoningTool callingJSON modeTextReasoning
API Details

Copy-ready connection details

Use these values in your client or SDK.
Base URL
https://api.kilo.ai/api/gateway
Model ID
nvidia/nemotron-3-super-120b-a12b:free
API format OpenAI-style
Technical Details

nemotron-3-super-120b-a12b specifications

Provider and model catalog
Context window 262K
Max output 262K
Status Online
Family nemotron
Released Mar 11, 2026
Last updated Aug 6, 2026
Free listing since Mar 11, 2026
Input text
Output text
Capabilities reasoning, tool calling, structured output, temperature control
Open weights Yes
AI Recommendation

Should you use nemotron-3-super-120b-a12b?

nemotron-3-super-120b-a12b is listed for chat, reasoning workloads and supports a 262K context window.

Use it when Kilo Code's free tier is enough for evaluation, demos, or light production traffic.

Best for
  • Chat
  • Reasoning
Strengths & Weaknesses

Strengths

  • Strong reasoning profile
  • Long context window
  • Reasoning mode listed
  • Tool calling support

Watch outs

  • Free-tier rate limits apply
  • Vision support is not listed
Benchmark Overview

Benchmark signals for nemotron-3-super-120b-a12b

Measured data
Intelligence General reasoning and instruction following
25.4/100
Coding Programming and code generation
37.7/100
Agentic Tool use and multi-step tasks
8.7/100
Context Maximum listed context window
262K
Pricing

nemotron-3-super-120b-a12b pricing per 1M tokens

Free tier listed
Input $0 per 1M tokens
Output $0 per 1M tokens
Free access Available Kilo Code
Rate limit ~200 req/hr provider policy
Availability

nemotron-3-super-120b-a12b availability by provider

2 alternatives

We found 3 provider listings for nemotron-3-super-120b-a12b. Check model ID, quota, pricing, and API format before switching providers.

Provider Model listing Access Context API Limits
Kilo Code nvidia/nemotron-3-super-120b-a12b:free Free tier 262K OpenAI-style ~200 req/hr
OpenRouter NVIDIA: Nemotron 3 Super (free) Free tier 262K OpenAI-style 200 req/day (free tier)
NVIDIA NIM Nemotron 3 Super 120B A12B Free tier 262K Native Varies
View Kilo Code setup guide →
Typical Use Cases

nemotron-3-super-120b-a12b use cases

Chat

nemotron-3-super-120b-a12b is tagged for chat in this catalog and works with OpenAI-compatible client libraries.

Reasoning

nemotron-3-super-120b-a12b is tagged for reasoning in this catalog and works with OpenAI-compatible client libraries.

FAQ

nemotron-3-super-120b-a12b free API FAQ

Is nemotron-3-super-120b-a12b free to use?

nemotron-3-super-120b-a12b is listed with free API access on Kilo Code, subject to the provider's quota and account policy.

What is the nemotron-3-super-120b-a12b model ID?

The model ID shown in this catalog is nvidia/nemotron-3-super-120b-a12b:free.

What are the nemotron-3-super-120b-a12b free tier rate limits on Kilo Code?

The listed free tier limit is ~200 req/hr. Limits can change per account tier, so confirm against the provider dashboard.

What context window does nemotron-3-super-120b-a12b support?

The listed context window is 262K tokens with up to 262K output tokens.

More about nemotron-3-super-120b-a12b

NVIDIA Nemotron 3 Super (120B MoE with 12 active experts) is NVIDIA's flagship reasoning model, available free through Kilo Code's OpenAI-compatible API gateway. With 262K context and strong reasoning performance, it is well-suited for complex analytical tasks, multi-step problem-solving, and technical writing. The 32K output cap provides room for detailed responses. As a gateway proxy through Kilo Code, rate limits depend on upstream availability rather than a published quota.

For API keys, setup steps, and provider-level limits, see the Kilo Code provider page.