NVIDIA NIM logo

Nemotron 3 Super 120B A12B Free API on NVIDIA NIM

Free API Verified
★★★★★★★★★★ 3.5 Benchmark-backed score

Nemotron middle tier for collaborative agents and high-volume reasoning workloads

Free APIOpenAI compatibleReasoningTool callingJSON modeReasoning
API Details

Copy-ready connection details

Use these values in your client or SDK.
Base URL
https://integrate.api.nvidia.com/v1
Model ID
nvidia/nemotron-3-super-120b-a12b
API format OpenAI Chat Completions
Technical Details

Nemotron 3 Super 120B A12B specifications

Provider and model catalog
Context window 262K
Max output 262K
Status Online
Family nemotron
Released Mar 11, 2026
Last updated Mar 11, 2026
Free listing since Mar 11, 2026
Input text
Output text
Capabilities reasoning, tool calling, structured output, temperature control
Open weights Yes
AI Recommendation

Should you use Nemotron 3 Super 120B A12B?

Nemotron 3 Super 120B A12B is listed for chat workloads and supports a 262K context window.

Use it when NVIDIA NIM's free tier is enough for evaluation, demos, or light production traffic.

Best for
  • Chat
Strengths & Weaknesses

Strengths

  • Long context window
  • Reasoning mode listed
  • Tool calling support
  • Structured JSON output

Watch outs

  • Vision support is not listed
Benchmark Overview

Benchmark signals for Nemotron 3 Super 120B A12B

Measured data
Intelligence General reasoning and instruction following
25.4/100
Coding Programming and code generation
37.7/100
Agentic Tool use and multi-step tasks
8.7/100
Context Maximum listed context window
262K
Pricing

Nemotron 3 Super 120B A12B pricing per 1M tokens

Free tier listed
Input $0.2 per 1M tokens
Output $0.8 per 1M tokens
Free access Available NVIDIA NIM
Availability

Nemotron 3 Super 120B A12B availability by provider

2 alternatives

We found 3 provider listings for nemotron-3-super-120b-a12b. Check model ID, quota, pricing, and API format before switching providers.

Provider Model listing Access Context API Limits
NVIDIA NIM Nemotron 3 Super 120B A12B Free tier 262K Native Varies
OpenRouter NVIDIA: Nemotron 3 Super (free) Free tier 262K OpenAI-style 200 req/day (free tier)
Kilo Code nvidia/nemotron-3-super-120b-a12b:free Free tier 262K OpenAI-style ~200 req/hr
View NVIDIA NIM setup guide →
Typical Use Cases

Nemotron 3 Super 120B A12B use cases

Chat

Nemotron 3 Super 120B A12B is tagged for chat in this catalog and works with OpenAI-compatible client libraries.

FAQ

Nemotron 3 Super 120B A12B free API FAQ

Is Nemotron 3 Super 120B A12B free to use?

Nemotron 3 Super 120B A12B is listed with free API access on NVIDIA NIM, subject to the provider's quota and account policy.

What is the Nemotron 3 Super 120B A12B model ID?

The model ID shown in this catalog is nvidia/nemotron-3-super-120b-a12b.

What context window does Nemotron 3 Super 120B A12B support?

The listed context window is 262K tokens with up to 262K output tokens.

More about Nemotron 3 Super 120B A12B

NVIDIA Nemotron 3 Super 120B A12B is an open hybrid MoE model with 120B total parameters (12B active), available free on OpenRouter. Built on a hybrid Mamba-Transformer Mixture-of-Experts architecture with multi-token prediction (MTP), it delivers over 50% higher token generation speed compared to leading open models. Designed for multi-agent applications — long-term agent coherence, cross-document reasoning, and multi-step task planning. Trained with multi-environment RL across 10+ environments. Latent MoE calls 4 experts for the cost of one. Up to 1M context window. Fully open: weights, datasets, and recipes. OpenAI-compatible. Free tier: 200 RPD.

For API keys, setup steps, and provider-level limits, see the NVIDIA NIM provider page.