NVIDIA NIM logo

llama-3.3-nemotron-super-49b-v1.5 Free API on NVIDIA NIM

Free API Verified
★★★★★★★★★★ 3.5 Benchmark-backed score

Nemotron model for efficient reasoning, coding, and specialized AI agents

Free APIOpenAI compatibleReasoningTool callingJSON modeTextReasoning
API Details

Copy-ready connection details

Use these values in your client or SDK.
Base URL
https://integrate.api.nvidia.com/v1
Model ID
nvidia/llama-3.3-nemotron-super-49b-v1.5
API format OpenAI Chat Completions
Technical Details

llama-3.3-nemotron-super-49b-v1.5 specifications

Provider and model catalog
Context window 131K
Max output 16K
Status Online
Family nemotron
Released Jul 25, 2025
Last updated Jun 30, 2026
Free listing since Jul 25, 2025
Input text
Output text
Capabilities reasoning, tool calling, structured output, temperature control
Open weights Yes
AI Recommendation

Should you use llama-3.3-nemotron-super-49b-v1.5?

llama-3.3-nemotron-super-49b-v1.5 is listed for chat, reasoning workloads and supports a 131K context window.

Use it when NVIDIA NIM's free tier is enough for evaluation, demos, or light production traffic.

Best for
  • Chat
  • Reasoning
Strengths & Weaknesses

Strengths

  • Strong reasoning profile
  • Long context window
  • Reasoning mode listed
  • Tool calling support

Watch outs

  • Free-tier rate limits apply
  • Vision support is not listed
Benchmark Overview

Benchmark signals for llama-3.3-nemotron-super-49b-v1.5

Measured data
Intelligence General reasoning and instruction following
18.7/100
Coding Programming and code generation
15.1/100
Agentic Tool use and multi-step tasks
9.4/100
Context Maximum listed context window
131K
Pricing

llama-3.3-nemotron-super-49b-v1.5 pricing per 1M tokens

Free tier listed
Input $0 per 1M tokens
Output $0 per 1M tokens
Free access Available NVIDIA NIM
Rate limit Up to 40 RPM provider policy
Availability

llama-3.3-nemotron-super-49b-v1.5 availability by provider

Current provider only

We only found this llama-3-3-nemotron-super-49b-v1-5 listing on NVIDIA NIM in the current catalog. Use the related models section for nearby alternatives.

Provider Model listing Access Context API Limits
NVIDIA NIM nvidia/llama-3.3-nemotron-super-49b-v1.5 Free tier 131K OpenAI-style Up to 40 RPM
View NVIDIA NIM setup guide →
Typical Use Cases

llama-3.3-nemotron-super-49b-v1.5 use cases

Chat

llama-3.3-nemotron-super-49b-v1.5 is tagged for chat in this catalog and works with OpenAI-compatible client libraries.

Reasoning

llama-3.3-nemotron-super-49b-v1.5 is tagged for reasoning in this catalog and works with OpenAI-compatible client libraries.

FAQ

llama-3.3-nemotron-super-49b-v1.5 free API FAQ

Is llama-3.3-nemotron-super-49b-v1.5 free to use?

llama-3.3-nemotron-super-49b-v1.5 is listed with free API access on NVIDIA NIM, subject to the provider's quota and account policy.

What is the llama-3.3-nemotron-super-49b-v1.5 model ID?

The model ID shown in this catalog is nvidia/llama-3.3-nemotron-super-49b-v1.5.

What are the llama-3.3-nemotron-super-49b-v1.5 free tier rate limits on NVIDIA NIM?

The listed free tier limit is Up to 40 RPM. Limits can change per account tier, so confirm against the provider dashboard.

What context window does llama-3.3-nemotron-super-49b-v1.5 support?

The listed context window is 131K tokens with up to 16K output tokens.

More about llama-3.3-nemotron-super-49b-v1.5

NVIDIA Nemotron Super 49B is a mid-size custom model by NVIDIA, free on NVIDIA NIM with up to 40 RPM and no daily token cap. Built on Llama 3.3 architecture with NVIDIA's training enhancements, it offers balanced performance for general-purpose tasks. OpenAI-compatible API. Requires free NVIDIA Developer Program membership and phone verification.

For API keys, setup steps, and provider-level limits, see the NVIDIA NIM provider page.