OpenRouter logo

NVIDIA: Nemotron 3 Nano Omni (free) Free API on OpenRouter

Free API Verified
★★★★★★★★★★ 3.5 Benchmark-backed score

Open Nemotron omni model combining reasoning with text, vision, and audio

Free APIOpenAI compatibleReasoningTool callingTextImageAudio
API Details

Copy-ready connection details

Use these values in your client or SDK.
Base URL
https://openrouter.ai/api/v1
Model ID
nvidia/nemotron-3-nano-omni-30b-a3b-reasoning:free
API format OpenAI Chat Completions + OpenAI Responses
Technical Details

NVIDIA: Nemotron 3 Nano Omni (free) specifications

Provider and model catalog
Context window 256K
Max output 66K
Status Online
Family nemotron
Released Apr 28, 2026
Last updated Aug 6, 2026
Free listing since Apr 28, 2026
Input text, image, video, audio
Output text
Capabilities reasoning, tool calling, file attachments, temperature control
Open weights Yes
AI Recommendation

Should you use NVIDIA: Nemotron 3 Nano Omni (free)?

NVIDIA: Nemotron 3 Nano Omni (free) is listed for chat, vision, reasoning workloads and supports a 256K context window.

Use it when OpenRouter's free tier is enough for evaluation, demos, or light production traffic.

Best for
  • Chat
  • Vision
  • Reasoning
Strengths & Weaknesses

Strengths

  • Strong reasoning profile
  • Long context window
  • Reasoning mode listed
  • Tool calling support

Watch outs

  • Free-tier rate limits apply
Benchmark Overview

Benchmark signals for NVIDIA: Nemotron 3 Nano Omni (free)

Measured data
Intelligence General reasoning and instruction following
21.4/100
Coding Programming and code generation
13.8/100
Agentic Tool use and multi-step tasks
23.9/100
Speed Observed generation speed
67 tok/s
Context Maximum listed context window
256K
Pricing

NVIDIA: Nemotron 3 Nano Omni (free) pricing per 1M tokens

Free tier listed
Input $0 per 1M tokens
Output $0 per 1M tokens
Free access Available OpenRouter
Rate limit 200 req/day (free tier) provider policy
Availability

NVIDIA: Nemotron 3 Nano Omni (free) availability by provider

2 alternatives

We found 3 provider listings for nemotron-3-nano-omni-30b-a3b-reasoning. Check model ID, quota, pricing, and API format before switching providers.

Provider Model listing Access Context API Limits
OpenRouter NVIDIA: Nemotron 3 Nano Omni (free) Free tier 256K OpenAI-style 200 req/day (free tier)
Kilo Code nvidia/nemotron-3-nano-omni-30b-a3b-reasoning:free Free tier 256K OpenAI-style ~200 req/hr
NVIDIA NIM Nemotron 3 Nano Omni 30B A3B Reasoning Free tier 256K Native Varies
View OpenRouter setup guide →
Typical Use Cases

NVIDIA: Nemotron 3 Nano Omni (free) use cases

Chat

NVIDIA: Nemotron 3 Nano Omni (free) is tagged for chat in this catalog and works with OpenAI-compatible client libraries.

Vision

NVIDIA: Nemotron 3 Nano Omni (free) is tagged for vision in this catalog and works with OpenAI-compatible client libraries.

Reasoning

NVIDIA: Nemotron 3 Nano Omni (free) is tagged for reasoning in this catalog and works with OpenAI-compatible client libraries.

FAQ

NVIDIA: Nemotron 3 Nano Omni (free) free API FAQ

Is NVIDIA: Nemotron 3 Nano Omni (free) free to use?

NVIDIA: Nemotron 3 Nano Omni (free) is listed with free API access on OpenRouter, subject to the provider's quota and account policy.

What is the NVIDIA: Nemotron 3 Nano Omni (free) model ID?

The model ID shown in this catalog is nvidia/nemotron-3-nano-omni-30b-a3b-reasoning:free.

What are the NVIDIA: Nemotron 3 Nano Omni (free) free tier rate limits on OpenRouter?

The listed free tier limit is 200 req/day (free tier). Limits can change per account tier, so confirm against the provider dashboard.

What context window does NVIDIA: Nemotron 3 Nano Omni (free) support?

The listed context window is 256K tokens with up to 66K output tokens.

More about NVIDIA: Nemotron 3 Nano Omni (free)

NVIDIA Nemotron 3 Nano Omni 30B A3B Reasoning is an open multimodal model with 30B total parameters (3B active via MoE), available free on OpenRouter. Accepts text, image, video, and audio input — built as a perception and context sub-agent for enterprise agent systems. Uses a Hybrid MoE Transformer-Mamba architecture with Conv3D video layers and Efficient Video Sampling (EVS) for ~2x higher throughput and 2.5x lower compute on video tasks vs. separate vision+speech pipelines. Up to 300K context, 16K reasoning budget. OpenAI-compatible. Free tier: 200 RPD.

For API keys, setup steps, and provider-level limits, see the OpenRouter provider page.