OpenRouter logo

NVIDIA: Llama Nemotron Embed VL 1B V2 (free) API status on OpenRouter

Check provider
Catalog profile Reasoning provider catalog metadata

NVIDIA Llama Nemotron Embed VL 1B V2 is a compact multimodal retrieval model available free on OpenRouter.

Check providerOpenAI compatibleReasoningTextImageEmbedding
API Details

Copy-ready connection details

Use these values in your client or SDK.
Base URL
https://openrouter.ai/api/v1
Model ID
nvidia/llama-nemotron-embed-vl-1b-v2-20260224:free
API format OpenAI Chat Completions + OpenAI Responses
Technical Details

NVIDIA: Llama Nemotron Embed VL 1B V2 (free) specifications

Provider catalog
Context window 131K
Max output 8K
Status Check provider
Released Feb 25, 2026
Last updated Jun 14, 2026
Free listing since Feb 25, 2026
Input text, image, embedding
Output text
Capabilities reasoning
AI Recommendation

Should you use NVIDIA: Llama Nemotron Embed VL 1B V2 (free)?

NVIDIA: Llama Nemotron Embed VL 1B V2 (free) is listed for chat, reasoning, embedding workloads and supports a 131K context window.

Check OpenRouter's current endpoint status before integrating this listing. Use the alternatives table if the provider listing is unavailable.

Best for
  • Chat
  • Reasoning
  • Embedding
Strengths & Weaknesses

Strengths

  • Strong reasoning profile
  • Long context window
  • Reasoning mode listed
  • Works with OpenAI-style SDKs

Watch outs

  • Current endpoint availability needs provider confirmation
  • Free-tier rate limits apply
  • Tool calling is not confirmed
Availability

NVIDIA: Llama Nemotron Embed VL 1B V2 (free) availability by provider

Current provider only

We only found this NVIDIA: Llama Nemotron Embed VL 1B V2 (free) listing on OpenRouter in the current catalog. Use the related models section for nearby alternatives.

Provider Model listing Access Context API Limits
OpenRouter NVIDIA: Llama Nemotron Embed VL 1B V2 (free) Check provider 131K OpenAI-style 200 req/day (free tier)
View OpenRouter setup guide →
Typical Use Cases

NVIDIA: Llama Nemotron Embed VL 1B V2 (free) use cases

Chat

NVIDIA: Llama Nemotron Embed VL 1B V2 (free) is tagged for chat in this catalog and works with OpenAI-compatible client libraries.

Reasoning

NVIDIA: Llama Nemotron Embed VL 1B V2 (free) is tagged for reasoning in this catalog and works with OpenAI-compatible client libraries.

Embedding

NVIDIA: Llama Nemotron Embed VL 1B V2 (free) is tagged for embedding in this catalog and works with OpenAI-compatible client libraries.

FAQ

NVIDIA: Llama Nemotron Embed VL 1B V2 (free) free API FAQ

Is NVIDIA: Llama Nemotron Embed VL 1B V2 (free) free to use?

NVIDIA: Llama Nemotron Embed VL 1B V2 (free) appears in the free model catalog for OpenRouter, but its current endpoint availability should be confirmed with the provider before use.

What is the NVIDIA: Llama Nemotron Embed VL 1B V2 (free) model ID?

The model ID shown in this catalog is nvidia/llama-nemotron-embed-vl-1b-v2-20260224:free.

What are the NVIDIA: Llama Nemotron Embed VL 1B V2 (free) free tier rate limits on OpenRouter?

The listed free tier limit is 200 req/day (free tier). Limits can change per account tier, so confirm against the provider dashboard.

What context window does NVIDIA: Llama Nemotron Embed VL 1B V2 (free) support?

The listed context window is 131K tokens with up to 8K output tokens.

More about NVIDIA: Llama Nemotron Embed VL 1B V2 (free)

NVIDIA Llama Nemotron Embed VL 1B V2 is a compact multimodal retrieval model available free on OpenRouter. Optimized for multimodal question-answering retrieval — it embeds documents as image, text, or both, retrievable via text query. Supports images containing text, tables, charts, and infographics. 131K context window, 1B parameters. Not a chat model — purpose-built for embedding and retrieval pipelines. OpenAI-compatible via OpenRouter. Free tier: 200 RPD.

For API keys, setup steps, and provider-level limits, see the OpenRouter provider page.