OpenRouter logo

NVIDIA: Llama Nemotron Rerank VL 1B V2 (free) API status on OpenRouter

Check provider
Catalog metadata matched Reasoning models.dev metadata

Reranking model for improving retrieval quality in search and recommendation systems

Check providerOpenAI compatibleReasoningTextImageRerank
API Details

Copy-ready connection details

Use these values in your client or SDK.
Base URL
https://openrouter.ai/api/v1
Model ID
nvidia/llama-nemotron-rerank-vl-1b-v2:free
API format OpenAI Chat Completions + OpenAI Responses
Technical Details

NVIDIA: Llama Nemotron Rerank VL 1B V2 (free) specifications

Provider and model catalog
Context window 10K
Max output 8K
Status Check provider
Family nemotron
Released Mar 31, 2026
Last updated Jun 14, 2026
Free listing since Jun 9, 2026
Input text, image
Output text
Capabilities reasoning, file attachments
Open weights Yes
AI Recommendation

Should you use NVIDIA: Llama Nemotron Rerank VL 1B V2 (free)?

NVIDIA: Llama Nemotron Rerank VL 1B V2 (free) is listed for chat, reasoning, embedding workloads and supports a 10K context window.

Check OpenRouter's current endpoint status before integrating this listing. Use the alternatives table if the provider listing is unavailable.

Best for
  • Chat
  • Reasoning
  • Embedding
Strengths & Weaknesses

Strengths

  • Strong reasoning profile
  • Reasoning mode listed
  • Open weights available
  • Works with OpenAI-style SDKs

Watch outs

  • Current endpoint availability needs provider confirmation
  • Free-tier rate limits apply
  • Tool calling is not confirmed
Availability

NVIDIA: Llama Nemotron Rerank VL 1B V2 (free) availability by provider

Current provider only

We only found this NVIDIA: Llama Nemotron Rerank VL 1B V2 (free) listing on OpenRouter in the current catalog. Use the related models section for nearby alternatives.

Provider Model listing Access Context API Limits
OpenRouter NVIDIA: Llama Nemotron Rerank VL 1B V2 (free) Check provider 10K OpenAI-style 200 req/day (free tier)
View OpenRouter setup guide →
Typical Use Cases

NVIDIA: Llama Nemotron Rerank VL 1B V2 (free) use cases

Chat

NVIDIA: Llama Nemotron Rerank VL 1B V2 (free) is tagged for chat in this catalog and works with OpenAI-compatible client libraries.

Reasoning

NVIDIA: Llama Nemotron Rerank VL 1B V2 (free) is tagged for reasoning in this catalog and works with OpenAI-compatible client libraries.

Embedding

NVIDIA: Llama Nemotron Rerank VL 1B V2 (free) is tagged for embedding in this catalog and works with OpenAI-compatible client libraries.

FAQ

NVIDIA: Llama Nemotron Rerank VL 1B V2 (free) free API FAQ

Is NVIDIA: Llama Nemotron Rerank VL 1B V2 (free) free to use?

NVIDIA: Llama Nemotron Rerank VL 1B V2 (free) appears in the free model catalog for OpenRouter, but its current endpoint availability should be confirmed with the provider before use.

What is the NVIDIA: Llama Nemotron Rerank VL 1B V2 (free) model ID?

The model ID shown in this catalog is nvidia/llama-nemotron-rerank-vl-1b-v2:free.

What are the NVIDIA: Llama Nemotron Rerank VL 1B V2 (free) free tier rate limits on OpenRouter?

The listed free tier limit is 200 req/day (free tier). Limits can change per account tier, so confirm against the provider dashboard.

What context window does NVIDIA: Llama Nemotron Rerank VL 1B V2 (free) support?

The listed context window is 10K tokens with up to 8K output tokens.

More about NVIDIA: Llama Nemotron Rerank VL 1B V2 (free)

Llama Nemotron Rerank VL 1B V2 is a 1.7B multimodal reranking model from NVIDIA. It evaluates the relevance of document images and text against user queries, designed for vision RAG pipelines handling charts, tables, infographics, and mixed-media documents. Functions as a cross-encoder that accepts text queries paired with image, text, or combined document inputs, delivering approximately 6-7% recall improvements over embedding-only baselines on visual document retrieval benchmarks.

For API keys, setup steps, and provider-level limits, see the OpenRouter provider page.