NVIDIA NIM logo

deepseek-v4-pro API status on NVIDIA NIM

Check provider
★★★★★★★★★★ 3.5 Benchmark-backed score

Open MoE flagship with million-token context for coding and long agent runs

Check providerOpenAI compatibleReasoningTool callingJSON modeTextReasoning
API Details

Copy-ready connection details

Use these values in your client or SDK.
Base URL
https://integrate.api.nvidia.com/v1
Model ID
deepseek-ai/deepseek-v4-pro
API format OpenAI Chat Completions
Technical Details

deepseek-v4-pro specifications

Provider and model catalog
Context window 1.0M
Max output 384K
Status Check provider
Family deepseek-thinking
Knowledge cutoff 2025-05
Released Apr 24, 2026
Last updated Aug 6, 2026
Free listing since Apr 24, 2026
Input text
Output text
Capabilities reasoning, tool calling, structured output, temperature control
Open weights Yes

External benchmark references

Benchmark Score Metric Date
SWE-Bench Verified 80.6 resolved Not listed
Artificial Analysis Coding Agent Index 50.1 average pass@1 Not listed
SWE-Atlas Codebase QnA 67.8 pass@1 Not listed
SWE-Bench Pro 18 pass@1 Not listed
AI Recommendation

Should you use deepseek-v4-pro?

deepseek-v4-pro is listed for chat workloads and supports a 1.0M context window.

Check NVIDIA NIM's current endpoint status before integrating this listing. Use the alternatives table if the provider listing is unavailable.

Best for
  • Chat
Strengths & Weaknesses

Strengths

  • Long context window
  • Reasoning mode listed
  • Tool calling support
  • Structured JSON output

Watch outs

  • Current endpoint availability needs provider confirmation
  • Free-tier rate limits apply
  • Vision support is not listed
Benchmark Overview

Benchmark signals for deepseek-v4-pro

Measured data
Intelligence General reasoning and instruction following
44.3/100
Coding Programming and code generation
59.4/100
Agentic Tool use and multi-step tasks
36.4/100
Speed Observed generation speed
64 tok/s
Context Maximum listed context window
1.0M
Pricing

deepseek-v4-pro pricing per 1M tokens

Check provider
Input $0.43 per 1M tokens
Output $0.87 per 1M tokens
Free access Check provider NVIDIA NIM
Rate limit Up to 40 RPM provider policy
Typical Use Cases

deepseek-v4-pro use cases

Chat

deepseek-v4-pro is tagged for chat in this catalog and works with OpenAI-compatible client libraries.

FAQ

deepseek-v4-pro free API FAQ

Is deepseek-v4-pro free to use?

deepseek-v4-pro appears in the free model catalog for NVIDIA NIM, but its current endpoint availability should be confirmed with the provider before use.

What is the deepseek-v4-pro model ID?

The model ID shown in this catalog is deepseek-ai/deepseek-v4-pro.

What are the deepseek-v4-pro free tier rate limits on NVIDIA NIM?

The listed free tier limit is Up to 40 RPM. Limits can change per account tier, so confirm against the provider dashboard.

What context window does deepseek-v4-pro support?

The listed context window is 1.0M tokens with up to 384K output tokens.

More about deepseek-v4-pro

DeepSeek V4 Pro is available free on NVIDIA NIM with up to 40 RPM and no daily token cap. DeepSeek's most capable model with 1M context and 384K output — one of the highest output ceilings on any free endpoint, making it practical for long-form generation and full-document analysis. NVIDIA's OpenAI-compatible API; requires free NVIDIA Developer Program membership and phone verification.

For API keys, setup steps, and provider-level limits, see the NVIDIA NIM provider page.