NVIDIA NIM logo

phi-4-multimodal-instruct API status on NVIDIA NIM

Check provider
★★★★★★★★★★ 3.5 Benchmark-backed score

General-purpose chat model for instruction following, writing, and analysis

Check providerOpenAI compatibleText
API Details

Copy-ready connection details

Use these values in your client or SDK.
Base URL
https://integrate.api.nvidia.com/v1
Model ID
microsoft/phi-4-multimodal-instruct
API format OpenAI Chat Completions
Technical Details

phi-4-multimodal-instruct specifications

Provider catalog
Context window 131K
Max output 8K
Status Check provider
Released Feb 26, 2025
Last updated Jul 7, 2026
Free listing since Feb 26, 2025
Input text
Output text
AI Recommendation

Should you use phi-4-multimodal-instruct?

phi-4-multimodal-instruct is listed for chat workloads and supports a 131K context window.

Check NVIDIA NIM's current endpoint status before integrating this listing. Use the alternatives table if the provider listing is unavailable.

Best for
  • Chat
Strengths & Weaknesses

Strengths

  • Long context window
  • Works with OpenAI-style SDKs

Watch outs

  • Current endpoint availability needs provider confirmation
  • Free-tier rate limits apply
  • Vision support is not listed
  • Tool calling is not confirmed
Benchmark Overview

Benchmark signals for phi-4-multimodal-instruct

Measured data
Intelligence General reasoning and instruction following
4.5/100
Speed Observed generation speed
17 tok/s
Context Maximum listed context window
131K
Pricing

phi-4-multimodal-instruct pricing per 1M tokens

Check provider
Input $0 per 1M tokens
Output $0 per 1M tokens
Free access Check provider NVIDIA NIM
Rate limit Up to 40 RPM provider policy
Availability

phi-4-multimodal-instruct availability by provider

Current provider only

We only found this phi-4-multimodal-instruct listing on NVIDIA NIM in the current catalog. Use the related models section for nearby alternatives.

Provider Model listing Access Context API Limits
NVIDIA NIM microsoft/phi-4-multimodal-instruct Check provider 131K OpenAI-style Up to 40 RPM
View NVIDIA NIM setup guide →
Typical Use Cases

phi-4-multimodal-instruct use cases

Chat

phi-4-multimodal-instruct is tagged for chat in this catalog and works with OpenAI-compatible client libraries.

FAQ

phi-4-multimodal-instruct free API FAQ

Is phi-4-multimodal-instruct free to use?

phi-4-multimodal-instruct appears in the free model catalog for NVIDIA NIM, but its current endpoint availability should be confirmed with the provider before use.

What is the phi-4-multimodal-instruct model ID?

The model ID shown in this catalog is microsoft/phi-4-multimodal-instruct.

What are the phi-4-multimodal-instruct free tier rate limits on NVIDIA NIM?

The listed free tier limit is Up to 40 RPM. Limits can change per account tier, so confirm against the provider dashboard.

What context window does phi-4-multimodal-instruct support?

The listed context window is 131K tokens with up to 8K output tokens.

More about phi-4-multimodal-instruct

General-purpose chat model for instruction following, writing, and analysis

For API keys, setup steps, and provider-level limits, see the NVIDIA NIM provider page.