NVIDIA NIM logo

llama-guard-4-12b Free API on NVIDIA NIM

Free API Verified
Catalog profile API ready provider catalog metadata

Llama Guard 4 12B is Meta's safety content moderation model, free on NVIDIA NIM.

Free APIOpenAI compatibleJSON modeTextImage
API Details

Copy-ready connection details

Use these values in your client or SDK.
Base URL
https://integrate.api.nvidia.com/v1
Model ID
meta/llama-guard-4-12b
API format OpenAI Chat Completions
Technical Details

llama-guard-4-12b specifications

Provider catalog
Context window 1.0M
Max output 16K
Status Online
Family llama-guard-4-12b
Released Apr 5, 2025
Last updated Aug 6, 2026
Free listing since Apr 5, 2025
Input text, image
Output text
Capabilities structured output
AI Recommendation

Should you use llama-guard-4-12b?

llama-guard-4-12b is listed for chat workloads and supports a 1.0M context window.

Use it when NVIDIA NIM's free tier is enough for evaluation, demos, or light production traffic.

Best for
  • Chat
Strengths & Weaknesses

Strengths

  • Long context window
  • Structured JSON output
  • Works with OpenAI-style SDKs
  • Live API verification available

Watch outs

  • Free-tier rate limits apply
  • Tool calling is not confirmed
Pricing

llama-guard-4-12b pricing per 1M tokens

Free tier listed
Input $0 per 1M tokens
Output $0 per 1M tokens
Free access Available NVIDIA NIM
Rate limit Up to 40 RPM provider policy
Availability

llama-guard-4-12b availability by provider

Current provider only

We only found this llama-guard-4-12b listing on NVIDIA NIM in the current catalog. Use the related models section for nearby alternatives.

Provider Model listing Access Context API Limits
NVIDIA NIM meta/llama-guard-4-12b Free tier 1.0M OpenAI-style Up to 40 RPM
View NVIDIA NIM setup guide →
Typical Use Cases

llama-guard-4-12b use cases

Chat

llama-guard-4-12b is tagged for chat in this catalog and works with OpenAI-compatible client libraries.

FAQ

llama-guard-4-12b free API FAQ

Is llama-guard-4-12b free to use?

llama-guard-4-12b is listed with free API access on NVIDIA NIM, subject to the provider's quota and account policy.

What is the llama-guard-4-12b model ID?

The model ID shown in this catalog is meta/llama-guard-4-12b.

What are the llama-guard-4-12b free tier rate limits on NVIDIA NIM?

The listed free tier limit is Up to 40 RPM. Limits can change per account tier, so confirm against the provider dashboard.

What context window does llama-guard-4-12b support?

The listed context window is 1.0M tokens with up to 16K output tokens.

More about llama-guard-4-12b

Llama Guard 4 12B is Meta's safety content moderation model, free on NVIDIA NIM. Unlike general-purpose chat models, it is specifically designed to detect unsafe content, policy violations, and harmful outputs in AI systems — a specialized tool for building content safety pipelines. Up to 40 RPM, no daily token cap. OpenAI-compatible; requires NVIDIA Developer Program membership and phone verification.

For API keys, setup steps, and provider-level limits, see the NVIDIA NIM provider page.