Groq logo

llama-4-maverick-17b-128e-instruct API status on Groq

Check provider
Catalog profile API ready provider catalog metadata

Llama 4 Maverick 17B on Groq is Meta's highest-expert-count MoE model with 128 active experts, running on Groq's LPU hardware for fast inference.

Check providerOpenAI compatibleText
API Details

Copy-ready connection details

Use these values in your client or SDK.
Base URL
https://api.groq.com/openai/v1
Model ID
llama-4-maverick-17b-128e-instruct
API format OpenAI Chat Completions + OpenAI Responses
Technical Details

llama-4-maverick-17b-128e-instruct specifications

Provider catalog
Context window 131K
Max output 8K
Status Check provider
Last updated May 10, 2026
Free listing since May 10, 2026
Input text
Output text
AI Recommendation

Should you use llama-4-maverick-17b-128e-instruct?

llama-4-maverick-17b-128e-instruct is listed for chat workloads and supports a 131K context window.

Check Groq's current endpoint status before integrating this listing. Use the alternatives table if the provider listing is unavailable.

Best for
  • Chat
Strengths & Weaknesses

Strengths

  • Long context window
  • Works with OpenAI-style SDKs

Watch outs

  • Current endpoint availability needs provider confirmation
  • Free-tier rate limits apply
  • Vision support is not listed
  • Tool calling is not confirmed
Availability

llama-4-maverick-17b-128e-instruct availability by provider

Current provider only

We only found this llama-4-maverick-17b-128e-instruct listing on Groq in the current catalog. Use the related models section for nearby alternatives.

Provider Model listing Access Context API Limits
Groq llama-4-maverick-17b-128e-instruct Check provider 131K OpenAI-style 15 RPM, 500 RPD
View Groq setup guide →
Typical Use Cases

llama-4-maverick-17b-128e-instruct use cases

Chat

llama-4-maverick-17b-128e-instruct is tagged for chat in this catalog and works with OpenAI-compatible client libraries.

FAQ

llama-4-maverick-17b-128e-instruct free API FAQ

Is llama-4-maverick-17b-128e-instruct free to use?

llama-4-maverick-17b-128e-instruct appears in the free model catalog for Groq, but its current endpoint availability should be confirmed with the provider before use.

What is the llama-4-maverick-17b-128e-instruct model ID?

The model ID shown in this catalog is llama-4-maverick-17b-128e-instruct.

What are the llama-4-maverick-17b-128e-instruct free tier rate limits on Groq?

The listed free tier limit is 15 RPM, 500 RPD. Limits can change per account tier, so confirm against the provider dashboard.

What context window does llama-4-maverick-17b-128e-instruct support?

The listed context window is 131K tokens with up to 8K output tokens.

More about llama-4-maverick-17b-128e-instruct

Llama 4 Maverick 17B on Groq is Meta's highest-expert-count MoE model with 128 active experts, running on Groq's LPU hardware for fast inference. The large expert count gives it broader knowledge and stronger instruction-following than the Scout variant, making it the better Groq option for complex tasks. Rate limits are notably tighter than Groq's other endpoints at 15 RPM and 500 requests per day, so it is best reserved for evaluations and high-value queries rather than high-volume traffic. OpenAI SDK compatible; registration required, no credit card.

For API keys, setup steps, and provider-level limits, see the Groq provider page.