GitHub Models logo

Llama-4-Maverick-17B-128E API status on GitHub Models

Check provider
Catalog profile Tool use provider catalog metadata

Llama 4 Maverick 17B is Meta's high-expert-count MoE model with 128 active experts, available for free on GitHub Models.

Check providerOpenAI compatibleTool callingTextImage
API Details

Copy-ready connection details

Use these values in your client or SDK.
Base URL
https://models.github.ai/inference
Model ID
llama-4-maverick-17b-128e
API format OpenAI Chat Completions
Technical Details

Llama-4-Maverick-17B-128E specifications

Provider catalog
Context window 256K
Max output 4K
Status Check provider
Family llama
Knowledge cutoff 2024-08
Released Apr 5, 2025
Last updated Jun 28, 2026
Free listing since Apr 5, 2025
Input text, image
Output text
Capabilities tool calling
Open weights Yes
AI Recommendation

Should you use Llama-4-Maverick-17B-128E?

Llama-4-Maverick-17B-128E is listed for chat workloads and supports a 256K context window.

Check GitHub Models's current endpoint status before integrating this listing. Use the alternatives table if the provider listing is unavailable.

Best for
  • Chat
Strengths & Weaknesses

Strengths

  • Long context window
  • Tool calling support
  • Works with OpenAI-style SDKs

Watch outs

  • Current endpoint availability needs provider confirmation
  • Free-tier rate limits apply
Availability

Llama-4-Maverick-17B-128E availability by provider

Current provider only

We only found this llama listing on GitHub Models in the current catalog. Use the related models section for nearby alternatives.

Provider Model listing Access Context API Limits
GitHub Models Llama-4-Maverick-17B-128E Check provider 256K OpenAI-style 10 RPM, 50 RPD
View GitHub Models setup guide →
Typical Use Cases

Llama-4-Maverick-17B-128E use cases

Chat

Llama-4-Maverick-17B-128E is tagged for chat in this catalog and works with OpenAI-compatible client libraries.

FAQ

Llama-4-Maverick-17B-128E free API FAQ

Is Llama-4-Maverick-17B-128E free to use?

Llama-4-Maverick-17B-128E appears in the free model catalog for GitHub Models, but its current endpoint availability should be confirmed with the provider before use.

What is the Llama-4-Maverick-17B-128E model ID?

The model ID shown in this catalog is llama-4-maverick-17b-128e.

What are the Llama-4-Maverick-17B-128E free tier rate limits on GitHub Models?

The listed free tier limit is 10 RPM, 50 RPD. Limits can change per account tier, so confirm against the provider dashboard.

What context window does Llama-4-Maverick-17B-128E support?

The listed context window is 256K tokens with up to 4K output tokens.

More about Llama-4-Maverick-17B-128E

Llama 4 Maverick 17B is Meta's high-expert-count MoE model with 128 active experts, available for free on GitHub Models. Despite its compact 17B active parameter footprint, the large expert pool gives it broad knowledge coverage — making it a strong general-purpose chat and instruction-following model that rivals much larger dense architectures. The free tier limits output to 4K tokens per request with 10 RPM and 50 requests per day, so it is best suited for interactive chat and short-form generation rather than long-form writing. Fully OpenAI SDK-compatible; requires only a GitHub account to start using.

For API keys, setup steps, and provider-level limits, see the GitHub Models provider page.