GitHub Models logo

Llama-4-Maverick-17B-128E-Instruct-FP8 Free API on GitHub Models

Free API
Catalog profile Tool use provider catalog metadata

Llama-4-Maverick-17B-128E-Instruct-FP8 — free model from GitHub Models.

Free APIOpenAI compatibleTool callingText
API Details

Copy-ready connection details

Use these values in your client or SDK.
Base URL
https://models.github.ai/inference
Model ID
llama-4-maverick-17b-128e-instruct-fp8
API format OpenAI Chat Completions
Technical Details

Llama-4-Maverick-17B-128E-Instruct-FP8 specifications

Provider catalog
Context window 256K
Max output 4K
Status Online
Last updated Aug 5, 2026
Free listing since Jan 31, 2025
Input text
Output text
Capabilities tool calling
AI Recommendation

Should you use Llama-4-Maverick-17B-128E-Instruct-FP8?

Llama-4-Maverick-17B-128E-Instruct-FP8 is listed for chat workloads and supports a 256K context window.

Use it when GitHub Models's free tier is enough for evaluation, demos, or light production traffic.

Best for
  • Chat
Strengths & Weaknesses

Strengths

  • Long context window
  • Tool calling support
  • Works with OpenAI-style SDKs

Watch outs

  • Free-tier rate limits apply
  • Vision support is not listed
Pricing

Llama-4-Maverick-17B-128E-Instruct-FP8 pricing per 1M tokens

Free tier listed
Input $0 per 1M tokens
Output $0 per 1M tokens
Free access Available GitHub Models
Rate limit 10 RPM, 50 RPD provider policy
Availability

Llama-4-Maverick-17B-128E-Instruct-FP8 availability by provider

Current provider only

We only found this Llama-4-Maverick-17B-128E-Instruct-FP8 listing on GitHub Models in the current catalog. Use the related models section for nearby alternatives.

Provider Model listing Access Context API Limits
GitHub Models Llama-4-Maverick-17B-128E-Instruct-FP8 Free tier 256K OpenAI-style 10 RPM, 50 RPD
View GitHub Models setup guide →
Typical Use Cases

Llama-4-Maverick-17B-128E-Instruct-FP8 use cases

Chat

Llama-4-Maverick-17B-128E-Instruct-FP8 is tagged for chat in this catalog and works with OpenAI-compatible client libraries.

FAQ

Llama-4-Maverick-17B-128E-Instruct-FP8 free API FAQ

Is Llama-4-Maverick-17B-128E-Instruct-FP8 free to use?

Llama-4-Maverick-17B-128E-Instruct-FP8 is listed with free API access on GitHub Models, subject to the provider's quota and account policy.

What is the Llama-4-Maverick-17B-128E-Instruct-FP8 model ID?

The model ID shown in this catalog is llama-4-maverick-17b-128e-instruct-fp8.

What are the Llama-4-Maverick-17B-128E-Instruct-FP8 free tier rate limits on GitHub Models?

The listed free tier limit is 10 RPM, 50 RPD. Limits can change per account tier, so confirm against the provider dashboard.

What context window does Llama-4-Maverick-17B-128E-Instruct-FP8 support?

The listed context window is 256K tokens with up to 4K output tokens.

More about Llama-4-Maverick-17B-128E-Instruct-FP8

Llama-4-Maverick-17B-128E-Instruct-FP8 — free model from GitHub Models.

For API keys, setup steps, and provider-level limits, see the GitHub Models provider page.