GitHub Models logo

Llama-4-Scout-17B-16E API status on GitHub Models

Check provider
Catalog profile Tool use provider catalog metadata

Llama 4 Scout 17B is Meta's efficient long-context model with 16 active experts (MoE), available free on GitHub Models.

Check providerOpenAI compatibleTool callingTextImage
API Details

Copy-ready connection details

Use these values in your client or SDK.
Base URL
https://models.github.ai/inference
Model ID
llama-4-scout-17b-16e
API format OpenAI Chat Completions
Technical Details

Llama-4-Scout-17B-16E specifications

Provider catalog
Context window 512K
Max output 4K
Status Check provider
Family llama
Knowledge cutoff 2024-08
Released Apr 5, 2025
Last updated Jun 28, 2026
Free listing since Apr 5, 2025
Input text, image
Output text
Capabilities tool calling
Open weights Yes
AI Recommendation

Should you use Llama-4-Scout-17B-16E?

Llama-4-Scout-17B-16E is listed for chat workloads and supports a 512K context window.

Check GitHub Models's current endpoint status before integrating this listing. Use the alternatives table if the provider listing is unavailable.

Best for
  • Chat
Strengths & Weaknesses

Strengths

  • Long context window
  • Tool calling support
  • Works with OpenAI-style SDKs

Watch outs

  • Current endpoint availability needs provider confirmation
  • Free-tier rate limits apply
Availability

Llama-4-Scout-17B-16E availability by provider

Current provider only

We only found this llama listing on GitHub Models in the current catalog. Use the related models section for nearby alternatives.

Provider Model listing Access Context API Limits
GitHub Models Llama-4-Scout-17B-16E Check provider 512K OpenAI-style 15 RPM, 150 RPD
View GitHub Models setup guide →
Typical Use Cases

Llama-4-Scout-17B-16E use cases

Chat

Llama-4-Scout-17B-16E is tagged for chat in this catalog and works with OpenAI-compatible client libraries.

FAQ

Llama-4-Scout-17B-16E free API FAQ

Is Llama-4-Scout-17B-16E free to use?

Llama-4-Scout-17B-16E appears in the free model catalog for GitHub Models, but its current endpoint availability should be confirmed with the provider before use.

What is the Llama-4-Scout-17B-16E model ID?

The model ID shown in this catalog is llama-4-scout-17b-16e.

What are the Llama-4-Scout-17B-16E free tier rate limits on GitHub Models?

The listed free tier limit is 15 RPM, 150 RPD. Limits can change per account tier, so confirm against the provider dashboard.

What context window does Llama-4-Scout-17B-16E support?

The listed context window is 512K tokens with up to 4K output tokens.

More about Llama-4-Scout-17B-16E

Llama 4 Scout 17B is Meta's efficient long-context model with 16 active experts (MoE), available free on GitHub Models. With a 512K context window — far beyond most free models — it can process entire novels, full code repositories, or multi-hour transcripts in a single request. The 17B total parameter footprint keeps inference fast and cost-effective, making it practical for retrieval-free document analysis and long-form summarization. Rate limits are 15 RPM and 150 requests per day with per-request output capped at 4K tokens. Fully OpenAI SDK-compatible and requires only a GitHub account.

For API keys, setup steps, and provider-level limits, see the GitHub Models provider page.