GitHub Models logo

gpt-4.1 API status on GitHub Models

Check provider
★★★★★★★★★★ 3.5 Benchmark-backed score

Long-lived GPT workhorse for coding, instruction following, and production apps

Check providerOpenAI compatibleTool callingJSON modePDF inputTextImage
API Details

Copy-ready connection details

Use these values in your client or SDK.
Base URL
https://models.github.ai/inference
Model ID
gpt-4-1
API format OpenAI Chat Completions
Technical Details

gpt-4.1 specifications

Provider and model catalog
Context window 1.0M
Max output 32K
Status Check provider
Family gpt
Knowledge cutoff 2024-04
Released Apr 14, 2025
Last updated Aug 6, 2026
Free listing since Apr 14, 2025
Input text, image, pdf
Output text
Capabilities tool calling, structured output, file attachments, temperature control

External benchmark references

Benchmark Score Metric Date
Aider Polyglot 52.4 percent correct 2025-04-14
AI Recommendation

Should you use gpt-4.1?

gpt-4.1 is listed for chat workloads and supports a 1.0M context window.

Check GitHub Models's current endpoint status before integrating this listing. Use the alternatives table if the provider listing is unavailable.

Best for
  • Chat
Strengths & Weaknesses

Strengths

  • Long context window
  • Tool calling support
  • Structured JSON output
  • Works with OpenAI-style SDKs

Watch outs

  • Current endpoint availability needs provider confirmation
  • Free-tier rate limits apply
Benchmark Overview

Benchmark signals for gpt-4.1

Measured data
Intelligence General reasoning and instruction following
12.7/100
Coding Programming and code generation
21.8/100
Agentic Tool use and multi-step tasks
27.3/100
Speed Observed generation speed
189 tok/s
Context Maximum listed context window
1.0M
Pricing

gpt-4.1 pricing per 1M tokens

Check provider
Input $2 per 1M tokens
Output $8 per 1M tokens
Access Check provider GitHub Models
Rate limit 10 RPM, 50 RPD provider policy
Availability

gpt-4.1 availability by provider

Current provider only

We only found this gpt-4-1 listing on GitHub Models in the current catalog. Use the related models section for nearby alternatives.

Provider Model listing Access Context API Limits
GitHub Models gpt-4.1 Check provider 1.0M OpenAI-style 10 RPM, 50 RPD
View GitHub Models setup guide →
Typical Use Cases

gpt-4.1 use cases

Chat

gpt-4.1 is tagged for chat in this catalog and works with OpenAI-compatible client libraries.

FAQ

gpt-4.1 free API FAQ

Is gpt-4.1 free to use?

gpt-4.1 appears in the free model catalog for GitHub Models, but its current endpoint availability should be confirmed with the provider before use.

What is the gpt-4.1 model ID?

The model ID shown in this catalog is gpt-4-1.

What are the gpt-4.1 free tier rate limits on GitHub Models?

The listed free tier limit is 10 RPM, 50 RPD. Limits can change per account tier, so confirm against the provider dashboard.

What context window does gpt-4.1 support?

The listed context window is 1.0M tokens with up to 32K output tokens.

More about gpt-4.1

GPT-4.1 is OpenAI's mid-tier model available for free through GitHub Models — Microsoft's AI playground for GitHub users. With a 1M-token context window, 32K max output, and OpenAI-native SDK compatibility, it is the most accessible way to use a GPT-4-class model without paying for an OpenAI API key. The free tier is rate-limited to 10 RPM and 50 requests per day, and per-request token caps apply (8K input / 4K output), so it is best suited for prototyping, testing prompts, and lightweight coding tasks rather than production deployment. Requires only a GitHub account — no separate API key signup needed.

For API keys, setup steps, and provider-level limits, see the GitHub Models provider page.