GitHub Models logo

gpt-4.1-mini Free API on GitHub Models

Free API
★★★★★★★★★★ 3.5 Benchmark-backed score

Affordable GPT-4.1 lane for fast coding help and structured extraction

Free APIOpenAI compatibleTool callingJSON modePDF inputTextImage
API Details

Copy-ready connection details

Use these values in your client or SDK.
Base URL
https://models.github.ai/inference
Model ID
gpt-4-1-mini
API format OpenAI Chat Completions
Technical Details

gpt-4.1-mini specifications

Provider and model catalog
Context window 1.0M
Max output 32K
Status Online
Family gpt-mini
Knowledge cutoff 2024-04
Released Apr 14, 2025
Last updated Aug 6, 2026
Free listing since Apr 14, 2025
Input text, image, pdf
Output text
Capabilities tool calling, structured output, file attachments, temperature control

External benchmark references

Benchmark Score Metric Date
Aider Polyglot 32.4 percent correct 2025-04-14
AI Recommendation

Should you use gpt-4.1-mini?

gpt-4.1-mini is listed for chat workloads and supports a 1.0M context window.

Use it when GitHub Models's free tier is enough for evaluation, demos, or light production traffic.

Best for
  • Chat
Strengths & Weaknesses

Strengths

  • Long context window
  • Tool calling support
  • Structured JSON output
  • Works with OpenAI-style SDKs

Watch outs

  • Free-tier rate limits apply
Benchmark Overview

Benchmark signals for gpt-4.1-mini

Measured data
Intelligence General reasoning and instruction following
14.8/100
Coding Programming and code generation
20.2/100
Agentic Tool use and multi-step tasks
1.7/100
Speed Observed generation speed
72 tok/s
Context Maximum listed context window
1.0M
Pricing

gpt-4.1-mini pricing per 1M tokens

Free tier listed
Input $0.4 per 1M tokens
Output $1.6 per 1M tokens
Free access Available GitHub Models
Rate limit 15 RPM, 150 RPD provider policy
Availability

gpt-4.1-mini availability by provider

Current provider only

We only found this gpt-4-1-mini listing on GitHub Models in the current catalog. Use the related models section for nearby alternatives.

Provider Model listing Access Context API Limits
GitHub Models gpt-4.1-mini Free tier 1.0M OpenAI-style 15 RPM, 150 RPD
View GitHub Models setup guide →
Typical Use Cases

gpt-4.1-mini use cases

Chat

gpt-4.1-mini is tagged for chat in this catalog and works with OpenAI-compatible client libraries.

FAQ

gpt-4.1-mini free API FAQ

Is gpt-4.1-mini free to use?

gpt-4.1-mini is listed with free API access on GitHub Models, subject to the provider's quota and account policy.

What is the gpt-4.1-mini model ID?

The model ID shown in this catalog is gpt-4-1-mini.

What are the gpt-4.1-mini free tier rate limits on GitHub Models?

The listed free tier limit is 15 RPM, 150 RPD. Limits can change per account tier, so confirm against the provider dashboard.

What context window does gpt-4.1-mini support?

The listed context window is 1.0M tokens with up to 32K output tokens.

More about gpt-4.1-mini

GPT-4.1 Mini is OpenAI's cost-optimized small model, available free through GitHub Models. With a 1M-token context window — matching the full GPT-4.1 — and 32K max output, it punches above its weight class for long-context tasks like document Q&A, full-codebase analysis, and multi-turn conversations. The trade-off versus full GPT-4.1 is lower reasoning depth on complex tasks. Rate-limited to 15 RPM and 150 requests per day with per-request caps (8K input / 4K output), it is generous enough for prototyping and personal projects. Fully OpenAI SDK-compatible; requires only a GitHub account.

For API keys, setup steps, and provider-level limits, see the GitHub Models provider page.