Google Gemini logo

Gemini 2.5 Flash Free API on Google Gemini

Free API Verified
★★★★★★★★★★ 3.5 Benchmark-backed score

Fast Gemini workhorse for multimodal apps where latency and price matter

Free APIOpenAI compatibleReasoningTool callingJSON modePDF inputText
API Details

Copy-ready connection details

Use these values in your client or SDK.
Base URL
https://generativelanguage.googleapis.com/v1beta
Model ID
gemini-2.5-flash
API format Google Gemini native generateContent + OpenAI Chat Completions
Technical Details

Gemini 2.5 Flash specifications

Provider and model catalog
Context window 1.0M
Max output 65K
Status Online
Family gemini-flash
Knowledge cutoff 2025-01
Released Jun 17, 2025
Last updated Aug 6, 2026
Free listing since May 20, 2025
Input text, image, audio, video, pdf
Output text
Capabilities reasoning, tool calling, structured output, file attachments, temperature control

External benchmark references

Benchmark Score Metric Date
Aider Polyglot 55.1 percent correct 2025-05-25
Artificial Analysis Coding Index 22.2 index 2026-06-02
SciCode 39.4 percent correct 2026-06-02
Terminal-Bench Hard 13.6 success rate 2026-06-02
AI Recommendation

Should you use Gemini 2.5 Flash?

Gemini 2.5 Flash is listed for chat workloads and supports a 1.0M context window.

Use it when Google Gemini's free tier is enough for evaluation, demos, or light production traffic.

Best for
  • Chat
Strengths & Weaknesses

Strengths

  • Long context window
  • Reasoning mode listed
  • Tool calling support
  • Structured JSON output

Watch outs

  • Free-tier rate limits apply
Benchmark Overview

Benchmark signals for Gemini 2.5 Flash

Measured data
Intelligence General reasoning and instruction following
27/100
Coding Programming and code generation
22.2/100
Agentic Tool use and multi-step tasks
18.8/100
Speed Observed generation speed
207 tok/s
Context Maximum listed context window
1.0M
Pricing

Gemini 2.5 Flash pricing per 1M tokens

Free tier listed
Input $0.3 per 1M tokens
Output $2.5 per 1M tokens
Free access Available Google Gemini
Rate limit 15 RPM, 1,500 RPD provider policy
Availability

Gemini 2.5 Flash availability by provider

Current provider only

We only found this gemini-2-5-flash listing on Google Gemini in the current catalog. Use the related models section for nearby alternatives.

Provider Model listing Access Context API Limits
Google Gemini Gemini 2.5 Flash Free tier 1.0M Native 15 RPM, 1,500 RPD
View Google Gemini setup guide →
Typical Use Cases

Gemini 2.5 Flash use cases

Chat

Gemini 2.5 Flash is tagged for chat in this catalog and works with OpenAI-compatible client libraries.

FAQ

Gemini 2.5 Flash free API FAQ

Is Gemini 2.5 Flash free to use?

Gemini 2.5 Flash is listed with free API access on Google Gemini, subject to the provider's quota and account policy.

What is the Gemini 2.5 Flash model ID?

The model ID shown in this catalog is gemini-2.5-flash.

What are the Gemini 2.5 Flash free tier rate limits on Google Gemini?

The listed free tier limit is 15 RPM, 1,500 RPD. Limits can change per account tier, so confirm against the provider dashboard.

What context window does Gemini 2.5 Flash support?

The listed context window is 1.0M tokens with up to 65K output tokens.

More about Gemini 2.5 Flash

Google Gemini 2.5 Flash is a fast multimodal model for chat, document analysis, coding assistance, and long-context workflows. The free Google AI Studio tier lists a 1 million token context window, up to 65,000 output tokens, and no credit card requirement, with quota limits that should be checked in the provider dashboard before production use.

For API keys, setup steps, and provider-level limits, see the Google Gemini provider page.