Google Gemini logo

Gemini 3.1 Flash-Lite Free API on Google Gemini

Free API Verified
★★★★★★★★★★ 3.5 Benchmark-backed score

Low-latency Gemini model for high-volume multimodal and agent workloads

Free APIOpenAI compatibleReasoningTool callingJSON modePDF inputText
API Details

Copy-ready connection details

Use these values in your client or SDK.
Base URL
https://generativelanguage.googleapis.com/v1beta
Model ID
gemini-3.1-flash-lite
API format Google Gemini native generateContent + OpenAI Chat Completions
Technical Details

Gemini 3.1 Flash-Lite specifications

Provider and model catalog
Context window 1.0M
Max output 65K
Status Online
Family gemini-flash-lite
Knowledge cutoff 2025-01
Released May 7, 2026
Last updated Aug 6, 2026
Free listing since Mar 3, 2026
Input text, image, video, audio, pdf
Output text
Capabilities reasoning, tool calling, structured output, file attachments, temperature control
AI Recommendation

Should you use Gemini 3.1 Flash-Lite?

Gemini 3.1 Flash-Lite is listed for chat workloads and supports a 1.0M context window.

Use it when Google Gemini's free tier is enough for evaluation, demos, or light production traffic.

Best for
  • Chat
Strengths & Weaknesses

Strengths

  • Long context window
  • Reasoning mode listed
  • Tool calling support
  • Structured JSON output

Watch outs

  • Free-tier rate limits apply
Benchmark Overview

Benchmark signals for Gemini 3.1 Flash-Lite

Measured data
Intelligence General reasoning and instruction following
25/100
Coding Programming and code generation
34.7/100
Agentic Tool use and multi-step tasks
6.2/100
Speed Observed generation speed
312 tok/s
Context Maximum listed context window
1.0M
Pricing

Gemini 3.1 Flash-Lite pricing per 1M tokens

Free tier listed
Input $0.25 per 1M tokens
Output $1.5 per 1M tokens
Free access Available Google Gemini
Rate limit 30 RPM, 1,500 RPD provider policy
Typical Use Cases

Gemini 3.1 Flash-Lite use cases

Chat

Gemini 3.1 Flash-Lite is tagged for chat in this catalog and works with OpenAI-compatible client libraries.

FAQ

Gemini 3.1 Flash-Lite free API FAQ

Is Gemini 3.1 Flash-Lite free to use?

Gemini 3.1 Flash-Lite is listed with free API access on Google Gemini, subject to the provider's quota and account policy.

What is the Gemini 3.1 Flash-Lite model ID?

The model ID shown in this catalog is gemini-3.1-flash-lite.

What are the Gemini 3.1 Flash-Lite free tier rate limits on Google Gemini?

The listed free tier limit is 30 RPM, 1,500 RPD. Limits can change per account tier, so confirm against the provider dashboard.

What context window does Gemini 3.1 Flash-Lite support?

The listed context window is 1.0M tokens with up to 65K output tokens.

More about Gemini 3.1 Flash-Lite

Low-latency Gemini model for high-volume multimodal and agent workloads

For API keys, setup steps, and provider-level limits, see the Google Gemini provider page.