Should you use Gemini 3.5 Flash-Lite?
Gemini 3.5 Flash-Lite is listed for chat workloads and supports a 1.0M context window.
Use it when Google Gemini's free tier is enough for evaluation, demos, or light production traffic.
https://generativelanguage.googleapis.com/v1beta gemini-3.5-flash-lite | Benchmark | Score | Metric | Date |
|---|---|---|---|
| SWE-Bench Pro | 54.2 | resolve rate | 2026-07-21 |
| Terminal-Bench | 54 | accuracy | 2026-07-21 |
| MLE-Bench | 39.2 | average position score | 2026-07-21 |
| GDPval-AA | 1140 | Elo | 2026-07-21 |
Gemini 3.5 Flash-Lite is listed for chat workloads and supports a 1.0M context window.
Use it when Google Gemini's free tier is enough for evaluation, demos, or light production traffic.
We only found this gemini-3-5-flash-lite listing on Google Gemini in the current catalog. Use the related models section for nearby alternatives.
| Provider | Model listing | Access | Context | API | Limits |
|---|---|---|---|---|---|
| | Gemini 3.5 Flash-Lite | Free tier | 1.0M | Native | 30 RPM, 1,500 RPD |
| Model | Provider | Context | Access |
|---|---|---|---|
| Gemini 3.6 Flash | Google Gemini | 1.0M | Free tier |
| Gemini 3.5 Flash | Google Gemini | 1.0M | Free tier |
| Gemini 3.1 Flash-Lite | Google Gemini | 1.0M | Free tier |
| Gemini 2.5 Flash | Google Gemini | 1.0M | Free tier |
| Gemini 2.5 Flash-Lite | Google Gemini | 1.0M | Free tier |
Gemini 3.5 Flash-Lite is tagged for chat in this catalog and works with OpenAI-compatible client libraries.
Gemini 3.5 Flash-Lite is listed with free API access on Google Gemini, subject to the provider's quota and account policy.
The model ID shown in this catalog is gemini-3.5-flash-lite.
The listed free tier limit is 30 RPM, 1,500 RPD. Limits can change per account tier, so confirm against the provider dashboard.
The listed context window is 1.0M tokens with up to 65K output tokens.
Fast Gemini model balancing multimodal reasoning, tool use, and cost
For API keys, setup steps, and provider-level limits, see the Google Gemini provider page.