Should you use Gemma 4 31B?
Gemma 4 31B is listed for chat workloads and supports a 262K context window.
Use it when Google Gemini's free tier is enough for evaluation, demos, or light production traffic.
https://generativelanguage.googleapis.com/v1beta gemma-4-31b-it Gemma 4 31B is listed for chat workloads and supports a 262K context window.
Use it when Google Gemini's free tier is enough for evaluation, demos, or light production traffic.
We only found this gemma listing on Google Gemini in the current catalog. Use the related models section for nearby alternatives.
| Provider | Model listing | Access | Context | API | Limits |
|---|---|---|---|---|---|
| | Gemma 4 31B | Free tier | 262K | Native | — |
| Model | Provider | Context | Access |
|---|---|---|---|
| Gemini 3.7 Flash | Google Gemini | 1.0M | Free tier |
| Gemini 3.6 Flash | Google Gemini | 1.0M | Free tier |
| Gemini 3.5 Flash | Google Gemini | 1.0M | Free tier |
| Gemini 3.5 Flash-Lite | Google Gemini | 1.0M | Free tier |
| Gemini 3.1 Flash-Lite | Google Gemini | 1.0M | Free tier |
Gemma 4 31B is tagged for chat in this catalog and works with OpenAI-compatible client libraries.
Gemma 4 31B is listed with free API access on Google Gemini, subject to the provider's quota and account policy.
The model ID shown in this catalog is gemma-4-31b-it.
The listed free tier limit is —. Limits can change per account tier, so confirm against the provider dashboard.
The listed context window is 262K tokens with up to 32K output tokens.
Gemma 4 31B — free model from Google Gemini.
For API keys, setup steps, and provider-level limits, see the Google Gemini provider page.