NVIDIA NIM logo

Gemma 4 31B IT Free API on NVIDIA NIM

Free API Verified
★★★★★★★★★★ 3.5 Benchmark-backed score

Largest Gemma 4 instruction model for open, self-hosted chat and reasoning

Free APIOpenAI compatibleReasoningTool callingJSON modeVisionReasoning
API Details

Copy-ready connection details

Use these values in your client or SDK.
Base URL
https://integrate.api.nvidia.com/v1
Model ID
google/gemma-4-31b-it
API format OpenAI Chat Completions
Technical Details

Gemma 4 31B IT specifications

Provider and model catalog
Context window 262K
Max output 33K
Status Online
Family gemma
Released Apr 2, 2026
Last updated Apr 2, 2026
Free listing since Apr 2, 2026
Input text, image
Output text
Capabilities reasoning, tool calling, structured output, file attachments, temperature control
Open weights Yes
AI Recommendation

Should you use Gemma 4 31B IT?

Gemma 4 31B IT is listed for chat workloads and supports a 262K context window.

Use it when NVIDIA NIM's free tier is enough for evaluation, demos, or light production traffic.

Best for
  • Chat
Benchmark Overview

Benchmark signals for Gemma 4 31B IT

Measured data
Intelligence General reasoning and instruction following
29.4/100
Coding Programming and code generation
43.4/100
Agentic Tool use and multi-step tasks
14.4/100
Context Maximum listed context window
262K
Pricing

Gemma 4 31B IT pricing per 1M tokens

Free tier listed
Input $0 per 1M tokens
Output $0 per 1M tokens
Free access Available NVIDIA NIM
Typical Use Cases

Gemma 4 31B IT use cases

Chat

Gemma 4 31B IT is tagged for chat in this catalog and works with OpenAI-compatible client libraries.

FAQ

Gemma 4 31B IT free API FAQ

Is Gemma 4 31B IT free to use?

Gemma 4 31B IT is listed with free API access on NVIDIA NIM, subject to the provider's quota and account policy.

What is the Gemma 4 31B IT model ID?

The model ID shown in this catalog is google/gemma-4-31b-it.

What context window does Gemma 4 31B IT support?

The listed context window is 262K tokens with up to 33K output tokens.

More about Gemma 4 31B IT

Google Gemma 4 31B is a dense 30.7B-parameter multimodal model from Google DeepMind, available free on OpenRouter. Supports text and image input with a 262K context window. Strong on coding, reasoning, and document understanding tasks. Includes native function calling, configurable thinking/reasoning mode, and multilingual support across 140+ languages. Apache 2.0 license. OpenAI-compatible via OpenRouter. Free tier: 200 RPD.

For API keys, setup steps, and provider-level limits, see the NVIDIA NIM provider page.