Groq logo

qwen3-32b API status on Groq

Check provider
★★★★★★★★★★ 3.5 Benchmark-backed score

Dense open Qwen model for self-hosted chat, reasoning, and coding

Check providerOpenAI compatibleReasoningTool callingJSON modeTextReasoning
API Details

Copy-ready connection details

Use these values in your client or SDK.
Base URL
https://api.groq.com/openai/v1
Model ID
qwen/qwen3-32b
API format OpenAI Chat Completions + OpenAI Responses
Technical Details

qwen3-32b specifications

Provider and model catalog
Context window 131K
Max output 131K
Status Check provider
Family qwen
Knowledge cutoff 2025-04
Released Apr 1, 2025
Last updated Jul 30, 2026
Free listing since Apr 28, 2025
Input text
Output text
Capabilities reasoning, tool calling, structured output, temperature control
Open weights Yes

External benchmark references

Benchmark Score Metric Date
Aider Polyglot 40 percent correct 2025-05-08
AI Recommendation

Should you use qwen3-32b?

qwen3-32b is listed for chat workloads and supports a 131K context window.

Check Groq's current endpoint status before integrating this listing. Use the alternatives table if the provider listing is unavailable.

Best for
  • Chat
Strengths & Weaknesses

Strengths

  • Long context window
  • Reasoning mode listed
  • Tool calling support
  • Structured JSON output

Watch outs

  • Current endpoint availability needs provider confirmation
  • Free-tier rate limits apply
  • Vision support is not listed
Benchmark Overview

Benchmark signals for qwen3-32b

Measured data
Intelligence General reasoning and instruction following
11.5/100
Coding Programming and code generation
15.3/100
Agentic Tool use and multi-step tasks
1.8/100
Speed Observed generation speed
98 tok/s
Context Maximum listed context window
131K
Pricing

qwen3-32b pricing per 1M tokens

Check provider
Input $0.7 per 1M tokens
Output $2.8 per 1M tokens
Free access Check provider Groq
Rate limit 30 RPM, 1,000 RPD provider policy
Availability

qwen3-32b availability by provider

1 alternative

We found 2 provider listings for qwen3-32b. Check model ID, quota, pricing, and API format before switching providers.

Provider Model listing Access Context API Limits
Groq qwen3-32b Check provider 131K OpenAI-style 30 RPM, 1,000 RPD
OVHcloud AI Endpoints Qwen3-32B Free tier 131K OpenAI-style 2 RPM (anonymous)
View Groq setup guide →
Typical Use Cases

qwen3-32b use cases

Chat

qwen3-32b is tagged for chat in this catalog and works with OpenAI-compatible client libraries.

FAQ

qwen3-32b free API FAQ

Is qwen3-32b free to use?

qwen3-32b appears in the free model catalog for Groq, but its current endpoint availability should be confirmed with the provider before use.

What is the qwen3-32b model ID?

The model ID shown in this catalog is qwen/qwen3-32b.

What are the qwen3-32b free tier rate limits on Groq?

The listed free tier limit is 30 RPM, 1,000 RPD. Limits can change per account tier, so confirm against the provider dashboard.

What context window does qwen3-32b support?

The listed context window is 131K tokens with up to 131K output tokens.

More about qwen3-32b

Qwen instruction model for multilingual chat, reasoning, and tool use

For API keys, setup steps, and provider-level limits, see the Groq provider page.