Cerebras logo

qwen-3-235b-a22b-instruct-2507 API status on Cerebras

Check provider
★★★★★★★★★★ 3.5 Benchmark-backed score

Qwen3-235B-A22B is Alibaba's massive 235B-parameter MoE model available for free on Cerebras Cloud's ultra-fast WSE inference hardware.

Check providerOpenAI compatibleTool callingJSON modeText
API Details

Copy-ready connection details

Use these values in your client or SDK.
Base URL
https://api.cerebras.ai/v1
Model ID
qwen-3-235b-a22b-instruct-2507
API format OpenAI Chat Completions
Technical Details

qwen-3-235b-a22b-instruct-2507 specifications

Provider catalog
Context window 131K
Max output 8K
Status Check provider
Family qwen3-235b-a22b-2507
Released Apr 28, 2025
Last updated Jun 15, 2026
Free listing since Apr 28, 2025
Input text
Output text
Capabilities tool calling, structured output
AI Recommendation

Should you use qwen-3-235b-a22b-instruct-2507?

qwen-3-235b-a22b-instruct-2507 is listed for chat workloads and supports a 131K context window.

Check Cerebras's current endpoint status before integrating this listing. Use the alternatives table if the provider listing is unavailable.

Best for
  • Chat
Strengths & Weaknesses

Strengths

  • Long context window
  • Tool calling support
  • Structured JSON output
  • Works with OpenAI-style SDKs

Watch outs

  • Current endpoint availability needs provider confirmation
  • Free-tier rate limits apply
  • Vision support is not listed
Benchmark Overview

Benchmark signals for qwen-3-235b-a22b-instruct-2507

Measured data
Intelligence General reasoning and instruction following
25/100
Coding Programming and code generation
22.1/100
Agentic Tool use and multi-step tasks
22.8/100
Speed Observed generation speed
60 tok/s
Context Maximum listed context window
131K
Pricing

qwen-3-235b-a22b-instruct-2507 pricing per 1M tokens

Check provider
Input $0.45 per 1M tokens
Output $1.8 per 1M tokens
Free access Check provider Cerebras
Rate limit 30 RPM, 14,400 RPD, 1M TPD provider policy
Typical Use Cases

qwen-3-235b-a22b-instruct-2507 use cases

Chat

qwen-3-235b-a22b-instruct-2507 is tagged for chat in this catalog and works with OpenAI-compatible client libraries.

FAQ

qwen-3-235b-a22b-instruct-2507 free API FAQ

Is qwen-3-235b-a22b-instruct-2507 free to use?

qwen-3-235b-a22b-instruct-2507 appears in the free model catalog for Cerebras, but its current endpoint availability should be confirmed with the provider before use.

What is the qwen-3-235b-a22b-instruct-2507 model ID?

The model ID shown in this catalog is qwen-3-235b-a22b-instruct-2507.

What are the qwen-3-235b-a22b-instruct-2507 free tier rate limits on Cerebras?

The listed free tier limit is 30 RPM, 14,400 RPD, 1M TPD. Limits can change per account tier, so confirm against the provider dashboard.

What context window does qwen-3-235b-a22b-instruct-2507 support?

The listed context window is 131K tokens with up to 8K output tokens.

More about qwen-3-235b-a22b-instruct-2507

Qwen3-235B-A22B is Alibaba's massive 235B-parameter MoE model available for free on Cerebras Cloud's ultra-fast WSE inference hardware. With 131K context and OpenAI-compatible API, it delivers strong general-purpose chat performance at speeds far exceeding GPU-based alternatives — Cerebras' wafer-scale engine means responses stream near-instantly even at this parameter scale. The free tier allows 14,400 requests per day at 30 RPM, making it viable for moderate production workloads. No credit card is required; the main trade-off versus running this model on other providers is the 8K output limit and lower RPM compared to Groq's Llama endpoints.

For API keys, setup steps, and provider-level limits, see the Cerebras provider page.