Cerebras logo

gpt-oss-120b Free API on Cerebras

Free API Verified
★★★★★★★★★★ 3.5 Benchmark-backed score

Open GPT reasoning model for self-hosted agents and controllable deployments

Free APIOpenAI compatibleReasoningTool callingJSON modeTextReasoning
API Details

Copy-ready connection details

Use these values in your client or SDK.
Base URL
https://api.cerebras.ai/v1
Model ID
gpt-oss-120b
API format OpenAI Chat Completions
Technical Details

gpt-oss-120b specifications

Provider and model catalog
Context window 131K
Max output 32K
Status Online
Family gpt-oss
Released Aug 5, 2025
Last updated Aug 6, 2026
Free listing since Aug 5, 2025
Input text
Output text
Capabilities reasoning, tool calling, structured output, temperature control
Open weights Yes
AI Recommendation

Should you use gpt-oss-120b?

gpt-oss-120b is listed for chat, coding workloads and supports a 131K context window.

Use it when Cerebras's free tier is enough for evaluation, demos, or light production traffic.

Best for
  • Chat
  • Coding
Strengths & Weaknesses

Strengths

  • Useful for coding workflows
  • Long context window
  • Reasoning mode listed
  • Tool calling support

Watch outs

  • Free-tier rate limits apply
  • Vision support is not listed
Benchmark Overview

Benchmark signals for gpt-oss-120b

Measured data
Intelligence General reasoning and instruction following
23.8/100
Coding Programming and code generation
30.4/100
Agentic Tool use and multi-step tasks
13.2/100
Speed Observed generation speed
180 tok/s
Context Maximum listed context window
131K
Pricing

gpt-oss-120b pricing per 1M tokens

Free tier listed
Input $0.15 per 1M tokens
Output $0.6 per 1M tokens
Free access Available Cerebras
Rate limit 5 RPM, 30K TPM, 1M TPD provider policy
Availability

gpt-oss-120b availability by provider

Current provider only

We only found this gpt-oss-120b listing on Cerebras in the current catalog. Use the related models section for nearby alternatives.

Provider Model listing Access Context API Limits
Cerebras gpt-oss-120b Free tier 131K OpenAI-style 5 RPM, 30K TPM, 1M TPD
View Cerebras setup guide →
Typical Use Cases

gpt-oss-120b use cases

Chat

gpt-oss-120b is tagged for chat in this catalog and works with OpenAI-compatible client libraries.

Coding

gpt-oss-120b is tagged for coding in this catalog and works with OpenAI-compatible client libraries.

FAQ

gpt-oss-120b free API FAQ

Is gpt-oss-120b free to use?

gpt-oss-120b is listed with free API access on Cerebras, subject to the provider's quota and account policy.

What is the gpt-oss-120b model ID?

The model ID shown in this catalog is gpt-oss-120b.

What are the gpt-oss-120b free tier rate limits on Cerebras?

The listed free tier limit is 5 RPM, 30K TPM, 1M TPD. Limits can change per account tier, so confirm against the provider dashboard.

What context window does gpt-oss-120b support?

The listed context window is 131K tokens with up to 32K output tokens.

More about gpt-oss-120b

Cerebras offers GPT-oss-120b, a powerful text-based LLM ideal for chat and coding, generating up to 8,000 tokens from 128,000-token contexts at 30 RPM, 14,400 RPD, and 1M TPD, without requiring a credit card and compatible with OpenAI.

For API keys, setup steps, and provider-level limits, see the Cerebras provider page.