Groq logo

GPT OSS 120B pricing and alternatives

Paid / limited
★★★★★★★★★★ 3.5 Benchmark-backed score

Open GPT reasoning model for self-hosted agents and controllable deployments

Paid / limitedOpenAI compatibleReasoningTool callingJSON modeReasoningText
API Details

Copy-ready connection details

Use these values in your client or SDK.
Base URL
https://api.groq.com/openai/v1
Model ID
openai/gpt-oss-120b
API format OpenAI Chat Completions + OpenAI Responses
No current free listing. This page is kept for comparison and SEO continuity. Check the alternatives below before integrating.
Technical Details

GPT OSS 120B specifications

Provider and model catalog
Context window 131K
Max output 66K
Status Online
Family gpt-oss
Released Aug 5, 2025
Last updated Aug 5, 2026
Free listing since Aug 5, 2025
Input text
Output text
Capabilities reasoning, tool calling, structured output, temperature control
Open weights Yes
AI Recommendation

Should you use GPT OSS 120B?

GPT OSS 120B is listed for chat, coding workloads and supports a 131K context window.

Use the comparison and availability sections before choosing this listing, because it is not currently marked as free.

Best for
  • Chat
  • Coding
Strengths & Weaknesses

Strengths

  • Useful for coding workflows
  • Long context window
  • Reasoning mode listed
  • Tool calling support

Watch outs

  • No current free listing on this provider
  • Vision support is not listed
Benchmark Overview

Benchmark signals for GPT OSS 120B

Measured data
Intelligence General reasoning and instruction following
23.8/100
Coding Programming and code generation
30.4/100
Agentic Tool use and multi-step tasks
13.2/100
Speed Observed generation speed
306 tok/s
Context Maximum listed context window
131K
Pricing

GPT OSS 120B pricing per 1M tokens

Provider pricing
Input $0.15 per 1M tokens
Output $0.6 per 1M tokens
Free access Not listed Groq
Availability

GPT OSS 120B availability by provider

Current provider only

We only found this GPT OSS 120B listing on Groq in the current catalog. Use the related models section for nearby alternatives.

Provider Model listing Access Context API Limits
Groq GPT OSS 120B Paid 131K OpenAI-style Varies
View Groq setup guide →
Typical Use Cases

GPT OSS 120B use cases

Chat

GPT OSS 120B is tagged for chat in this catalog and works with OpenAI-compatible client libraries.

Coding

GPT OSS 120B is tagged for coding in this catalog and works with OpenAI-compatible client libraries.

FAQ

GPT OSS 120B free API FAQ

Is GPT OSS 120B free to use?

GPT OSS 120B is not currently marked as free on Groq.

What is the GPT OSS 120B model ID?

The model ID shown in this catalog is openai/gpt-oss-120b.

What context window does GPT OSS 120B support?

The listed context window is 131K tokens with up to 66K output tokens.

More about GPT OSS 120B

GPT-OSS 120B is OpenAI's open-weight 117B-parameter Mixture-of-Experts model with 5.1B active parameters per forward pass, available free on OpenRouter (also on Groq, Cerebras, and Cloudflare). Supports configurable reasoning depth, full chain-of-thought access, and native tool use — function calling, browsing, and structured outputs. Optimized to run on a single H100 GPU with native MXFP4 quantization. Built for reasoning-heavy and agentic tasks. 131K context window. Text-only. OpenAI-compatible. Free tier: 200 RPD on OpenRouter.

For API keys, setup steps, and provider-level limits, see the Groq provider page.